Concept

Formant — where it appears

A peak in the response of a resonator, fixed by its shape, that stays put while the notes played through it move. Which vowel is heard depends on where the peaks are relative to each other; who is speaking depends on the scale of the whole pattern.

Named by 9 essays across 3 fields — each of them below, with the objects they name alongside it.

The vowel in "hod", sung at 110 Hz. The partials of a 110 Hz note, each drawn at the amplitude the vocal tract's resonances give it. The peaks of the curve are the formants — 730 Hz and 1090 Hz — and they stay where they are when the pitch changes, because they are a property of the shape of the mouth and not of the note being sung.

A vowel is two resonances

The vowel in "heed" is the same vowel sung high or low, and nothing about it is a property of the note. It is two peaks in the response of the mouth, sitting at fixed frequencies while the partials of the voice slide underneath them.

timbre · Spectrum
Long-term average spectra: an orchestra, playing forte against a trained operatic soloist. Each source's mean spectrum over a long passage, in decibels below its own strongest region, on a logarithmic frequency axis. An orchestra, playing forte peaks at 250 Hz and is 30 dB down by 3,150 Hz; a trained operatic soloist peaks at 250 Hz and is 11 dB down by 3,150 Hz. The shapes are the same until about 1 kHz and separate above it: at 3153 Hz the difference is 19.0 decibels, which is the largest anywhere in the range. Nothing here is about level. Both curves are drawn against their own peaks, so what is being compared is shape.

One voice over ninety players

A soloist heard over a full orchestra is not louder than it and could not be. What the trained voice does instead is put a peak of energy at three kilohertz, which is where the orchestra's spectrum has already fallen away and where the ear's own threshold happens to be lowest. Nineteen decibels of advantage, in a place nobody is competing for.

timbre · The voice
The vowel in "hod", sung at 110 Hz. The partials of a 110 Hz note, each drawn at the amplitude the vocal tract's resonances give it. The peaks of the curve are the formants — 730 Hz and 1090 Hz — and they stay where they are when the pitch changes, because they are a property of the shape of the mouth and not of the note being sung.

The sound a listener knows best

A voice is recognisable across every vowel it says, across two octaves of pitch, down a bad telephone line and in a whisper where there is no pitch at all. Nothing that survives all of that can be a frequency. What survives is a ratio: the resonances of a vocal tract are set by its length, so a shorter tract multiplies every formant by the same factor, and identity is a scale on the spectral envelope rather than a position within it. Between an adult man and a child the whole pattern moves by a fifth, and the vowel does not change at all.

timbre · The voice
Where A has been. Documented pitch standards and surviving instruments, plotted as cents from A415. The extremes are 392 Hz and 465 Hz, which is 296 cents apart — 3.0 semitones, close enough to a minor third that a piece written at one and played at the other is in a different key. Nothing here is a preference; each is a decision somebody recorded.

The note that moved by a minor third

A has been anywhere between 392 and 465 hertz in the surviving record, which is 296 cents — a minor third short of six. The usual conclusion is that absolute pitch level is a convention and nothing musical depends on it. That is true of everything written on paper and false of everything that happens in a throat or a body: a singer's register break sits at a fixed frequency, so across the historical range it lands three semitones further down the written page. Transposing a piece is not a uniform operation, because the performer does not transpose.

tuning · Pitch standard
How many intervals a duet is smooth at. Every ordered pairing of 5 radiators, with the number of wells in its dissonance curve over an octave from 262 hertz. Rows are the instrument underneath and columns the one above, so the grid is not symmetric about its diagonal and that asymmetry is the result. The count runs from 1 to 7 across the grid, and a cell and its mirror need not agree: a violin under a clarinet has 1 and a clarinet under a violin has 2. The fifth is a well in every one of the 25 pairings and no other interval is.

Which instrument is underneath

Every roughness curve so far compares two tones of the same timbre, which is a duet nobody plays. Give the two notes different instruments and the sum stops being symmetric: the same written interval, on the same two players, is up to five times rougher depending on which of them takes the lower note. From G3 upward the fifth is a well in all twenty-five pairings and no other interval is; below it, three pairings lose even that, and all three have a clarinet on top.

timbre · Spectrum
What the climb costs, in tonnes and in inharmonicity. The total string tension a piano frame carries at each historical standard, computed from the string scaling used here. Tension goes as the square of the pitch, so the climb from 392 to 465 hertz is a rise from 11.5 to 16.2 tonnes on one instrument. The same ratio runs the other way through inharmonicity, which is inversely proportional to tension: the same wire at 465 hertz has 29 per cent less of it than at 392, so its partials are that much closer to a true series.

What the climb was a search for

The four-hundred-year rise in the pitch standard is usually explained as a search for brilliance, and an earlier essay ended by asking whether that was even coherent: does raising a string's tension change its spectrum as well as its pitch? Mostly it does not. What changes the timbre is something else entirely, and it is a mechanism that applies to instruments with no strings at all — every filter in the chain is fixed in absolute frequency while the notes move against it, so raising the standard is not a brightening but a reshuffling, and some notes lose.

tuning · Pitch standard
Which pairs blend is a question about the note. The level at which a doubled pair's composite changes owner, drawn for all 15 pairs of 6 radiators over 2.6 octaves from 131 to 784 hertz. A pair blends when that level is inside the shaded band, which is the twenty-four decibels either way two players can manage; a curve outside it, or absent, is a pair one instrument owns at every balance. 6 of 15 pairs blend at the bottom of the range and 12 at the top. Every filter in this collection is fixed in frequency and the fundamental is not, so a radiator's shape is a function of pitch and so is everything computed from two of them — the blend ranking at the bottom and at the top disagree on 70 of 105 comparisons, which is more than half, so the order has turned over rather than merely shuffled.

The blend table has a row for every note

Eight earlier essays sound their instruments at one note, and one of them says why that cannot be innocent: every filter here is fixed in frequency and the fundamental is not. Swept over four octaves, the number of pairs that blend doubles from six to twelve, the ranking turns over rather than shuffles — seventy of a hundred and five comparisons swap — and a clarinet with an oboe goes from the best pair in the collection to the eleventh.

timbre · Spectrum
Partials 3, 4, 12 of "hod", over one vibrato cycle. The level of three partials of a 220 hertz note on the vowel in "hod", each about its own mean, over one cycle of a vibrato of ±71 cents at 6.0 hertz. The pale curve is the frequency deviation itself, for phase reference. Partial 3 at 660 hertz swings 4.22 decibels and peaks with the frequency; Partial 4 at 880 hertz swings 0.41 decibels and peaks twice a cycle; Partial 12 at 2640 hertz swings 8.62 decibels and peaks against it. The formants of this vowel are at 730, 1090, 2440 hertz and do not move; a partial below one rises as the frequency rises and one above it falls, so the modulations of a single note run in opposite directions at the same instant.

The partial that gets louder as it goes sharp

Eleven earlier essays sweep a set of partials and hold their amplitudes still, and a real tract does not move with the fundamental. Put the formant account and the sweeping account together and every partial acquires an amplitude modulation at the vibrato rate: 0.07 decibels on the fundamental and 8.6 on the twelfth partial of the same note. They are in phase below a formant and anti-phase above one — not ninety degrees apart — and the whole note swings 0.82 decibels, because they cancel.

instruments · The voice
Where a section's fluctuation stops being a beat, on 220 hertz. Two rates up the spectrum of a 220 hertz note sung by a section whose voices are spread by 15 cents. The rising line is the beat rate between a typical pair of them, which grows with the partial because a mistuning in cents is a difference in hertz that scales with frequency; it reaches the 15 hertz at which a beat stops being a beat by partial 5.5, at 1217 hertz. The flat line is the amplitude modulation the vibrato imposes through the formants, which is 6.0 hertz at every partial because the vibrato modulates every partial by the same number of cents at the same rate. The two are equal at 487 hertz. Above 1217 hertz the beating has become roughness and the only fluctuation left is the vibrato's — and that frequency is the same one an octave up, where it is partial 2.8 instead.

The rate that does not rise with the partial

Twelve earlier essays give every vibrato the same six hertz, and the measured spread is 5.5 to 7.5. Putting the two fluctuations a choir contains on one axis shows why the rate matters: the beating between mistuned voices rises with the partial and leaves the range a listener follows as fluctuation at 1,217 hertz, while the vibrato's own modulation is six hertz at every partial. Above that frequency a section fluctuates by vibrato alone — and if every singer had the same rate, it would barely fluctuate at all.

instruments · The voice

Named alongside it

The objects these essays reach for when they reach for this one.

Source-filterResonanceCritical bandwidthRegisterRoughnessSpectral envelopeSpectrumVocal tractHistorical-performanceMaskingModulationOrchestration

All concepts