Generator

The vowel in "hod", sung at 110 Hz

The partials of a 110 Hz note, each drawn at the amplitude the vocal tract's resonances give it. The peaks of the curve are the formants — 730 Hz and 1090 Hz — and they stay where they are when the pitch changes, because they are a property of the shape of the mouth and not of the note being sung.
The vowel in "hod", sung at 110 HzThe partials of a 110 Hz note, each drawn at the amplitude the vocal tract's resonances give it. The peaks of the curve are the formants — 730 Hz and 1090 Hz — and they stay where they are when the pitch changes, because they are a property of the shape of the mouth and not of the note being sung.F1730 HzF21090 Hz050010001500200025003000hertzamplitudethe filter — the shape of the mouththe partials — the pitch being sungPeterson & Barney, 1952 — mean adult male values11 partials carry most of the identity

Drawn above with its standard settings, which is almost never how an essay draws it: an essay states the numbers it is arguing about, so the figure a reader meets there is about that argument rather than about the drawing in general. Every option a placement passes is checked against the ones this function actually reads, because an option it does not read is silently ignored and the figure quietly draws what is above instead.

It makes a noise. 18 of its 18 placements carry sound, built from the same numbers as the drawing, offering these buttons: "hawed" at the same pitch, "heed" an octave up, "heed" at 110 Hz, "heed" at 196 Hz, "heed" at 880 Hz — a soprano's problem, "heed" at the same pitch, "heed, a 17.5 cm tract" at the same pitch, "hod" an octave up and 13 more.

Called by 6 essays

the blast radius of changing it

The vowel in "hod", sung at 110 Hz. The partials of a 110 Hz note, each drawn at the amplitude the vocal tract's resonances give it. The peaks of the curve are the formants — 730 Hz and 1090 Hz — and they stay where they are when the pitch changes, because they are a property of the shape of the mouth and not of the note being sung.

A vowel is two resonances

The vowel in "heed" is the same vowel sung high or low, and nothing about it is a property of the note. It is two peaks in the response of the mouth, sitting at fixed frequencies while the partials of the voice slide underneath them.

timbre · Spectrum
A bowed string on 196 Hz, through a violin body. The source is a sawtooth at one over n; the filter is the body's measured response, with A0 at 275 Hz, B1− at 460 Hz, B1+ at 540 Hz, bridge hill at 2500 Hz. What is radiated is their product, drawn as the bars. The resonances stay where they are when the note changes, exactly as a vowel's formants do — which is why an instrument has a voice rather than a tone, and why the same argument that identifies a vowel identifies a violin. The body frequencies are measured means over full-size instruments rather than computed from a plate.

The body is the filter

A violin string radiates almost nothing. What reaches a room is the string's sawtooth multiplied by the body's response, and that response is a comb of measured resonances that stays put while the note moves. It is the same arithmetic that identifies a vowel, on wood instead of a mouth — which is why an instrument has a voice rather than a tone.

timbre · Spectrum
Two mechanisms, the notes both of them make, and the seam. The frequency range of each laryngeal mechanism for an adult male voice, on a logarithmic axis, with the band both can produce shaded. M1 — chest runs 82–349 Hz and M2 — falsetto runs 220–698 Hz, so 799 cents of the range — 8.0 semitones — can be sung either way. The two dots inside that band are the measured signature that this is a bifurcation rather than a threshold: the change upward happens at 330 Hz and the change downward at 294 Hz, 200 cents lower. A threshold is crossed at the same place in both directions and this is not.

Two mechanisms, and the seam between them

Every singer has a place in the range where the voice changes character, and eight semitones of it can be produced either way. The measurement that settles what kind of a place it is takes ten seconds: the change upward happens two hundred cents higher than the change downward, and a threshold cannot do that.

instruments · The voice
Long-term average spectra: an orchestra, playing forte against a trained operatic soloist. Each source's mean spectrum over a long passage, in decibels below its own strongest region, on a logarithmic frequency axis. An orchestra, playing forte peaks at 250 Hz and is 30 dB down by 3,150 Hz; a trained operatic soloist peaks at 250 Hz and is 11 dB down by 3,150 Hz. The shapes are the same until about 1 kHz and separate above it: at 3153 Hz the difference is 19.0 decibels, which is the largest anywhere in the range. Nothing here is about level. Both curves are drawn against their own peaks, so what is being compared is shape.

One voice over ninety players

A soloist heard over a full orchestra is not louder than it and could not be. What the trained voice does instead is put a peak of energy at three kilohertz, which is where the orchestra's spectrum has already fallen away and where the ear's own threshold happens to be lowest. Nineteen decibels of advantage, in a place nobody is competing for.

timbre · The voice
The vowel in "hod", sung at 110 Hz. The partials of a 110 Hz note, each drawn at the amplitude the vocal tract's resonances give it. The peaks of the curve are the formants — 730 Hz and 1090 Hz — and they stay where they are when the pitch changes, because they are a property of the shape of the mouth and not of the note being sung.

The sound a listener knows best

A voice is recognisable across every vowel it says, across two octaves of pitch, down a bad telephone line and in a whisper where there is no pitch at all. Nothing that survives all of that can be a frequency. What survives is a ratio: the resonances of a vocal tract are set by its length, so a shorter tract multiplies every formant by the same factor, and identity is a scale on the spectral envelope rather than a position within it. Between an adult man and a child the whole pattern moves by a fifth, and the vowel does not change at all.

timbre · The voice
Partials 3, 4, 12 of "hod", over one vibrato cycle. The level of three partials of a 220 hertz note on the vowel in "hod", each about its own mean, over one cycle of a vibrato of ±71 cents at 6.0 hertz. The pale curve is the frequency deviation itself, for phase reference. Partial 3 at 660 hertz swings 4.22 decibels and peaks with the frequency; Partial 4 at 880 hertz swings 0.41 decibels and peaks twice a cycle; Partial 12 at 2640 hertz swings 8.62 decibels and peaks against it. The formants of this vowel are at 730, 1090, 2440 hertz and do not move; a partial below one rises as the frequency rises and one above it falls, so the modulations of a single note run in opposite directions at the same instant.

The partial that gets louder as it goes sharp

Eleven earlier essays sweep a set of partials and hold their amplitudes still, and a real tract does not move with the fundamental. Put the formant account and the sweeping account together and every partial acquires an amplitude modulation at the vibrato rate: 0.07 decibels on the fundamental and 8.6 on the twelfth partial of the same note. They are in phase below a formant and anti-phase above one — not ninety degrees apart — and the whole note swings 0.82 decibels, because they cancel.

instruments · The voice

All figures · What can be heard