Concept

Source-filter — where it appears

The model that treats a sound as an excitation passed through a resonator, the two being specified independently of each other. It is what makes a vowel separable from a pitch, and a speaker's identity separable from what they said.

Named by 14 essays across 2 fields — each of them below, with the objects they name alongside it.

The vowel in "hod", sung at 110 Hz. The partials of a 110 Hz note, each drawn at the amplitude the vocal tract's resonances give it. The peaks of the curve are the formants — 730 Hz and 1090 Hz — and they stay where they are when the pitch changes, because they are a property of the shape of the mouth and not of the note being sung.

A vowel is two resonances

The vowel in "heed" is the same vowel sung high or low, and nothing about it is a property of the note. It is two peaks in the response of the mouth, sitting at fixed frequencies while the partials of the voice slide underneath them.

timbre · Spectrum
Three attacks, the first 50 ms. How loudness changes over the life of a note, for plucked, bowed and struck, drawn over the first 50 milliseconds. By the right-hand edge the plucked note is at 91%, the bowed note is at 36%, the struck note is at 97% — attack times of 4 ms, 140 ms, 2 ms, a spread of 70 to one, and the part a listener uses to tell them apart. Remove the attack from a recorded piano and it stops sounding like a piano, which is the shortest demonstration that the envelope carries as much identity as the spectrum.

The first fifty milliseconds

A spectrum is supposed to be what makes a trumpet a trumpet. Cut the first fifty milliseconds off a recorded note and listeners stop being able to name the instrument — while the spectrum they are hearing is unchanged. Identity is in the part of the sound that ends before the note has properly started.

timbre · Envelope
Where each family's tone-hole lattice stops reflecting. The cutoff frequency of an open tone-hole lattice, from Benade's formula, for four woodwind geometries: clarinet 1824 Hz, oboe 2990 Hz, flute 1690 Hz, bassoon 506 Hz. Below its cutoff a note's wave turns round at the first open hole and the instrument is a tube of that length; above it the wave passes through the whole lattice and radiates from the far end, so the upper part of every note's spectrum leaves the instrument from the same place whichever note is fingered. That is what gives a family one recognisable voice across its range.

Above a certain note the holes stop working

A row of open tone holes reflects the wave and makes the tube shorter — up to a frequency. Above it the wave runs straight past the whole lattice and leaves from the bell, so the top of every note's spectrum radiates from the same place whichever note is fingered. That cutoff is computable, it differs by family, and it is most of what makes an oboe sound like an oboe.

instruments · Tone holes
A bowed string on 196 Hz, through a violin body. The source is a sawtooth at one over n; the filter is the body's measured response, with A0 at 275 Hz, B1− at 460 Hz, B1+ at 540 Hz, bridge hill at 2500 Hz. What is radiated is their product, drawn as the bars. The resonances stay where they are when the note changes, exactly as a vowel's formants do — which is why an instrument has a voice rather than a tone, and why the same argument that identifies a vowel identifies a violin. The body frequencies are measured means over full-size instruments rather than computed from a plate.

The body is the filter

A violin string radiates almost nothing. What reaches a room is the string's sawtooth multiplied by the body's response, and that response is a comb of measured resonances that stays put while the note moves. It is the same arithmetic that identifies a vowel, on wood instead of a mouth — which is why an instrument has a voice rather than a tone.

timbre · Spectrum
A string mode swept through a body resonance at 460 Hz. What the string plays against what comes out. Away from the resonance the two are the same and the line is the diagonal. Near it the mode splits into a pair, and the note warbles at the difference between them — 12.9 Hz at the centre, which is slow enough to be counted and far too fast to be a tremolo. The splitting is a coupled oscillator and has nothing to do with the wolf fifth of a tuning system, which is a twenty-three cent arithmetic residue and shares only the word.

The other wolf

A cellist's wolf note is a string mode landing on a body resonance, at which point the two stop being separable and start exchanging energy — the mode splits in two and the note warbles at the difference. It is a coupled oscillator. The tuning system's wolf is twelve fifths failing to close by 23.5 cents. They share a word and nothing else.

timbre · Beating
Where each frequency goes, from a source 18 cm across. Polar response of a circular radiator of radius 9 cm at 200 Hz (ka = 0.3), 800 Hz (ka = 1.3), 2000 Hz (ka = 3.3), 5000 Hz (ka = 8.2). Zero degrees is straight ahead. The low frequency is a circle — it goes everywhere — and the high one is a narrow lobe with nulls either side of it, so a listener off to the side hears the same note with its top missing.

An instrument points

A source radiates evenly while it is small compared with the wavelength and beams once it is not, and the crossover is one number. So the same instrument is omnidirectional in its bottom octave and a searchlight in its top one — which means its spectrum depends on where the listener is standing, and a microphone position is a choice about what the instrument sounds like.

timbre · Room acoustics
The flow through the larynx, over two periods of a 110 Hz note. Volume flow against time, in Rosenberg's two-half-cosine model of the glottal pulse — a slow opening, a faster closing, and a closed phase during which no air passes at all. M1 — chest is open for 50 per cent of each period and opens 2.4 times as slowly as it closes. Nothing here is a displacement: the folds are a valve on a steady stream of air, and the flat stretches are the moments they are shut. At 110 Hz each period lasts 9.1 milliseconds, of which 4.5 is silence.

The other instrument with a reed

The folds do not vibrate the way a string does. They open and shut across a steady stream of air, once per period, and what leaves the larynx is a train of flow pulses with a closed phase in it. Everything said about the voice's tone is a statement about the shape of that pulse — and the shape has two numbers in it.

instruments · The voice
Long-term average spectra: an orchestra, playing forte against a trained operatic soloist. Each source's mean spectrum over a long passage, in decibels below its own strongest region, on a logarithmic frequency axis. An orchestra, playing forte peaks at 250 Hz and is 30 dB down by 3,150 Hz; a trained operatic soloist peaks at 250 Hz and is 11 dB down by 3,150 Hz. The shapes are the same until about 1 kHz and separate above it: at 3153 Hz the difference is 19.0 decibels, which is the largest anywhere in the range. Nothing here is about level. Both curves are drawn against their own peaks, so what is being compared is shape.

One voice over ninety players

A soloist heard over a full orchestra is not louder than it and could not be. What the trained voice does instead is put a peak of energy at three kilohertz, which is where the orchestra's spectrum has already fallen away and where the ear's own threshold happens to be lowest. Nineteen decibels of advantage, in a place nobody is competing for.

timbre · The voice
The vowel in "hod", sung at 110 Hz. The partials of a 110 Hz note, each drawn at the amplitude the vocal tract's resonances give it. The peaks of the curve are the formants — 730 Hz and 1090 Hz — and they stay where they are when the pitch changes, because they are a property of the shape of the mouth and not of the note being sung.

The sound a listener knows best

A voice is recognisable across every vowel it says, across two octaves of pitch, down a bad telephone line and in a whisper where there is no pitch at all. Nothing that survives all of that can be a frequency. What survives is a ratio: the resonances of a vocal tract are set by its length, so a shorter tract multiplies every formant by the same factor, and identity is a scale on the spectral envelope rather than a position within it. Between an adult man and a child the whole pattern moves by a fifth, and the vowel does not change at all.

timbre · The voice
One instrument, five spectra. The first 12 partials of a bowed string at 5 pitches, each passed through the same fixed body response and normalised to its own loudest partial. The resonances stay where they are and the partials slide under them, so the pattern is different at every note: the power-weighted centroid runs from 1.02 to 1.71 across the compass and is not monotonic in pitch. A source with no body at all would give 2.35 at every pitch, which is the single number every roughness figure here uses for a string.

An instrument is not one timbre

Every roughness verdict so far uses one partial list per instrument, and a violin does not have one. Its body's resonances stay where they are while the fundamental moves, so the radiated spectrum is different at every pitch: the power-weighted centroid runs from 1.02 to 1.64 across the compass, and it is not monotonic. Which means the result that a spectrum chooses its own scale gives a different scale at every note on the same instrument — three minima in the dissonance curve at the bottom of a violin's range, two in the middle and four at the top.

timbre · Spectrum
How many intervals a duet is smooth at. Every ordered pairing of 5 radiators, with the number of wells in its dissonance curve over an octave from 262 hertz. Rows are the instrument underneath and columns the one above, so the grid is not symmetric about its diagonal and that asymmetry is the result. The count runs from 1 to 7 across the grid, and a cell and its mirror need not agree: a violin under a clarinet has 1 and a clarinet under a violin has 2. The fifth is a well in every one of the 25 pairings and no other interval is.

Which instrument is underneath

Every roughness curve so far compares two tones of the same timbre, which is a duet nobody plays. Give the two notes different instruments and the sum stops being symmetric: the same written interval, on the same two players, is up to five times rougher depending on which of them takes the lower note. From G3 upward the fifth is a well in all twenty-five pairings and no other interval is; below it, three pairings lose even that, and all three have a clarinet on top.

timbre · Spectrum
Which pairs blend is a question about the note. The level at which a doubled pair's composite changes owner, drawn for all 15 pairs of 6 radiators over 2.6 octaves from 131 to 784 hertz. A pair blends when that level is inside the shaded band, which is the twenty-four decibels either way two players can manage; a curve outside it, or absent, is a pair one instrument owns at every balance. 6 of 15 pairs blend at the bottom of the range and 12 at the top. Every filter in this collection is fixed in frequency and the fundamental is not, so a radiator's shape is a function of pitch and so is everything computed from two of them — the blend ranking at the bottom and at the top disagree on 70 of 105 comparisons, which is more than half, so the order has turned over rather than merely shuffled.

The blend table has a row for every note

Eight earlier essays sound their instruments at one note, and one of them says why that cannot be innocent: every filter here is fixed in frequency and the fundamental is not. Swept over four octaves, the number of pairs that blend doubles from six to twelve, the ranking turns over rather than shuffles — seventy of a hundred and five comparisons swap — and a clarinet with an oboe goes from the best pair in the collection to the eleventh.

timbre · Spectrum
Partials 3, 4, 12 of "hod", over one vibrato cycle. The level of three partials of a 220 hertz note on the vowel in "hod", each about its own mean, over one cycle of a vibrato of ±71 cents at 6.0 hertz. The pale curve is the frequency deviation itself, for phase reference. Partial 3 at 660 hertz swings 4.22 decibels and peaks with the frequency; Partial 4 at 880 hertz swings 0.41 decibels and peaks twice a cycle; Partial 12 at 2640 hertz swings 8.62 decibels and peaks against it. The formants of this vowel are at 730, 1090, 2440 hertz and do not move; a partial below one rises as the frequency rises and one above it falls, so the modulations of a single note run in opposite directions at the same instant.

The partial that gets louder as it goes sharp

Eleven earlier essays sweep a set of partials and hold their amplitudes still, and a real tract does not move with the fundamental. Put the formant account and the sweeping account together and every partial acquires an amplitude modulation at the vibrato rate: 0.07 decibels on the fundamental and 8.6 on the twelfth partial of the same note. They are in phase below a formant and anti-phase above one — not ninety degrees apart — and the whole note swings 0.82 decibels, because they cancel.

instruments · The voice
The attack is the balance dial, turned by the clock. The level of a violin against a clarinet on one note at 392 hertz, moment by moment through the attack, with both players starting together. Two envelopes rising at different rates are a balance, so this axis is the same dial a conductor turns — and its whole travel is 6.02 decibels, which is twenty times the log of the ratio of the two attack times, 45 against 90 milliseconds, and nothing else. The pair does not begin as one player alone: both envelopes leave zero at the same slope ratio, so the dial starts at a finite offset rather than at silence. The dashed line is the balance at which the composite changes owner, -3.48 decibels — inside the travel, so the note belongs to a clarinet for its first 29 milliseconds and to a violin for the rest of its life.

The blend arrives before the note does

Nine essays on spectrum draw a steady state, and the strongest cue that two instruments are two instruments is that they do not start together. Two envelopes rising at different rates turn out to be a balance — the same dial an earlier essay swept — so the attack is that dial moved by the clock, and its whole travel is fixed at twenty times the log of the two attack times. It is six decibels for a clarinet with a violin against a crossing twelve to twenty-two decibels out, so one pair in ten changes hands during its own attack, and which one depends on a convention rather than on the instruments.

timbre · Spectrum

Named alongside it

The objects these essays reach for when they reach for this one.

SpectrumTimbreFormantPartialResonanceRoughnessRegisterSpectral envelopeStanding waveAttack transientCritical bandwidthEnvelope

All concepts