Concept

Timbre — where it appears

Whatever distinguishes two sounds of the same pitch, loudness and duration, which is largely spectrum and how it changes over time. It is defined negatively because no single measurement covers it, and the attack carries much of it.

Named by 36 essays across 5 fields — each of them below, with the objects they name alongside it.

Four spectra of the same note. The amplitude of each partial for 4 timbres at the same pitch — pure, string, clarinet, bell. These are the exact lists the sound buttons here synthesise from, so the picture and the sound are the same data.

The ear hears the list, not the shape

Two sounds with the same partials and different phases have completely different waveforms and sound identical. What the ear extracts is a list of frequencies and strengths, and everything else is discarded.

timbre · Spectrum
Three attacks, the first 50 ms. How loudness changes over the life of a note, for plucked, bowed and struck, drawn over the first 50 milliseconds. By the right-hand edge the plucked note is at 91%, the bowed note is at 36%, the struck note is at 97% — attack times of 4 ms, 140 ms, 2 ms, a spread of 70 to one, and the part a listener uses to tell them apart. Remove the attack from a recorded piano and it stops sounding like a piano, which is the shortest demonstration that the envelope carries as much identity as the spectrum.

The first fifty milliseconds

A spectrum is supposed to be what makes a trumpet a trumpet. Cut the first fifty milliseconds off a recorded note and listeners stop being able to name the instrument — while the spectrum they are hearing is unchanged. Identity is in the part of the sound that ends before the note has properly started.

timbre · Envelope
Partial 3, mistuned by 3%. A 10-partial tone on 220 Hz with one partial treated differently from the rest. Mistuning it moves it off the harmonic grid by 3.0 per cent, which is 19.8 Hz — slow enough to be heard as a beat rather than as a separate pitch, and enough for the partial to be heard out of the note as a whistle of its own. An onset difference does the same to a partial that is exactly in tune.

What makes two partials one note

A note is a stack of ten or twenty simultaneous tones and is heard as one thing. The obvious explanation is that they are whole-number multiples of a fundamental — and the obvious explanation is not sufficient. Mistune one partial by three per cent and it leaves the note; give a perfectly harmonic partial a thirty-millisecond head start and it leaves too. Shared behaviour beats arithmetic.

timbre · Auditory scene
Where each family's tone-hole lattice stops reflecting. The cutoff frequency of an open tone-hole lattice, from Benade's formula, for four woodwind geometries: clarinet 1824 Hz, oboe 2990 Hz, flute 1690 Hz, bassoon 506 Hz. Below its cutoff a note's wave turns round at the first open hole and the instrument is a tube of that length; above it the wave passes through the whole lattice and radiates from the far end, so the upper part of every note's spectrum leaves the instrument from the same place whichever note is fingered. That is what gives a family one recognisable voice across its range.

Above a certain note the holes stop working

A row of open tone holes reflects the wave and makes the tube shorter — up to a frequency. Above it the wave runs straight past the whole lattice and leaves from the bell, so the top of every note's spectrum radiates from the same place whichever note is fingered. That cutoff is computable, it differs by family, and it is most of what makes an oboe sound like an oboe.

instruments · Tone holes
A string struck at one 7th of its length. The amplitude of each partial of an ideal string excited at 0.1429 of its length. The mode shape is a sine, so a partial with a node at the excitation point cannot be set moving at all: partials 7, 14 are silent here. The envelope over the rest is one over n, a struck string's.

Where the hammer lands

Strike a string at exactly one over n and the nth partial is silent, because the hammer has landed on that mode's node and cannot move it. A piano's hammers strike between a seventh and a ninth of the way along, which puts the seventh partial — the most dissonant member of the series — at or near a null. That is a design decision made in wood, and it takes one line of trigonometry.

instruments · Excitation point
The same note, hit at a middling dynamic. The spectrum of a struck string with the hammer's own contact time applied as a low-pass. Contact lasts 1.60 ms at this force, against 2.26 ms at the softest and 0.95 ms at the loudest drawn — felt is a nonlinear spring, so a harder blow is a shorter contact and a brighter note. The spectral centroid moves from partial 1.5 to partial 2.2, which is a change of timbre and not of loudness.

A hammer is not an impulse

Contact lasts a couple of milliseconds, which low-passes the note — any partial whose half-period is shorter than the contact is barely excited. Piano felt is a spring that stiffens as it compresses, so a harder blow makes the contact shorter, the corner higher and the note brighter. A loud note is not a scaled-up quiet one, and no linear model gives that.

instruments · Excitation point
A bell tuned to 294 Hz, and the note it is heard at. The partials of a well-tuned church bell, as ratios to the prime, with the three that imply the strike note marked. The nominal, twelfth and double octave sit at 2, 3 and 4, which is a harmonic series on 1 — so they imply a fundamental at 294 Hz, an octave below the loudest partial the bell has. The tierce at 1.2 is a MINOR third above the prime, which is why a bell has a minor quality by construction rather than by choice.

A bell has no fundamental

The note a listener names when a church bell is struck is not any partial the bell has. Its nominal, twelfth and double octave sit at 2, 3 and 4 times the prime, which is a harmonic series on a pitch an octave below the loudest thing in the sound — and that pitch is supplied by the listener. Founders have been tuning it by ear since the fifteenth century.

instruments · Missing fundamental
A bowed string on 196 Hz, through a violin body. The source is a sawtooth at one over n; the filter is the body's measured response, with A0 at 275 Hz, B1− at 460 Hz, B1+ at 540 Hz, bridge hill at 2500 Hz. What is radiated is their product, drawn as the bars. The resonances stay where they are when the note changes, exactly as a vowel's formants do — which is why an instrument has a voice rather than a tone, and why the same argument that identifies a vowel identifies a violin. The body frequencies are measured means over full-size instruments rather than computed from a plate.

The body is the filter

A violin string radiates almost nothing. What reaches a room is the string's sawtooth multiplied by the body's response, and that response is a comb of measured resonances that stays put while the note moves. It is the same arithmetic that identifies a vowel, on wood instead of a mouth — which is why an instrument has a voice rather than a tone.

timbre · Spectrum
A string mode swept through a body resonance at 460 Hz. What the string plays against what comes out. Away from the resonance the two are the same and the line is the diagonal. Near it the mode splits into a pair, and the note warbles at the difference between them — 12.9 Hz at the centre, which is slow enough to be counted and far too fast to be a tremolo. The splitting is a coupled oscillator and has nothing to do with the wolf fifth of a tuning system, which is a twenty-three cent arithmetic residue and shares only the word.

The other wolf

A cellist's wolf note is a string mode landing on a body resonance, at which point the two stop being separable and start exchanging energy — the mode splits in two and the note warbles at the difference. It is a coupled oscillator. The tuning system's wolf is twelve fifths failing to close by 23.5 cents. They share a word and nothing else.

timbre · Beating
Where each frequency goes, from a source 18 cm across. Polar response of a circular radiator of radius 9 cm at 200 Hz (ka = 0.3), 800 Hz (ka = 1.3), 2000 Hz (ka = 3.3), 5000 Hz (ka = 8.2). Zero degrees is straight ahead. The low frequency is a circle — it goes everywhere — and the high one is a narrow lobe with nulls either side of it, so a listener off to the side hears the same note with its top missing.

An instrument points

A source radiates evenly while it is small compared with the wavelength and beams once it is not, and the crossover is one number. So the same instrument is omnidirectional in its bottom octave and a searchlight in its top one — which means its spectrum depends on where the listener is standing, and a microphone position is a choice about what the instrument sounds like.

timbre · Room acoustics
Three spectra, and where each one's consonances fall. The positions of the roughness minima for 3 partial sets, over 12 semitones from the same fundamental, found by scanning each curve rather than by marking them. A harmonic tone — 498 cents at 3.1 per cent prominence, 702 cents at 37.1 per cent prominence, 884 cents at 6.2 per cent prominence. Partials stretched by 2.1 per octave — 533 cents at 3.7 per cent prominence, 751 cents at 40.1 per cent prominence, 947 cents at 7.3 per cent prominence. A bar's partials — 1, 2.76, 5.40, 8.93 — 694 cents at 5.6 per cent prominence, 870 cents at 14.2 per cent prominence, 982 cents at 1.0 per cent prominence, 1166 cents at 14.7 per cent prominence. The minima are computed by the same well-finder the dissonance curve uses, which requires a minimum to have one per cent of the curve's range on both sides of it before it counts.

A spectrum chooses its own scale

The roughness curve's dips are always described as landing on the small whole numbers. They do not land on numbers. They land where partials coincide, and stretching a spectrum by seven per cent moves every one of them by exactly seven per cent — so the scale a set of intervals belongs to is a property of the instrument's spectrum rather than of arithmetic.

intervals · Consonance
The vowel in "hod", sung at 110 Hz. The partials of a 110 Hz note, each drawn at the amplitude the vocal tract's resonances give it. The peaks of the curve are the formants — 730 Hz and 1090 Hz — and they stay where they are when the pitch changes, because they are a property of the shape of the mouth and not of the note being sung.

The sound a listener knows best

A voice is recognisable across every vowel it says, across two octaves of pitch, down a bad telephone line and in a whisper where there is no pitch at all. Nothing that survives all of that can be a frequency. What survives is a ratio: the resonances of a vocal tract are set by its length, so a shorter tract multiplies every formant by the same factor, and identity is a scale on the spectral envelope rather than a position within it. Between an adult man and a child the whole pattern moves by a fifth, and the vowel does not change at all.

timbre · The voice
C4, in every place it can be played. A guitar neck with the 4 places C4 can be stopped, drawn at the fret spacing a 64.8-centimetre scale actually has. The stave writes one note and the tablature writes one of these; each notation says exactly what the other leaves out. The speaking lengths run from 61.2 down to 27.2 centimetres, so a hand plucking 12 centimetres from the bridge meets between 20 and 44 per cent of the string.

What a tablature keeps

Middle C can be stopped in four places on a guitar. The speaking lengths run from 61 to 27 centimetres, so a hand plucking twelve centimetres from the bridge meets between a fifth and nearly a half of the string, and the comb of missing partials is different at every one: the second partial is thirteen decibels stronger in the best position than in the worst. A stave writes one note for all four. A tablature writes four different things and cannot say which note any of them is.

instruments · Notation
A note gets duller as it dies. Each partial of a string note against time, with the loss rising as the partial number to the power 1 — so the fundamental takes 6 seconds to fall sixty decibels and the 8th takes 0.75. The heavy line is the power-weighted centroid, falling from partial 1.77 toward the fundamental; it is halfway there after 0.15 seconds. A single-rate envelope would draw all of these as parallel lines and the centroid as a horizontal one, and a struck string does neither: what is left at the end of a long note is very nearly a sine.

The note that gets duller as it dies

Every envelope drawn so far is one curve applied to a whole sound, and no struck string behaves that way. A string loses energy to air, to internal friction and to the bridge, and all three losses rise with frequency — so a note with a six-second fundamental has a sixteenth partial that is gone in under half a second, and the sound moving toward the listener is a spectrum collapsing toward its own fundamental. Which means an instrument is identified twice: once by the fifty milliseconds of its attack, which the earlier essays measured, and again by how fast its colour drains, which they did not.

timbre · Envelope
One instrument, five spectra. The first 12 partials of a bowed string at 5 pitches, each passed through the same fixed body response and normalised to its own loudest partial. The resonances stay where they are and the partials slide under them, so the pattern is different at every note: the power-weighted centroid runs from 1.02 to 1.71 across the compass and is not monotonic in pitch. A source with no body at all would give 2.35 at every pitch, which is the single number every roughness figure here uses for a string.

An instrument is not one timbre

Every roughness verdict so far uses one partial list per instrument, and a violin does not have one. Its body's resonances stay where they are while the fundamental moves, so the radiated spectrum is different at every pitch: the power-weighted centroid runs from 1.02 to 1.64 across the compass, and it is not monotonic. Which means the result that a spectrum chooses its own scale gives a different scale at every note on the same instrument — three minima in the dissonance curve at the bottom of a violin's range, two in the middle and four at the top.

timbre · Spectrum
Six ways to put three players on three notes. The same chord — G3, B♭3, D4 — played by clarinet, oboe, voice in all 6 possible assignments, scored by the roughness each produces. Every bar is the same pitches and the same instruments; only who is on which note changes. The worst is 1.42 times the best, which is a factor a score can control and a chord symbol cannot express at all. Each row is labelled from the bottom note upward.

Which player on which note

An interval's roughness depends on which instrument is underneath, so the pair does not commute. Three players over three notes is the smallest thing that asymmetry has anywhere to go: six assignments, all of them the same chord, and across 450 of them the roughest averages half again the smoothest and reaches six times it. It is orchestration in the only form that can be computed here — not which chord, and not which voicing, but who is on which note.

timbre · Spectrum
The bow's window along each string, with the bow held still. Schelleng's window — the ratio of the greatest bow force a note tolerates to the least — for every written pitch on every string that can play it, with the bow 35 mm from the bridge and staying there. The window widens up each string because β is the bow's distance over the SOUNDING length and the sounding length is shortening, and it steps down at each new string because a heavier string carries a narrower one. The dashed lines are what was drawn earlier, which held β at 0.09 for every note on every string. The widest disagreement between two strings at one pitch is 1.48 to one, at A♭5.

Two dials the player turns together

Schelleng's window is a ratio of two bow forces, and an earlier essay found that the string's own impedance is in it twice. What it held still was the bowing point — as though a player kept β constant while moving up the fingerboard, which is the opposite of what a bow does. Let the bow stay where a bow stays and the window widens up every string, the spread between two strings at one pitch goes from 1.4 to 1.5 the other way round, and the string that is easiest changes.

instruments · Bowed string
How much of each spectrum a listener can assemble into one note. Each partial of each spectrum at the harmonic number it is nearest, against the whole-number series that fuses the most of them, with anything more than 1 per cent out marked as heard separately. an ideal string keeps 10 of 10; a piano string keeps 9 of 10; a bell keeps 7 of 8; a bar keeps 2 of 6; a kettledrum keeps 3 of 5. The fundamental is capped at a tenth of the top partial, and the cap is load-bearing rather than tidy: a bell's ratios are all whole multiples of a tenth, so an unconstrained search finds a fundamental twenty-five harmonics down, calls every partial exact, and reports that a bell fuses perfectly. Nothing that high is resolved and the low harmonics of it are not there.

The spectrum that will not fuse

A partial about one per cent off its harmonic is heard as a sound of its own rather than as part of a note. Apply that criterion to a whole spectrum instead of to one mistuned component and it becomes a count: a piano string keeps nine of its ten partials, a bell keeps seven of eight, a bar keeps two of six. The physics of inharmonicity has had an essay here for a long time. This is what it sounds like.

perception · Auditory scene
Every assignment, at equal levels and at its own best balance. The 6 ways of putting 3 players on a chord, each drawn twice: hollow at equal levels, which is what an assignment ranking sees, and filled at the levels that balance the parts and then minimise roughness. Every scoring here is at the same total loudness, 23.3 sones, so two points are comparable. Solving the discrete problem first picks violin · clarinet · oboe; solving both at once picks violin · oboe · clarinet, and the two-stage answer costs 9.0 per cent more roughness. The orderings do not keep their places between the two columns, which is the whole of the argument: a ranking taken at equal levels is not a ranking.

Who plays what and how loud is one question

Two lines of argument, one about spectrum and one about loudness, each stopped at the same wall and each said so. One of them can choose who plays which note and has every player at the same level; the other can choose how loud each part is and has nobody assigned to anything. Put together they are a single problem with two kinds of variable, and solving it in stages picks a different answer from solving it at once — nine per cent rougher, at the same loudness, on an ordinary triad.

form · Orchestration
Trumpet at three dynamics, as a spectrum rather than a level. The radiated partials of a trumpet at 45, 70, 95 decibels, each normalised to its own strongest partial so that only the SHAPE is compared. A linear source would give three identical pictures. This one does not: the spectral centroid moves from partial 2.19 to 6.41, a factor of 2.92, because the excitation is nonlinear and blowing harder steepens the pressure front rather than scaling it. The tilt used is 3 decibels per octave of partial number per ten decibels of level, referred to 70 dB — a stipulated, ordinal number, not a measurement of any instrument.

A dynamic mark changes what a note is

Every spectrum until now is a shape with a level in front of it, so that playing ten decibels louder raises every partial by ten. That is true of exactly one instrument in an orchestra. Everybody else steepens their own spectrum as they lean on it, and a trumpet's centre of gravity moves from the second partial to the sixth across a dynamic range while an organ flue pipe's does not move at all.

timbre · Orchestration
Two fusion cues, and they do not agree about a single spectrum. Each spectrum twice. Hollow is the harmonicity census — the fraction of partials near enough a whole multiple of one fundamental to fuse, which is harmonicity. Filled is the same fraction under common fate: how many partials decay at a rate within a factor of 2 of the strongest partial's. Ranked by harmonicity the order is an ideal string, a piano string, a bell, a kettledrum, a bar; ranked by common fate it is a kettledrum, a bar, an ideal string, a piano string, a bell. The two orderings are nearly reversed. An ideal string is perfect on the first cue and 20 per cent on the second, and a kettledrum — the worst spectrum in this collection for fitting a series — is the only one whose partials all die together.

The partials that do not die together

The fusion census is a still photograph: it asks whether a set of partials fits one harmonic series and has no term for time. Put time in and the ranking reverses. An ideal string is perfect on harmonicity and holds a fifth of its partials together by decay; a kettledrum — the worst spectrum in this collection for fitting a series — is the only one whose modes all die at one rate.

perception · Auditory scene
Three ways a note starts, on one pair of axes. Every instrument in this collection, with how long its note takes to speak measured twice: across, in periods of the note itself; up, in milliseconds. The three clusters are three different pieces of physics. A wind instrument accumulates energy in a resonance, so its wait is the resonance's Q — 27, 26, 26, 19, 37 periods here. A bowed string is at full amplitude the moment Helmholtz motion begins and what takes time is the bow reaching a force inside Schelleng's window, which is 3.0, 3.8, 5.1, 7.6 periods. A plucked or struck string is at full amplitude at once and the only duration in it is the exciter's own contact, under a tenth of a period. The two axes do not agree: the slowest instrument in milliseconds is an alto saxophone at 159, and in periods it is an alto saxophone at 37.

A note takes a number of periods to speak

Two separate accounts stopped at the same object and said so. A wind instrument's note takes as long to arrive as its resonance takes to build, and that time is the resonance's Q divided by its frequency — so the number of periods is the Q and has no pitch in it, while the number of milliseconds does. The two readings order the instruments differently, and both orderings are wanted.

instruments · Onset time
A family resemblance, in the heights rather than in the frequencies. The peak heights of trumpet, F horn, tenor trombone, plotted against peak number rather than against frequency. The three differ in length by a factor of 2.4 and their frequency series cannot be made to overlap; their heights agree to 3.2 decibels on average and their Qs to a factor of 1.30. The agreement improves up the series — 6.9 decibels at the first peak and 1.6 at the 8th — which is an earlier claim arriving as a measurement: a family has one voice because it has one filter, and the filter is visible in what the bore pushes back with and not in where its resonances are.

A family resemblance in the heights

Trumpet, horn and trombone differ in length by a factor of two and a half, so their frequency series cannot be laid over one another. Their impedance peaks agree to three decibels in height and to thirty per cent in Q, peak for peak, and the agreement improves with peak number. An earlier essay inferred that a family has one voice because it has one filter; the solver can now be asked directly.

instruments · Bore profile
How far out of tune an interval has to be before its beat is usable. An earlier essay produced a modulation index for every interval on a stated timbre; whether a fluctuation of that index at that rate can be detected is a published function of both. Running one against the other turns a table of decibels into a window in cents. The bar is the mistuning over which the beat is both deep enough to notice and at a rate a tuner can use — not so slow that a beat takes half a minute to complete, not so fast that it has stopped being a beat. the major third gives the widest window, 1.3 to 60.0 cents, and the minor sixth the narrowest, 0.8 to 38.9. The ordering is the opposite of the dip's: the octave has the shallowest dip on a string spectrum and the widest usable window, because its coincidence sits at a low partial and a given mistuning therefore produces a slower beat. Depth and rate pull opposite ways, and it is the rate that decides.

The beat a tuner can actually use

There is a modulation index for every interval on every timbre, and the published threshold for detecting a fluctuation is a function of exactly that and its rate. Running one against the other turns a table of decibels into a window in cents — and reverses the ordering, because the interval with the shallowest dip has the widest window.

intervals · Beating
The violin is the one where the bow's width catches up. Three separate limits on how near the bridge a bow can go, for the four bowed instruments. The ribbon's near edge reaches the bridge at 5, 6, 6, 8 millimetres; the ribbon covers half the corner's bridge-side excursion at 10, 11, 12, 15; and Schelleng's force window narrows to a factor of 3 at 10, 12, 21, 32. On the viola, cello and double bass the force window binds first, by a margin that widens to a factor of two on the bass. On the violin the ribbon binds first, at 10 millimetres against 9.8. A bow's hair is as wide as a hand can control and a string is as long as its pitch requires, so the ratio between them is a fact about the violin rather than about bowing.

The bow is not a point either

Seven earlier essays set the bowing point to a number between 0.02 and 0.3 and drew it as a point. A violin bow's hair is ten millimetres wide on a string of three hundred and twenty-five, so near the bridge the ribbon is wider than its own distance from it. In the spectrum that turns out to be worth nothing. In the geometry it is the reason sul ponticello has a floor, and the violin is the one instrument of the four where it arrives before the force does.

instruments · Bowed string
Intonation is a unison problem and nothing else. The roughness between two instruments on one note, against how far apart they are in cents, drawn for a unison and for the intervals beside it. A perfect unison is 0.0007 — the partials coincide and there is nothing to beat. Five cents apart it is 0.0465, 65 times as rough, and ten cents apart it is rougher than a major third played exactly. The mechanism is that partial n of a note mistuned by c cents is mistuned by c cents as well, which is n times as many hertz — so the top of the spectrum enters the critical band long before the fundamental does. The other curves are flat, because a third's roughness is set by which partials nearly coincide and a few cents does not change which.

Two players on one note

Six essays have put one instrument on each note of a chord, and the commonest thing an orchestrator actually does is put two on the same note. Two independent sources add in power, so the composite is neither of them — except that it nearly always is one of them, because the level at which ownership changes hands is rarely at zero. And a unison ten cents out is rougher than a major third dead in tune.

timbre · Spectrum
A woodwind with holes graduated 12 mm to 6 mm, drilled so that every fingering is in tune. A cylindrical bore 567 millimetres of acoustic length and 15 across, with twelve tone holes through a 4-millimetre wall. Opening them one at a time from the far end takes it up a chromatic scale from D3 to D4. The stations are not copied from a maker's drawing: each was solved so that its own fingering sounds its equal-tempered note in this model, one hole at a time down the tube with every hole below it already open, which is what a reamer and a tuning fork do. The worst fingering is 11.9 cents out. The diameters run 12.0 millimetres at the bell end to 6.0 at the top, and the spacings close from 30 millimetres to 19.

The cutoff that is a list

Five earlier essays have quoted one number for a woodwind's cutoff — 1,824 hertz for a clarinet — from a formula written for an infinite lattice of identical holes. Solve a whole twelve-hole chart instead and the number is eleven different numbers, running from 2,193 hertz down to 1,574, which is 574 cents. The lowest fingering has no cutoff at all, and which way the list runs turns out to be a design decision rather than a fact about woodwinds.

instruments · Tone holes
Who owns clarinet, oboe, voice at every balance. The composite of three players at 392 hertz belongs to whichever of them it is nearest in log-spectral distance, and here that is drawn over the whole plane of balances a conductor could set — the second and third players from 24 decibels below the first to 24 above. voice owns 79 per cent of the square. The three regions meet where all three distances are equal, which is the only balance at which the composite belongs to nobody: it is at -0.3 and -19.1 decibels, inside the square and therefore a balance an ensemble could actually be asked for. A trio has a colour of its own at one point, not over a region.

A section has a loudest member, not a colour

Two players on one note have a balance at which the composite belongs to neither, and that is what blending means. Three should have three such balances and no reason for them to agree — a trio with a rock-paper-scissors ownership would have no strongest member at all. Twenty trios, sixty pairwise comparisons, and not one disagreement: the possibility is real, arbitrary spectra do it once in twenty, and instruments never do.

timbre · Spectrum
Which pairs blend is a question about the note. The level at which a doubled pair's composite changes owner, drawn for all 15 pairs of 6 radiators over 2.6 octaves from 131 to 784 hertz. A pair blends when that level is inside the shaded band, which is the twenty-four decibels either way two players can manage; a curve outside it, or absent, is a pair one instrument owns at every balance. 6 of 15 pairs blend at the bottom of the range and 12 at the top. Every filter in this collection is fixed in frequency and the fundamental is not, so a radiator's shape is a function of pitch and so is everything computed from two of them — the blend ranking at the bottom and at the top disagree on 70 of 105 comparisons, which is more than half, so the order has turned over rather than merely shuffled.

The blend table has a row for every note

Eight earlier essays sound their instruments at one note, and one of them says why that cannot be innocent: every filter here is fixed in frequency and the fundamental is not. Swept over four octaves, the number of pairs that blend doubles from six to twelve, the ranking turns over rather than shuffles — seventy of a hundred and five comparisons swap — and a clarinet with an oboe goes from the best pair in the collection to the eleventh.

timbre · Spectrum
Four players on three notes, every arrangement. The 36 ways of putting 4 players on a 3-note chord so that every note is covered, ranked by roughness, all at one total loudness of 26.9 sones. Each row is shaded by which note carries the pair. The best is flue | clarinet+violin | oboe and the worst is clarinet | oboe+flue | violin, a factor of 2.21. Every earlier essay puts exactly one player on each note, which is a permutation; a doubling makes the arrangement a surjection instead, and the doubled note sounds neither of its two players but the composite they make. Which note gets the pair explains 7 per cent of the spread here and which players sit on the lowest note explains 89: the fourth player is a much smaller decision than the three that were already there.

The fourth player is a spectrum, not a decision

Six earlier essays put exactly one instrument on each note, which makes an arrangement a permutation — and the commonest operation in orchestration is a doubling, which does not. Four players on three notes give thirty-six arrangements instead of six, and the extra choice turns out to be the smallest thing on the page: which note carries the pair explains three per cent of the spread and which players sit on the bass explains eighty-nine. A doubled note can be priced as one player, and which one is not the one a spectral account would have named.

instruments · Orchestration
A struck note's two ends are the same for every loss law. The partial levels of a string spectrum struck at 80 decibels on 130.8 hertz, and what is left of it when the fundamental itself falls under the threshold of hearing, for three laws relating a partial's decay rate to its number. The left panel is every one of them: a loss law cannot change the spectrum at the instant of the strike, because no time has passed. The other three are every one of them too: whatever the law, the note ends with nothing above the threshold. So both ends of the slide are shared, and everything that distinguishes an exponent of 0.5 from an exponent of 1 from an exponent of 2 is in the middle.

The middle nobody could have guessed

A struck note has no steady state, only a slide from one spectrum to another — so the question is what the middle carries that the ends do not. The answer is exact rather than statistical: every loss law in the family leaves the strike with the same spectrum and ends in the same silence, so both endpoints carry precisely nothing about which of them it is. The whole difference is 41.3 decibels, and it peaks 0.38 seconds in, seven per cent of the way through the note.

timbre · Envelope
Scored the way these figures score it, a clarinet is the worst of the six. One close triad at 70 decibels through 17 registers, drawn once for each of the 6 spectra to hand. The score is the share of every partial written, which is the quantity the register figure published. pure 100 per cent at best, string 79 per cent at best, clarinet 54 per cent at best, reed 71 per cent at best, bell 52 per cent at best, organ 72 per cent at best. A clarinet's four even partials are twenty-eight decibels below its odd ones and are inaudible beside their own neighbours before any chord is built, so counting them in the denominator makes the spectrum that survives its own masking best look like the one that survives it worst.

A clarinet keeps what a string loses

Every masker, probe, chord, line and texture until now is eight partials falling as 1/n, and it was not even an option a placement could pass. Sweeping the six spectra to hand says the clarinet is the worst of them — 54 per cent of itself at best against a string's 79 — and that answer is an artefact of the score. Counted against what each note keeps on its own, the clarinet keeps 100 per cent where the string keeps 79, because its components stand a twelfth apart rather than an octave. The missing parameter was the spectrum; the second missing parameter was the denominator.

perception · Masking
The attack is the balance dial, turned by the clock. The level of a violin against a clarinet on one note at 392 hertz, moment by moment through the attack, with both players starting together. Two envelopes rising at different rates are a balance, so this axis is the same dial a conductor turns — and its whole travel is 6.02 decibels, which is twenty times the log of the ratio of the two attack times, 45 against 90 milliseconds, and nothing else. The pair does not begin as one player alone: both envelopes leave zero at the same slope ratio, so the dial starts at a finite offset rather than at silence. The dashed line is the balance at which the composite changes owner, -3.48 decibels — inside the travel, so the note belongs to a clarinet for its first 29 milliseconds and to a violin for the rest of its life.

The blend arrives before the note does

Nine essays on spectrum draw a steady state, and the strongest cue that two instruments are two instruments is that they do not start together. Two envelopes rising at different rates turn out to be a balance — the same dial an earlier essay swept — so the attack is that dial moved by the clock, and its whole travel is fixed at twenty times the log of the two attack times. It is six decibels for a clarinet with a violin against a crossing twelve to twenty-two decibels out, so one pair in ten changes hands during its own attack, and which one depends on a convention rather than on the instruments.

timbre · Spectrum
A doubled pizzicato gives its note away while it is still the louder. The power of a violin plucked, against a flue pipe holding the same note at 392 hertz, through the first 600 milliseconds of the pluck, with the pluck starting 12 decibels up and its fundamental decaying over 1 second. With each partial losing level in proportion to its number, the composite stops resembling the pluck at 70 ms, when the pluck is still 5.2 decibels the louder. With every partial fading together it would keep the note until 543 ms. The dashed line is the balance at which the steady-state doubling changes owner, minus 20.6 decibels: the release crosses the owner long before its balance gets there, because what hands the note over is the pluck's upper partials going, not its level.

A doubled pizzicato gives its note away early

The attack turns the balance between two players on one note by a few decibels and stops. A pluck does not stop — every partial of it decays, so a pizzicato doubled by a held instrument walks the balance for the whole note, and the expectation was a handover as slow as the decay. It is fast. A one-second pizzicato over a flute loses its note in 70 milliseconds, while it is still five decibels the louder, because what hands the note over is its upper partials going first. A uniform fade would have kept it eight times as long.

timbre · Spectrum
How long a doubled pizzicato keeps its note, seat by seat, in two rooms. How long a doubled violin pizzicato on 392 hertz keeps its note against the metres from the players, the pluck starting 12 dB up and decaying over 1 s with a loss exponent of 1. a concert hall, a flue pipe: 1 → 86 ms, 1.5 → 123 ms, 2 → 226 ms, 3 → 359 ms, 5 → 445 ms, 7 → 481 ms, 10 → 506 ms, 15 → 522 ms, 20 → 528 ms, 30 → 532 ms; a concert hall, an oboe: 1 → 52 ms, 1.5 → 55 ms, 2 → 59 ms, 3 → 77 ms, 5 → 149 ms, 7 → 195 ms, 10 → 224 ms, 15 → 242 ms, 20 → 248 ms, 30 → 254 ms; a concert hall, a clarinet: 1 → 44 ms, 1.5 → 45 ms, 2 → 45 ms, 3 → 47 ms, 5 → 53 ms, 7 → 65 ms, 10 → 86 ms, 15 → 105 ms, 20 → 112 ms, 30 → 118 ms; a large stone church, a flue pipe: 1 → 440 ms, 1.5 → 578 ms, 2 → 651 ms, 3 → 728 ms, 5 → 784 ms, 7 → 803 ms, 10 → 814 ms, 15 → 820 ms, 20 → 822 ms, 30 → 824 ms; a large stone church, an oboe: 1 → 56 ms, 1.5 → 65 ms, 2 → 89 ms, 3 → 207 ms, 5 → 281 ms, 7 → 303 ms, 10 → 315 ms, 15 → 322 ms, 20 → 325 ms, 30 → 326 ms; a large stone church, a clarinet: 1 → 44 ms, 1.5 → 43 ms, 2 → 43 ms, 3 → 44 ms, 5 → 53 ms, 7 → 67 ms, 10 → 80 ms, 15 → 89 ms, 20 → 92 ms, 30 → 94 ms. The mid-band critical distance is 5.3 m in a concert hall and 2.3 m in a large stone church. In none of the 60 cases does the note return to the pluck once it has left.

A room keeps a pizzicato from giving its note away

Doubled by a flute, a one-second pizzicato loses its note in 70 milliseconds dry, because its upper partials go first. The question left open was whether a room, whose reverberation keeps those partials alive, gives the note back afterwards. It does not give it back. It stops the note going: ten metres into a concert hall the pluck keeps it for 506 milliseconds, in a stone church for 814, and the room's own uneven decay takes back between a quarter and two fifths of that. In a room the loss law that decided everything dry matters a tenth as much, because the room's decay has become the clock.

timbre · Spectrum
A room pulls the compass apart rather than evening it out. How long a pizzicato entering 6 decibels above a held note keeps the composite spectrum, at eight pitches across two and a half octaves, heard 15 metres from the stage. no room: 0.07, 0.05, 0.06, 0.06, 0.02, 0.05, 0.05, 0.04 seconds; a concert hall: 0.61, 0.27, 0.50, 0.38, 0.01, 0.25, 0.27, 0.19 seconds; a large stone church: 1.03, 0.39, 0.79, 0.56, never, 0.33, 0.35, 0.23 seconds. Dry the figures barely move — a spread of 3.0 across the whole compass — because a room is the thing that varies with frequency and there is none. In a hall the spread is 44. The room does not scale the dry answer by a constant: it multiplies it by between four and nine times depending on the note, and at C6 it makes the pluck's position worse rather than better, because the two instruments' spectra nearly coincide there and the pluck starts only 4.0 decibels ahead instead of twelve.

One note in the compass loses its pizzicato

Dry, how long a pluck keeps the composite spectrum barely depends on which note it plays: three-hundredths of a second at the worst pitch and seven at the best, a spread of three. In a concert hall the same eight notes spread by a factor of forty-three, and in a stone church one of them never gets the note at all. The room does not scale the dry answer by a constant — it multiplies it by between four and nine times depending on the pitch, and at the one note where the two instruments' spectra nearly coincide it makes the pluck's position worse instead of better.

timbre · Spectrum

Named alongside it

The objects these essays reach for when they reach for this one.

SpectrumPartialRoughnessSource-filterOrchestrationBrightnessEnvelopeAttack transientRegisterCritical bandwidthDecayExcitation point

All concepts