Concept

Spectrum — where it appears

The list of a sound's partials together with their amplitudes. It is nearly all of what the ear uses to identify a sound, and the phase relations between those partials — which change the waveform completely — are nearly inaudible.

Named by 48 essays across 7 fields — each of them below, with the objects they name alongside it.

A note with its first partial removed. The spectrum of a 220 Hz tone with the lowest partial deleted, and the wave that remains. The wave still repeats 220 times a second, because the repeat rate of a sum of harmonics is fixed by the spacing between them and the spacing has not changed. The pitch heard is the one that is no longer in the sound.

The note that is not there

A telephone reproduces nothing below about three hundred hertz, and a bass line down at eighty comes through it perfectly. The pitch that is heard is not a frequency present in the sound, and one nineteenth-century experiment settles what it is instead.

intervals · Missing fundamental
The critical band, measured in semitones. The width of the ear's frequency-analysis band at each pitch, converted from hertz into semitones. Two intervals drawn as horizontal lines cross the curves: below the crossing the interval fits inside one band and its notes are not resolved from each other, and above it they are.

A third is rougher in the bass

Consonance is usually presented as a property an interval has. It is not. The same major third is muddy two octaves below middle C and clean two octaves above it, the ratio never changed, and the frequency where it stops being muddy can be solved for.

intervals · Consonance
Four spectra of the same note. The amplitude of each partial for 4 timbres at the same pitch — pure, string, clarinet, bell. These are the exact lists the sound buttons here synthesise from, so the picture and the sound are the same data.

The ear hears the list, not the shape

Two sounds with the same partials and different phases have completely different waveforms and sound identical. What the ear extracts is a list of frequencies and strengths, and everything else is discarded.

timbre · Spectrum
The vowel in "hod", sung at 110 Hz. The partials of a 110 Hz note, each drawn at the amplitude the vocal tract's resonances give it. The peaks of the curve are the formants — 730 Hz and 1090 Hz — and they stay where they are when the pitch changes, because they are a property of the shape of the mouth and not of the note being sung.

A vowel is two resonances

The vowel in "heed" is the same vowel sung high or low, and nothing about it is a property of the note. It is two peaks in the response of the mouth, sitting at fixed frequencies while the partials of the voice slide underneath them.

timbre · Spectrum
Roughness across an octave. Sensory dissonance between two complex tones as the upper one is swept through 19.02 semitones, computed by summing the roughness between every pair of their partials. Nothing here is placed by hand. The wells this spectrum produces sit on 7/5, on 11/7, on 5/3, on 9/5, on 11/5, on 7/3, on 13/5, found by scanning the curve rather than by marking them. The wells are where this spectrum's partials coincide, and for a spectrum with no even partials they are not where an octave-based scale would put them.

A scale without an octave, and the spectrum that asks for it

Every scale so far repeats at the 2:1. That looks like a law and it is a consequence — the octave fuses because every partial of the upper note lands on one of the lower. Take a spectrum with no even partials and the 2:1 stops doing that, while the 3:1 starts.

scales · Beyond twelve
Where each family's tone-hole lattice stops reflecting. The cutoff frequency of an open tone-hole lattice, from Benade's formula, for four woodwind geometries: clarinet 1824 Hz, oboe 2990 Hz, flute 1690 Hz, bassoon 506 Hz. Below its cutoff a note's wave turns round at the first open hole and the instrument is a tube of that length; above it the wave passes through the whole lattice and radiates from the far end, so the upper part of every note's spectrum leaves the instrument from the same place whichever note is fingered. That is what gives a family one recognisable voice across its range.

Above a certain note the holes stop working

A row of open tone holes reflects the wave and makes the tube shorter — up to a frequency. Above it the wave runs straight past the whole lattice and leaves from the bell, so the top of every note's spectrum radiates from the same place whichever note is fingered. That cutoff is computable, it differs by family, and it is most of what makes an oboe sound like an oboe.

instruments · Tone holes
A string struck at one 7th of its length. The amplitude of each partial of an ideal string excited at 0.1429 of its length. The mode shape is a sine, so a partial with a node at the excitation point cannot be set moving at all: partials 7, 14 are silent here. The envelope over the rest is one over n, a struck string's.

Where the hammer lands

Strike a string at exactly one over n and the nth partial is silent, because the hammer has landed on that mode's node and cannot move it. A piano's hammers strike between a seventh and a ninth of the way along, which puts the seventh partial — the most dissonant member of the series — at or near a null. That is a design decision made in wood, and it takes one line of trigonometry.

instruments · Excitation point
The same note, hit at a middling dynamic. The spectrum of a struck string with the hammer's own contact time applied as a low-pass. Contact lasts 1.60 ms at this force, against 2.26 ms at the softest and 0.95 ms at the loudest drawn — felt is a nonlinear spring, so a harder blow is a shorter contact and a brighter note. The spectral centroid moves from partial 1.5 to partial 2.2, which is a change of timbre and not of loudness.

A hammer is not an impulse

Contact lasts a couple of milliseconds, which low-passes the note — any partial whose half-period is shorter than the contact is barely excited. Piano felt is a spring that stiffens as it compresses, so a harder blow makes the contact shorter, the corner higher and the note brighter. A loud note is not a scaled-up quiet one, and no linear model gives that.

instruments · Excitation point
Helmholtz motion, bowed at 9% of the way from the bridge. Above: the string at 5 instants of one period. It is two straight lines meeting at a corner, and the corner travels round the string rather than the string swinging. Below: the resulting force on the bridge, a sawtooth whose two segments are in the ratio 0.09 to 0.91 — the bow's own position. A sawtooth contains every harmonic at exactly one over n, so the spectrum barely changes with bow position even though the waveform plainly does.

The bow makes a corner

A bowed string is not a string swinging. It is two straight lines meeting at a single sharp corner that travels round the string once per period, triggering the slip that keeps it going. The force on the bridge is therefore a sawtooth, and a sawtooth is every harmonic at exactly one over n — which is why a bowed string is the most nearly perfect harmonic series in the orchestra.

instruments · Bowed string
A bowed string on 196 Hz, through a violin body. The source is a sawtooth at one over n; the filter is the body's measured response, with A0 at 275 Hz, B1− at 460 Hz, B1+ at 540 Hz, bridge hill at 2500 Hz. What is radiated is their product, drawn as the bars. The resonances stay where they are when the note changes, exactly as a vowel's formants do — which is why an instrument has a voice rather than a tone, and why the same argument that identifies a vowel identifies a violin. The body frequencies are measured means over full-size instruments rather than computed from a plate.

The body is the filter

A violin string radiates almost nothing. What reaches a room is the string's sawtooth multiplied by the body's response, and that response is a comb of measured resonances that stays put while the note moves. It is the same arithmetic that identifies a vowel, on wood instead of a mouth — which is why an instrument has a voice rather than a tone.

timbre · Spectrum
Where each frequency goes, from a source 18 cm across. Polar response of a circular radiator of radius 9 cm at 200 Hz (ka = 0.3), 800 Hz (ka = 1.3), 2000 Hz (ka = 3.3), 5000 Hz (ka = 8.2). Zero degrees is straight ahead. The low frequency is a circle — it goes everywhere — and the high one is a narrow lobe with nulls either side of it, so a listener off to the side hears the same note with its top missing.

An instrument points

A source radiates evenly while it is small compared with the wavelength and beams once it is not, and the crossover is one number. So the same instrument is omnidirectional in its bottom octave and a searchlight in its top one — which means its spectrum depends on where the listener is standing, and a microphone position is a choice about what the instrument sounds like.

timbre · Room acoustics
Three spectra, and where each one's consonances fall. The positions of the roughness minima for 3 partial sets, over 12 semitones from the same fundamental, found by scanning each curve rather than by marking them. A harmonic tone — 498 cents at 3.1 per cent prominence, 702 cents at 37.1 per cent prominence, 884 cents at 6.2 per cent prominence. Partials stretched by 2.1 per octave — 533 cents at 3.7 per cent prominence, 751 cents at 40.1 per cent prominence, 947 cents at 7.3 per cent prominence. A bar's partials — 1, 2.76, 5.40, 8.93 — 694 cents at 5.6 per cent prominence, 870 cents at 14.2 per cent prominence, 982 cents at 1.0 per cent prominence, 1166 cents at 14.7 per cent prominence. The minima are computed by the same well-finder the dissonance curve uses, which requires a minimum to have one per cent of the curve's range on both sides of it before it counts.

A spectrum chooses its own scale

The roughness curve's dips are always described as landing on the small whole numbers. They do not land on numbers. They land where partials coincide, and stretching a spectrum by seven per cent moves every one of them by exactly seven per cent — so the scale a set of intervals belongs to is a property of the instrument's spectrum rather than of arithmetic.

intervals · Consonance
Every voicing of a major triad, least rough first. All 27 arrangements of the same three pitch classes within 3 octaves from 131 Hz, scored for roughness. The best is spaced 19 then 9 semitones — wide below, close above — and the worst is the chord in close position at the bottom of the range, 6.6 times rougher with exactly the same notes in it.

Where to put the third

Take three pitch classes, four octaves to put them in, and score all twenty-seven arrangements. The smoothest is root, fifth an octave up, third two octaves up — which is partials one, three and five of the harmonic series — and the roughest, at every register tried, is the chord in close root position at the bottom of the range. Every orchestration manual states that rule and none of them derives it.

harmony · Consonance
Six reverberation times for one room. Sabine's arithmetic evaluated in each octave band from the published absorption coefficients of the surfaces. a large stone church runs from 6.3 seconds at 125 Hz to 2.4 at 4 kHz — a bass ratio of 1.38, where concert halls are specified between 1.1 and 1.25.

A room does not decay evenly

Sabine's arithmetic gives one number and absorption is a strong function of frequency, so a room has six reverberation times rather than one. A stone church rings for 6.3 seconds at 125 hertz and 2.4 at 4 kilohertz, which means a chord left in it does not fade — it changes shape, losing its top before it loses its bottom, and arriving at the listener as a different sonority from the one played.

timbre · Room acoustics
Roughness and loudness of one interval against level. A minor third on C4 evaluated at every level from 20 to 100 dB SPL per note, with both quantities drawn relative to their own value at 60 dB. Roughness is the Plomp–Levelt sum over the partials that are above ISO 226's threshold at that level; loudness is the same partials in sones. Between 50 and 95 dB the interval grows 3.2 × 10⁴ times rougher and 24.9 times louder, so roughness grows like loudness raised to the power 3.2.

The same chord is harsher when it is louder

Every roughness number so far was computed at a level nobody stated. Roughness is the product of two partial amplitudes, so it is quadratic in pressure, while loudness is compressive — which makes a minor third at middle C thirty-two thousand times rougher at fortissimo than at pianissimo and only twenty-five times louder. A chord has no single consonance to report.

intervals · Consonance
The roughness curve for ideal bar. The partials are at 1 : 2.756 : 5.404 : 8.933 : 13.340 times the fundamental, which is not a harmonic series, so nothing about where the wells fall can be read off the small whole numbers. Sensory dissonance between two complex tones as the upper one is swept through 13.00 semitones, computed by summing the roughness between every pair of their partials. Nothing here is placed by hand. The wells this spectrum produces sit 7 cents below 3/2, 14 cents below 5/3, 13 cents above 7/4, found by scanning the curve rather than by marking them. The wells are where this spectrum's partials coincide, and for a spectrum with no even partials they are not where an octave-based scale would put them.

The spectrum that was supposed to explain the gamelan

Run the model that designed the Bohlen–Pierce scale on a bar's partials and it asks for a compressed pseudo-octave at 1166 cents and wells at 694 and 870. A measured slendro's degrees are at 231, 474, 717 and 955, and only one of the four is near anything the model wants. Sweep the partials and a spectrum wanting a 240-cent step can be found — sitting sixty per cent of the way up its own roughness curve, which is what a scale-level test says about the whole comparison.

timbre · Beyond twelve
The flow through the larynx, over two periods of a 110 Hz note. Volume flow against time, in Rosenberg's two-half-cosine model of the glottal pulse — a slow opening, a faster closing, and a closed phase during which no air passes at all. M1 — chest is open for 50 per cent of each period and opens 2.4 times as slowly as it closes. Nothing here is a displacement: the folds are a valve on a steady stream of air, and the flat stretches are the moments they are shut. At 110 Hz each period lasts 9.1 milliseconds, of which 4.5 is silence.

The other instrument with a reed

The folds do not vibrate the way a string does. They open and shut across a steady stream of air, once per period, and what leaves the larynx is a train of flow pulses with a closed phase in it. Everything said about the voice's tone is a statement about the shape of that pulse — and the shape has two numbers in it.

instruments · The voice
Two mechanisms, the notes both of them make, and the seam. The frequency range of each laryngeal mechanism for an adult male voice, on a logarithmic axis, with the band both can produce shaded. M1 — chest runs 82–349 Hz and M2 — falsetto runs 220–698 Hz, so 799 cents of the range — 8.0 semitones — can be sung either way. The two dots inside that band are the measured signature that this is a bifurcation rather than a threshold: the change upward happens at 330 Hz and the change downward at 294 Hz, 200 cents lower. A threshold is crossed at the same place in both directions and this is not.

Two mechanisms, and the seam between them

Every singer has a place in the range where the voice changes character, and eight semitones of it can be produced either way. The measurement that settles what kind of a place it is takes ten seconds: the change upward happens two hundred cents higher than the change downward, and a threshold cannot do that.

instruments · The voice
How near the node a hammer has to land. The level of partial 7, relative to the loudest partial of the same note, against how far the hammer is from that partial's node — for a contact 2.6 per cent of the speaking length wide, which is 16 mm on a 620 mm string. At the node exactly the partial is absent whatever the width, because a mode shape is odd about its own node and integrating an odd function symmetrically gives zero. Twenty decibels of suppression needs the hammer within 2.82 mm, and thirty needs 0.89 mm. The width itself barely matters in the musical range: its sinc reaches its first zero only at partial 77.

The hammer is not a point either

The comb of missing partials is computed throughout for a contact of no width, and a piano hammer is sixteen millimetres of felt on a string of six hundred. Integrating the mode shape across the felt turns out not to fill the null in: a mode is odd about its own node, so a symmetric contact centred on it gives an exact zero however wide it is. What destroys the null is being in the wrong place, and the tolerance is 2.8 millimetres for twenty decibels of suppression and 0.89 for thirty. The design decision found at the outset is a manufacturing tolerance.

instruments · Excitation point
One instrument, five spectra. The first 12 partials of a bowed string at 5 pitches, each passed through the same fixed body response and normalised to its own loudest partial. The resonances stay where they are and the partials slide under them, so the pattern is different at every note: the power-weighted centroid runs from 1.02 to 1.71 across the compass and is not monotonic in pitch. A source with no body at all would give 2.35 at every pitch, which is the single number every roughness figure here uses for a string.

An instrument is not one timbre

Every roughness verdict so far uses one partial list per instrument, and a violin does not have one. Its body's resonances stay where they are while the fundamental moves, so the radiated spectrum is different at every pitch: the power-weighted centroid runs from 1.02 to 1.64 across the compass, and it is not monotonic. Which means the result that a spectrum chooses its own scale gives a different scale at every note on the same instrument — three minima in the dissonance curve at the bottom of a violin's range, two in the middle and four at the top.

timbre · Spectrum
Six ways to put three players on three notes. The same chord — G3, B♭3, D4 — played by clarinet, oboe, voice in all 6 possible assignments, scored by the roughness each produces. Every bar is the same pitches and the same instruments; only who is on which note changes. The worst is 1.42 times the best, which is a factor a score can control and a chord symbol cannot express at all. Each row is labelled from the bottom note upward.

Which player on which note

An interval's roughness depends on which instrument is underneath, so the pair does not commute. Three players over three notes is the smallest thing that asymmetry has anywhere to go: six assignments, all of them the same chord, and across 450 of them the roughest averages half again the smoothest and reaches six times it. It is orchestration in the only form that can be computed here — not which chord, and not which voicing, but who is on which note.

timbre · Spectrum
How much of each spectrum a listener can assemble into one note. Each partial of each spectrum at the harmonic number it is nearest, against the whole-number series that fuses the most of them, with anything more than 1 per cent out marked as heard separately. an ideal string keeps 10 of 10; a piano string keeps 9 of 10; a bell keeps 7 of 8; a bar keeps 2 of 6; a kettledrum keeps 3 of 5. The fundamental is capped at a tenth of the top partial, and the cap is load-bearing rather than tidy: a bell's ratios are all whole multiples of a tenth, so an unconstrained search finds a fundamental twenty-five harmonics down, calls every partial exact, and reports that a bell fuses perfectly. Nothing that high is resolved and the low harmonics of it are not there.

The spectrum that will not fuse

A partial about one per cent off its harmonic is heard as a sound of its own rather than as part of a note. Apply that criterion to a whole spectrum instead of to one mistuned component and it becomes a count: a piano string keeps nine of its ten partials, a bell keeps seven of eight, a bar keeps two of six. The physics of inharmonicity has had an essay here for a long time. This is what it sounds like.

perception · Auditory scene
Trumpet at three dynamics, as a spectrum rather than a level. The radiated partials of a trumpet at 45, 70, 95 decibels, each normalised to its own strongest partial so that only the SHAPE is compared. A linear source would give three identical pictures. This one does not: the spectral centroid moves from partial 2.19 to 6.41, a factor of 2.92, because the excitation is nonlinear and blowing harder steepens the pressure front rather than scaling it. The tilt used is 3 decibels per octave of partial number per ten decibels of level, referred to 70 dB — a stipulated, ordinal number, not a measurement of any instrument.

A dynamic mark changes what a note is

Every spectrum until now is a shape with a level in front of it, so that playing ten decibels louder raises every partial by ten. That is true of exactly one instrument in an orchestra. Everybody else steepens their own spectrum as they lean on it, and a trumpet's centre of gravity moves from the second partial to the sixth across a dynamic range while an organ flue pipe's does not move at all.

timbre · Orchestration
Which of the four terms is doing the removing, note by note. Each of three terms has a corner — a partial number above which it is the one taking the energy away — and the lowest corner at each pitch is the term that binds. The comb is flat at 8, because where the hammer lands is a fraction of the length and does not change with the note. The contact time's corner falls from 34 at A0 to 0.5 at A7. They cross at about E3, 165 hertz: below it the strike point decides the spectrum and above it the contact time does. The dispersion's corner is above both at every pitch on the instrument — 5.8 at the top against a contact corner of 0.5 — so that earlier term is real and is nowhere the binding one. That is a statement no one of them could make on its own.

Four terms, and only one of them binds

Six earlier essays each computed one term of a piano's excitation spectrum and handed it to the next, and nobody had multiplied them. Multiplied, three of the four have a corner — a partial above which they are the term doing the removing — and putting the three corners on one axis says which is in charge at each pitch. Below E3 the strike point decides; above it the contact time does; and the dispersion, which is the newest and most laborious term, is nowhere the binding one.

timbre · Excitation point
Which interval gives a tuner the deepest null. A tuner listening to an interval p:q is listening to the lower note's p-th partial against the upper note's q-th, and how deep the beat's trough goes is decided by those two amplitudes rather than by the interval. For a string spectrum, whose partials fall as one over n, the minor third pairs partial 6 against partial 5 at a ratio of 1.18 for a dip of 21.8 decibels; the major third pairs partial 5 against partial 4 at a ratio of 1.25 for a dip of 19.1 decibels; the fourth pairs partial 4 against partial 3 at a ratio of 1.32 for a dip of 17.2 decibels; the fifth pairs partial 3 against partial 2 at a ratio of 1.52 for a dip of 13.8 decibels; the major sixth pairs partial 5 against partial 3 at a ratio of 1.65 for a dip of 12.2 decibels; the minor sixth pairs partial 8 against partial 5 at a ratio of 1.67 for a dip of 12.0 decibels; the octave pairs partial 2 against partial 1 at a ratio of 2.00 for a dip of 9.5 decibels. The best is the minor third at 21.8 and the worst is the octave at 9.5, which is the reverse of the order a tuner is usually taught to trust: the deepest null in the list is on the interval whose coincidence sits highest in the spectrum, where adjacent partials are nearly equal in strength.

A beat has a depth, and six essays held it at one

Every beat figure so far adds two tones of equal amplitude, which is the single ratio at which the trough of a beat is a true null — and a null is what a tuner is actually listening for. Vary the ratio and the picture changes: at two to one the dip is nine and a half decibels, at ten to one it is under two, and the interval that gives the shallowest null of all is the octave.

intervals · Beating
Intonation is a unison problem and nothing else. The roughness between two instruments on one note, against how far apart they are in cents, drawn for a unison and for the intervals beside it. A perfect unison is 0.0007 — the partials coincide and there is nothing to beat. Five cents apart it is 0.0465, 65 times as rough, and ten cents apart it is rougher than a major third played exactly. The mechanism is that partial n of a note mistuned by c cents is mistuned by c cents as well, which is n times as many hertz — so the top of the spectrum enters the critical band long before the fundamental does. The other curves are flat, because a third's roughness is set by which partials nearly coincide and a few cents does not change which.

Two players on one note

Six essays have put one instrument on each note of a chord, and the commonest thing an orchestrator actually does is put two on the same note. Two independent sources add in power, so the composite is neither of them — except that it nearly always is one of them, because the level at which ownership changes hands is rarely at zero. And a unison ten cents out is rougher than a major third dead in tune.

timbre · Spectrum
How much of each instrument is narrow enough to steepen a wave. For each bore, the quantity that decides how nonlinear it is: the narrowest radius divided by the radius at each station, integrated along the tube. The shading is that integrand, so a bar that stays dark is a tube still doing damage to the wave and a bar that fades is a flare that has thinned it out. Divided by the instrument's own length the integral is a pure number: a plain cylinder 1.00, a tenor trombone 0.88, a trumpet 0.86, an F horn 0.59, a cone of a trumpet's length 0.24. A plain cylinder is 1 by construction, a cone of the same length and mouth is 0.24, and the ordering across the brass family is the one players give when asked which of them can be made to blare.

The partials the tube makes itself

Eleven earlier essays compute a passive linear resonator, and none of them ever says so. At a real fortissimo the air in a brass instrument is not linear: a compression outruns a rarefaction, the wave leans forward as it travels, and the fourth partial of a loud trumpet note is seventy decibels louder than a scaled-up quiet one — generated in the tube rather than at the lips. How much of it happens is an integral over the bore, and it is why a flugelhorn cannot be blown into being a trumpet.

instruments · Air column
A plucking point 50 millimetres from the nut, across a harpsichord's compass. The jacks stand in a rail and the rail is one object, so a register plucks at a fixed distance from the nut while the strings shorten by a factor of ten from the bass to the treble. At 50 millimetres that is one 36th of the string at C2 and one 3.5th at C6 — a plucking fraction that changes by a factor of 10.1 without the maker moving anything. The comb corner is one over that fraction and falls with it, from 36 to 3.5; the dispersion corner has the fraction under a cube root and falls only from 106 to 21. They never meet. Setting one over p equal to the cube root of two over three times the inharmonicity and the fraction gives a plucking point at p = √(1.5B), which on this instrument's iron wire is between 3.4 and 9.9 millimetres from the nut — nearer to it than any register a harpsichord has ever carried, the lute stop included. So the maker's one free choice is the binding term at every pitch and every register position, which is exactly what the piano's is not: on a piano the contact time takes the decision away above E3.

What the second register is for

Nine earlier essays take the plucking fraction as given — an eighth, a sixth, a seventh — and a harpsichord's jacks stand in a rail, so what is given is a distance and the fraction is a consequence: 50 millimetres is one thirty-sixth of the string in the bass and one three-and-a-halfth in the treble. Engage two registers on one string and the amplitudes add exactly, the near one fills the far one's missing partials, the pair digs holes of its own that neither has, and the combination comes out louder and rounder than either — including the bright one.

instruments · Excitation point
Who owns clarinet, oboe, voice at every balance. The composite of three players at 392 hertz belongs to whichever of them it is nearest in log-spectral distance, and here that is drawn over the whole plane of balances a conductor could set — the second and third players from 24 decibels below the first to 24 above. voice owns 79 per cent of the square. The three regions meet where all three distances are equal, which is the only balance at which the composite belongs to nobody: it is at -0.3 and -19.1 decibels, inside the square and therefore a balance an ensemble could actually be asked for. A trio has a colour of its own at one point, not over a region.

A section has a loudest member, not a colour

Two players on one note have a balance at which the composite belongs to neither, and that is what blending means. Three should have three such balances and no reason for them to agree — a trio with a rock-paper-scissors ownership would have no strongest member at all. Twenty trios, sixty pairwise comparisons, and not one disagreement: the possibility is real, arbitrary spectra do it once in twenty, and instruments never do.

timbre · Spectrum
Which pairs blend is a question about the note. The level at which a doubled pair's composite changes owner, drawn for all 15 pairs of 6 radiators over 2.6 octaves from 131 to 784 hertz. A pair blends when that level is inside the shaded band, which is the twenty-four decibels either way two players can manage; a curve outside it, or absent, is a pair one instrument owns at every balance. 6 of 15 pairs blend at the bottom of the range and 12 at the top. Every filter in this collection is fixed in frequency and the fundamental is not, so a radiator's shape is a function of pitch and so is everything computed from two of them — the blend ranking at the bottom and at the top disagree on 70 of 105 comparisons, which is more than half, so the order has turned over rather than merely shuffled.

The blend table has a row for every note

Eight earlier essays sound their instruments at one note, and one of them says why that cannot be innocent: every filter here is fixed in frequency and the fundamental is not. Swept over four octaves, the number of pairs that blend doubles from six to twelve, the ranking turns over rather than shuffles — seventy of a hundred and five comparisons swap — and a clarinet with an oboe goes from the best pair in the collection to the eleventh.

timbre · Spectrum
Four players on three notes, every arrangement. The 36 ways of putting 4 players on a 3-note chord so that every note is covered, ranked by roughness, all at one total loudness of 26.9 sones. Each row is shaded by which note carries the pair. The best is flue | clarinet+violin | oboe and the worst is clarinet | oboe+flue | violin, a factor of 2.21. Every earlier essay puts exactly one player on each note, which is a permutation; a doubling makes the arrangement a surjection instead, and the doubled note sounds neither of its two players but the composite they make. Which note gets the pair explains 7 per cent of the spread here and which players sit on the lowest note explains 89: the fourth player is a much smaller decision than the three that were already there.

The fourth player is a spectrum, not a decision

Six earlier essays put exactly one instrument on each note, which makes an arrangement a permutation — and the commonest operation in orchestration is a doubling, which does not. Four players on three notes give thirty-six arrangements instead of six, and the extra choice turns out to be the smallest thing on the page: which note carries the pair explains three per cent of the spread and which players sit on the bass explains eighty-nine. A doubled note can be priced as one player, and which one is not the one a spectral account would have named.

instruments · Orchestration
one measured slendro is the scale least committed to a spectrum. Each scale drawn, placed by how smooth it is against 2000 random scales of the same size, under 4 spectra. Low is smooth. Every one of them lands in the smoothest third under every spectrum, so the account that a spectrum chooses a scale survives as a statement about levels. What separates them is the length of each row's line, which is how much the answer moves when the instrument changes: one measured slendro spans 2.9 percentile points and Rast, Turkish theory spans 17.6, a factor of 6.0. The scale usually offered as the case for spectral matching is the one whose standing barely depends on the spectrum, and Rast, Turkish theory is the one that depends on it most.

The scale least committed to its own instrument

The essays on scales beyond twelve draw scales without spectra and spectra without scales, and none of them had put a tradition's own degrees through the roughness model. Doing it for six scales and five spectra says the account survives — every tradition lands between the eighth and the twenty-sixth percentile of random scales of its size, and under a pure tone not one of them does. But the scale whose standing moves least when the instrument changes is the gamelan's, at 2.9 percentile points, and the one that moves most is the five-limit diatonic at 17.4. The scale the argument is usually made about is the one it explains least.

scales · Beyond twelve
Partials 3, 4, 12 of "hod", over one vibrato cycle. The level of three partials of a 220 hertz note on the vowel in "hod", each about its own mean, over one cycle of a vibrato of ±71 cents at 6.0 hertz. The pale curve is the frequency deviation itself, for phase reference. Partial 3 at 660 hertz swings 4.22 decibels and peaks with the frequency; Partial 4 at 880 hertz swings 0.41 decibels and peaks twice a cycle; Partial 12 at 2640 hertz swings 8.62 decibels and peaks against it. The formants of this vowel are at 730, 1090, 2440 hertz and do not move; a partial below one rises as the frequency rises and one above it falls, so the modulations of a single note run in opposite directions at the same instant.

The partial that gets louder as it goes sharp

Eleven earlier essays sweep a set of partials and hold their amplitudes still, and a real tract does not move with the fundamental. Put the formant account and the sweeping account together and every partial acquires an amplitude modulation at the vibrato rate: 0.07 decibels on the fundamental and 8.6 on the twelfth partial of the same note. They are in phase below a formant and anti-phase above one — not ninety degrees apart — and the whole note swings 0.82 decibels, because they cancel.

instruments · The voice
The partials the air makes, on the series the bore actually has. The input impedance of a trumpet, with the partials wave steepening manufactures from its A4 at 428 hertz drawn on it as vertical marks. The steepening is a distortion of one periodic waveform, so it makes its partials at exact integer multiples — 856, 1285, 1713, 2141, 2569 hertz. The bore's own resonances are not at integer multiples of anything: the peaks the manufactured partials aim at sit at 886, 1346, 1812, 2278, 2741. So every manufactured partial lands flat of the peak it might have used, by 59, 81, 98, 108, 112 cents — 2.6, 2.7, 3.0, 4.0, 4.9 half-widths of the peaks in question, which is outside the half-power point of every one of them. The dashed line is where the bell stops reflecting, at 2945 hertz; above it there is no peak to land on or miss.

Which notes go brassy first

Wave steepening manufactures partials at exact integer multiples of the note being played. A brass instrument's resonances are not at integer multiples of anything, and twelve earlier essays have measured how far off they are without ever putting the two on one axis. Put them there and a manufactured partial never lands on a peak — never once, at any note, by a miss that is the same number of hertz every time. So the alignment cannot be what decides which notes blare, and the thing that does turns out to be the bell.

instruments · Air column
Two registers 50 microseconds apart on one string. The partials of a 355-millimetre string plucked at 32 and 50 millimetres from the nut, drawn twice: once with the two quills releasing together, once with the far one releasing 0.05 milliseconds later, which is 0.026 of this string's period. A delayed release turns partial n through 2·pi·f_n·dt, so the rotation is proportional to the partial number: the fundamental is turned 9 degrees and partial 24 is turned 226. Simultaneous, the pair is missing partials 17 and 20 — holes it digs for itself where the two combs are equal and opposite. Staggered, it is missing none of them: a rotation of anything at all takes two amplitudes out of opposition. The partials neither comb can excite at all — 7, 11, 22 — are filled either way, because where one comb is zero the sum is the other one whatever its phase. The fundamental's gain over the far register alone falls from 4.36 decibels to 4.34.

The interval between two quills

Two jacks on one key are voiced separately and do not let go at the same instant. That interval turns each partial of the later pluck through an angle proportional to its number — so it leaves the fundamental alone and inverts the twentieth partial, and the holes the pair digs for itself vanish at a hundredth of a period. What a regulator can tolerate turns out to be one fixed fraction of a period at every pitch, which on a five-octave instrument is a factor of sixteen in milliseconds.

instruments · Excitation point
Scored the way these figures score it, a clarinet is the worst of the six. One close triad at 70 decibels through 17 registers, drawn once for each of the 6 spectra to hand. The score is the share of every partial written, which is the quantity the register figure published. pure 100 per cent at best, string 79 per cent at best, clarinet 54 per cent at best, reed 71 per cent at best, bell 52 per cent at best, organ 72 per cent at best. A clarinet's four even partials are twenty-eight decibels below its odd ones and are inaudible beside their own neighbours before any chord is built, so counting them in the denominator makes the spectrum that survives its own masking best look like the one that survives it worst.

A clarinet keeps what a string loses

Every masker, probe, chord, line and texture until now is eight partials falling as 1/n, and it was not even an option a placement could pass. Sweeping the six spectra to hand says the clarinet is the worst of them — 54 per cent of itself at best against a string's 79 — and that answer is an artefact of the score. Counted against what each note keeps on its own, the clarinet keeps 100 per cent where the string keeps 79, because its components stand a twelfth apart rather than an octave. The missing parameter was the spectrum; the second missing parameter was the denominator.

perception · Masking
One doubling, held down a phrase. Where a single held arrangement of 4 players on 3 notes stands among the 36 at each chord of a 5-chord passage, best at the top, with what each chord would rather have named along the bottom. The held answer is flue pipe · trumpet · clarinet+violin, and it is the chord's own first choice at 4 of 5 of them. Holding it costs 16.2 per cent of the passage's roughness against re-scoring every chord — which is 2.4 per cent of the range the choice actually spans, since the arrangements at one chord differ by a factor of 7.8 on average. The cost is not spread over the passage: 1 chord carries nearly all of it.

An orchestrator doubles a line, not a chord

Three earlier essays made the objective a functional over a passage and a later one went back to holding one chord still. Put the doubling back into time and the retreat turns out to have been cheap: one arrangement held down a five-chord phrase is that phrase's own best answer at four of its five chords and costs 2.4 per cent of the range the choice spans — while the forward mask named earlier as the third temporal constant reaches for twenty milliseconds rather than two hundred, and cannot change the answer at any pace at all.

instruments · Orchestration
The attack is the balance dial, turned by the clock. The level of a violin against a clarinet on one note at 392 hertz, moment by moment through the attack, with both players starting together. Two envelopes rising at different rates are a balance, so this axis is the same dial a conductor turns — and its whole travel is 6.02 decibels, which is twenty times the log of the ratio of the two attack times, 45 against 90 milliseconds, and nothing else. The pair does not begin as one player alone: both envelopes leave zero at the same slope ratio, so the dial starts at a finite offset rather than at silence. The dashed line is the balance at which the composite changes owner, -3.48 decibels — inside the travel, so the note belongs to a clarinet for its first 29 milliseconds and to a violin for the rest of its life.

The blend arrives before the note does

Nine essays on spectrum draw a steady state, and the strongest cue that two instruments are two instruments is that they do not start together. Two envelopes rising at different rates turn out to be a balance — the same dial an earlier essay swept — so the attack is that dial moved by the clock, and its whole travel is fixed at twenty times the log of the two attack times. It is six decibels for a clarinet with a violin against a crossing twelve to twenty-two decibels out, so one pair in ten changes hands during its own attack, and which one depends on a convention rather than on the instruments.

timbre · Spectrum
Every beat in a family is the same depth, and none is near the threshold. The 8 members of the beat family a octave mistuned by 6.0 cents makes at A3 on a real string, each placed at the rate it beats and at the modulation index a listener's filter delivers there. The rising curve is the published detection threshold for amplitude modulation, which is flat at 0.03 below about fifty fluctuations a second and rises above it. The flat dashed line at 0.667 is the depth the pair has in isolation, and it is the same for every member: on a spectrum falling as one over n, the k-th member pairs partials 2k and 1k, whose ratio is 2.00 whatever k is. What the filled points show is the smaller effect that does depend on the member — the partials on either side of the coincidence leak into the same filter, add level without adding fluctuation, and dilute the index from 0.662 to 0.405. Inside the countable rate window the narrowest margin over the threshold is a factor of 16.7, on member 4. The depth criterion removes nothing.

Every member of a beat family is the same depth

A mistuned octave's beats have been counted on their rate and their place, with the depth recorded as the thing left out and a prediction that the shallow upper members would take the count from four to two. The depth turns out not to fall at all: on any power-law spectrum every member of a family has exactly the modulation index its interval's own ratio gives, at every register, on every wire. The count does fall to two, and the thing that takes it there is the criterion that essay was already using.

intervals · Beating
An entrance stops being a loudness event and never stops being a colour one. The same oboe entering on the same note at the same level, against how many players were already sounding. Its contribution to the loudness falls from 15.5 phons to 0.32 — a factor of 48 — and crosses the one-phon difference limen at 5 players already playing. Its contribution to the roughness rises by a factor of 12.3 over the same range, because roughness is a sum over pairs and the entrant makes one new pair with everybody. Both curves are drawn as a share of their own largest value, since a phon and a squared pascal have no exchange rate. The claim is the two directions, not the crossing point of two units.

An entrance is a change of colour

Eight essays on orchestration move the assignment and hold the ensemble still, and a score does the opposite: it brings players in and takes them out. Loudness is a sum over parts and roughness is a sum over pairs, so the player who joins adds one term to the first and one to the second for everybody already there. What the entrance is worth in phons falls by a factor of forty-eight across the range an ensemble spans and crosses the difference limen at five players; what it is worth in roughness rises by twelve, and by a further factor of ten for every ten decibels the passage is played at.

form · Orchestration
Which chord of a passage has room for the part that is entering. An oboe entering on one note, tried at each chord of a five-chord passage, scored by how far its own partials sit above the threshold the ensemble already sounding puts over them. The best moment gives it 9.0 decibels of margin and the worst 0.3, a spread of 8.7 — and the best moment is not the quietest chord, which is vi, close below. Room for an entrance is spectral rather than dynamic. A chord with a hole in its written spacing need not have one in its spectrum, because the partials of its bass fill the middle whatever the notes above it do.

The chord that has room for an entrance

Three essays have made the ensemble something a score can change and none of them has asked when. The ensemble already sounding puts a masked threshold over whatever register an entering part takes, and that threshold is set by the voicing rather than by the dynamic — so the five chords of one passage differ by 8.7 decibels in how much of an entering oboe survives them, and the quietest chord of the five is the worst place in the passage to bring somebody in. Swept over the entrant's own pitch, the choice of moment is worth as much as the choice of register.

form · Orchestration
A doubled pizzicato gives its note away while it is still the louder. The power of a violin plucked, against a flue pipe holding the same note at 392 hertz, through the first 600 milliseconds of the pluck, with the pluck starting 12 decibels up and its fundamental decaying over 1 second. With each partial losing level in proportion to its number, the composite stops resembling the pluck at 70 ms, when the pluck is still 5.2 decibels the louder. With every partial fading together it would keep the note until 543 ms. The dashed line is the balance at which the steady-state doubling changes owner, minus 20.6 decibels: the release crosses the owner long before its balance gets there, because what hands the note over is the pluck's upper partials going, not its level.

A doubled pizzicato gives its note away early

The attack turns the balance between two players on one note by a few decibels and stops. A pluck does not stop — every partial of it decays, so a pizzicato doubled by a held instrument walks the balance for the whole note, and the expectation was a handover as slow as the decay. It is fast. A one-second pizzicato over a flute loses its note in 70 milliseconds, while it is still five decibels the louder, because what hands the note over is its upper partials going first. A uniform fade would have kept it eight times as long.

timbre · Spectrum
How long a doubled pizzicato keeps its note, seat by seat, in two rooms. How long a doubled violin pizzicato on 392 hertz keeps its note against the metres from the players, the pluck starting 12 dB up and decaying over 1 s with a loss exponent of 1. a concert hall, a flue pipe: 1 → 86 ms, 1.5 → 123 ms, 2 → 226 ms, 3 → 359 ms, 5 → 445 ms, 7 → 481 ms, 10 → 506 ms, 15 → 522 ms, 20 → 528 ms, 30 → 532 ms; a concert hall, an oboe: 1 → 52 ms, 1.5 → 55 ms, 2 → 59 ms, 3 → 77 ms, 5 → 149 ms, 7 → 195 ms, 10 → 224 ms, 15 → 242 ms, 20 → 248 ms, 30 → 254 ms; a concert hall, a clarinet: 1 → 44 ms, 1.5 → 45 ms, 2 → 45 ms, 3 → 47 ms, 5 → 53 ms, 7 → 65 ms, 10 → 86 ms, 15 → 105 ms, 20 → 112 ms, 30 → 118 ms; a large stone church, a flue pipe: 1 → 440 ms, 1.5 → 578 ms, 2 → 651 ms, 3 → 728 ms, 5 → 784 ms, 7 → 803 ms, 10 → 814 ms, 15 → 820 ms, 20 → 822 ms, 30 → 824 ms; a large stone church, an oboe: 1 → 56 ms, 1.5 → 65 ms, 2 → 89 ms, 3 → 207 ms, 5 → 281 ms, 7 → 303 ms, 10 → 315 ms, 15 → 322 ms, 20 → 325 ms, 30 → 326 ms; a large stone church, a clarinet: 1 → 44 ms, 1.5 → 43 ms, 2 → 43 ms, 3 → 44 ms, 5 → 53 ms, 7 → 67 ms, 10 → 80 ms, 15 → 89 ms, 20 → 92 ms, 30 → 94 ms. The mid-band critical distance is 5.3 m in a concert hall and 2.3 m in a large stone church. In none of the 60 cases does the note return to the pluck once it has left.

A room keeps a pizzicato from giving its note away

Doubled by a flute, a one-second pizzicato loses its note in 70 milliseconds dry, because its upper partials go first. The question left open was whether a room, whose reverberation keeps those partials alive, gives the note back afterwards. It does not give it back. It stops the note going: ten metres into a concert hall the pluck keeps it for 506 milliseconds, in a stone church for 814, and the room's own uneven decay takes back between a quarter and two fifths of that. In a room the loss law that decided everything dry matters a tenth as much, because the room's decay has become the clock.

timbre · Spectrum
A room pulls the compass apart rather than evening it out. How long a pizzicato entering 6 decibels above a held note keeps the composite spectrum, at eight pitches across two and a half octaves, heard 15 metres from the stage. no room: 0.07, 0.05, 0.06, 0.06, 0.02, 0.05, 0.05, 0.04 seconds; a concert hall: 0.61, 0.27, 0.50, 0.38, 0.01, 0.25, 0.27, 0.19 seconds; a large stone church: 1.03, 0.39, 0.79, 0.56, never, 0.33, 0.35, 0.23 seconds. Dry the figures barely move — a spread of 3.0 across the whole compass — because a room is the thing that varies with frequency and there is none. In a hall the spread is 44. The room does not scale the dry answer by a constant: it multiplies it by between four and nine times depending on the note, and at C6 it makes the pluck's position worse rather than better, because the two instruments' spectra nearly coincide there and the pluck starts only 4.0 decibels ahead instead of twelve.

One note in the compass loses its pizzicato

Dry, how long a pluck keeps the composite spectrum barely depends on which note it plays: three-hundredths of a second at the worst pitch and seven at the best, a spread of three. In a concert hall the same eight notes spread by a factor of forty-three, and in a stone church one of them never gets the note at all. The room does not scale the dry answer by a constant — it multiplies it by between four and nine times depending on the pitch, and at the one note where the two instruments' spectra nearly coincide it makes the pluck's position worse instead of better.

timbre · Spectrum
The tempo moves it further than the touch does. One measured slendro, scored among random scales of its size under a free bar, 4 s, at five tempi and under each touch. Left to ring it runs from 29 at 0.15 seconds a note to 83 at 2.4 — a span of 54 percentile points, where the two touches differ by at most 17. So the scale is smoother than most of its size when the music is fast and rougher than most when it is slow, and how the bar is damped is the smaller decision. The two touches converge at the slow end because a bar that has died before its successor is sounding against nothing whatever the player does.

The tempo moves a scale further than the touch

A gamelan is played two ways on the same bars: a saron's are damped as the next is struck and a gendèr's ring over their resonators. That decision moves a slendro's standing among random scales of its size by up to seventeen percentile points, which is real. Over the tempo levels a piece actually moves through it moves by fifty-four — from the twenty-ninth percentile at a fast elaboration to the eighty-third at a slow one. The same five pitches on the same bars are a smoother-than-average scale and a rougher-than-average one, and which depends on how fast they are played.

scales · Beyond twelve
The scale's standing belongs to the ringing instrument. Where a measured slendro sits among random five-note scales, at three density ratios, scored three ways: the two instruments together, the fast ringing part on its own, and the slow damped part on its own. At one slow note to 2 fast ones the ensemble is at the 7th percentile, the ringing part alone at the 6th, and the damped part alone at the 95th; At one slow note to 4 fast ones the ensemble is at the 5th percentile, the ringing part alone at the 7th, and the damped part alone at the 95th; At one slow note to 8 fast ones the ensemble is at the 5th percentile, the ringing part alone at the 5th, and the damped part alone at the 95th. A low percentile is a scale smoother than most of its size. The damped part on its own is rougher than nineteen random scales in twenty, because it has almost no simultaneity for its intervals to be smooth in; the ensemble is at the ringing part's figure throughout. And the cross pairs, which are a quarter of the roughness, do not make the ensemble worse than either part alone.

The scale belongs to the ringing instrument

A gamelan plays two instruments on one scale at once — a saron damped at every stroke and a gendèr several times faster with its bars left to ring — and a quarter of the roughness a listener receives falls on pairs that cross between them, which no figure had computed. The cross pairs turn out to be no worse than either stream's own. What decides the scale's standing is the fast part: the ensemble sits at the sixth percentile among random five-note scales and so does the ringing instrument alone, while the damped one alone sits at the ninety-sixth.

scales · Beyond twelve
The arch belongs to hearing, and the spacing only moves it. The share of a close major triad's twenty-four components that stand above what the rest of the chord masks, at 70 dB, with the root from C1 to C7, for three spectra given the same amplitude law and different frequencies: the harmonic series, a founder's bell, and a stiff string with B = 0.01. harmonic series: 0.04 at C1, peaking at 0.79 on E3, 0.42 at C7; a founder's bell: 0.04 at C1, peaking at 0.75 on C4, 0.38 at C7; a stiff string: 0.04 at C1, peaking at 0.71 on E3, 0.46 at C7. Only one of the three is a harmonic series, and all three rise out of the bass, peak in the middle of the compass and fall in the treble.

The arch belongs to hearing, not to the series

A chord delivers most of its partials in the middle of the compass and loses them in the bass and the treble, and every spectrum that showed that arch was built on whole multiples of a fundamental. Give the same amplitudes to a bell's eight modes and to a stiff string's stretched partials and the arch is still there, peaking within a major third of where the harmonic series peaks. What the spacing changes is the detail: a bell crowds its tierce and quint into a quarter of a critical band in the bass and loses them, and a stiff string's stretch buys the bass back.

perception · Masking
A wrong bar costs the same whichever instrument it is on. What moving one degree of a measured slendro by 10 cents, flat or sharp, adds to the roughness per second of a two-instrument texture — a ringing part at 0.15 s a note over a damped one four times slower — on the ringing instrument and on the damped one. The ensemble in tune scores 7.47. Degree 1 (0¢): ringing 0.142 flat and 0.202 sharp, damped 0.163 and 0.185, of which beating 0.178; Degree 2 (231¢): ringing 0.119 flat and 0.068 sharp, damped 0.096 and 0.090, of which beating 0.093; Degree 3 (474¢): ringing 0.106 flat and 0.021 sharp, damped 0.069 and 0.057, of which beating 0.063; Degree 5 (717¢): ringing 0.079 flat and 0.072 sharp, damped 0.075 and 0.078, of which beating 0.077; Degree 6 (955¢): ringing 0.023 flat and 0.018 sharp, damped 0.019 and 0.022, of which beating 0.020. Over all ten errors the ringing instrument's cost 0.85 and the damped one's 0.85, and the wrong bar beating against the other instrument's right one comes to 0.86 on either — as much as the whole, because the intervals the error changes add as often as they save. One error costs about 1.1% of the texture's roughness.

A wrong bar beats the same on either instrument

The ringing instrument carries a gamelan scale's standing, so a tuning error on it was predicted to cost more than the same error on the damped instrument. Put in the beating the model lacked, and the prediction fails: moved ten cents, a degree costs the same 0.85 summed over the scale on either instrument, because almost all of the cost is the wrong bar beating against the right one on the other instrument, and a beat belongs to both of the bars that make it. Moving a whole instrument costs exactly what its five degrees cost separately, and it takes nothing from the scale's standing.

scales · Beyond twelve

Named alongside it

The objects these essays reach for when they reach for this one.

RoughnessTimbrePartialCritical bandwidthOrchestrationRegisterSource-filterBrightnessExcitation pointInharmonicityDecayMasking

All concepts