Timbre and acoustics
A string does everything at once
A plucked string does not vibrate at one frequency. It vibrates at all the whole-number multiples of one frequency simultaneously, and nearly everything in music theory is downstream of that fact.
The ear hears the list, not the shape
Two sounds with the same partials and different phases have completely different waveforms and sound identical. What the ear extracts is a list of frequencies and strengths, and everything else is discarded.
The shape of a note, which is most of what an instrument is
Cut the first fifty milliseconds off a recorded piano and listeners stop calling it a piano. The attack carries more identity than the steady tone it leads into, and it is the part every spectrum plot leaves out.
The room is part of the instrument
A room has frequencies it supports and frequencies it will not. In a small one those frequencies are far apart, so some bass notes are loud in one corner and absent in another — and no equipment fixes it.
A vowel is two resonances
The vowel in "heed" is the same vowel sung high or low, and nothing about it is a property of the note. It is two peaks in the response of the mouth, sitting at fixed frequencies while the partials of the voice slide underneath them.
The piano is tuned wrong on purpose
Every well-tuned piano has a sharp treble and a flat bass, by up to a third of a semitone at the extremes. It is not an error, it is not a compromise about keys, and it follows from one property of a steel wire that can be computed from its diameter and its length.
How long a room rings, and where the formula stops
Sabine's reverberation time is one line of arithmetic — volume over absorption — and it built the modern concert hall. It also predicts that a room whose walls absorb everything still rings, which is a room with no reverberation at all, and the error is largest in exactly the rooms most music is now made in.
The first fifty milliseconds
A spectrum is supposed to be what makes a trumpet a trumpet. Cut the first fifty milliseconds off a recorded note and listeners stop being able to name the instrument — while the spectrum they are hearing is unchanged. Identity is in the part of the sound that ends before the note has properly started.
What makes two partials one note
A note is a stack of ten or twenty simultaneous tones and is heard as one thing. The obvious explanation is that they are whole-number multiples of a fundamental — and the obvious explanation is not sufficient. Mistune one partial by three per cent and it leaves the note; give a perfectly harmonic partial a thirty-millisecond head start and it leaves too. Shared behaviour beats arithmetic.
The body is the filter
A violin string radiates almost nothing. What reaches a room is the string's sawtooth multiplied by the body's response, and that response is a comb of measured resonances that stays put while the note moves. It is the same arithmetic that identifies a vowel, on wood instead of a mouth — which is why an instrument has a voice rather than a tone.
The other wolf
A cellist's wolf note is a string mode landing on a body resonance, at which point the two stop being separable and start exchanging energy — the mode splits in two and the note warbles at the difference. It is a coupled oscillator. The tuning system's wolf is twelve fifths failing to close by 23.5 cents. They share a word and nothing else.
An instrument points
A source radiates evenly while it is small compared with the wavelength and beams once it is not, and the crossover is one number. So the same instrument is omnidirectional in its bottom octave and a searchlight in its top one — which means its spectrum depends on where the listener is standing, and a microphone position is a choice about what the instrument sounds like.
The room chooses the harmonic rhythm
A chord in a cathedral is still sounding, seven decibels down, when the next one arrives — and the one after that, and the one after that. Reverberation is linear in decibels, so the number of chords audible at once is one number divided by another, and it puts a hard ceiling on how fast a composer writing for that building can change harmony. The ceiling is computable, and the music written for those rooms sits under it.
Where a room stops being a room
A small room has frequencies it supports and frequencies it will not, and a hall has a reverberation time. Those are two separate accounts and they are descriptions of the same building at different frequencies. The crossover is one formula, and in a bedroom it lands at about two hundred hertz — in the middle of the bass register, and above nothing at all in a concert hall.
A room does not decay evenly
Sabine's arithmetic gives one number and absorption is a strong function of frequency, so a room has six reverberation times rather than one. A stone church rings for 6.3 seconds at 125 hertz and 2.4 at 4 kilohertz, which means a chord left in it does not fade — it changes shape, losing its top before it loses its bottom, and arriving at the listener as a different sonority from the one played.
The first eighty milliseconds are a different room
Draw a line across a room's decay and the energy on either side is two opposite verdicts about one building — clarity before it, reverberation after. The line is a property of the ear, not of the room, and putting it at 50 milliseconds and at 80 makes the two published design targets fall out — a room for speech at one second, a room for music at 1.6.
The spectrum that was supposed to explain the gamelan
Run the model that designed the Bohlen–Pierce scale on a bar's partials and it asks for a compressed pseudo-octave at 1166 cents and wells at 694 and 870. A measured slendro's degrees are at 231, 474, 717 and 955, and only one of the four is near anything the model wants. Sweep the partials and a spectrum wanting a 240-cent step can be found — sitting sixty per cent of the way up its own roughness curve, which is what a scale-level test says about the whole comparison.
One voice over ninety players
A soloist heard over a full orchestra is not louder than it and could not be. What the trained voice does instead is put a peak of energy at three kilohertz, which is where the orchestra's spectrum has already fallen away and where the ear's own threshold happens to be lowest. Nineteen decibels of advantage, in a place nobody is competing for.
The bell decides what gets out
A tube resonates because the wave turns round at the open end, and it is audible because some of the wave does not. Those are the same number with opposite signs. One quantity — the size of the opening against a wavelength — decides how loud an instrument is, how bright it is and how directional it is, and a bell moves the boundary rather than removing it.
The sound a listener knows best
A voice is recognisable across every vowel it says, across two octaves of pitch, down a bad telephone line and in a whisper where there is no pitch at all. Nothing that survives all of that can be a frequency. What survives is a ratio: the resonances of a vocal tract are set by its length, so a shorter tract multiplies every formant by the same factor, and identity is a scale on the spectral envelope rather than a position within it. Between an adult man and a child the whole pattern moves by a fifth, and the vowel does not change at all.
What a choir does that a soloist cannot
Two singers on one note produce one beat and it can be counted. Sixteen produce a hundred and twenty at once, and the amplitude still fluctuates by as much as it did — a choir is no steadier than a duet. What has gone is not the fluctuation but its rate: the modulation energy that sat in a single line at two voices is spread across a band at sixteen, with no line in it. That is the choral sound, and it is also why the just-intonation drift this site measured describes only ensembles that hold their pitch still.
The mark that is not a level
There are six of them, they carry no units, and a performer has to turn one into a number before it means anything. What they instruct is not loudness. On a struck string a harder blow shortens the hammer's contact from 2.26 milliseconds to 0.95, which moves the first null of its own pulse from the third partial to the sixth: the partials between those are not quieter at pianissimo, they are gone. A fortissimo is a different sound, and the page has one word for both things it changes.
The note that gets duller as it dies
Every envelope drawn so far is one curve applied to a whole sound, and no struck string behaves that way. A string loses energy to air, to internal friction and to the bridge, and all three losses rise with frequency — so a note with a six-second fundamental has a sixteenth partial that is gone in under half a second, and the sound moving toward the listener is a spectrum collapsing toward its own fundamental. Which means an instrument is identified twice: once by the fifty milliseconds of its attack, which the earlier essays measured, and again by how fast its colour drains, which they did not.
An instrument is not one timbre
Every roughness verdict so far uses one partial list per instrument, and a violin does not have one. Its body's resonances stay where they are while the fundamental moves, so the radiated spectrum is different at every pitch: the power-weighted centroid runs from 1.02 to 1.64 across the compass, and it is not monotonic. Which means the result that a spectrum chooses its own scale gives a different scale at every note on the same instrument — three minima in the dissonance curve at the bottom of a violin's range, two in the middle and four at the top.
Which instrument is underneath
Every roughness curve so far compares two tones of the same timbre, which is a duet nobody plays. Give the two notes different instruments and the sum stops being symmetric: the same written interval, on the same two players, is up to five times rougher depending on which of them takes the lower note. From G3 upward the fifth is a well in all twenty-five pairings and no other interval is; below it, three pairings lose even that, and all three have a clarinet on top.
Which player on which note
An interval's roughness depends on which instrument is underneath, so the pair does not commute. Three players over three notes is the smallest thing that asymmetry has anywhere to go: six assignments, all of them the same chord, and across 450 of them the roughest averages half again the smoothest and reaches six times it. It is orchestration in the only form that can be computed here — not which chord, and not which voicing, but who is on which note.
A section against another section
The choir has been treated as a unison, and no choir sings only unisons. Two sections an interval apart beat between partials rather than between fundamentals — the third brings the fifth partial of one against the fourth of the other — and equal temperament puts that coincidence fourteen cents out. So two sections singing a tempered third beat at nearly nine per second with every singer in both of them perfectly in tune, and the same temperament is inaudible on a fifth.
The pulse that was assumed
Every figure until now low-passes the string's excitation with the spectrum of a half-sine, which is what a hammer would deliver against a rigid wall. An earlier essay said so and declined to do better. Doing better takes forty lines and refuses the prediction that came with it: the corner's round trips govern the spectrum as expected, and the contact time is governed by something else entirely — the mass ratio discovered one essay earlier.
What the mouthpiece is actually for
A cup and a throat look like a comfort fitting and are a component. Put a real one on the front of a real bore in the horn equation and the whole mode series slides sideways by about a seventh of its own spacing — which is exactly the distance between a series whose second resonance is 1.85 times the spacing and one whose second resonance is the second harmonic. The flare decides whether the modes are evenly spaced; the mouthpiece decides which harmonic each one is.
The corner does not come back a corner
Stepping a hammer forward against the string's own returning wave showed that the coupling matters most in the bass. It assumed the wave that came back was the shape that left. On a stiff string it is not: partials travel at different speeds, and by the top octave of a piano the sixth partial is the first to arrive out of place. But a periodic wave can only see phase modulo a cycle, and counted that way twenty-one of forty-eight partials are still in step — which is why the returning wave is degraded rather than destroyed.
A dynamic mark changes what a note is
Every spectrum until now is a shape with a level in front of it, so that playing ten decibels louder raises every partial by ten. That is true of exactly one instrument in an orchestra. Everybody else steepens their own spectrum as they lean on it, and a trumpet's centre of gravity moves from the second partial to the sixth across a dynamic range while an organ flue pipe's does not move at all.
Four terms, and only one of them binds
Six earlier essays each computed one term of a piano's excitation spectrum and handed it to the next, and nobody had multiplied them. Multiplied, three of the four have a corner — a partial above which they are the term doing the removing — and putting the three corners on one axis says which is in charge at each pitch. Below E3 the strike point decides; above it the contact time does; and the dispersion, which is the newest and most laborious term, is nowhere the binding one.
The flare that makes a series harmonic
Every figure until now treats a partial as n times a fundamental, and the whole of brass instrument design is the business of making that true. A plain cylinder is 127 cents from a harmonic series and a plain cone is 21; the best flare found by sweeping is 4.6, and only two per cent of the swept surface comes within five cents of it. The shape is forced rather than chosen.
The crossing belongs to the felt
On a piano the strike point decides the top of the spectrum below E3 and the hammer's contact decides it above. The merger's own bookkeeping asked whether that crossing is a fact about pianos or about hammers. A plectrum never crosses at all and a hard beater crosses two octaves higher, and the boundary is a contact time of about an eighth of a millisecond — ten times shorter than felt.
The bell is tuned for the cup
Sweeping a brass bell's two flare parameters found a shape that makes the resonance series harmonic to four and a half cents. Bolt the mouthpiece that instrument is actually played with onto it and the series is twenty-five cents out — worse than a plain cone. The flare and the cup are not two independent choices, and the best bell for a maker to build is one that is deliberately wrong.
Two players on one note
Six essays have put one instrument on each note of a chord, and the commonest thing an orchestrator actually does is put two on the same note. Two independent sources add in power, so the composite is neither of them — except that it nearly always is one of them, because the level at which ownership changes hands is rarely at zero. And a unison ten cents out is rougher than a major third dead in tune.
A section has a loudest member, not a colour
Two players on one note have a balance at which the composite belongs to neither, and that is what blending means. Three should have three such balances and no reason for them to agree — a trio with a rock-paper-scissors ownership would have no strongest member at all. Twenty trios, sixty pairwise comparisons, and not one disagreement: the possibility is real, arbitrary spectra do it once in twenty, and instruments never do.
The blend table has a row for every note
Eight earlier essays sound their instruments at one note, and one of them says why that cannot be innocent: every filter here is fixed in frequency and the fundamental is not. Swept over four octaves, the number of pairs that blend doubles from six to twelve, the ranking turns over rather than shuffles — seventy of a hundred and five comparisons swap — and a clarinet with an oboe goes from the best pair in the collection to the eleventh.
One cup and seven lengths
A trumpet is seven tubes with one mouthpiece serving all of them, and the harmonicity of its resonance series is a different number in every position — 19.5 cents open, 24.8 with all three valves down. The bare bore is flat across the same seven lengths to within eight-tenths of a cent, so none of it is a length problem. It is the cup, standing still while the series walks past it, and pressing all three valves does to the series exactly what fitting a mouthpiece 56 per cent of the catalogue depth would do.
The top that falls while the note lasts
Four earlier essays have asked where the harmonic series stops, and all four answered with a number computed for a tone that never ends. A struck C3 has 152 audible partials at the strike and eight after two-thirds of a second, so the ear's own resolution limit governs the first ten per cent of the note and the decay governs the rest. Playing ten decibels louder buys the ear a further tenth of a second, and doubling its reign would take fifty-eight.
The middle nobody could have guessed
A struck note has no steady state, only a slide from one spectrum to another — so the question is what the middle carries that the ends do not. The answer is exact rather than statistical: every loss law in the family leaves the strike with the same spectrum and ends in the same silence, so both endpoints carry precisely nothing about which of them it is. The whole difference is 41.3 decibels, and it peaks 0.38 seconds in, seven per cent of the way through the note.
The collapse belongs to the bass
Every envelope drawn until now has no pitch in it: the decay model counts partials rather than measuring them in hertz. Put a fundamental in and the collapse it describes shrinks monotonically up the keyboard, from 25.2 semitones of colour drain at A0 to 6.6 at C8, because a partial has to fit under twenty kilohertz to exist and the top note has four of them. Above C6 a struck note has fewer partials than the ear could have resolved, and there is nothing left for a collapse to take away.
A fifth on a piano is not a fifth a second later
Nine essays draw one note in silence, and a listener is given a texture. Put two struck notes an interval apart and consonance comes apart into two quantities that had agreed while nothing moved: the partial coincidence that names the interval outlives the strike in exactly the order common-practice theory ranks its intervals, and the roughness that scores it reorders itself inside a third of a second, with the fifth overtaken by four intervals every treatise calls harsher.
The blend arrives before the note does
Nine essays on spectrum draw a steady state, and the strongest cue that two instruments are two instruments is that they do not start together. Two envelopes rising at different rates turn out to be a balance — the same dial an earlier essay swept — so the attack is that dial moved by the clock, and its whole travel is fixed at twenty times the log of the two attack times. It is six decibels for a clarinet with a violin against a crossing twelve to twenty-two decibels out, so one pair in ten changes hands during its own attack, and which one depends on a convention rather than on the instruments.
An inversion lasts as long as its outer sixth
A struck interval keeps the partial coincidence that names it for a time set by its ratio, and a chord is three intervals at once. Voiced over one bass and struck on a piano, a triad keeps the evidence of all three only as long as its weakest one lasts, and for an inversion that is the sixth on the outside: a major sixth lasts as long as a major third, a minor sixth dies first. So the major six-four and the minor sixth chord are the most durable voicings of their triads and the major sixth chord and the minor six-four the least — and unlike a dyad, a triad's inversions keep their order by roughness through almost the whole decay.
A doubled pizzicato gives its note away early
The attack turns the balance between two players on one note by a few decibels and stops. A pluck does not stop — every partial of it decays, so a pizzicato doubled by a held instrument walks the balance for the whole note, and the expectation was a handover as slow as the decay. It is fast. A one-second pizzicato over a flute loses its note in 70 milliseconds, while it is still five decibels the louder, because what hands the note over is its upper partials going first. A uniform fade would have kept it eight times as long.
A room keeps a pizzicato from giving its note away
Doubled by a flute, a one-second pizzicato loses its note in 70 milliseconds dry, because its upper partials go first. The question left open was whether a room, whose reverberation keeps those partials alive, gives the note back afterwards. It does not give it back. It stops the note going: ten metres into a concert hall the pluck keeps it for 506 milliseconds, in a stone church for 814, and the room's own uneven decay takes back between a quarter and two fifths of that. In a room the loss law that decided everything dry matters a tenth as much, because the room's decay has become the clock.
A bow holds the number a blow hides
Two struck notes with different loss laws are identical at the strike and identical at the end, which is why separating them at all meant looking in the middle. Drive the same two strings continuously and the loss law stops being a rate and becomes a slope: each partial settles at its drive over its own loss, so the exponent adds to the source's roll-off and sits in the spectrum for as long as the bow moves. It is 18.7 decibels of separation available from the first instant, against 41.3 that a blow delivers after four tenths of a second and then takes away.
The room is the slower of the two
A reverberant field is the source convolved with the room, so a partial's tail falls at the slower of the two rates rather than at their sum — and the room is slower for exactly the partials the string is losing fastest. Half a note's colour is gone in 0.163 seconds in no room at all, 0.313 in a concert hall and 1.441 in a stone church. The destination is identical in all three, because a room cannot hold a partial up above the fundamental it is also holding. What a hall takes away is the rate, and the rate was the whole of the identity cue.
A damper changes the clock, not the colour
A damper is an extra loss on the string rather than a second decay, so it adds the same number of nepers a second to every partial — and adding a constant to every rate leaves every difference between rates exactly where it was. The damped spectrum at any instant is the ringing spectrum at that instant shifted bodily down, to machine precision. The colour goes on draining at its own rate; the note simply runs out of seconds, and how many it gets is written on the page as a note value and a tempo.
One note in the compass loses its pizzicato
Dry, how long a pluck keeps the composite spectrum barely depends on which note it plays: three-hundredths of a second at the worst pitch and seven at the best, a spread of three. In a concert hall the same eight notes spread by a factor of forty-three, and in a stone church one of them never gets the note at all. The room does not scale the dry answer by a constant — it multiplies it by between four and nine times depending on the pitch, and at the one note where the two instruments' spectra nearly coincide it makes the pluck's position worse instead of better.
A damper cannot reach into the room
The essay before this one proposed the arithmetic for a damped note in a hall: take the slower of the string's rate and the room's, then add the damper's to whichever won. The composition is wrong, and it is wrong in the one place that decides the answer. A damper is a loss on the string, so it belongs inside the minimum where a room can overrule it — and past about three seconds of reverberation it is overruled on every partial, so the damper removes no audible seconds of note at all.
A seventh chord cannot be spaced to last like a triad
A struck triad keeps the partial coincidences of all its intervals for at most 1.26 seconds, whichever way it is spaced. Run the same census over every inversion and spacing of five kinds of seventh chord and none gets near: the dominant, minor and half-diminished sevenths top out at about two thirds of a second, the major seventh at 0.42. The diminished seventh, which theory calls the least stable of them, lasts longest at 0.86 — because it is the only one with no tone or semitone among its pitch-class distances, and the worst distance a chord contains sets a ceiling no spacing can lift.
Only the player hears a staccato end
A damper stops a string in a seventh of a second, and in a hall the room goes on for two. A listener hears both, mixed in proportion to how close they sit, and the question was at what distance the short part stops mattering. The answer is closer than any seat. A damped note's twenty-decibel fall has doubled in length by a seventh of a hall's critical distance — 77 centimetres in a two-second concert hall — and by a quarter of it in a jazz club. The end of a staccato is something the pianist hears and the front row does not.