Concept

Spectral balance — where it appears

How a sound's energy is distributed across frequency, which a room, an instrument body or an audience can change without changing the notes. It is what a long-term average spectrum measures, and it is how a soloist stands clear of an orchestra.

Named by 11 essays across 5 fields — each of them below, with the objects they name alongside it.

Every point on one curve sounds equally loud. The equal-loudness contours of ISO 226:2003, evaluated from the standard's own parameters. The lowest curve is the threshold of hearing. Because the curves are not parallel — they crowd together in the bass and spread apart in the middle — the same change in decibels is a different change in loudness at every frequency, and a spectrum that was balanced at one level is not balanced at another.

The quietest thing audible, and why the volume knob is a tone control

A decibel is a fact about air. A phon is a fact about a listener, and the two do not line up — the map between them bends with frequency, and it bends differently at every level. One consequence is that turning a piece of music down does not turn all of it down equally, and the amount by which it does not is a number.

perception · Loudness
Six reverberation times for one room. Sabine's arithmetic evaluated in each octave band from the published absorption coefficients of the surfaces. a large stone church runs from 6.3 seconds at 125 Hz to 2.4 at 4 kHz — a bass ratio of 1.38, where concert halls are specified between 1.1 and 1.25.

A room does not decay evenly

Sabine's arithmetic gives one number and absorption is a strong function of frequency, so a room has six reverberation times rather than one. A stone church rings for 6.3 seconds at 125 hertz and 2.4 at 4 kilohertz, which means a chord left in it does not fade — it changes shape, losing its top before it loses its bottom, and arriving at the listener as a different sonority from the one played.

timbre · Room acoustics
The most a 6 cm cone can make of a low note. The maximum sound pressure level at one metre from a circular radiator of effective radius 3.2 cm, moving 1.5 mm at its limit, in a system resonating at 250 Hz. Below resonance the cone is already at that limit and the pressure a piston makes goes as the square of frequency, so the curve falls at twelve decibels an octave: 66 dB at 40 Hz, 78 dB at 80 Hz, 94 dB at 200 Hz. The 40 Hz figure is 28 decibels below the 200 Hz one, and that gap is arithmetic about a radius and a displacement rather than a property of any particular loudspeaker. The dots are the harmonics of a 41.2 Hz note with a one-over-n source spectrum; the loudest of them is the 6th.

The bass a small loudspeaker does not make

A three-inch cone at its excursion limit produces sixty-six decibels at forty hertz, which the ear converts to seventeen phons — barely above nothing. The note is heard anyway, because its harmonics are radiated and its fundamental is supplied by the listener. Computing what arrives turns the residue from a curiosity into a design decision, and finds that the fundamental of a low note is not the loudest part of it on any system a listener is likely to own.

instruments · Missing fundamental
Long-term average spectra: an orchestra, playing forte against a trained operatic soloist. Each source's mean spectrum over a long passage, in decibels below its own strongest region, on a logarithmic frequency axis. An orchestra, playing forte peaks at 250 Hz and is 30 dB down by 3,150 Hz; a trained operatic soloist peaks at 250 Hz and is 11 dB down by 3,150 Hz. The shapes are the same until about 1 kHz and separate above it: at 3153 Hz the difference is 19.0 decibels, which is the largest anywhere in the range. Nothing here is about level. Both curves are drawn against their own peaks, so what is being compared is shape.

One voice over ninety players

A soloist heard over a full orchestra is not louder than it and could not be. What the trained voice does instead is put a peak of energy at three kilohertz, which is where the orchestra's spectrum has already fallen away and where the ear's own threshold happens to be lowest. Nineteen decibels of advantage, in a place nobody is competing for.

timbre · The voice
What gets out of an opening, for 4 openings. The fraction of the wave's energy radiated at an open end against frequency, in the baffled-piston model — the radiation resistance of a circular piston, normalised to the tube's own impedance. Each curve runs from nothing at the bottom, where the opening is far smaller than a wavelength and the wave simply turns round, to everything above ka ≈ 2. Half the energy leaves at 6364 Hz for a flute's embouchure end (radius 10 mm), 2015 Hz for a clarinet's bell (radius 30 mm), 975 Hz for a trumpet's bell (radius 62 mm), 403 Hz for a horn's bell (radius 150 mm). The crossover goes as one over the radius, so the widest and narrowest here are 15.8 times apart in frequency. The same number decides how strongly the tube resonates and how much sound it makes, which is why a bell cannot brighten an instrument without also weakening its own resonances.

The bell decides what gets out

A tube resonates because the wave turns round at the open end, and it is audible because some of the wave does not. Those are the same number with opposite signs. One quantity — the size of the opening against a wavelength — decides how loud an instrument is, how bright it is and how directional it is, and a bell moves the boundary rather than removing it.

timbre · Air column
C4, in every place it can be played. A guitar neck with the 4 places C4 can be stopped, drawn at the fret spacing a 64.8-centimetre scale actually has. The stave writes one note and the tablature writes one of these; each notation says exactly what the other leaves out. The speaking lengths run from 61.2 down to 27.2 centimetres, so a hand plucking 12 centimetres from the bridge meets between 20 and 44 per cent of the string.

What a tablature keeps

Middle C can be stopped in four places on a guitar. The speaking lengths run from 61 to 27 centimetres, so a hand plucking twelve centimetres from the bridge meets between a fifth and nearly a half of the string, and the comb of missing partials is different at every one: the second partial is thirteen decibels stronger in the best position than in the worst. A stave writes one note for all four. A tablature writes four different things and cannot say which note any of them is.

instruments · Notation
A note gets duller as it dies. Each partial of a string note against time, with the loss rising as the partial number to the power 1 — so the fundamental takes 6 seconds to fall sixty decibels and the 8th takes 0.75. The heavy line is the power-weighted centroid, falling from partial 1.77 toward the fundamental; it is halfway there after 0.15 seconds. A single-rate envelope would draw all of these as parallel lines and the centroid as a horizontal one, and a struck string does neither: what is left at the end of a long note is very nearly a sine.

The note that gets duller as it dies

Every envelope drawn so far is one curve applied to a whole sound, and no struck string behaves that way. A string loses energy to air, to internal friction and to the bridge, and all three losses rise with frequency — so a note with a six-second fundamental has a sixteenth partial that is gone in under half a second, and the sound moving toward the listener is a spectrum collapsing toward its own fundamental. Which means an instrument is identified twice: once by the fifty milliseconds of its attack, which the earlier essays measured, and again by how fast its colour drains, which they did not.

timbre · Envelope
What a hand in the bell buys, and what it costs. How far the instrument flattens and how much radiation it loses, against the share of the bell's mouth the hand blocks, for an F horn's 15 centimetre mouth on a 3.7 metre acoustic length. The two curves are the same aperture radius read twice: a narrower mouth adds inertance, which lengthens the tube, and is acoustically smaller, which stops radiating. The mark is the 25.6 cents left owed earlier on the eleventh partial — it needs 62 per cent of the mouth blocked and costs 3.8 decibels of radiated power. Fully stopping is a semitone and nine decibels.

The hand that changes the bore

A conjecture refuted here left a question with a number on it: the natural trumpet's eleventh partial is 48.7 cents flat, the lips can move it 23.1, and 25.6 cents are owed by something that is not an embouchure. There is exactly one thing a player can change about the bore while playing, and putting the hand in the bell buys those 25.6 cents at a cost of 3.8 decibels — because the aperture that tunes the instrument is the aperture that radiates it.

tuning · Air column
Every assignment, at equal levels and at its own best balance. The 6 ways of putting 3 players on a chord, each drawn twice: hollow at equal levels, which is what an assignment ranking sees, and filled at the levels that balance the parts and then minimise roughness. Every scoring here is at the same total loudness, 23.3 sones, so two points are comparable. Solving the discrete problem first picks violin · clarinet · oboe; solving both at once picks violin · oboe · clarinet, and the two-stage answer costs 9.0 per cent more roughness. The orderings do not keep their places between the two columns, which is the whole of the argument: a ranking taken at equal levels is not a ranking.

Who plays what and how loud is one question

Two lines of argument, one about spectrum and one about loudness, each stopped at the same wall and each said so. One of them can choose who plays which note and has every player at the same level; the other can choose how loud each part is and has nobody assigned to anything. Put together they are a single problem with two kinds of variable, and solving it in stages picks a different answer from solving it at once — nine per cent rougher, at the same loudness, on an ordinary triad.

form · Orchestration
Trumpet at three dynamics, as a spectrum rather than a level. The radiated partials of a trumpet at 45, 70, 95 decibels, each normalised to its own strongest partial so that only the SHAPE is compared. A linear source would give three identical pictures. This one does not: the spectral centroid moves from partial 2.19 to 6.41, a factor of 2.92, because the excitation is nonlinear and blowing harder steepens the pressure front rather than scaling it. The tilt used is 3 decibels per octave of partial number per ten decibels of level, referred to 70 dB — a stipulated, ordinal number, not a measurement of any instrument.

A dynamic mark changes what a note is

Every spectrum until now is a shape with a level in front of it, so that playing ten decibels louder raises every partial by ten. That is true of exactly one instrument in an orchestra. Everybody else steepens their own spectrum as they lean on it, and a trumpet's centre of gravity moves from the second partial to the sixth across a dynamic range while an organ flue pipe's does not move at all.

timbre · Orchestration
The top voice arrives whole and the bottom one arrives as a sine. 4 parts sounding together, each of 8 partials, with every partial tested against the summed masked threshold of every component in the texture. A filled mark is a partial the listener receives and an open one is a partial the part would have had alone and does not have here. The bass at C3 keeps 1 of 8, the tenor at C4 keeps 2 of 8, the alto at E4 keeps 5 of 8, the soprano at G5 keeps 8 of 8, every one of them at 70 decibels. Every part is at the same level and the difference is entirely where each one sits: masking spreads upward, so the part at the top of the texture has nothing above it to be masked by and the part at the bottom has everything.

The listener is given the top voice, and the bass as a sine

Four earlier essays put the masker and the probe in the same voice. Put them in different voices — a four-part texture at one level — and the soprano arrives with all eight of its partials, the alto with five, the tenor with two and the bass with one. Balancing the loudness, which is the constraint a scoring is solved under, changes none of that: equal loudness is not equal spectrum and cannot be made so.

perception · Masking

Named alongside it

The objects these essays reach for when they reach for this one.

OrchestrationBrightnessRadiation efficiencyTimbreLoudnessMaskingPartialBoreCutoffDecayEnd correctionEqual-loudness contour

All concepts