Concept

Register — where it appears

The span of notes an instrument produces from one mode of its resonator, before it moves to the next one. Crossing between registers changes the spectrum audibly, which is why the crossing points are the difficult places on most instruments.

Named by 46 essays across 9 fields — each of them below, with the objects they name alongside it.

What a 60 cm tube supports, by how its ends are closed. The first 6 modes of an open cylinder and a stopped cylinder, all of the same acoustic length. An open tube supports every whole multiple of the fundamental and reaches the second mode an octave up. A cylinder stopped at one end supports only the odd multiples and reaches its second mode 1902 cents up, which is a twelfth. Its fundamental is also an octave below the others, because it fits a quarter of a wavelength where they fit a half.

A tube that skips every other partial

Stop one end of a cylinder and half its modes vanish. That single fact about where the pressure has to be decides that a clarinet sounds hollow, that it plays an octave below its length suggests, and that it must cover nineteen semitones with fingers before it can overblow — while every other woodwind covers twelve.

instruments · Air column
What a 60 cm tube supports, by how its ends are closed. The first 6 modes of a stopped cylinder and a cone, all of the same acoustic length. A cone supports every whole multiple of the fundamental and reaches the second mode an octave up. A cylinder stopped at one end supports only the odd multiples and reaches its second mode 1902 cents up, which is a twelfth. Its fundamental is also an octave below the others, because it fits a quarter of a wavelength where they fit a half.

A cone is not a cylinder

A saxophone has a reed at a closed end, exactly as a clarinet does, and it overblows at the octave rather than the twelfth. If the previous essay's argument were about reeds that would refute it. It is about geometry, and a cone closed at its apex has the complete harmonic series for a reason that takes one line of algebra and is genuinely surprising.

instruments · Air column
The end correction, for a bore of radius 7.5 mm. How flat a tube sounds against what its physical length alone would predict, because the wave carries on past the opening before it turns round. The correction is 4.6 mm at every note — Levine and Schwinger's 0.6133 times the radius for an unflanged end — and the error it causes is 13 cents on a 60 cm sounding length and 52 cents on 15 cm. It is the same millimetres in both cases.

The tube ends after it ends

A wave does not turn round at the opening. It carries on into the room for about six-tenths of the bore radius and reflects there, so every tube is acoustically longer than it is. The correction is a fixed number of millimetres against a wavelength that halves every octave — a rounding error at the bottom of an instrument's range and most of a semitone at the top.

instruments · Air column
One register vent, 12 fingerings. Where the second mode's pressure node sits for each fingering of a stopped tube, against a single register hole drilled 12 cm from the mouthpiece. The node is a third of the way along the sounding length, so it moves every time a hole is opened, and the vent's error runs from -38 to -1 cents across the range. A perfect register system would need one hole per fingering. The number of holes actually fitted is one, and the leftover is a design decision rather than a fault.

One hole doing a dozen jobs

A register key works by forcing a pressure node where the second mode already has one, which kills the fundamental and leaves the mode above. The node sits a fixed fraction along the sounding length — and the sounding length changes with every fingering, while the hole stays where it was drilled. The leftover error is computable, and it is why the throat notes are the ones players complain about.

instruments · Tone holes
The least rough 7 notes of the twelve, at 262 Hz. Every 7-note selection from the twelve that contains C, scored for total Plomp–Levelt roughness at a root of 262 Hz, ranked. The best 8 are shown with the spread between them. The major scale ranks 5 of 462; the first selection that is a mode of the diatonic set ranks 1. Change the root and the ranking changes, because roughness is a fact about frequencies and a scale is not.

Roughness cannot choose a scale

Score all four hundred and sixty-two seven-note selections from the twelve for roughness and the answer at middle C is a mode of the diatonic set, first out of four hundred and sixty-two. Ask the same question an octave lower and the same selection ranks a hundred and forty-first. The model is not wrong; it is answering a question about frequencies, and a scale is not one.

intervals · Consonance
Every voicing of a major triad, least rough first. All 27 arrangements of the same three pitch classes within 3 octaves from 131 Hz, scored for roughness. The best is spaced 19 then 9 semitones — wide below, close above — and the worst is the chord in close position at the bottom of the range, 6.6 times rougher with exactly the same notes in it.

Where to put the third

Take three pitch classes, four octaves to put them in, and score all twenty-seven arrangements. The smoothest is root, fifth an octave up, third two octaves up — which is partials one, three and five of the harmonic series — and the roughest, at every register tried, is the chord in close root position at the bottom of the range. Every orchestration manual states that rule and none of them derives it.

harmony · Consonance
Two mechanisms, the notes both of them make, and the seam. The frequency range of each laryngeal mechanism for an adult male voice, on a logarithmic axis, with the band both can produce shaded. M1 — chest runs 82–349 Hz and M2 — falsetto runs 220–698 Hz, so 799 cents of the range — 8.0 semitones — can be sung either way. The two dots inside that band are the measured signature that this is a bifurcation rather than a threshold: the change upward happens at 330 Hz and the change downward at 294 Hz, 200 cents lower. A threshold is crossed at the same place in both directions and this is not.

Two mechanisms, and the seam between them

Every singer has a place in the range where the voice changes character, and eight semitones of it can be produced either way. The measurement that settles what kind of a place it is takes ten seconds: the change upward happens two hundred cents higher than the change downward, and a threshold cannot do that.

instruments · The voice
Long-term average spectra: an orchestra, playing forte against a trained operatic soloist. Each source's mean spectrum over a long passage, in decibels below its own strongest region, on a logarithmic frequency axis. An orchestra, playing forte peaks at 250 Hz and is 30 dB down by 3,150 Hz; a trained operatic soloist peaks at 250 Hz and is 11 dB down by 3,150 Hz. The shapes are the same until about 1 kHz and separate above it: at 3153 Hz the difference is 19.0 decibels, which is the largest anywhere in the range. Nothing here is about level. Both curves are drawn against their own peaks, so what is being compared is shape.

One voice over ninety players

A soloist heard over a full orchestra is not louder than it and could not be. What the trained voice does instead is put a peak of energy at three kilohertz, which is where the orchestra's spectrum has already fallen away and where the ear's own threshold happens to be lowest. Nineteen decibels of advantage, in a place nobody is competing for.

timbre · The voice
The same intervals, played higher and higher. Sensory roughness for four fixed intervals as the pair is transposed up five octaves, computed from the same Plomp–Levelt model as the dissonance curve. Every one of them falls as it rises, and the small intervals fall furthest — so how consonant an interval is depends on where it is played, not only on what it is.

A chord is a register

The same three pitch classes are five times rougher at the bottom of a piano than in the middle, and the arrangement that minimises the roughness over a low bass turns out to be the bass's own fifth and sixth partials. The orchestration rule about low thirds is not a convention. It falls out of the width of a critical band, exactly.

intervals · The triad
Where A has been. Documented pitch standards and surviving instruments, plotted as cents from A415. The extremes are 392 Hz and 465 Hz, which is 296 cents apart — 3.0 semitones, close enough to a minor third that a piece written at one and played at the other is in a different key. Nothing here is a preference; each is a decision somebody recorded.

The note that moved by a minor third

A has been anywhere between 392 and 465 hertz in the surviving record, which is 296 cents — a minor third short of six. The usual conclusion is that absolute pitch level is a convention and nothing musical depends on it. That is true of everything written on paper and false of everything that happens in a throat or a body: a singer's register break sits at a fixed frequency, so across the historical range it lands three semitones further down the written page. Transposing a piece is not a uniform operation, because the performer does not transpose.

tuning · Pitch standard
Where a twelfth comes from. The range of a walk with no walls, against how many notes it runs for, at three settings of the one parameter it has. The parameter is fitted to the post-skip reversal rate and to nothing else; the range is then read off. With no central tendency at all the walk passes two octaves by 60 notes and keeps going. At the setting that reproduces 70 per cent reversal — κ = 0.78 — the range is 11.9 semitones at thirty notes and 17.9 at a hundred and twenty. It grows logarithmically, so over the whole plausible length of a tune it sits between an octave and a fifteenth, and a twelfth is the middle of that. The three tunes carried here are marked and all three fall below the curve.

The twelfth, and where it comes from

Melodies occupy about an octave and a fifth, and an earlier essay set out to explain that by the singer's register break and found that it does not: the chest mechanism alone spans two octaves and a semitone. The answer is in a parameter the essay on leaps fitted and then put down. A walk with no walls whose central tendency reproduces the post-skip reversal rate has a range that grows logarithmically — six semitones at eight notes, twelve at thirty, eighteen at a hundred and twenty — so across every length a tune plausibly has, the span is between an octave and a fifteenth.

form · Melody
One key, nineteen notes, one right answer. A register hole disturbs a mode in proportion to the square of that mode's pressure at the hole, so the place that spoils the fundamental and leaves the third harmonic alone is the third harmonic's own pressure node — a third of the way down whatever length is sounding. The length changes with every fingering and the key does not move, so the two curves cross at one note. At A3 the key sits almost exactly on the node and the twelfth speaks cleanly; at A♭4 it is at 62 per cent of the tube and disturbs the third harmonic by 96 per cent of what an antinode would, which is the throat of the instrument and is exactly where players say the notes are worst.

The hole that spoils a note

A tone hole shortens the tube. A register hole does the opposite job: it is small enough to shorten nothing and is placed where it will wreck the fundamental's resonance and leave the third harmonic's alone, so the note jumps a twelfth instead of retuning. The place that does both is a pressure node of the harmonic being kept — a third of the way along whatever length is sounding — and the length changes with every fingering while the key does not. One key is at the right place for exactly one note, and the note it is worst for is in the throat of the instrument, which is where players say the instrument is worst.

instruments · Tone holes
How much of the writing the prohibitions forbid. I – vi – ii – V – I in C major, written in 3, 4, 5, 6 parts, with every ensemble covering the same total compass. A triad has three pitch classes, so n parts double n − 3 of them and a duet cannot state one at all. Of every ordered pair of complete voicings of V and I, the share with no parallel fifth or octave between any pair of voices: 91.2% at 3, 75.0% at 4, 44.3% at 5, 17.9% at 6.

Why the exercise is in four parts

Eight earlier essays move four voices, and nothing here ever chose four. Cut one choir's compass into three parts instead and the two most famous prohibitions in music cost exactly nothing — the cheapest realisation already obeys them. Cut it into six and they cost more than half a semitone per voice per chord change, because a parallel octave needs a doubled note and six voices sharing three notes can hardly avoid one. Four is the smallest number of parts at which the rules have a price at all, and it is the largest at which the price is small.

harmony · Voice-leading
One instrument, five spectra. The first 12 partials of a bowed string at 5 pitches, each passed through the same fixed body response and normalised to its own loudest partial. The resonances stay where they are and the partials slide under them, so the pattern is different at every note: the power-weighted centroid runs from 1.02 to 1.71 across the compass and is not monotonic in pitch. A source with no body at all would give 2.35 at every pitch, which is the single number every roughness figure here uses for a string.

An instrument is not one timbre

Every roughness verdict so far uses one partial list per instrument, and a violin does not have one. Its body's resonances stay where they are while the fundamental moves, so the radiated spectrum is different at every pitch: the power-weighted centroid runs from 1.02 to 1.64 across the compass, and it is not monotonic. Which means the result that a spectrum chooses its own scale gives a different scale at every note on the same instrument — three minima in the dissonance curve at the bottom of a violin's range, two in the middle and four at the top.

timbre · Spectrum
The bow's window along each string, with the bow held still. Schelleng's window — the ratio of the greatest bow force a note tolerates to the least — for every written pitch on every string that can play it, with the bow 35 mm from the bridge and staying there. The window widens up each string because β is the bow's distance over the SOUNDING length and the sounding length is shortening, and it steps down at each new string because a heavier string carries a narrower one. The dashed lines are what was drawn earlier, which held β at 0.09 for every note on every string. The widest disagreement between two strings at one pitch is 1.48 to one, at A♭5.

Two dials the player turns together

Schelleng's window is a ratio of two bow forces, and an earlier essay found that the string's own impedance is in it twice. What it held still was the bowing point — as though a player kept β constant while moving up the fingerboard, which is the opposite of what a bow does. Let the bow stay where a bow stays and the window widens up every string, the spread between two strings at one pitch goes from 1.4 to 1.5 the other way round, and the string that is easiest changes.

instruments · Bowed string
The heard moment against the pitch, on an instrument whose own attack is 8 ms. A note cannot establish an amplitude in less than 4 of its own cycles, so the attack has a floor of 4 periods — 145 milliseconds at A0 and 1.9 at C7. Below A4 the floor is longer than the instrument's own attack and the pitch decides the heard moment; above it the instrument does. The lag runs from 46.0 milliseconds at the bottom to 2.5 at the top, a spread of 44 milliseconds that no player can play their way out of.

A low note cannot start on time

Three earlier essays have held the pitch at one value. A note cannot establish an amplitude in less than a few of its own cycles, so the attack has a floor that rises as the pitch falls — 146 milliseconds at the bottom of a piano and three at the top. On an instrument whose action takes eight milliseconds everywhere, that is a forty-three millisecond spread across the keyboard from the period alone, and no player can do anything about it.

rhythm · Perceptual-centre
The fingerboard, with the hardest place on it. Every written pitch from G3 to G6 on every string that can reach it, shaded by the width of Schelleng's bow-force window there — dark is narrow, which is a note that is hard to start. Two effects are multiplied and neither earlier figure could show the other: the body's admittance, which depends on the frequency, and the bowing fraction and string impedance, which depend on where the hand is. The worst place is C♯4 on the G3 string, 6 semitones up it, at a window of 12.2 against 471 at the easiest — a factor of 39 across the instrument. It sits where the body's A0 resonance crosses the heaviest string played high, which is the compounding this figure was drawn to find: neither variable alone puts a minimum there.

The hardest place on the fingerboard

Two things narrow a bow's window and each has been drawn alone. The bridge's admittance is a function of frequency; the bowing fraction is a function of where the left hand is. They meet on a real fingerboard, and multiplying the two curves gives a map with a worst place on it — C♯ on the G string, sixth position, where the body's air resonance crosses the heaviest string played short. The window there is twelve, against three hundred and eighty at the easiest.

instruments · Bowed string
Which of the four terms is doing the removing, note by note. Each of three terms has a corner — a partial number above which it is the one taking the energy away — and the lowest corner at each pitch is the term that binds. The comb is flat at 8, because where the hammer lands is a fraction of the length and does not change with the note. The contact time's corner falls from 34 at A0 to 0.5 at A7. They cross at about E3, 165 hertz: below it the strike point decides the spectrum and above it the contact time does. The dispersion's corner is above both at every pitch on the instrument — 5.8 at the top against a contact corner of 0.5 — so that earlier term is real and is nowhere the binding one. That is a statement no one of them could make on its own.

Four terms, and only one of them binds

Six earlier essays each computed one term of a piano's excitation spectrum and handed it to the next, and nobody had multiplied them. Multiplied, three of the four have a corner — a partial above which they are the term doing the removing — and putting the three corners on one axis says which is in charge at each pitch. Below E3 the strike point decides; above it the contact time does; and the dispersion, which is the newest and most laborious term, is nowhere the binding one.

timbre · Excitation point
Which notes of a scored chord have to be played early. Four parts of one chord, each with its own instrument, its own pitch and its own dynamic, and the perceptual centre that comes out of all three. piano, sforzando on E1: an attack family of 8 milliseconds against a pitch floor of 97, so the pitch is what limits it, shortened by the dynamic to 65, heard 20.5 after it starts and needing to be played 12.0 early; flute, quiet on A5: an attack family of 60 milliseconds against a pitch floor of 5, so the instrument is, shortened by the dynamic to 69, heard 21.8 after it starts and needing to be played 13.2 early; violin, mezzo forte on E4: an attack family of 90 milliseconds against a pitch floor of 12, so the instrument is, shortened by the dynamic to 90, heard 28.5 after it starts and needing to be played 19.9 early; trumpet, forte on A3: an attack family of 30 milliseconds against a pitch floor of 18, so the instrument is, shortened by the dynamic to 27, heard 8.6 after it starts and needing to be played 0.0 early. The spread is 19.9 milliseconds, which is well above the two or three a listener resolves, so a conductor asking for these four to sound together is asking for four different physical onsets.

Which notes have to be played early

There are three separate contributions to one quantity — the instrument's attack family, the dynamic it is played at, and the note's own period — and every figure so far varies one and holds the others. Added together for a real scoring they do not add: a sforzando low piano note is pitch-limited to a hundred-millisecond attack and the sforzando shortens it back to sixty-five, so flattening the dynamics makes the ensemble's spread larger rather than smaller.

rhythm · Perceptual-centre
Up a trumpet, the two clocks disagree. Every impedance peak of a trumpet, with the settling time each implies. The Q rises up the ladder and so does the frequency, and the settling time is their ratio — so in milliseconds the wait falls from 153 at the C♯2 to 24 at the B5, while in periods it rises from 11 to 24. Both curves are monotone and they point opposite ways, which means a player going up the instrument gets notes that arrive sooner and take longer in their own terms. Nothing here is a measurement of an instrument: it is the transmission-line solve of a trumpet-shaped bore, whose peak Qs are sensitive to how finely the sweep is sampled at about five per cent.

The higher note speaks sooner and takes longer

Up a brass instrument the settling time in milliseconds falls by a factor of seven and the settling time in periods rises by a factor of three. Both curves are read off the same impedance sweep, both are monotone over most of the compass, and they point in opposite directions — so the slowest note of the instrument depends entirely on which clock is used to time it.

tuning · Onset time
One vent on a cone, over the notes it has to serve. A cone's modes are a full harmonic series, so the mode a register vent has to keep is the octave rather than the twelfth, and its pressure node sits at half the sounding length measured from the virtual apex. The apex is a fixed point of the instrument and the bell end is not, so every fingering moves the ideal position and there is one hole. The upper line is how much the vent spoils the fundamental, which is what it is for; the lower is how much it damages the octave, which is the cost. Their ratio runs from 451.2 at F♯4 down to 1.9 at B♭3 — a factor of 235 across a single register. The vent is placed at 37 per cent of the longest sounding length, which is where the worst fingering is least badly served.

The vent a cone cannot place

A clarinet's register key has to spoil a fundamental and leave a twelfth. A saxophone's has to spoil a fundamental and leave an octave, whose pressure node sits at half the sounding length from the virtual apex — and the apex is a fixed point while the bell end is not. Over one register the ideal position moves by a factor of two, and one hole is right for one note.

instruments · Tone holes
Seven holes that all sound 196 hertz, and none of them agrees about the twelfth. Each dot is a hole radius, placed at the station that makes the first resonance 196 hertz. The stations run from 411 millimetres for a 7.5-millimetre hole to 307 for a 1.4-millimetre one, which is a fifth of the tube. Up the axis is what the second resonance does: a cylinder's should be three times the first, and it is -2 cents from it for the widest hole and -453 for the narrowest. The hole's inertance rises with frequency, so a narrow hole lengthens the tube more for the twelfth than for the fundamental — and two holes that are interchangeable in the first register are a fourth apart in the second.

A hole is a short tube

Four earlier essays have treated an open tone hole as a point where the pressure is released. It is not: the air in a hole has mass, and a hole with mass does not end the bore, it loads it. Seven holes drilled at seven stations all sound the same G — and their twelfths are spread over a fourth. Cross-fingering falls out of the same arithmetic, and it is not made of what everybody says it is.

instruments · Tone holes
A woodwind with holes graduated 12 mm to 6 mm, drilled so that every fingering is in tune. A cylindrical bore 567 millimetres of acoustic length and 15 across, with twelve tone holes through a 4-millimetre wall. Opening them one at a time from the far end takes it up a chromatic scale from D3 to D4. The stations are not copied from a maker's drawing: each was solved so that its own fingering sounds its equal-tempered note in this model, one hole at a time down the tube with every hole below it already open, which is what a reamer and a tuning fork do. The worst fingering is 11.9 cents out. The diameters run 12.0 millimetres at the bell end to 6.0 at the top, and the spacings close from 30 millimetres to 19.

The cutoff that is a list

Five earlier essays have quoted one number for a woodwind's cutoff — 1,824 hertz for a clarinet — from a formula written for an infinite lattice of identical holes. Solve a whole twelve-hole chart instead and the number is eleven different numbers, running from 2,193 hertz down to 1,574, which is 574 cents. The lowest fingering has no cutoff at all, and which way the list runs turns out to be a design decision rather than a fact about woodwinds.

instruments · Tone holes
A plucking point 50 millimetres from the nut, across a harpsichord's compass. The jacks stand in a rail and the rail is one object, so a register plucks at a fixed distance from the nut while the strings shorten by a factor of ten from the bass to the treble. At 50 millimetres that is one 36th of the string at C2 and one 3.5th at C6 — a plucking fraction that changes by a factor of 10.1 without the maker moving anything. The comb corner is one over that fraction and falls with it, from 36 to 3.5; the dispersion corner has the fraction under a cube root and falls only from 106 to 21. They never meet. Setting one over p equal to the cube root of two over three times the inharmonicity and the fraction gives a plucking point at p = √(1.5B), which on this instrument's iron wire is between 3.4 and 9.9 millimetres from the nut — nearer to it than any register a harpsichord has ever carried, the lute stop included. So the maker's one free choice is the binding term at every pitch and every register position, which is exactly what the piano's is not: on a piano the contact time takes the decision away above E3.

What the second register is for

Nine earlier essays take the plucking fraction as given — an eighth, a sixth, a seventh — and a harpsichord's jacks stand in a rail, so what is given is a distance and the fraction is a consequence: 50 millimetres is one thirty-sixth of the string in the bass and one three-and-a-halfth in the treble. Engage two registers on one string and the amplitudes add exactly, the near one fills the far one's missing partials, the pair digs holes of its own that neither has, and the combination comes out louder and rounder than either — including the bright one.

instruments · Excitation point
The roughest chord on the page is at C1 and the roughest one heard is at A♭2. One close triad at 70 decibels through 17 registers, with its roughness computed twice: over every partial in the score, and over only the partials that stand above what the chord itself masks. The written curve rises all the way down and its maximum is the lowest register drawn, C1, which is the low-interval rule as it has always been computed here. The delivered curve turns over at A♭2 and falls to nothing below A♭1: a close triad down there is not rough, because it is not arriving as a chord — 1 of its 24 partials survives at C1 and there is almost nothing left to beat against anything.

A low chord stops being rough by stopping being a chord

Every count of audible partials until now is of a chord at middle C, and the three registers it did compare span C3 to C5 — a third of the range a chord is written in. Move the same triad down and the count collapses: 79 per cent of its partials arrive at E3 and 4 per cent at C1. So the roughest chord on the page is the lowest one and the roughest chord a listener receives is at G2, and where that maximum sits moves nearly two octaves with the dynamic.

perception · Masking
Which pairs blend is a question about the note. The level at which a doubled pair's composite changes owner, drawn for all 15 pairs of 6 radiators over 2.6 octaves from 131 to 784 hertz. A pair blends when that level is inside the shaded band, which is the twenty-four decibels either way two players can manage; a curve outside it, or absent, is a pair one instrument owns at every balance. 6 of 15 pairs blend at the bottom of the range and 12 at the top. Every filter in this collection is fixed in frequency and the fundamental is not, so a radiator's shape is a function of pitch and so is everything computed from two of them — the blend ranking at the bottom and at the top disagree on 70 of 105 comparisons, which is more than half, so the order has turned over rather than merely shuffled.

The blend table has a row for every note

Eight earlier essays sound their instruments at one note, and one of them says why that cannot be innocent: every filter here is fixed in frequency and the fundamental is not. Swept over four octaves, the number of pairs that blend doubles from six to twelve, the ranking turns over rather than shuffles — seventy of a hundred and five comparisons swap — and a clarinet with an oboe goes from the best pair in the collection to the eleventh.

timbre · Spectrum
There is less to collapse the higher the note is. The same string model struck at 80 decibels on nine fundamentals, an octave apart. The heavy line is how far the spectral centroid falls between the strike and the note's death, in semitones — the whole of the colour drain, in the unit a musician has for pitch. It runs from 25.2 semitones at A0 to 6.6 at C8, a factor of 3.8, and it falls monotonically. The reason is the dashed line, which is how many partials exist at all: 727 under the audio ceiling at A0 and 4 at C8. From C7 upward a struck note has fewer partials than the ear could have resolved — 9 against 10 — so there is nothing left for a collapse to take away.

The collapse belongs to the bass

Every envelope drawn until now has no pitch in it: the decay model counts partials rather than measuring them in hertz. Put a fundamental in and the collapse it describes shrinks monotonically up the keyboard, from 25.2 semitones of colour drain at A0 to 6.6 at C8, because a partial has to fit under twenty kilohertz to exist and the top note has four of them. Above C6 a struck note has fewer partials than the ear could have resolved, and there is nothing left for a collapse to take away.

timbre · Envelope
Where the register break falls on a tenor's page. The two measured laryngeal crossings — 330 hertz going up and 294 coming down — read as WRITTEN notes, against the pitch standard the part is performed at. The crossings are frequencies and do not move; the notation does, so the seam slides down the stave by exactly the interval the standard rises. At A392 the upward crossing is written F♯4, at A415 it is F4, at A440 E4 and at A465 E♭4 — a minor third of movement across four centuries, on a part nobody rewrote. Across the range drawn the seam passes 4 written semitones. The shaded horizontal band is the tenor's written compass, C3 to A4; the seam is inside it at 7 of the 7 documented standards drawn.

A standard moves the page, and not the seam

Every earlier essay has priced a pitch standard against something with a fixed length in it. A voice has none, so nothing about it changes at all — what changes is where the written note falls against a break in the larynx that is a frequency and stays put. At A415 that break is written F4, at A440 it is E4 and at Chorton it is E♭4: a minor third of movement across four centuries, on a part nobody rewrote.

tuning · Pitch standard
Two registers 50 microseconds apart on one string. The partials of a 355-millimetre string plucked at 32 and 50 millimetres from the nut, drawn twice: once with the two quills releasing together, once with the far one releasing 0.05 milliseconds later, which is 0.026 of this string's period. A delayed release turns partial n through 2·pi·f_n·dt, so the rotation is proportional to the partial number: the fundamental is turned 9 degrees and partial 24 is turned 226. Simultaneous, the pair is missing partials 17 and 20 — holes it digs for itself where the two combs are equal and opposite. Staggered, it is missing none of them: a rotation of anything at all takes two amplitudes out of opposition. The partials neither comb can excite at all — 7, 11, 22 — are filled either way, because where one comb is zero the sum is the other one whatever its phase. The fundamental's gain over the far register alone falls from 4.36 decibels to 4.34.

The interval between two quills

Two jacks on one key are voiced separately and do not let go at the same instant. That interval turns each partial of the later pluck through an angle proportional to its number — so it leaves the fundamental alone and inverts the twentieth partial, and the holes the pair digs for itself vanish at a hundredth of a period. What a regulator can tolerate turns out to be one fixed fraction of a period at every pitch, which on a five-octave instrument is a factor of sixteen in milliseconds.

instruments · Excitation point
A 20-decibel crescendo is 27 phons on a bass note and 20 on a high one. The same change of level, from 60 to 80 decibels, converted to loudness at each register through ISO 226's equal-loudness contours rather than at one kilohertz. The heavy curve gives each note a string spectrum, so its partials are converted in their own bands and summed; the pale one is the fundamental alone. On the spectrum-aware curve the crescendo is worth 27.3 phons at C1 and 20.3 at C7. On the fundamental alone it is 77 at C1, which is not a finding but an artefact: a 60-decibel tone at 33 hertz sits 1.8 decibels above the threshold of hearing and is very nearly nothing. The honest correction is the smaller one, and it is still a difference of 7.1 phons across the compass for a mark written in the same ink.

A subito piano is four seconds longer in the bass

Every loudness figure with time in it converts level to loudness at one kilohertz, and the equal-loudness contours say that no other frequency works that way. Joining the two sorts the published numbers into those that were about the treble and those that were not. Three move a great deal — a twenty-decibel crescendo is worth 27 phons on a bass note and 20 on a high one, and the seven seconds a subito piano takes becomes eleven and a third. Three do not move at all, and the reason they do not is the same reason in every case.

perception · Loudness
Scored the way these figures score it, a clarinet is the worst of the six. One close triad at 70 decibels through 17 registers, drawn once for each of the 6 spectra to hand. The score is the share of every partial written, which is the quantity the register figure published. pure 100 per cent at best, string 79 per cent at best, clarinet 54 per cent at best, reed 71 per cent at best, bell 52 per cent at best, organ 72 per cent at best. A clarinet's four even partials are twenty-eight decibels below its odd ones and are inaudible beside their own neighbours before any chord is built, so counting them in the denominator makes the spectrum that survives its own masking best look like the one that survives it worst.

A clarinet keeps what a string loses

Every masker, probe, chord, line and texture until now is eight partials falling as 1/n, and it was not even an option a placement could pass. Sweeping the six spectra to hand says the clarinet is the worst of them — 54 per cent of itself at best against a string's 79 — and that answer is an artefact of the score. Counted against what each note keeps on its own, the clarinet keeps 100 per cent where the string keeps 79, because its components stand a twelfth apart rather than an octave. The missing parameter was the spectrum; the second missing parameter was the denominator.

perception · Masking
Eighty-one chords the expectation model cannot tell apart. Every voicing of a dominant seventh on G inside the three octaves above its own root, placed by how rough it is and how far its outer voices are apart, and coloured by which member of the chord is at the bottom. The roughness runs from 0.269 to 1.449, a factor of 5.4, computed from each voicing's own spectrum under Plomp and Levelt's roughness model. The identity surprise the expectation model assigns is 3.51 bits for every one of the 81, because it is a function of a scale degree and its predecessor and there is no register anywhere in it. What separates them is spacing rather than inversion: roughness falls as the outer voices spread apart, correlating -0.42 with the span, and is indifferent to which member of the chord is at the bottom at 0.02. The seventh in the bass is not what makes a chord rough; a fourth and a third packed together at the bottom of the range is.

Eighty-one chords, one number

A dominant seventh has eighty-one arrangements inside three octaves and their roughness spans a factor of five and a half. The tonal-expectation model gives every one of them the same 3.51 bits, because its states are scale degrees and there is no register anywhere in them. Conditioning the surprise on the voicing costs no corpus — and the arithmetic says the conditioning belongs beside the probability rather than inside it, for three reasons that can each be computed.

harmony · Tonal-expectation
Which chord of a passage has room for the part that is entering. An oboe entering on one note, tried at each chord of a five-chord passage, scored by how far its own partials sit above the threshold the ensemble already sounding puts over them. The best moment gives it 9.0 decibels of margin and the worst 0.3, a spread of 8.7 — and the best moment is not the quietest chord, which is vi, close below. Room for an entrance is spectral rather than dynamic. A chord with a hole in its written spacing need not have one in its spectrum, because the partials of its bass fill the middle whatever the notes above it do.

The chord that has room for an entrance

Three essays have made the ensemble something a score can change and none of them has asked when. The ensemble already sounding puts a masked threshold over whatever register an entering part takes, and that threshold is set by the voicing rather than by the dynamic — so the five chords of one passage differ by 8.7 decibels in how much of an entering oboe survives them, and the quietest chord of the five is the worst place in the passage to bring somebody in. Swept over the entrant's own pitch, the choice of moment is worth as much as the choice of register.

form · Orchestration
The schedule that hears every entrance best holds the high parts back. Six parts waiting to enter a five-chord passage over four sounding players, each entering once and staying: the schedule under which the least audible entrance is as audible as it can be made. brass on E3 enters at I, open with -1.4 decibels of mean margin over the mask; oboe on E4 enters at vi, close below with -2.6 decibels of mean margin over the mask; clarinet on G4 enters at I, open with -1.4 decibels of mean margin over the mask; voice on C5 enters at IV, close above with 2.1 decibels of mean margin over the mask; violin on G5 enters at V, bracketing with 6.9 decibels of mean margin over the mask; flue pipe on C6 enters at I, hollow with 4.0 decibels of mean margin over the mask. The least audible entrance is at -2.6 decibels and the margins sum to 7.6; of all 15625 schedules 0 have a better least audible entrance and 576 a larger sum.

Room is used up by whoever enters first

The chord with the most room for a part entering alone is a fact about that chord. It stops being a fact the moment two parts want it, because each part that comes in raises the mask over everybody after it. Given six parts waiting to enter a five-chord passage, choosing each part's moment the way one part's moment is chosen puts three of them into the same chord and lands in the bottom fifth of all 15,625 schedules. Placing them one at a time does no better. The schedule under which the least audible entrance is heard best is unique, and it brings the low and middle parts in while the texture is thin and holds the three highest back for the last three chords — because a high part keeps its room over a full texture and a middle part does not.

form · Orchestration
Level does not dilute the register's roughness, it multiplies it. The mean roughness of the I – vi – IV – V – I arrivals at 4 registers, each relative to the register as written, read three ways. Level-free, the bass is 8.6 times rougher than the treble. With every note at 70 dB it is 8.6 times, the same factor, because one level rescales every pair alike. With each chord played at the level that makes it as loud as the written register's chords — 82.8 dB −2 octaves, 75.5 dB −1 octave, 70.0 dB as written, 66.8 dB +1 octave — the bass is 343 times rougher than the treble, because roughness grows with the square of the pressure and the bass needs more of it to be heard at the same loudness.

A rough arrival is rough because of its spacing

The pair the expectation essays report for every chord — how surprising it was, how rough its voicing is — has no level in it. Putting level back in answers the question it left open, and not the way it was framed. At one written dynamic the arrivals keep their order from 40 to 90 dB at three registers of four, and the bass stays 8.6 times rougher than the treble. Made equally loud, the bass has to be played 12.8 dB harder, and it is 343 times rougher: level does not explain the register's roughness away, it multiplies it.

harmony · Tonal-expectation
From partial 3 the room is the slower of the two. Decay rates in nepers a second for each partial of a note on 130.8 hertz, in a concert hall. The rising curve is the string's own loss, which grows as the partial number to the power 1. The flat-ish curve is the room's, from its reverberation time at that partial's frequency. A reverberant field is the source convolved with the room, so a partial's tail falls at the SLOWER of the two — the heavy line — and the room keeps returning energy the string has stopped making. From partial 3, at 392 hertz, the room is in charge: 6 of the note's 8 partials are held up by the room rather than let go by the string. Those are exactly the partials the string was losing fastest, which is why the room does not merely lengthen the note.

The room is the slower of the two

A reverberant field is the source convolved with the room, so a partial's tail falls at the slower of the two rates rather than at their sum — and the room is slower for exactly the partials the string is losing fastest. Half a note's colour is gone in 0.163 seconds in no room at all, 0.313 in a concert hall and 1.441 in a stone church. The destination is identical in all three, because a room cannot hold a partial up above the fundamental it is also holding. What a hall takes away is the rate, and the rate was the whole of the identity cue.

timbre · Envelope
The drain does not stop when the key does. The note does.. Semitones of colour gone, against time, for a note on 130.8 hertz left to ring and for the same note released after 0.4 seconds onto a damper of 0.15 seconds. The two curves lie on each other until the key comes up, and the damped one then ends: the note is inaudible at 0.52 seconds with 8.7 of the free note's 9.9 semitones delivered. A damper adds one loss to every partial alike, so it adds the same number to every decay rate and leaves every DIFFERENCE between rates exactly as it was — the spectrum at each instant is the ringing spectrum shifted bodily down by 400 decibels a second. The colour goes on draining at its own rate the whole time. What the damper takes away is not the drain but the seconds.

A damper changes the clock, not the colour

A damper is an extra loss on the string rather than a second decay, so it adds the same number of nepers a second to every partial — and adding a constant to every rate leaves every difference between rates exactly where it was. The damped spectrum at any instant is the ringing spectrum at that instant shifted bodily down, to machine precision. The colour goes on draining at its own rate; the note simply runs out of seconds, and how many it gets is written on the page as a note value and a tempo.

timbre · Envelope
A room pulls the compass apart rather than evening it out. How long a pizzicato entering 6 decibels above a held note keeps the composite spectrum, at eight pitches across two and a half octaves, heard 15 metres from the stage. no room: 0.07, 0.05, 0.06, 0.06, 0.02, 0.05, 0.05, 0.04 seconds; a concert hall: 0.61, 0.27, 0.50, 0.38, 0.01, 0.25, 0.27, 0.19 seconds; a large stone church: 1.03, 0.39, 0.79, 0.56, never, 0.33, 0.35, 0.23 seconds. Dry the figures barely move — a spread of 3.0 across the whole compass — because a room is the thing that varies with frequency and there is none. In a hall the spread is 44. The room does not scale the dry answer by a constant: it multiplies it by between four and nine times depending on the note, and at C6 it makes the pluck's position worse rather than better, because the two instruments' spectra nearly coincide there and the pluck starts only 4.0 decibels ahead instead of twelve.

One note in the compass loses its pizzicato

Dry, how long a pluck keeps the composite spectrum barely depends on which note it plays: three-hundredths of a second at the worst pitch and seven at the best, a spread of three. In a concert hall the same eight notes spread by a factor of forty-three, and in a stone church one of them never gets the note at all. The room does not scale the dry answer by a constant — it multiplies it by between four and nine times depending on the pitch, and at the one note where the two instruments' spectra nearly coincide it makes the pluck's position worse instead of better.

timbre · Spectrum
The breath is the looser ceiling nearly everywhere. How long a trained singer can hold a phrase on one breath, across a compass and at four dynamics, against the 8-second ceiling the psychological present puts on the same phrase. The flow through the folds rises with pitch and with loudness, so the breath ceiling falls both ways: at 60 decibels it runs 32.6 seconds at the bottom of the compass to 21.2 at the top; at 70 decibels it runs 23.1 seconds at the bottom of the compass to 15.0 at the top; at 80 decibels it runs 16.4 seconds at the bottom of the compass to 10.6 at the top; at 90 decibels it runs 11.6 seconds at the bottom of the compass to 7.5 at the top. The shaded line is the listener's ceiling and it does not move. The breath binds only where the two lines cross — 1 of the 40 cells drawn, all of them loud and high. So the constraint everybody names when asked why a phrase is the length it is, is almost never the constraint that decides it.

The ceiling everybody names is the loose one

Ask why phrases are the length they are and the answer given is the breath. It is arithmetic — usable lung volume over the air a note costs per second — and it comes out between fifteen and twenty-three seconds at a comfortable dynamic and between seven and twelve at a loud one. The ceiling the present moment imposes, the two-to-eight seconds inside which a stretch is heard as one thing rather than as a series, is two to three times tighter at almost every note and dynamic. A singer in an adagio is not running out of breath at the phrase end. They are running out of present.

form · Phrase
Given the bar in octaves, the cue gets the degree back. How often the reading names the right scale degree, against how much the bass cue is worth, under three rules for what a bass note rewards. the bass is the root: 68 per cent at best; the bass is some chord tone: 67 per cent at best; the bar in octaves: 92 per cent at best; the bass is the root, on a root line: 89 per cent at best. The published cue rewards the triad rooted on the bass, which is right only when the chord is in root position; rewarding any triad containing the bass is right always and rewards three chords a bar instead of one. The third rule is not a bass cue at all — given the bar voiced in octaves a listener knows which three pitch classes are sounding, and that names the chord outright. It reaches 92 per cent on a real line against the published cue's 68 on the same line, and the dashed curve is what that cue manages on a line made entirely of roots — 89 per cent.

Given the bar in octaves, the degree comes back

A bass note is a bare pitch class in the usual figure, and a register in any realisation anybody plays. Voice each bar in octaves and a listener knows which three pitch classes are sounding and which is at the bottom, which names the chord outright — and the rule that uses it reads the right scale degree in 92 per cent of bars on a real bass line, against 68 for the published cue on the same line and 89 for that cue on a line made entirely of roots. The question was whether the register recovers the 89. It recovers it and passes it.

scales · Key-relations
A sharper cue is worth nothing to a reading that will follow it anywhere. How often the reading names the right scale degree, against how much it costs to change key, at a bass worth 1.5 on a real bass line. the bass is the root: 20 per cent at a key cost of 0.5 and 68 at 4; the bass is some chord tone: 47 per cent at a key cost of 0.5 and 59 at 4; the bar in octaves: 46 per cent at a key cost of 0.5 and 91 at 4; roots, and a line of roots: 49 per cent at a key cost of 0.5 and 91 at 4. Where a key change is cheap the rule that names the chord outright reads no better than the rule that names three — 46 against 47 per cent — because a reading that will move key for one bar's evidence follows a sharp cue wherever it points. The sharper cue's whole advantage appears only once the reading is reluctant enough to stay put, and by a key cost of 2.2 it is 31 points ahead.

A sharper cue is worth nothing to a reading that moves

The rule that names the chord outright reads 92 per cent of scale degrees right where the published bass cue reads 68 — at the key cost these readings have always been run at. Sweep that cost and the advantage is not a property of the cue. Where a change of key is cheap the sharp rule reads 46 per cent and the vaguest rule 47, because a reading that will move key on one bar's evidence follows a sharp cue wherever it points. The cue's whole value is borrowed from the model's reluctance to be moved.

scales · Key-relations
The two parameters are not one, and the reason is a ceiling. The plane of the two parameters these readings have been swept one at a time: how much weight the bass cue carries, against what a change of key costs. Every cell is how often the reading names the right scale degree, and the lines are the contours of equal share. along the 50 per cent contour the product of the two coordinates runs from 0.10 to 0.50; along the 60 per cent contour the product of the two coordinates runs from 0.43 to 1.84; along the 70 per cent contour the product of the two coordinates runs from 0.66 to 2.56; along the 80 per cent contour the product of the two coordinates runs from 1.85 to 4.91; along the 90 per cent contour the product of the two coordinates runs from 2.71 to 15.20. If the two multiplied cleanly those products would be constant and the contours would be hyperbolae. They are not: every contour turns upward and then vertical, because past a bass weight of about 3 more of the cue buys nothing at all and only reluctance is left to buy anything with. The key cost has an interior best, at 3 on this grid, where the reading names 93 per cent of degrees — so a reading that will not change key at all is worse than one that will, which no sweep of a single parameter had found.

The two parameters turn out to have a ceiling between them

The essay before this one asked whether the bass cue's weight and the cost of changing key are one quantity with two names, and said the test was a contour: if they multiply, the curves of equal degree share are hyperbolae. They are not. Along the ninety per cent contour the product of the two runs from 2.7 to 15.2, because past a bass weight of about one and a half the reading saturates and more cue buys nothing. And the sweep finds something no single-parameter sweep here could: the key cost has a best value, and a reading that will never change key is worse than one that will.

scales · Key-relations
With 4 of six required at the end, the best schedule hands parts over. Six parts entering, leaving and re-entering a five-chord passage over four sounding players, in the walk through all 64 sets of sounding parts that makes the least audible entrance as audible as possible, with no memory of the chord before and at least 4 of the six sounding at the last chord. I, open: clarinet on G4 enters at 0.6 dB; vi, close below: oboe on E4 enters at 0.3 dB, and clarinet leaves; IV, close above: brass on E3 enters at -0.7 dB, voice on C5 enters at 2.2 dB, and oboe leaves; V, bracketing: violin on G5 enters at 7.2 dB; I, hollow: flue pipe on C6 enters at 4.0 dB. The least audible entrance is -0.72 dB against -2.65 for the best schedule in which nobody leaves; the walk has 2 exits and 6 entrances.

An exit is worth nothing until the tutti is given up

Six parts entering a five-chord passage have a best schedule when each enters once and stays, and letting parts leave and come back was supposed to improve it. Searched over every set of sounding parts at every chord, it improves it by exactly nothing, with or without the chord before still masking — as long as all six must be playing at the end. Let one part be missing from the final chord and the weakest entrance gains 1.4 decibels; let two be missing and it gains 1.9, by a relay in which the parts with least room come in, are heard for one chord, and give way.

form · Orchestration
The content errs slow and the bass errs fast. Passages of eight bars built at three harmonic rhythms, each with a bass that states every new chord's root and moves to another chord tone on a beat 50 per cent of the time. Two readings are asked the chord rate: one from how much the pitch-class content changes across a grid, one from how completely the bass's moves land on it. At two chords a bar the content reading names the rate 63 per cent of the time, too fast 0 and too slow 37; the bass reading 100, 0 and 0. At one chord a bar the content reading names the rate 67 per cent of the time, too fast 0 and too slow 33; the bass reading 18, 80 and 2. At a chord every two bars the content reading names the rate 98 per cent of the time, too fast 2 and too slow 0; the bass reading 0, 100 and 0. The two readings miss in opposite directions and at opposite ends of the tempo range: the content reading at fast harmonic rhythms, by naming a multiple, and the bass reading at slow ones, by naming its own arpeggiation.

The bass errs fast where the content errs slow

Asked how often the chords change, a reading built on pitch-class content names a slower multiple and never a faster rate. Give the passage a bass that states each new root and moves between chord tones inside a chord, and a reading built on the bass's moves errs the other way: it names a faster grid and never a slower one. At two chords a bar the bass is right every time; at a chord every two bars it is never right. Six ways of combining the two readings each trade one end of the range for the other.

harmony · Progression
The arch belongs to hearing, and the spacing only moves it. The share of a close major triad's twenty-four components that stand above what the rest of the chord masks, at 70 dB, with the root from C1 to C7, for three spectra given the same amplitude law and different frequencies: the harmonic series, a founder's bell, and a stiff string with B = 0.01. harmonic series: 0.04 at C1, peaking at 0.79 on E3, 0.42 at C7; a founder's bell: 0.04 at C1, peaking at 0.75 on C4, 0.38 at C7; a stiff string: 0.04 at C1, peaking at 0.71 on E3, 0.46 at C7. Only one of the three is a harmonic series, and all three rise out of the bass, peak in the middle of the compass and fall in the treble.

The arch belongs to hearing, not to the series

A chord delivers most of its partials in the middle of the compass and loses them in the bass and the treble, and every spectrum that showed that arch was built on whole multiples of a fundamental. Give the same amplitudes to a bell's eight modes and to a stiff string's stretched partials and the arch is still there, peaking within a major third of where the harmonic series peaks. What the spacing changes is the detail: a bell crowds its tierce and quint into a quarter of a critical band in the bass and loses them, and a stiff string's stretch buys the bass back.

perception · Masking
Counted over what arrives, the balanced bass is not the roughest register. The mean roughness of the I – vi – IV – V – I arrivals with each chord played as loud as the written register's, relative to the written register, counted over every partial and over the partials that stand above what the rest of the chord masks. Every partial: 70 −2 octaves, 7.48 −1 octave, 1.00 as written, 0.20 +1 octave. Delivered partials only: 6e-9 −2 octaves, 2.83 −1 octave, 1.00 as written, 0.19 +1 octave. Over every partial the lowest register is 343 times rougher than the highest; over what arrives it is the smoothest of the four, and the roughest is −1 octave, 2.8 times the written register.

A bass chord low enough to balance has already hidden its tenor

Played as loud as the written register, a progression two octaves down is 343 times rougher than the same progression an octave up — if every partial on the page is counted. Count only the partials that stand above what the rest of the chord masks and that register is the smoothest of the four, with nothing left that beats. The balance is not what does it: the extra thirteen decibels move no voice by more than two partials. The register had already buried the tenor at the written dynamic.

perception · Tonal-expectation

Named alongside it

The objects these essays reach for when they reach for this one.

Critical bandwidthSpectrumPartialRoughnessBoreMaskingOrchestrationResonanceBoundary conditionBrightnessOverblowingTimbre

All concepts