Notation hides things
The same seven, started later
A mode is not a new scale. It is the same seven notes with a different one treated as home, and the entire change in character comes from reassigning which degree the semitones fall next to.
What a raised seventh is for
The harmonic minor is usually taught as a scale with an odd gap in it. It is better understood as a repair to a single chord — raising the seventh degree turns the dominant triad from minor to major, and everything else about the scale is the bill for that.
A scale is not a set of pitches
Two ragas can have identical pitch sets and be different ragas. Two national theories of one maqam put its third degree thirty-five cents apart. Both facts are fatal to the idea that a mode is a collection of notes, and both are ordinary in the traditions concerned.
Rhythm is a circle, and the bar line is a choice
Draw a rhythm as a line and it looks like a sequence of decisions. Draw it as a cycle and the same pattern turns out to be shared across continents, differing only in where somebody decided to start counting.
Beats of unequal length
A bar of nine in Balkan practice is not nine of anything. It is four beats, three short and one long, and the inequality is at the beat level rather than inside it — which is a thing no single division of a bar can produce.
The milliseconds that are the groove
A Viennese orchestra plays the second beat of a waltz about fifty milliseconds early, every bar. A jazz soloist sits thirty behind the ride cymbal. Neither deviation can be notated, and not because notation is coarse — because it measures the wrong quantity.
The shape of a note, which is most of what an instrument is
Cut the first fifty milliseconds off a recorded piano and listeners stop calling it a piano. The attack carries more identity than the steady tone it leads into, and it is the part every spectrum plot leaves out.
A piece is mostly itself again
Take a piece of music, encode each bar as the notes sounding in it, and compare every bar with every other bar. The picture that comes out has blocks and stripes in it, and those blocks and stripes are the form — arrived at by arithmetic that has never heard of an exposition, a chorus or a refrain.
The boundary is where the neighbourhood changes
A section boundary can be found by an operator that never sees a section. It walks the diagonal of a similarity matrix asking one local question — do the bars behind me resemble each other, do the bars ahead resemble each other, and do the two groups resemble each other — and where the answer is yes, yes, no, there is an edge. What it cannot find turns out to say more than what it can.
How much of this is new
Repetition can be counted rather than looked at. Feed a piece's bars to a compressor and the bits it needs are a measure of how much of the piece is a repeat of an earlier part of itself. The measurement works, the number is real, and it turns out to be a statement about the description rather than about the music — which is the most useful thing it has to say.
A phrase is a number of seconds
Musical phrases are described in bars, and four is the number everybody names. But the constraint that fixes a phrase is a property of the listener and is measured in seconds, so the bar count is whatever the tempo makes it. Across seven ordinary tempos the bar count that lands inside the window moves by a factor of eight, while the window itself does not move at all.
An ending that exists so a bigger one can
Half of the cadences in tonal music are built to fail. A phrase that stopped convincingly at bar four would be a piece four bars long, so the ending at bar four is engineered to arrive and not to settle — and the components it withholds are exactly the ones its partner at bar eight supplies. Closure is nested, and the nesting is what turns two phrases into one thing.
The bar above the bar
A four-bar group is a bar whose beats are bars. That is not an analogy — it is the same computation, and the same metre-induction model produces one when it is handed bars instead of beats, unchanged. What decides where the hierarchy of levels stops is not in the arithmetic at all, and it is a number the phrase essay already measured.
How long until it comes back
A self-similarity matrix has a second reading that nobody looks for. Add up each diagonal instead of walking along one, and out falls repetition as a function of how long ago — a period, in bars, with no segmentation, no kernel width and no bar numbers anywhere in the answer. Five of the six schemes here report the length a listener would have named. The sixth reports something better.
The same thing somewhere else
A measure built on which notes are sounding calls a passage that comes back a fifth higher a stranger. There is a dial that fixes this, and turning it is supposed to be a trade — more sensitivity to a transposed return, less specificity against a coincidental one. It is not that trade. Two different statistics answer opposite ways, and the setting that would compromise between them is the worst one available.
One player is not two clocks
Two accounts of a pianist playing three against two, simulated from the same noise. With a timekeeper in each hand the hands drift apart by a hundred milliseconds inside two dozen bars. With one timekeeper they never drift at all. The measurement everybody cites as evidence for a timekeeper turns out to be the same number under both accounts.
The right period at the wrong phase
Finding the beat is two problems, not one. How far apart the beats are, and where the first one is. This site's rule set answers the first confidently on a tresillo and returns a three-way tie on the second — the same score for the downbeat on step one, step four and step seven — which is not a near miss but no answer at all.
Syncopation is a number about the metre
Longuet-Higgins and Lee price a syncopation at the metrical weight a note skips over. That makes it computable, and it makes it a property of a pair rather than of a rhythm — the son clave scores 4 read from step one, 2 from step three and 8 from step four, and not one onset has moved. Worse, the induction rules pick very nearly the reading that scores lowest.
No term for an unequal beat
A Balkan bar of nine is four beats, three short and one long. The induction rules lay their strong positions at a fixed period, so the only readings of nine it can offer are nine, three and one — and for a bar of seven, where seven is prime, the only readings are seven and one. The right answer is not among the candidates.
The deviations are not noise
Two published timing profiles, split into the quantities they actually carry. A jazz soloist thirty milliseconds behind the ride is almost pure offset and has no pattern at all. A Viennese second beat is almost pure pattern and has no offset. They are different things, they are reported under one word, and no figure of total deviation tells them apart.
Seven rotations that are not seven modes
Rotate harmonic minor and the result is not a family the way the diatonic modes are a family. Four of its six generic intervals come in three specific sizes rather than two, so a fourth might be four, five or six semitones; three of its seven rotations have no fifth above their own tonic; and four have a tonic inside a tritone. The word mode has been doing two different jobs.
A tuning is not a table of cents
Two instruments of a pair are tuned deliberately apart, and the tuner has to choose between a fixed number of cents and a fixed number of beats — which differ by a factor of sixteen across four octaves and cannot both be held. A list of pitches records the first, cannot record the second, and cannot record at all that there are two instruments.
The circle is a circle, and the map is not
Among the twelve major keys, notes in common is a strict function of distance round the circle of fifths — one value for each step count, no exceptions — so there is nothing else to measure and the map really is one-dimensional. A second axis appears only when the minor keys are added, and it is a different kind of move: the relative shares all seven notes and the parallel is one semitone away.
An ending is a deceleration
Every performance slows down at the end and the slowing has a shape. Tempo read against score position is the velocity of a body stopping — a square root rather than a straight line — and the three candidate curves agree at both ends by construction, so the whole audible difference is in the middle, where they part by fifteen per cent of the passage's length.
A silence long enough to be an ending
Four of the five closure components are present or absent. Silence is the one with a continuous scale, so it is the one that can be given a threshold — and the threshold arrives from two literatures that were not chosen to agree, landing between three and a half and five and a half seconds. In a large hall, the room's own decay uses up most of it.
A degree is where it goes next
The first essay on scales beyond twelve said a mode is not a set and listed four things a set cannot record. This one makes the sharpest of them arithmetic. Specify a mode as an ascent and a descent — which is how a living tradition specifies one — and it becomes a directed graph on its degrees; sixty-three such graphs collapse onto a single pentatonic set, sixty-two of them with an ascent that is not the descent reversed, and every standard measure of a scale returns the same value for all of them.
The rotation the necklace cannot see
Ask Bjorklund's algorithm for the world's timelines and the usual answer is that it produces them. Search every rotation of each Euclidean pattern for a match and the answer is more interesting: the tresillo is E(3,8) exactly, the bossa-nova and the standard bell pattern are rotations of theirs, and the son and rumba claves — the two best-known timelines in the world — are not Euclidean at any rotation whatever. Where the match does hold it holds up to a starting position, and a starting position is the one thing a timeline is.
The same distance, under two names
Four hundred cents is a major third or a diminished fourth, and on a keyboard nothing in the sound distinguishes them. An earlier essay was about the boundary between two categories; this is about two categories at one acoustic value, and the surprise is where the ambiguity comes from. In quarter-comma meantone a major third is 386 cents and a diminished fourth is 427 — two names, two pitches, forty-one cents apart. Equal temperament collapsed them, and what a listener now supplies from context used to be in the sound.
Which notes are the chord
A progression is a list of chords, and before there is a list something has to decide which of the notes sounding are chord tones and which are passing. Take the eight notes of a scale as eight quavers and score every triad and seventh at every root: barred as written the best reading is C major seventh, with the barline moved by one quaver it is D minor seventh, and with no metre at all three readings tie exactly and the passage has no best analysis. Same eight notes in all three. Harmonic analysis is a function of a variable that is not harmony.
The stave is not a ruler
A hundred and eighty essays here draw pitch against an axis somebody computed. The one axis every reader already owns is the five lines, and it is not a pitch axis at all: it counts letters. Seven positions carry twelve pitches, so the same vertical distance is two intervals before an accidental is allowed and six after — and the accidental is not an extra symbol on a complete scale but the repair for a scale with five values missing.
Two names for one key
A keyboard has one key between G and A and the page has two names for it. That looks like redundancy and it is not: the two names are twelve fifths apart on a chain, and in every tuning anybody played before the nineteenth century they are two different pitches. The size of the difference is twelve fifths against seven octaves and nothing else — twenty-three cents one way in Pythagorean, forty-one the other way in meantone, and zero at exactly one point in between.
The time signature is a claim
A bar line is not a measurement. It is a claim about where the accents are, made before the sound exists, and there is a metre-induction model here that can be handed the same onsets and asked whether it agrees. On a hemiola it does not: the page says three and the model says six, by a margin of twelve against minus four. And on a bar of seven the signature is not even a candidate — what decides the reading is the beaming, which the signature does not contain.
The mark that is not a level
There are six of them, they carry no units, and a performer has to turn one into a number before it means anything. What they instruct is not loudness. On a struck string a harder blow shortens the hammer's contact from 2.26 milliseconds to 0.95, which moves the first null of its own pulse from the third partial to the sixth: the partials between those are not quieter at pianissimo, they are gone. A fortissimo is a different sound, and the page has one word for both things it changes.
What a tablature keeps
Middle C can be stopped in four places on a guitar. The speaking lengths run from 61 to 27 centimetres, so a hand plucking twelve centimetres from the bridge meets between a fifth and nearly a half of the string, and the comb of missing partials is different at every one: the second partial is thirteen decibels stronger in the best position than in the worst. A stave writes one note for all four. A tablature writes four different things and cannot say which note any of them is.
Three notations, one progression
A figured bass, a Roman numeral and a chord symbol are three professional notations for the same four chords, and the number of four-part realisations each of them admits can be counted exactly rather than argued about. With no parallel fifths or octaves the counts are sixteen billion, fifty-nine million and two million: a factor of eight thousand between the loosest and the tightest. What each one collapses is what its tradition thought a chord was, and the three do not agree.
Where a phrase ends
Run a boundary detector over the three tunes used throughout and it agrees with the notated phrasing on one of them perfectly and on another almost not at all. The reason is which cue each tune uses: Twinkle's phrases all end on a long note, so a duration-weighted detector finds five of five with no false alarms; Ode to Joy's run on in crotchets and its phrasing is in the intervals, where a duration detector finds one of three and a pitch detector finds all three and eight others. No fixed weighting serves both, and the published one is worse on each tune than the single cue that tune uses.
The repeat that is not in the notes
Eight earlier essays compare bars by writing each one as a bag of pitch classes and taking a cosine. Nothing chose that encoding — the first used it and the other seven inherited it. Encode the same six schemes four other ways and one of the three findings survives untouched, one survives with different numbers, and one turns out to have been a statement about the encoding all along: the boundary operator finds none of the section edges in three schemes as a bag of pitch classes and every one of them as tonic, subdominant and dominant.
Three decisions that constrain each other
Every model here decides one thing at a time — the key from the pitch classes, the metre from the onsets, the chords from the metre — and an earlier essay ended by saying a listener does all three at once. Resolving them jointly costs a hundred and fifty-seven times the search and changes the reading of two passages in five. It never once changes the key, and the reason it cannot is the reason the whole account is built the way it is.
The clef is an integer
The first essay on notation found that the staff's vertical axis counts letters rather than pitch, and named the clef as a question it was leaving open. Paid, it is arithmetic: a staff holds eleven letters, no voice or instrument is that narrow, and the eight clefs of European practice step through the axis in thirds — a spacing that buys everything a set of fifteen would buy on a wide range, for eight.
A boundary at a stated level
A boundary detector run over three tunes agreed with the notation on one and barely at all on another, and left two things owing: a version with a scale parameter, and a version run on performance timings. Both are paid here, and they pay differently — the scale removes a free parameter from the comparison and does not rescue the hard case, while two per cent of rubato does.
Playing louder is playing earlier
An accent has two effects on when its note is heard and neither is a timing decision. A harder-driven instrument has a shorter attack, and a criterion set by the surrounding music is crossed sooner by a bigger rise — so a twelve-decibel accent on a bowed note is heard twenty-four milliseconds early with no change whatever in when the bow was put down. It is also the measurement that tells the two competing models apart.
A quantity resting against a wall
The short note of a swung pair was found sitting exactly on the fast edge of the tempo window. If that edge is a floor rather than a fitted number, then every swing statistic computed until now was computed on a censored sample — and a censored sample has a mean that is 0.80 standard deviations too high, a spread that is 40 per cent too low, and a correction gain that can come out twice what the players actually have.
Which player on which note
An interval's roughness depends on which instrument is underneath, so the pair does not commute. Three players over three notes is the smallest thing that asymmetry has anywhere to go: six assignments, all of them the same chord, and across 450 of them the roughest averages half again the smoothest and reaches six times it. It is orchestration in the only form that can be computed here — not which chord, and not which voicing, but who is on which note.
A key-finder that keeps the order
Every key-finding model until now begins by throwing order away — a histogram over a window, correlated against twenty-four profiles — and the thing that most obviously declares a key is an ordered pair of chords. A model that keeps the order costs 84 states and 7,056 transitions against 24 hypotheses and none, and what it buys is the one test a histogram was said not to pass: two keys in turn and two keys at once have histograms 98 per cent alike, and are two different sequences.
The cadence as evidence
Every model so far infers a key from a bag of notes, and the thing that most obviously declares a key is a cadence — an ordered pair of chords in which the order is the whole content. The essays on closure built a five-component cadence vector long before this and nobody has used it for this. Two passages differing in one note, one cadencing in C and one in G, get opposite answers from the ordered pairs and the same answer from the profile.
The level the tempo chooses
The boundary detector has a scale parameter and an earlier essay left it free, ending with the sentence that names this one: the scale is in notes and the psychological present is in seconds. The psychological present is two to eight seconds, a tempo converts one to the other, and the level a listener reads then stops being a parameter at all — which is a prediction with teeth, because the same tune at two tempos should be phrased differently at levels the arithmetic names in advance.
The notations invented for the overflow
Every proposal to replace the staff since the seventeenth century is a response to a specific overflow, and the commonest one — a chromatic staff with twelve positions per octave instead of seven — makes a trade computable from the same two integers the clef essay counted. It buys the accidentals outright and pays 40 per cent of the range, three ledger positions for one, and twenty-three enharmonic distinctions that are not merely absent from the page but unrecoverable.
A low note cannot start on time
Three earlier essays have held the pitch at one value. A note cannot establish an amplitude in less than a few of its own cycles, so the attack has a floor that rises as the pitch falls — 146 milliseconds at the bottom of a piano and three at the top. On an instrument whose action takes eight milliseconds everywhere, that is a forty-three millisecond spread across the keyboard from the period alone, and no player can do anything about it.
The shape a wall leaves behind
A floor on execution biases every statistic computed on swing timing, and the bias was priced with Pearson's truncated-normal formulas. Those are the formulas for a sample with everything below the wall thrown away. A player who cannot execute a short gap does not throw the attempt away — it comes out at the floor. That is a censored sample, its bias is exactly half, and its skew is two thirds larger.
The dynamics are in the score already
Count the parts in each bar, realise them in their ranges, put every partial in its critical band, sum the loudnesses and run the result through the two smoothers built earlier. What comes out is a dynamic curve for a piece with no performance in it anywhere — and it says that doubling the number of parts inside a fixed register adds three decibels of power and about one of loudness, because the extra parts land in bands that were already occupied. Let the register widen with the parts and the same arithmetic gives eight phon, which is what a tutti actually is.
A final chord is not made loud by adding to it
An earlier essay on closure said the loudest cue an ending has needs a corpus rather than an arithmetic. The arithmetic was built one essay ago, so it does not. Run four ending textures through it and two things come out backwards: a final chord three parts thicker than the rest arrives *quieter* against the running impression than the passage it ends, and a texture that drops a part a bar does not get quieter at all until the bar where there is one part left.
The axis that is not a time axis
Eight earlier essays have measured the staff's vertical axis to a position. Its horizontal one has never been asked about, and the answer is that it is proportional to nothing: measured off the typesetter used here, a note gets 58 points before its duration is considered at all and 15 points a crotchet after — so the constant is 93 per cent of what the shortest note gets, and a note four times as long is not four times as wide.
A dynamic mark changes what a note is
Every spectrum until now is a shape with a level in front of it, so that playing ten decibels louder raises every partial by ten. That is true of exactly one instrument in an orchestra. Everybody else steepens their own spectrum as they lean on it, and a trumpet's centre of gravity moves from the second partial to the sixth across a dynamic range while an organ flue pipe's does not move at all.
A detector whose resolution the performance sets
The boundary detector lost its free parameter when the psychological present became a number of notes at a stated tempo, and what that held still was named at the time: a performance slows into a phrase end, so the number of notes inside the present is not the same everywhere in a tune — it falls exactly where a boundary is. Making the width follow the performance recovers half of what removing the parameter cost, and honestly leaves the other half.
The passage built to make them disagree
A cadence count and a key-profile correlation have been put on one scale and run on a passage where the two agree, which tells nobody anything. Building one where they disagree — the notes of one key and the cadences of another — measures the exchange rate at about six bars per cadence, and finds something the agreement case concealed: the two statistics cannot be varied independently, because lengthening the profile evidence adds cadence evidence too, for a third key.
The margin the dynamic program already had
Nine earlier essays produce a single best reading, and the passages worth arguing about are the ones where two readings are nearly equally good. What is needed for that has been inside the model from early on: a dynamic program that finds a best path has, by construction, the best score into every state at every bar — so the gap between the best reading and the best reading in any other key is already computed, and printing it turns every analysis here into a measurement of ambiguity.
How much music a page holds
Nine earlier essays have measured notations, and every one of them is a page — a two-dimensional object read in a fixed order by a reader who has to turn it. One measured the vertical axis and another the horizontal, and multiplying them gives the one design constraint on notation that is not about legibility at all: a chromatic staff turns pages a third more often than an ordinary one, and a proportional spacing rule turns them nearly twice as often as a columnar one.
A subito piano is a rate, not a level
All three earlier essays score one chord held still. An orchestration is a succession, and the running impression carries a chord into the one after it — so a written dynamic is an instruction to the listener's impression rather than to the instantaneous sound, and there are markings that cannot be produced at all. The correction runs to silence and the impression still sits above the target.
A family of readings
Removing the detector's free parameter, and then its constant tempo, cost persistence both times — the property that made its boundaries ordered rather than merely found. Recovering it means a family of adaptive readings rather than one, indexed by the width of the psychological present, which is the one parameter that cannot be removed, because it is a fact about listeners.
The page is read by an eye
A sight-reader's eye sits a fixed number of notes ahead of the sounding one and a fixation takes in a fixed number of millimetres, and the spacing rule converts between them. Two bounds follow, from the reader rather than from the music — and the one everybody would expect to bind does not. The saccade rate has enormous headroom at any playable tempo, and what decides is acuity.
A hole is a short tube
Four earlier essays have treated an open tone hole as a point where the pressure is released. It is not: the air in a hole has mass, and a hole with mass does not end the bore, it loads it. Seven holes drilled at seven stations all sound the same G — and their twelfths are spread over a fourth. Cross-fingering falls out of the same arithmetic, and it is not made of what everybody says it is.
The instrument that cannot be moved
A string is regauged and a woodwind is scaled. An organ's pitch is the length of its pipes, and metal can be cut off and cannot be put back — so an organ is a ratchet that only goes sharp. The mechanical answer was to shift the keyboard against the pipes, and its cost is not the transposition. It is that the temperament's key colours rotate out from under the notation, by an amount measured in fifths rather than in semitones.
An equal note cannot be masked
Three earlier essays are about one instant, and forward masking lasts two hundred milliseconds — longer than a note at any brisk tempo. So a fast line should be a sequence of events hiding each other, and it is not: a note masks itself ten decibels harder than its predecessor can, at any speed. What does hide a line is dynamic contrast, and the boundary is fifteen decibels.
A page has two decibels
The account of closure called its dynamic component a corpus debt: it had no model of dynamics in a form. Loudness supplies one, and applying it answers the debt by refusing it. Adding parts to a final chord one at a time, each at the same level, moves the loudness by under two decibels and not monotonically — while a player has sixty. A texture that thins is not a diminuendo.
A reader does not read notes
Eleven earlier essays count notes, and the page cannot tell one line of eight quavers from another. A reader can: a scale of eight is one object where eight leaps are eight. Measured against the melodic interval distribution, the same eight notes are four times as much to read — and the eye–hand span, the best-measured quantity in the reading literature, is four notes of a tune and one of a leaping line.
A rest is a diminuendo
Three parameters swept over factors of ten, three and forty-eight returned the same reading to four significant figures. The one manipulation left unswept moves it by twenty-four decibels: a silence. A listener's running impression decays at the loudness smoother's two-second release, so a general pause is a diminuendo nobody wrote, and at the length that makes a gap an ending it is worth more than any marking a composer has.
The cutoff that is a list
Five earlier essays have quoted one number for a woodwind's cutoff — 1,824 hertz for a clarinet — from a formula written for an infinite lattice of identical holes. Solve a whole twelve-hole chart instead and the number is eleven different numbers, running from 2,193 hertz down to 1,574, which is 574 cents. The lowest fingering has no cutoff at all, and which way the list runs turns out to be a design decision rather than a fact about woodwinds.
The parent nobody measured
Every figure drawn behind a wall so far assumed a normal parent, and the assumption turns out to matter in exactly the wrong place. The censored bias is half the parent's mean absolute deviation — a theorem, not a coincidence — so it lands between 0.35 and 0.43 spreads for every distribution tried, and the truncated one runs from 0.58 to 1.00. But the third moment proposed as the test moves 2.9 across parents against 0.65 between the two rules, and a censored sample from a slightly left-skewed parent has a skew of 1.007 where a truncated normal has 0.995. The statistic that does work is a count of ties.
A staccato is a dynamic mark
Every loudness figure in this collection is of a sound that has been going on long enough, and no note in music has. Run the running-loudness model on notes with lengths in them and an articulation turns out to command 1.5 phons at a slow tempo and 8.1 at a fast one — more than the 1.8 decibels a whole texture commands, on the same page, written down in the same ink, and counted by nobody.
The listener is given the top voice, and the bass as a sine
Four earlier essays put the masker and the probe in the same voice. Put them in different voices — a four-part texture at one level — and the soprano arrives with all eight of its partials, the alto with five, the tenor with two and the bass with one. Balancing the loudness, which is the constraint a scoring is solved under, changes none of that: equal loudness is not equal spectrum and cannot be made so.
A soft chord has to fade in
Forward masking sits between the two integration times already in play — two hundred milliseconds against a thirty-seven millisecond roughness window and a two-second loudness release — and it was owed as the term that might eat the dissonance contrast. It does not. It widens it, by three per cent at a chorale's pace and fifty-nine at four chords a second, because it can only ever remove partials and it can only reach the chord a dynamic has already made quiet. What it does instead is stranger: one chord in the passage is entirely inaudible for its first twelve milliseconds and takes a quarter of a second to arrive whole.
The key-finder with no tonic
Thirteen essays of key-finding have run over twelve major collections and not twenty-four keys, so a passage in A minor is read as C. Adding the missing twelve costs almost nothing and fixes almost nothing — because the model has no tonic in it at all, and below three raised sevenths in a passage the extra states are worth exactly zero.
Where the chord actually lands
Every timing surprise so far is evaluated on a downbeat: the curve spans a factor of 6.7 and every number is read at its peak. Charge the waiting as well as the arrival and the costs over a bar become a proper distribution, the spread between the best and worst beat rises to a factor of 19.5, and a third of the probability sits on a bar in which nothing changes at all.
The long note and the strong note
The segmentation that produces every object connected here has carried a free parameter since the day it was written: whether a note counts for its metrical weight or for how long it is held. Only the first has ever been drawn. The two name different chords on sixteen per cent of passages where the cues agree about the notes and forty-three per cent where they do not — and where they disagree most sharply a mixture of them picks a third chord neither one asks for.
A standard moves the page, and not the seam
Every earlier essay has priced a pitch standard against something with a fixed length in it. A voice has none, so nothing about it changes at all — what changes is where the written note falls against a break in the larynx that is a frequency and stays put. At A415 that break is written F4, at A440 it is E4 and at Chorton it is E♭4: a minor third of movement across four centuries, on a part nobody rewrote.
A rest needs a dry room
The twenty-four decibels a general pause is worth assume the sound stops when the players do. Put a hall under it and a bar of silence keeps 60 per cent of its value in a shoebox concert hall and 22 per cent in a cathedral, while a three-and-a-half-second one keeps 86 and 45 — because a hall's tail has a length and a rest either outlives it or does not. A written bar of silence is worth half its dry value at 2.76 seconds of reverberation, which falls between the concert hall and the stone church.
Four parts are easier to read than two
Twelve earlier essays read one line, and a score is several at once. Measured through the voice-leading model, a notehead of a four-part chorale asks a reader for 1.30 bits and a note of an independent line asks 1.89 — so twice the ink is less than three quarters of the load. The reason is a boundary those essays already established: a duet has no complete voicing of any triad at all, so two parts cannot be read from their harmony and have to be read as two melodies.
A subito piano is four seconds longer in the bass
Every loudness figure with time in it converts level to loudness at one kilohertz, and the equal-loudness contours say that no other frequency works that way. Joining the two sorts the published numbers into those that were about the treble and those that were not. Three move a great deal — a twenty-decibel crescendo is worth 27 phons on a bass note and 20 on a high one, and the seven seconds a subito piano takes becomes eleven and a third. Three do not move at all, and the reason they do not is the same reason in every case.
A tonic bought with the function
Putting the measured probe-tone profile inside the ordered key-finder is one term, and it does what was predicted: the natural-minor passage is named A minor at every bar instead of C major. It also does two things nobody predicted. It renames a scheme that has been read in the wrong key at every bar since the day it was written, and it destroys the chord's function while it is buying the key.
A fourth decision, and two that were never made
The joint search resolves key, metre and segmentation together and holds the segmentation's cue mixture at zero. Adding the mixture is one loop, and reading the search in order to add it turns up something worse than a missing axis: on the passages it is drawn on, the key it reads is the same key at all forty-eight of its hypotheses and the metre scores every barline identically. The fourth axis then cannot be ranked at all until each reading is measured against its own null, because a mixture changes the ruler and not only the answer.
Eighty-one chords, one number
A dominant seventh has eighty-one arrangements inside three octaves and their roughness spans a factor of five and a half. The tonal-expectation model gives every one of them the same 3.51 bits, because its states are scale degrees and there is no register anywhere in them. Conditioning the surprise on the voicing costs no corpus — and the arithmetic says the conditioning belongs beside the probability rather than inside it, for three reasons that can each be computed.
A general pause is spent by the note after it
Whether a composer should write the pause before the diminuendo or after it looks like a question about how big the ensemble is. It is not. Forty decibels of ensemble are worth one decibel of silence, and the order is worth seven — because a running impression rises twenty times faster than it falls, so half a general pause is spent by ninety-seven milliseconds of sound.
Where the note is costs more than which note it is
Thirteen earlier essays measure a page, and the three that price a reader price only its pitches — every line they measure is a run of equal notes. A metrical weight normalised by its own sum is a probability, and minus its logarithm is bits — the same substitution made earlier for intervals. Measured over the tunes used throughout it comes out at 2.23 bits a note against the pitches' 1.89, so the larger half of a reader's load is where the note is.
A part that leaves is not a part that arrives
The same player, the same note, the same level, and the only difference is which way round it happens. A listener's loudness reading takes 1.43 seconds to receive half of a departure and 90 milliseconds to receive half of an arrival — a factor of sixteen with nothing asymmetric in the sound at all, since both readings integrate the same two states in the same order. The colour reading receives the two identically, because a window has no direction, so a departure is a change whose grain arrives at once and whose level takes most of two seconds.
The chord that has room for an entrance
Three essays have made the ensemble something a score can change and none of them has asked when. The ensemble already sounding puts a masked threshold over whatever register an entering part takes, and that threshold is set by the voicing rather than by the dynamic — so the five chords of one passage differ by 8.7 decibels in how much of an entering oboe survives them, and the quietest chord of the five is the worst place in the passage to bring somebody in. Swept over the entrant's own pitch, the choice of moment is worth as much as the choice of register.
A note lasts until the next one starts
Pricing where a note is against which note it is left duration as the term it had not, with a prediction that it would be small. Measured on the three tunes these readings are built on, it is exactly zero — and it is zero by construction, because those tunes are stored as pitches and lengths with no rests in them, so every duration is its own inter-onset interval. The prediction cannot be tested on the corpus that produced it. Priced directly, a rest costs 0.67 bits a note where a tenth of the notes have one, which is not well under half a bit.
The notehead that is not a note
Every quantity so far is charged per notehead, and a tie is the one mark on the staff that puts a notehead on the page carrying no event. Its cost is not the decision that identifies it — that is half a bit where a tenth of the noteheads are continuations. It is the decision plus the whole reading of a notehead that turns out to have been unnecessary, which is 1.05 bits, twice the decision and a fifth of what a note of music costs. Set beside a dot and a longer note value, the tie is five times the price of either and is the only one of the three that can cross a barline.
Leaps do not fall where offbeats do
Every reading load computed so far is a sum of two terms priced as though the axes were independent, and an earlier essay named the interaction it could not reach. Measured on the same hundred and one notes every other essay uses, the mutual information between how far a note moves and where it falls in the bar is 0.31 bits — a fifth of the smaller axis, and a sixth of a note's total load. Every reading load published so far is high by that amount, and the quantity saturates at exactly the grid the tunes are notated on, which is the check that it is measuring the music rather than the grid.
One number for a page
Four terms and an interaction give a single bit rate per note, and with it the exchange rate a long run of essays has been pointing at. Six kinds of line span 5.26 bits and seven kinds of rhythm span 4.46, so a composer trading a harder tune against a harder rhythm is trading quantities within eighteen per cent of each other — and pages that look nothing alike sit on the same contour. The hardest page on the grid costs 13.96 bits a note and the easiest 4.24, a factor of three and a half, and the subject closes there.
Two cues meet in a corner
The profile finds a key's tonic and the bass finds a chord's degree, and until now each was swept with the other held at nothing. Swept together across 42 settings, the plane they make is not the ridge that was predicted. The tonic is a step in one direction, at a profile share of 0.55, and the bass cannot move it; the degree is a slope in the other, rising to 89 per cent as the bass is weighted, and the profile barely touches it. Every question is answered only in a corner of the plane — and the one place the two cues overlap is the one piece of music both can rescue.
A bass line is not a list of roots
Every bass note the key-finder has been given was its chord's root, and under that line a bass cue reads 89 per cent of scheme bars on the right degree. Give the same chords an economical bass that moves to the nearest chord tone, as a keyboard reduction would, and nearly half of them are inverted. The cue that rewards the triad rooted on the bass then reads 68 per cent at best and worse as it is trusted more; the cue that rewards any triad containing the bass cannot be fooled and stops at 61. The same inverted line does one thing the roots never did: it puts the leading note of each new key at the bottom, and finds the rondo's modulations.
The smoothness is in the skips
Traditional scales come out smoother than random scales of their size because every pair of their degrees is counted once. Count only the pairs a melody on the scale actually sounds next to each other, and every seven-note scale here is rougher than ninety per cent of random ones, under every spectrum, and under a pure tone as well. The smoothness is carried by the pairs a melody reaches by skipping — its thirds and above all its fifths, which are smoother than all but two random scales in two thousand. Counted by their own ascents and descents, the two ragas that share one set of notes stand in different places at last: Deshkar's skipped Re buys a smoother ascent and costs it the fifths.
The chords never move the barline
Every hypothesis the joint search had drawn was one bar long, and on one bar with a note in every slot the metre cannot choose a barline at all. Four bars with rests in them make the barline a decision the metre and the chords both have an opinion about, and the prediction was that the chords would move the barline more often than the barline moves the chords. It is the other way round, completely: whenever the two prefer different barlines the search takes the metre's, on up to 72 per cent of passages, and in fifteen hundred passages the chords never once move it. What the chords decide is the one thing the metre cannot see — whether the bar starts on the downbeat or half a bar later — and they decide it right a little over two times in three at best.
The chords are a weak witness to the barline
Scaled by its own range, the metre overrules the chords every time the two disagree about where a bar begins. The obvious repair is to score each reading against its own chance — the metre against the same number of notes placed at random, the chords against the passage's notes shuffled across its bars — and add the standard scores. It changes very little: the search finds the barline within six points of where the product found it, and the chords gain the power to move the barline only on passages whose rhythm says nothing, where random notes move it nearly as often. The null's real result is the size of the two witnesses. At the written barline the metre stands up to 5.9 standard deviations above chance, and the chords, with every note a tone of its bar's chord, stand 1.55 above it at best.
Room is used up by whoever enters first
The chord with the most room for a part entering alone is a fact about that chord. It stops being a fact the moment two parts want it, because each part that comes in raises the mask over everybody after it. Given six parts waiting to enter a five-chord passage, choosing each part's moment the way one part's moment is chosen puts three of them into the same chord and lands in the bottom fifth of all 15,625 schedules. Placing them one at a time does no better. The schedule under which the least audible entrance is heard best is unique, and it brings the low and middle parts in while the texture is thin and holds the three highest back for the last three chords — because a high part keeps its room over a full texture and a middle part does not.
A proportion is only as fine as its two durations
Analyses of form measure proportions in bars and report them to three figures — a climax at 0.618, a section in the ratio 3 : 2. A listener has each part only as an estimate of how long it lasted, and a ratio of two estimates is blurred by both. Timed as well as anyone times a single second, eleven proportions fit between 1 : 1 and 3 : 1; timed from memory over minutes, two do. The golden section is told from 3 : 2 only below a Weber fraction of 5.4 per cent.
A final chord stands out for a twentieth of a second
A general pause drives a listener's running impression down, and the final chord that follows is supposed to cash the fall in. It cashes in at most two thirds of it. The note's own loudness rises with a 22-millisecond constant and the impression with a 99-millisecond one, so after a bar of silence the chord stands furthest above the impression 54 milliseconds in, by 5.1 of the 8.1 decibels the silence bought, and after 206 milliseconds the two are within a phon of each other. A short stamp spends most of its life standing out; a chord held a second and a half spends a seventh of it.
The chords mark the barline by changing there
Read bar by bar, the chords stood barely above chance at the barline and broke the metre's half-bar tie two times in three at best. Read instead by where they change — how different the chords are across a candidate's barlines against how different they are across the middle of its bars — the same notes break the tie right on 81 to 96 per cent of passages, and added to the metre they find the barline on up to 89 per cent against 61. The weakness was the question the old reading asked, not the harmony.
A scale is committed to how long its instrument rings
A traditional scale's standing on the roughness model moves when the instrument changes, and that movement was read as a commitment to the instrument's spectrum. Weight every pair of notes a melody sounds by how much of the earlier note is still ringing when the later one begins, and the movement grows — the tempered diatonic's by 37 percentile points at a note every 0.6 seconds — but almost none of it is the spectrum. Put four instruments' rings on one spectrum and the scale moves 38.6 points; put four spectra under one ring and it moves 9.4. And the partials that make a fifth smooth are the first to stop sounding.
A damper changes the clock, not the colour
A damper is an extra loss on the string rather than a second decay, so it adds the same number of nepers a second to every partial — and adding a constant to every rate leaves every difference between rates exactly where it was. The damped spectrum at any instant is the ringing spectrum at that instant shifted bodily down, to machine precision. The colour goes on draining at its own rate; the note simply runs out of seconds, and how many it gets is written on the page as a note value and a tempo.
A form is sharp at the bottom and vague at the top
A movement is divided into sections, each into phrases, each into bars, and every level is a ratio of two estimates. Timed, the hierarchy is almost uniformly blunt — 2.1 distinguishable proportions at the top and 5.2 at the bottom, because a Weber fraction is a step function of duration and six of a piece's seven levels fall in one step of it. Counted, the same hierarchy runs from 2.2 to 12.2 and sharpens monotonically downward. At no level of either does timing separate 3 : 2 from the golden section.
A golden section is a coin toss with six coins
An analysis that reports a climax at 0.618 of a piece has not tested one prediction; it has looked at a piece with several defensible boundaries and reported whichever landed nearest. The rate at which that happens under no hypothesis is one line of arithmetic, and the tolerance it needs is not a number chosen on the page — it is the blur a listener's own timing puts on the judgement. Over a stretch of minutes that blur covers everything from 0.492 to 0.730 of the piece, which contains the halfway point, and six candidate boundaries produce a hit eighty per cent of the time.
The change reading follows the chords, not the bar
Every passage read until now changes chord exactly at the barline, which is the one harmonic rhythm at which 'the chords change here' and 'the bar starts here' are the same sentence. Pull them apart and the reading goes with the chords: at one chord a bar it stands 1.52 standard units above the other candidates and finds the barline half the time, at two chords a bar it stands 0.01 above them and is at chance, and at a chord every two bars its margin is exactly half — because half the barlines then carry no change at all.
Asked for the rate, it answers a multiple
A reading that follows the chord rate rather than the bar can be asked what the rate is, and the shape of its errors is the whole of why it looked like a barline detector. Given two chords a bar it returns the right period a third of the time and something slower two thirds; given a chord every two bars it is right nine times in ten. It errs slow and essentially never fast, because a change every four slots also falls on every eighth slot and a slower grid inherits a faster rate's evidence — which is the same asymmetry that makes a pitch detector report an octave too low.
Given the bar in octaves, the degree comes back
A bass note is a bare pitch class in the usual figure, and a register in any realisation anybody plays. Voice each bar in octaves and a listener knows which three pitch classes are sounding and which is at the bottom, which names the chord outright — and the rule that uses it reads the right scale degree in 92 per cent of bars on a real bass line, against 68 for the published cue on the same line and 89 for that cue on a line made entirely of roots. The question was whether the register recovers the 89. It recovers it and passes it.
A sharper cue is worth nothing to a reading that moves
The rule that names the chord outright reads 92 per cent of scale degrees right where the published bass cue reads 68 — at the key cost these readings have always been run at. Sweep that cost and the advantage is not a property of the cue. Where a change of key is cheap the sharp rule reads 46 per cent and the vaguest rule 47, because a reading that will move key on one bar's evidence follows a sharp cue wherever it points. The cue's whole value is borrowed from the model's reluctance to be moved.
The two parameters turn out to have a ceiling between them
The essay before this one asked whether the bass cue's weight and the cost of changing key are one quantity with two names, and said the test was a contour: if they multiply, the curves of equal degree share are hyperbolae. They are not. Along the ninety per cent contour the product of the two runs from 2.7 to 15.2, because past a bass weight of about one and a half the reading saturates and more cue buys nothing. And the sweep finds something no single-parameter sweep here could: the key cost has a best value, and a reading that will never change key is worse than one that will.
A louder final chord is a deeper silence and a brighter sound
A final chord marked a step louder than the passage was supposed to stand above a listener's running impression for as long as it sounded, since the impression can climb no higher than the chord. It climbs exactly that high, and the stand closes in half a second as it always did. What a louder mark actually buys is depth — about seven phons, the same depth a second of silence buys — and a spectrum whose balance point sits most of a whole tone higher, which, unlike the stand, lasts for the whole chord.
The bass errs fast where the content errs slow
Asked how often the chords change, a reading built on pitch-class content names a slower multiple and never a faster rate. Give the passage a bass that states each new root and moves between chord tones inside a chord, and a reading built on the bass's moves errs the other way: it names a faster grid and never a slower one. At two chords a bar the bass is right every time; at a chord every two bars it is never right. Six ways of combining the two readings each trade one end of the range for the other.
The ceiling is thirteen bars with names
The key-finder's two cues stop buying anything at about 93 per cent of scheme bars read right, and the obvious suspect was the chord segmentation feeding it. The reading was never given a segmentation: every bar arrives with its true chord. What the ceiling is made of can be listed instead, and it is thirteen bars of 188 — seven in a bridge of secondary dominants, four in a minor episode whose chords C major also owns, and two at the edges of a modulation. No weight of any cue moves one of them.
A bass that holds through a change marks the barline
At two chords a bar the chords change on the barline and on the half-bar alike, so a reading of where they change is at chance, and the metre ties the two. The bass has one more piece of evidence: a change inside the bar can be voiced over the note already sounding, and a change on the barline is voiced over its root. Hold the bass through one mid-bar change in six and the metre and bass together place the barline in 85 per cent of eight-bar passages; one in three, 97. The convention cannot be stronger than that, and the chord changes, asked first, only get in the way.