What the listener supplies
The ear builds objects, and sometimes offers a choice
What arrives at an ear is one pressure signal. What a listener gets is a set of separate things — a violin, a voice, a car outside. The assignment is a construction, and the clearest evidence is that it can be flipped by changing nothing but the speed: one sequence of tones is a single line when slow and two lines when fast, with a wide region in between where the listener may choose.
A memory for the note itself, and it is dated
Absolute pitch is usually described as a rare perceptual gift. It is better described as a memory for a convention — and conventions have dates. Possessors trained on A=440 mis-name Baroque pitch by a semitone, their own labels drift sharp with age, and meanwhile most listeners without it start familiar songs within a semitone of the record.
The chord is still major, and that is why temperament works
A major third can be seventeen cents wrong and still be a major third. That tolerance is not a failure of hearing — it is the reason the whole subject of tuning is a discussion rather than a catastrophe. Every temperament ever proposed moves intervals around inside their categories, and the one thing none of them may do is push one across a boundary.
Counting produced the hierarchy
Ask listeners how well each of the twelve notes fits after a passage in C major and the answers are not a smooth gradient. They fall into four groups with no overlap at all: the tonic, then the rest of the tonic triad, then the rest of the scale, then everything else — categories the subject had names for centuries before anybody ran the experiment.
Consonance is half learned, and this is the half
This site's founding claim is that consonance is small whole numbers. Two computable models say so and they disagree about which chords — which is already awkward. The cross-cultural evidence is worse: listeners with little exposure to Western music discriminate roughness exactly as anyone does, match octaves exactly as anyone does, and rate consonant and dissonant chords as equally pleasant.
The beat has a preferred rate, and it is not the notation's
Metre is a hierarchy of levels and only one of them is tapped. Which one is decided by the listener's own clock rather than by the time signature: the window in which a series of events can be a beat at all runs from about a tenth of a second to two seconds, with a marked preference near half a second — so at the extremes of tempo the notated beat and the felt one part company, predictably.
A bell has no fundamental
The note a listener names when a church bell is struck is not any partial the bell has. Its nominal, twelfth and double octave sit at 2, 3 and 4 times the prime, which is a harmonic series on a pitch an octave below the loudest thing in the sound — and that pitch is supplied by the listener. Founders have been tuning it by ear since the fifteenth century.
What a drum is doing instead
An ideal membrane's modes are zeros of Bessel functions — 1, 1.59, 2.14, 2.30 — which support no common fundamental, so a drum rings without a note. A timpani is that problem solved — the kettle's air and the radiation load drag four modes onto 1, 1.5, 2, 2.5, and the pitch a timpanist tunes is the fundamental those four imply and none of them is.
The boundary is where the neighbourhood changes
A section boundary can be found by an operator that never sees a section. It walks the diagonal of a similarity matrix asking one local question — do the bars behind me resemble each other, do the bars ahead resemble each other, and do the two groups resemble each other — and where the answer is yes, yes, no, there is an edge. What it cannot find turns out to say more than what it can.
A phrase is a number of seconds
Musical phrases are described in bars, and four is the number everybody names. But the constraint that fixes a phrase is a property of the listener and is measured in seconds, so the bar count is whatever the tempo makes it. Across seven ordinary tempos the bar count that lands inside the window moves by a factor of eight, while the window itself does not move at all.
One of these eight-bar phrases accelerates
The sentence and the period both occupy eight bars, both end with a cadence, and both are recognised by ear rather than counted. What separates them is arithmetic. One halves its unit halfway through and the other does not, and the difference comes out as a single ratio — 0.67 against 1.00 — computed from nothing but the lengths of the parts.
What makes an ending an ending
Cadences are ranked. The authentic one is strong, the plagal weaker, the deceptive weaker still — and the solver here measured the quantity that ranking is usually explained by and found it says something else entirely. What survives is not a weaker version of the ranking but a different kind of object, with five components and no total.
An ending that exists so a bigger one can
Half of the cadences in tonal music are built to fail. A phrase that stopped convincingly at bar four would be a piece four bars long, so the ending at bar four is engineered to arrive and not to settle — and the components it withholds are exactly the ones its partner at bar eight supplies. Closure is nested, and the nesting is what turns two phrases into one thing.
The bar above the bar
A four-bar group is a bar whose beats are bars. That is not an analogy — it is the same computation, and the same metre-induction model produces one when it is handed bars instead of beats, unchanged. What decides where the hierarchy of levels stops is not in the arithmetic at all, and it is a number the phrase essay already measured.
The chord that did not come
A deceptive cadence is described as a surprise, and the explanation offered is that the wrong chord arrived. Measured against the tonal hierarchy already in use, the wrong chord is the second best-fitting triad in the key — and two of its three voices do exactly what they would have done in the right one. The surprise is not statistical. It is one voice, and it is the bass.
How often the chord changes
Two pieces can use the same chords in the same order and be nothing alike, because a progression says which chords and not how fast. Harmonic rhythm is the second variable, it runs over a factor of thirty between the styles that use it, and both of its limits are set by things that are not harmony — a listener's memory at the slow end and a building at the fast one.
A cycle cannot cadence
Every component of closure is defined by a first time and a last time. Music built on a repeating cycle has neither, so the whole apparatus returns zero on it — not a small value, zero, at every setting. What such music uses instead is how many things are playing, and that is a curve which can be computed from the onsets and nothing else.
Every interval a different number of times
Count the intervals inside a major scale and the six answers are 2, 5, 4, 3, 6 and 1 — six different numbers, no two alike. That is not decoration. It means the number of notes a key shares with a transposition of itself identifies the distance uniquely, so a listener who can only count common tones can still tell exactly how far a modulation went.
The form a first hearing cannot have
Every figure so far was computed with the whole piece in hand. Run the same methods over only the bars already heard and one of the two methods survives intact — the boundary operator turns out to be causal at a fixed delay of a few bars — while the other collapses. The period of a piece is not knowable until the piece is nearly over, and in two of the six schemes here not until its last bar.
A return has to be remembered
A stripe four bars off the diagonal and a stripe twenty-four bars off it are the same ink and are not the same experience. Convert the lag axis to seconds, discount every comparison by how long ago it was, and the ranking of these six schemes by how repetitive they are changes — and the decay constant and the tempo turn out to enter the arithmetic as one number rather than two.
Two pipes for a note neither makes
The resultant stop sounds a sixteen-foot pipe and a ten-and-two-thirds-foot pipe and asks the listener for a thirty-two-foot note. It is the missing fundamental built on purpose from the fewest partials that can imply anything — and counting how many fundamentals two partials actually imply is the arithmetic behind three centuries of builders disagreeing about whether it works.
The bass a small loudspeaker does not make
A three-inch cone at its excursion limit produces sixty-six decibels at forty hertz, which the ear converts to seventeen phons — barely above nothing. The note is heard anyway, because its harmonics are radiated and its fundamental is supplied by the listener. Computing what arrives turns the residue from a curiosity into a design decision, and finds that the fundamental of a low note is not the loudest part of it on any system a listener is likely to own.
A dissonance is what has to be resolved
The perfect fourth is a consonance between two upper voices and a dissonance against the bass, and the same three notes are involved either way. Score both arrangements with the roughness model and the one the rules call a dissonance comes out thirty per cent smoother. Whatever the rule is tracking, it is not the sound.
Which of the two is the beat
A polyrhythm is notated as a bar of two with three laid across it. Run the metre rules used here over the composite and they choose the three — at every ratio tried, without exception. Add the one other rule the rules have, and the answer changes hands at a bar of 1,013 milliseconds, which is 118 to the minute.
Syncopation is a number about the metre
Longuet-Higgins and Lee price a syncopation at the metrical weight a note skips over. That makes it computable, and it makes it a property of a pair rather than of a rhythm — the son clave scores 4 read from step one, 2 from step three and 8 from step four, and not one onset has moved. Worse, the induction rules pick very nearly the reading that scores lowest.
The beat that is never sounded
A listener who has heard four bars of a groove and then hears two bars with the downbeats taken out does not move the downbeat. This site's rule set does, every time, on every pattern tried — and the direction it moves in says exactly what kind of model would be needed instead.
Nothing in the census knows which note is home
All four properties six earlier essays are about are invariant under rotation and transposition — one property tuple over all eighty-four rotations and transpositions of the diatonic set. So the census that separates 349 shapes cannot separate a major scale from its own Aeolian mode, and everything that makes one note a tonic is outside it.
What a tonic costs in seconds
The standard key-finding algorithm cannot be run on a pitch-class set at all — a flat histogram has no variance and the correlation is undefined. Give it durations and it answers with the parent key for all seven modes identically, and it takes between 15.8 and 30.0 per cent of the total time spent on one note before it names that note instead.
How much evidence a modulation needs
Run a key-finder bar by bar over a progression that moves to the dominant at bar six. With four bars of history the answer becomes the new key at bar seven and holds. With three bars it never gets there at all, and reports E minor and B minor on the way. The window decides the lag as much as the music does.
One voice over ninety players
A soloist heard over a full orchestra is not louder than it and could not be. What the trained voice does instead is put a peak of energy at three kilohertz, which is where the orchestra's spectrum has already fallen away and where the ear's own threshold happens to be lowest. Nineteen decibels of advantage, in a place nobody is competing for.
The root an ear supplies
A major triad's notes fit 4:5:6 with nothing missing and a fundamental two octaves below the bass. A minor triad's fit two different series with two different answers a major sixth apart, and the model cannot choose between them. The ambiguity theorists argued about for two centuries is a computable quantity, and the spectrum invented to remove it is one no object produces.
The leap that pays itself back
Every melody textbook teaches that a leap should be followed by a step in the opposite direction, and every corpus that has been counted agrees — around seven leaps in ten are answered that way. A random walk with two walls, no memory of the leap and no rule of any kind reverses after 61 per cent of them, and the residue is not a rule either. What is left when the walls are accounted for is a prediction the rule does not make, and it is the prediction that decides between them.
The shape that survives everything else
Throw away a melody's key, its tuning, its instrument and the sizes of its intervals, and what is left is a string of pluses and minuses. That string is what a listener who cannot name a note still has, and it costs 37 per cent of the tune to keep. The arch that melodic shape is famous for is not in it as a preference: enumerate every six-note sequence that begins and ends on the lowest degree it uses and 99.6 per cent of them are arches, because a melody that comes home from below has nowhere to go first but up.
The sound a listener knows best
A voice is recognisable across every vowel it says, across two octaves of pitch, down a bad telephone line and in a whisper where there is no pitch at all. Nothing that survives all of that can be a frequency. What survives is a ratio: the resonances of a vocal tract are set by its length, so a shorter tract multiplies every formant by the same factor, and identity is a scale on the spectral envelope rather than a position within it. Between an adult man and a child the whole pattern moves by a fifth, and the vowel does not change at all.
Which notes are the chord
A progression is a list of chords, and before there is a list something has to decide which of the notes sounding are chord tones and which are passing. Take the eight notes of a scale as eight quavers and score every triad and seventh at every root: barred as written the best reading is C major seventh, with the barline moved by one quaver it is D minor seventh, and with no metre at all three readings tie exactly and the passage has no best analysis. Same eight notes in all three. Harmonic analysis is a function of a variable that is not harmony.
The time signature is a claim
A bar line is not a measurement. It is a claim about where the accents are, made before the sound exists, and there is a metre-induction model here that can be handed the same onsets and asked whether it agrees. On a hemiola it does not: the page says three and the model says six, by a margin of twelve against minus four. And on a bar of seven the signature is not even a candidate — what decides the reading is the beaming, which the signature does not contain.
The part of the tune that is kept
Contour survives transposition, retuning, a change of instrument and a doubling of every interval, and the usual explanation is that it is what a listener retains. That can be counted rather than assumed. A six-note melody over eight degrees carries eighteen bits; its contour carries 6.59 — not the 7.92 the number of distinct shapes suggests, because the shapes are wildly unequal — and the effective alphabet is ninety-six out of two hundred and forty-three. Each further note adds 1.28 bits of shape against three of melody, and at about nine notes a contour is specific enough to pick one tune out of a thousand.
The pitch that moves the wrong distance
Take three partials two hundred hertz apart and move every one of them up by forty. The spacing has not changed, so anything reading the pitch off how often the waveform repeats must give the same answer as before. The pitch moves to 204 — the shift divided by the harmonic number — which is what a harmonic template predicts and what listeners report. Push the shift to a hundred and a second reading overtakes the first, so there are two pitches and neither is the spacing. This is the measurement that closes the question, and it closes it by ruling out one mechanism rather than by choosing between the two that are left.
The one note that decides the mode
A key signature names the seven and not the rotation, so the mode has to be heard. Enumerate every set of degrees that lies inside at least one of the seven modes — 510 of them — and count how many modes each leaves standing. Fifteen minimal sets decide, every one has two members, and every one is the mode's own tritone. The only degree all seven share is the tonic; no degree belongs to a single mode; and averaged over the orders the degrees could arrive in, 5.67 of the seven have to have been heard first.
Where a phrase ends
Run a boundary detector over the three tunes used throughout and it agrees with the notated phrasing on one of them perfectly and on another almost not at all. The reason is which cue each tune uses: Twinkle's phrases all end on a long note, so a duration-weighted detector finds five of five with no false alarms; Ode to Joy's run on in crotchets and its phrasing is in the intervals, where a duration detector finds one of three and a pitch detector finds all three and eight others. No fixed weighting serves both, and the published one is worse on each tune than the single cue that tune uses.
What the onsets left out
Eight essays induce a metre from a list of ones and zeros, and every failure they recorded was argued about as a failure of the rules. Two of the three are not. Note length is already in that list and the scoring throws it away: put it back and the son clave's two-way tie resolves to the notated downbeat. But the groove with its beat removed cannot be repaired by any cue at any strength, and the reason is arithmetic rather than empirical — the true phase has no onset on any of its strong positions, so there is no credit for a cue to multiply and its line is exactly flat.
Two keys at once
Every key figure so far assumes one key is sounding, and the standard key-finder has no value it can return that means two. Play one progression in C and the same progression a major third away at the same time, and the model does not report uncertainty: it reports E minor, at a correlation of 0.886, against 0.959 for the same progression in one key. It is as confident as it ever is, and neither key it names is being played. Give it the missing hypothesis — pairs of key profiles rather than single ones — and it recovers both keys at every separation, all seven of seven.
The boundary that barely moves
Every identification figure here has fixed category centres, and the essay before this one ended by admitting that real boundaries are supposed to move with context. Three mechanisms could move one, and their predictions are an order of magnitude apart and in two different directions. Expectation on its own — a listener who thinks one interval twenty times more likely than the other — is worth three and a half cents.
The alternation a key-finder cannot follow
Two keys sounding together are not in the key-finder's vocabulary, and an earlier essay ended by pointing at the other case and saying what was missing: a passage whose alternation rate can be varied while everything else is held still. Built, it gives a rule with three numbers in it — the second key is never named below a block of three bars, the tracking clears ninety per cent above twice the window, and the reading is late by half the window throughout.
Three decisions that constrain each other
Every model here decides one thing at a time — the key from the pitch classes, the metre from the onsets, the chords from the metre — and an earlier essay ended by saying a listener does all three at once. Resolving them jointly costs a hundred and fifty-seven times the search and changes the reading of two passages in five. It never once changes the key, and the reason it cannot is the reason the whole account is built the way it is.
The note that sounds twice
A triad has three notes and a four-part texture has four voices, so one note is doubled — and the voicing model used here leaves the choice free because the rules have an opinion about it. Asked properly, the arithmetic agrees with the treatises for the first time in nine essays: root, then fifth, then third. For a minor triad it does not agree, and for a symmetric chord it correctly has nothing to say.
A note is heard after it starts
Every rhythm essay until now has treated a note's onset as the moment it happens. It is not: the instant a listener aligns a note with a beat is later than its physical start by an amount the note's own attack decides, and for a sung or bowed note that amount is about thirty milliseconds — the size of the whole quantity six essays on microtiming set out to measure.
How much an anchor would have to be worth
Two pitch errors added in quadrature assume an independence nobody measured — a listener inside a key hears a note as a scale degree, and a shared reference is exactly a correlated error. Turning the dial is not evidence. What is evidence is that the dial is not free: a shared error cancels out of a difference completely, so a listener's single-note limen and their interval limen give the two components with nothing left over, and halving the interval limen needs a correlation of exactly 0.75.
The listener the model was never run for
Five earlier essays rest on one number — how finely a listener resolves a pitch — and every one of them used a trained listener's eleven cents. The model's dependence on it is not gentle: the capacity goes as its reciprocal and the expectation shift as its square, so an untrained listener at thirty-five cents has two nameable categories per octave rather than six, and a foreign tuning system is not mis-transcribed by them but absorbed.
The spectrum that will not fuse
A partial about one per cent off its harmonic is heard as a sound of its own rather than as part of a note. Apply that criterion to a whole spectrum instead of to one mistuned component and it becomes a count: a piano string keeps nine of its ten partials, a bell keeps seven of eight, a bar keeps two of six. The physics of inharmonicity has had an essay here for a long time. This is what it sounds like.
The dynamics are in the score already
Count the parts in each bar, realise them in their ranges, put every partial in its critical band, sum the loudnesses and run the result through the two smoothers built earlier. What comes out is a dynamic curve for a piece with no performance in it anywhere — and it says that doubling the number of parts inside a fixed register adds three decibels of power and about one of loudness, because the extra parts land in bands that were already occupied. Let the register widen with the parts and the same arithmetic gives eight phon, which is what a tutti actually is.
A final chord is not made loud by adding to it
An earlier essay on closure said the loudest cue an ending has needs a corpus rather than an arithmetic. The arithmetic was built one essay ago, so it does not. Run four ending textures through it and two things come out backwards: a final chord three parts thicker than the rest arrives *quieter* against the running impression than the passage it ends, and a texture that drops a part a bar does not get quieter at all until the bar where there is one part left.
The notes in between
Every figure until now is about two notes with nothing between them, and a melody is notes with other notes between them. Two published accounts of what the intervening material does predict opposite signs — one says the key is restated and the shared reference is refreshed, the other says each note competes for the same memory and it decays. By eight notes they differ by four and a half cents, which is nearly half the limen the interval would have with no anchor at all.
A modulation and a borrowing are one number apart
The key-finder that keeps the order has one tuned parameter, and said so. Sweep it and the parameter turns out to be the whole verdict: below a threshold the model hears two keys alternating, above it one key with borrowed chords. The threshold rises with how slowly the keys alternate — and the value it chose sits above every threshold in range, so its finding about fast alternation was a consequence of the tuning.
A count and a correlation
Cadence evidence is a count of ordered pairs and profile evidence is a correlation with a template, and the cadence essay refused to total them because they are not in the same units. Score each against its own chance and they are — standard deviations above chance are the same unit whatever produced them. Then the trouble moves: the obvious null does nothing at all to one of the two, because shuffling the bars leaves a pitch-class histogram exactly as it was.
Surprise is a number
The chord that did not come was described rather than measured. Its measure is the information content of what did arrive, and a model of the probability has been to hand since the key-finding essays — eight root-motion weights, ordinal and stipulated. Reading them as a distribution prices a deceptive cadence at 2.71 bits against a perfect one's 1.85, and turns up the fact that the largest of the eight had never been read by anything.
The hardest place on the fingerboard
Two things narrow a bow's window and each has been drawn alone. The bridge's admittance is a function of frequency; the bowing fraction is a function of where the left hand is. They meet on a real fingerboard, and multiplying the two curves gives a map with a worst place on it — C♯ on the G string, sixth position, where the body's air resonance crosses the heaviest string played short. The window there is twelve, against three hundred and eighty at the easiest.
The partials that do not die together
The fusion census is a still photograph: it asks whether a set of partials fits one harmonic series and has no term for time. Put time in and the ranking reverses. An ideal string is perfect on harmonicity and holds a fifth of its partials together by decay; a kettledrum — the worst spectrum in this collection for fitting a series — is the only one whose modes all die at one rate.
The quantity a rival account says is not there
Two earlier essays measure how much two notes' errors are correlated through a shared anchor, and price what that correlation would be worth. There is a rival account in which a listener refers each note to a key and never forms the distance at all — under which the correlation is not small, it is a description of something that is not happening. The two accounts agree on almost everything and disagree on one manipulation, and the manipulation costs an afternoon.
A detector whose resolution the performance sets
The boundary detector lost its free parameter when the psychological present became a number of notes at a stated tempo, and what that held still was named at the time: a performance slows into a phrase end, so the number of notes inside the present is not the same everywhere in a tune — it falls exactly where a boundary is. Making the width follow the performance recovers half of what removing the parameter cost, and honestly leaves the other half.
A chord, given a key and a predecessor
A chord's improbability has been priced from its root motion alone, which left one multiplication unmade: a chord is also improbable because its notes do not fit the key, and that number has been available since the probe-tone profile. Multiplied and renormalised, the two give a conditional distribution — and a distribution has an entropy, which is the quantity a surprise has to be read against and which a list of preferences cannot supply.
A cycle that says where it is
Euclidean timelines were asked how quickly they tell a listener where in the cycle they are, and answered it with rotational asymmetry: a symmetric pattern never locates at all. A colotomic cycle answers the same question with nothing asymmetric in it. Several isochronous layers at nested periods — a gong every sixteen, a kempul every eight, a kenong every four — put the position in which instruments sound, and the position is legible from a single stroke.
The exchange rate nobody has
There are now two cues that disagree about how a spectrum divides, and every figure so far reports them separately because there is no principled way here to weigh one against the other. The published apparatus is a competition between grouping hypotheses with a cost per cue — and the useful thing it produces is not the winner but how much of the exchange rate the winner survives.
Expectation is a curve, not a list
Every quantity so far is attached to a chord change: a list of surprises, one per event. A listener's expectation is continuous, sharpening through a bar and collapsing when the change arrives — and the two ingredients for it were already here, in two other accounts. What comes out is a second surprise, for when a chord arrives rather than for which one it is.
A family of readings
Removing the detector's free parameter, and then its constant tempo, cost persistence both times — the property that made its boundaries ordered rather than merely found. Recovering it means a family of adaptive readings rather than one, indexed by the width of the psychological present, which is the one parameter that cannot be removed, because it is a fact about listeners.
How many bars an ensemble needs
The map of required leads is something nobody tells the players, because the leads are a property of the instruments' attacks. So an ensemble has to find them, and the mechanism is the one the microtiming essays describe: each player hears sounds rather than onsets and moves toward the others. It converges on the map to within a fifth of a millisecond, in five beats, and there is a best correction gain.
How much of the reading arrives late
Every margin reported earlier is two-sided: the best path through a key at a bar is the score into it plus the score onward from it, and the second half uses bars a listener has not heard. Dropping that term is one line. On a thirty-two-bar song it removes more than half the certainty, and on a passage built to be ambiguous it changes the key named at nine bars out of eleven.
The pitch that does not wobble
Three earlier essays have treated a vibrato as a modulation of roughness. The reason singers use one is what it does to the note, and there is an extractor here that turns a set of partials into a pitch and has never been asked what it does with partials that will not hold still. The period survives, at a cost that rises with the extent — and the practice stops within a hair of where the cost becomes total.
The dissonance arrives and the dynamic does not
A scoring decides two things at once and both of them have to be integrated by a listener before they exist. The loudness smoother's release is two seconds and the roughness window is thirty-seven milliseconds, and that ratio of fifty decides which of the two survives at the pace music is actually played. Nothing anybody performs is fast enough to blur a dissonance, and a great deal of it is fast enough to average a dynamic.
The cue that settles it
Arbitrating between two grouping cues meant sweeping an exchange rate nobody could supply. The cue it had no term for at all is the one every account calls strongest, and its strength is computable: a struck string's partials start together to within a tenth of a millisecond against a threshold of twenty. Put that into the competition and three of five verdicts change, all the same way — and a bell becomes one sound.
A page has two decibels
The account of closure called its dynamic component a corpus debt: it had no model of dynamics in a form. Loudness supplies one, and applying it answers the debt by refusing it. Adding parts to a final chord one at a time, each at the same level, moves the loudness by under two decibels and not monotonically — while a player has sixty. A texture that thins is not a diminuendo.
The listener who forgets
Setting an analyst's reading of a key against a listener's measures what arrives late. Both passes assume perfect recall of their own half — which is as wrong going forward as knowing the future is going back. Put a decay on the forward pass and a listener with a memory of a few bars keeps a quarter of the confidence and nine tenths of the answers.
Two surprises and one event
A chord change is surprising twice over — in which chord it is, and in when it comes — and a listener meets one event. Adding two surprises needs to know how much they share, which is a fact about a repertoire nobody has. Sweeping it instead: the total moves by half, the ordering does not move at all, and the most surprising moment in a passage is the same one whatever the answer turns out to be.
A reader does not read notes
Eleven earlier essays count notes, and the page cannot tell one line of eight quavers from another. A reader can: a scale of eight is one object where eight leaps are eight. Measured against the melodic interval distribution, the same eight notes are four times as much to read — and the eye–hand span, the best-measured quantity in the reading literature, is four notes of a tune and one of a leaping line.
A rest is a diminuendo
Three parameters swept over factors of ten, three and forty-eight returned the same reading to four significant figures. The one manipulation left unswept moves it by twenty-four decibels: a silence. A listener's running impression decays at the loudness smoother's two-second release, so a general pause is a diminuendo nobody wrote, and at the length that makes a gap an ending it is worth more than any marking a composer has.
A blown note does not start late, it starts slowly
Computing the onset cue removed a free parameter and turned out to be unanimous, and it predicted that a wind instrument would put it back, because a blown note's partials arrive over tens of milliseconds. They do — 121 on a clarinet — and it is not an asynchrony: every partial begins the instant the reed does and they differ in rate, not in time. Read at a tenth of the steady amplitude the spread is 5.5 milliseconds against a threshold of twenty, so the cue is still unanimous, and the missing number is no longer the exchange rate but the criterion.
Eleven partials is one too many
Six earlier essays census exactly ten partials and no figure has ever passed another number. At eleven, the harmonicity census stops finding a perfect harmonic series' own fundamental, takes the octave above it, calls every odd partial inharmonic, and the competition cuts an ideal string in two. It is not the arbitration — the cost of a second stream was swept over a factor of fifty and every verdict came back identical — it is a cap that exists for a good reason and turns out to be the same number as the count.
The cycle that outruns the memory
A timeline and a colotomy were compared at equal strokes and the comparison had no clock in it. A memory span is a number of seconds and a cycle is a number of steps, so the two only meet through a tempo — and at a clave's two seconds a listener's memory covers twenty-eight steps and forgets nothing, while at a gong cycle's forty it covers 1.4 and forgets almost everything. The single line is the better locator up to twenty-three seconds a cycle and the layered code is better after it, which is very close to where each is actually used.
A part entering is not a change of level
Five earlier essays measure a sonority that is already sounding. The one thing an orchestrator actually controls is the entry, and priced at every semitone it comes to 0.89 phons — under the difference limen — with only 29 of 61 available entries clearing it and none below A♭4. What an entry is instead is an object arriving, and the reason is that its own rise time is faster than the listener's.
The bars a key is made of
A discount that treats every bar alike is the wrong shape for a memory, so take each bar away in turn and see what it was worth. The bars an analysis points at carry seven times as much of the answer as the ordinary ones and slightly less of the certainty — and a memory built only of them reads the passage worse than a memory with no structure in it at all.
The surprise of nothing happening
A hazard charges a listener twice — once when the chord changes and once, quietly, at every beat it does not. The second term was expected to dominate, because there are more beats than changes. It is a third of the bill at one chord a bar and never reaches a half at any rate a metre survives, because the cost of a beat where nothing happens is second-order small.
Where the chord actually lands
Every timing surprise so far is evaluated on a downbeat: the curve spans a factor of 6.7 and every number is read at its peak. Charge the waiting as well as the arrival and the costs over a bar become a proper distribution, the spread between the best and worst beat rises to a factor of 19.5, and a third of the probability sits on a bar in which nothing changes at all.
The part of the error a key cannot touch
Four earlier essays turn one dial — the correlation between two notes' pitch errors — and apply it to the whole of a note's limen. Half of that limen is not the listener's: a note of finite length does not carry its frequency more finely than 1/2T, and no context can put information into a signal that is not there. So the correlation has a ceiling, it is 0.21 at a quarter-second note at A4 and 0.07 at A2, and the figure that prices a correlation of one half is drawn where one half is unavailable.
The middle nobody could have guessed
A struck note has no steady state, only a slide from one spectrum to another — so the question is what the middle carries that the ends do not. The answer is exact rather than statistical: every loss law in the family leaves the strike with the same spectrum and ends in the same silence, so both endpoints carry precisely nothing about which of them it is. The whole difference is 41.3 decibels, and it peaks 0.38 seconds in, seven per cent of the way through the note.
The rate that does not rise with the partial
Twelve earlier essays give every vibrato the same six hertz, and the measured spread is 5.5 to 7.5. Putting the two fluctuations a choir contains on one axis shows why the rate matters: the beating between mistuned voices rises with the partial and leaves the range a listener follows as fluctuation at 1,217 hertz, while the vibrato's own modulation is six hertz at every partial. Above that frequency a section fluctuates by vibrato alone — and if every singer had the same rate, it would barely fluctuate at all.
Four parts are easier to read than two
Twelve earlier essays read one line, and a score is several at once. Measured through the voice-leading model, a notehead of a four-part chorale asks a reader for 1.30 bits and a note of an independent line asks 1.89 — so twice the ink is less than three quarters of the load. The reason is a boundary those essays already established: a duet has no complete voicing of any triad at all, so two parts cannot be read from their harmony and have to be read as two melodies.
Where a wrong head gives itself away
Every claim so far maps a delay to a direction through one fixed geometry, and the listener acquires that map while the geometry grows under them by seventy per cent. So the map can be wrong — and the essay before this one said the error would be largest on the median plane, where the delay curve is steepest. It is exactly zero there. The steepness is in the error and in the threshold and cancels between them, which leaves a listener whose internal head is 1.3 millimetres out with one place to catch it: hard to the side, where nobody localises well.
A short note is heard more in tune than it is
Four earlier essays treat a key as something that reduces the noise in a pitch judgement. Treat it instead as a prior and the prediction changes kind: not a smaller error but a systematic bias, pulling a short note toward the nearest scale degree by an amount the Fourier bound sets. Thirty cents out of tune on an eighth-of-a-second note is heard as eight. And the part the debt got wrong is the part that matters — the bias does not vanish on a long note. It stops at 17 per cent at A4 and at 48 per cent at A2, because the likelihood's width has a floor that no duration removes.
A tonic bought with the function
Putting the measured probe-tone profile inside the ordered key-finder is one term, and it does what was predicted: the natural-minor passage is named A minor at every bar instead of C major. It also does two things nobody predicted. It renames a scheme that has been read in the wrong key at every bar since the day it was written, and it destroys the chord's function while it is buying the key.
A fourth decision, and two that were never made
The joint search resolves key, metre and segmentation together and holds the segmentation's cue mixture at zero. Adding the mixture is one loop, and reading the search in order to add it turns up something worse than a missing axis: on the passages it is drawn on, the key it reads is the same key at all forty-eight of its hypotheses and the metre scores every barline identically. The fourth axis then cannot be ranked at all until each reading is measured against its own null, because a mixture changes the ruler and not only the answer.
Eighty-one chords, one number
A dominant seventh has eighty-one arrangements inside three octaves and their roughness spans a factor of five and a half. The tonal-expectation model gives every one of them the same 3.51 bits, because its states are scale degrees and there is no register anywhere in them. Conditioning the surprise on the voicing costs no corpus — and the arithmetic says the conditioning belongs beside the probability rather than inside it, for three reasons that can each be computed.
A twenty-five is a nine until its last unit
Every account of long additive metres says they are heard as groups of shorter ones, and the metre-induction model had never been pointed at the claim. Pointed at it, the model does not prefer the group — it prefers the long bar outright, and would go on preferring it more the longer anybody listened. What it cannot do is start: the evidence that separates a bar of twenty-five from a bar of nine does not exist until the whole bar has been heard, and at the tempo an unequal metre is best played at the psychological present holds sixteen units.
A general pause is spent by the note after it
Whether a composer should write the pause before the diminuendo or after it looks like a question about how big the ensemble is. It is not. Forty decibels of ensemble are worth one decibel of silence, and the order is worth seven — because a running impression rises twenty times faster than it falls, so half a general pause is spent by ninety-seven milliseconds of sound.
Where the note is costs more than which note it is
Thirteen earlier essays measure a page, and the three that price a reader price only its pitches — every line they measure is a run of equal notes. A metrical weight normalised by its own sum is a probability, and minus its logarithm is bits — the same substitution made earlier for intervals. Measured over the tunes used throughout it comes out at 2.23 bits a note against the pitches' 1.89, so the larger half of a reader's load is where the note is.
An interval is two posteriors subtracted
Treating a key as a prior over one note predicts that an interval's pull is not the single-note pull doubled, because the two degrees are not equally weighted. Half of that is wrong: splitting a mistuning between the two notes gives exactly the mean of what each end gives alone, to a thousandth, at every one of the twenty-one intervals in the scale. What is not the mean is which end carries it — and the pull turns out to be largest not on the shortest notes but on notes of about an eighth of a second, where the likelihood is a quarter of a semitone wide.
Which end the mistuning is on
Twenty-one intervals in the major scale, each with two ends, and the same twenty-five cents reaches a listener at anywhere between 32 and 79 per cent of its size depending on which of the two notes carries it. The most lopsided is the tonic to the leading note, where a departure on the upper note arrives two and a half times as strongly as the same departure on the lower. Two mechanisms produce it and they can be separated by one flag: two thirds of the asymmetry is register and one third is the key.
The tone on the root changes hands at the fifth
Every combination tone of a just interval is a harmonic of a fundamental neither note contains, and which harmonic is fixed by the ratio. The difference tone lands on that fundamental for every interval up to the fifth; the cubic product lands on it for the fifth and every interval above except the minor sixth. So the loud product names the root of a narrow interval and the quiet one names the root of a wide one — and a just major seventh's difference tone is a note seven harmonics up that no keyboard has.
A major triad's combination tones are its own notes
Play a just major triad of pure tones and two of the ear's cubic products land exactly on its root and its fifth. The reason is a condition rather than a coincidence — a chord's cubic products fall on its own notes when its middle note is the mean of the outer two in hertz — and it holds for the major triad in root position and in the six-four, and for no minor triad in any position or tuning. Equal temperament misses the landing by one number, 5.6 hertz on middle C, which is a beat that belongs to no pair of notes in the chord.
The bass line under a passage in thirds
A major scale harmonised in parallel thirds gives the ear a difference tone under every pair, and in five-limit just intonation those tones are a diatonic bass line — C, A, C, F, G, F, G, C — made of the scale's own notes. Tempered, the same line moves only by whole tones, a neutral third and a fourth stretched to 650 cents, and wobbles by up to 84 cents from note to note. In sixths the bass is drawn by the other product, because the cubic product of a pair is the difference tone of the same pair inverted.
Two cues meet in a corner
The profile finds a key's tonic and the bass finds a chord's degree, and until now each was swept with the other held at nothing. Swept together across 42 settings, the plane they make is not the ridge that was predicted. The tonic is a step in one direction, at a profile share of 0.55, and the bass cannot move it; the degree is a slope in the other, rising to 89 per cent as the bass is weighted, and the profile barely touches it. Every question is answered only in a corner of the plane — and the one place the two cues overlap is the one piece of music both can rescue.
A bass line is not a list of roots
Every bass note the key-finder has been given was its chord's root, and under that line a bass cue reads 89 per cent of scheme bars on the right degree. Give the same chords an economical bass that moves to the nearest chord tone, as a keyboard reduction would, and nearly half of them are inverted. The cue that rewards the triad rooted on the bass then reads 68 per cent at best and worse as it is trusted more; the cue that rewards any triad containing the bass cannot be fooled and stops at 61. The same inverted line does one thing the roots never did: it puts the leading note of each new key at the bottom, and finds the rondo's modulations.
The chords never move the barline
Every hypothesis the joint search had drawn was one bar long, and on one bar with a note in every slot the metre cannot choose a barline at all. Four bars with rests in them make the barline a decision the metre and the chords both have an opinion about, and the prediction was that the chords would move the barline more often than the barline moves the chords. It is the other way round, completely: whenever the two prefer different barlines the search takes the metre's, on up to 72 per cent of passages, and in fifteen hundred passages the chords never once move it. What the chords decide is the one thing the metre cannot see — whether the bar starts on the downbeat or half a bar later — and they decide it right a little over two times in three at best.
The chords are a weak witness to the barline
Scaled by its own range, the metre overrules the chords every time the two disagree about where a bar begins. The obvious repair is to score each reading against its own chance — the metre against the same number of notes placed at random, the chords against the passage's notes shuffled across its bars — and add the standard scores. It changes very little: the search finds the barline within six points of where the product found it, and the chords gain the power to move the barline only on passages whose rhythm says nothing, where random notes move it nearly as often. The null's real result is the size of the two witnesses. At the written barline the metre stands up to 5.9 standard deviations above chance, and the chords, with every note a tone of its bar's chord, stand 1.55 above it at best.
The accent buys two units, however loud it is
A twenty-five cannot be told from a group of shorter bars on its onsets until more steps have gone by than a listener's present holds, and the obvious objection is that nobody plays an aksak bar as bare onsets: the long beat is louder, and the bar's first beat is marked. So how loud does an accent have to be? The question has a surprising answer. Loudness is not the variable. The existing accent cue changes nothing, and delays the answer where it changes anything. An accent that a reading has to predict works at any strength at all, and at no strength does more than a fixed amount: on the long beat it buys the two steps of a short beat, and on the downbeat it takes every arrangement to one floor — the bar less its last beat — which no cue carried by the notes can break. The longest bar that can be heard as one moves from sixteen units to eighteen.
A proportion is only as fine as its two durations
Analyses of form measure proportions in bars and report them to three figures — a climax at 0.618, a section in the ratio 3 : 2. A listener has each part only as an estimate of how long it lasted, and a ratio of two estimates is blurred by both. Timed as well as anyone times a single second, eleven proportions fit between 1 : 1 and 3 : 1; timed from memory over minutes, two do. The golden section is told from 3 : 2 only below a Weber fraction of 5.4 per cent.
A return is shorter than its first hearing
A rondo's refrain takes three fifths of the clock and a verse-and-chorus song is balanced to the bar. Count instead the bars a listener could not have predicted when they arrived, and the returns shrink to between a tenth and a quarter of what is kept — so a song equal by the clock is between three and seven times heavier in its first half. A coder that learns repeats one bar at a time says the halves are equal, and the two memories disagree by more than any proportion a listener could confuse.
A final chord stands out for a twentieth of a second
A general pause drives a listener's running impression down, and the final chord that follows is supposed to cash the fall in. It cashes in at most two thirds of it. The note's own loudness rises with a 22-millisecond constant and the impression with a 99-millisecond one, so after a bar of silence the chord stands furthest above the impression 54 milliseconds in, by 5.1 of the 8.1 decibels the silence bought, and after 206 milliseconds the two are within a phon of each other. A short stamp spends most of its life standing out; a chord held a second and a half spends a seventh of it.
The chords mark the barline by changing there
Read bar by bar, the chords stood barely above chance at the barline and broke the metre's half-bar tie two times in three at best. Read instead by where they change — how different the chords are across a candidate's barlines against how different they are across the middle of its bars — the same notes break the tie right on 81 to 96 per cent of passages, and added to the metre they find the barline on up to 89 per cent against 61. The weakness was the question the old reading asked, not the harmony.
A rough arrival is rough because of its spacing
The pair the expectation essays report for every chord — how surprising it was, how rough its voicing is — has no level in it. Putting level back in answers the question it left open, and not the way it was framed. At one written dynamic the arrivals keep their order from 40 to 90 dB at three registers of four, and the bass stays 8.6 times rougher than the treble. Made equally loud, the bass has to be played 12.8 dB harder, and it is 343 times rougher: level does not explain the register's roughness away, it multiplies it.
A combination-tone bass needs a forte
A scale in just thirds draws a diatonic bass line through its difference tones, and in sixths the cubic product draws one. Given the two published level laws, with their constants swept, the thirds' bass is not heard at all below primaries of about 66 dB and is heard whole only from 71 to 81. The cubic products are a different kind of object: the primaries mask them decibel for decibel as they rise, so no dynamic changes whether they are heard. Most of the thirds' inner line never is, and the sixths' bass needs a forte and a gentle law.
Knowing every metre is slower than knowing none
A long aksak bar cannot be told from its shorter cuts by induction before one step into its last beat, and no accent carried by the notes moves that floor. The obvious escape is a listener who knows the repertoire and recognises the metre instead. Recognition among all 1,820 arrangements of twos and threes never beats the floor, is never quicker than induction, and is slower for half the metres: a nine induced in 9 steps is recognised in 27. What breaks the floor is a small repertoire that leaves out the metre's own longest cut — with the cut known, no repertoire of any size does.
The bell pattern is slowest only to a perfect memory
Among the orders of its own gaps, a named timeline is usually both the most even and the slowest to locate — for a listener who never forgets. Give the listener a memory that halves and the result comes apart. Of six timelines slowest among their orders with perfect memory, only the fume-fume stays slowest for every forgetting listener, and the standard bell pattern, which is the fume-fume with onsets and rests exchanged and settles at exactly the same floors, is second of its three orders for every memory of half its cycle or less. The census ranking survives better, and in fourteen of twenty-one censuses it was the arithmetic of a pattern that repeats.
A dancer who comes in late needs the downbeat marked
Every window for recognising an aksak metre so far started at its written downbeat. A dancer joining a dance already going has not heard the downbeat, and the arithmetic of that is blunt: a metre entered part-way is, onset for onset, each of its own rotations heard from their downbeats, and the rotations are metres too — 2+2+3 and 3+2+2 are counted differently. So on onsets, and with the long beats accented, no metre is ever told from its rotations. Only an accented downbeat tells them apart, and with it a listener who knows thirty metres recognises 54 per cent of them inside the present from a random entry, against 1 per cent without.
Three harmonics of the bass arrive before the bass
Every product priced until now was between two pure tones, and nothing that plays thirds is pure. Give each note a spectrum and the ear receives every pair of partials — and for a just interval p:q every one of their products is an exact multiple of the same absent fundamental. That crowd lands where the threshold of hearing is tens of decibels cheaper, so it names the bass at 71 decibels where the component at the bass's own frequency needs 74, and at 75 against 85 an octave lower. Tempered, the crowd still forms and names a note seventy cents flat.
The ghost bass drops a twelfth at a forte
Both crowds arrive at once and every member of both is a multiple of the same absent fundamental, so a listener is never given a choice between them — only a different subset of one harmonic series at every dynamic. Softly, the subset is an exact gapless series on three times the fundamental. Loudly, the difference tones fill in the low harmonics and no template on the higher note survives them. Between 62 and 72 decibels, depending on the interval, the note the crowd names falls by an octave or a twelfth, and the two qualities of third cross at different levels.
A count is not an estimate
Both established routes to a proportion are estimates — a duration timed, blurred by a Weber fraction, and a duration stored, biased by what was new. A listener who has induced a hypermetre has a third, and it is exact until it fails. It fails two ways that pull opposite: a slip miscounts one unit and its relative cost falls as the section lengthens, while a lapse loses the count entirely and its chance compounds. The mixture has a floor at about four units, where counting is 3.7 times finer than timing, and it is worth almost nothing past a hundred.
Given the bar in octaves, the degree comes back
A bass note is a bare pitch class in the usual figure, and a register in any realisation anybody plays. Voice each bar in octaves and a listener knows which three pitch classes are sounding and which is at the bottom, which names the chord outright — and the rule that uses it reads the right scale degree in 92 per cent of bars on a real bass line, against 68 for the published cue on the same line and 89 for that cue on a line made entirely of roots. The question was whether the register recovers the 89. It recovers it and passes it.
The error that moves straight ahead
The essay before this one found that a listener whose internal head is the wrong size makes no error at all on the median plane, and has to look hard to the side to catch it. Every head drawn here has its ears at equal radii, which makes the delay curve odd and every error a factor — and a factor cannot move a zero. Real heads are not symmetric. A constant offset of twenty microseconds displaces a listener's straight ahead by two and a quarter degrees, and it displaces every other direction by the same number of just-noticeable steps, exactly.
An exit is worth nothing until the tutti is given up
Six parts entering a five-chord passage have a best schedule when each enters once and stays, and letting parts leave and come back was supposed to improve it. Searched over every set of sounding parts at every chord, it improves it by exactly nothing, with or without the chord before still masking — as long as all six must be playing at the end. Let one part be missing from the final chord and the weakest entrance gains 1.4 decibels; let two be missing and it gains 1.9, by a relay in which the parts with least room come in, are heard for one chord, and give way.
Checkpoints sharpen the middle of a form, not its top
A listener who counts bars loses the count somewhere in a long section and is thrown back on timing the whole of it. A listener who also counts phrases can mend the lapse at the last phrase. If the two counts fail independently, the level a minute long goes from four distinguishable proportions to nineteen; the whole eight-minute piece goes only from two to four, because thirty phrases are long enough to lose a count as well. And if a fifth of lapses take both counts at once, three quarters of the gain is gone.
The bass errs fast where the content errs slow
Asked how often the chords change, a reading built on pitch-class content names a slower multiple and never a faster rate. Give the passage a bass that states each new root and moves between chord tones inside a chord, and a reading built on the bass's moves errs the other way: it names a faster grid and never a slower one. At two chords a bar the bass is right every time; at a chord every two bars it is never right. Six ways of combining the two readings each trade one end of the range for the other.
The played notes already name the ghost bass
The products of a just third's partials, fitted on their own, name a note a twelfth above the bass when the interval is soft and drop to the bass when it is loud. Put the two played notes back beside them and the drop disappears: the notes and their products name the bass at every dynamic, because the notes' own partials are harmonics of it already. What the dynamic changes is not which note is implied but how complete its harmonic series is — nine holes from the notes alone, four when soft, none when loud.
The ceiling is thirteen bars with names
The key-finder's two cues stop buying anything at about 93 per cent of scheme bars read right, and the obvious suspect was the chord segmentation feeding it. The reading was never given a segmentation: every bar arrives with its true chord. What the ceiling is made of can be listed instead, and it is thirteen bars of 188 — seven in a bridge of secondary dominants, four in a minor episode whose chords C major also owns, and two at the edges of a modulation. No weight of any cue moves one of them.
A late dancer needs the landmarks, not the rhythm
A listener who joins an additive-metre dance part-way recognises it far more reliably with the downbeat accented. Strip the stream down to its landmarks — the onsets that begin a bar or a long beat, with every other onset removed — and the listener does as well or better: knowing a hundred metres, 23 per cent are recognised within the present against 14 with every onset. Unmarked, the landmarks still beat every onset with the long beats accented. Neither half does it alone; what identifies a metre from a late entry is where its long beats sit relative to its bar.
Against a pulse the bell pattern is the easiest to place
Heard alone, the standard bell pattern is not the quickest order of its own gaps to place in its cycle, for a listener with any memory. Heard against a pulse every three steps or every four — which is how anyone hears it — it is the quickest, at every memory and at every alignment of pulse and bell, and by a wide margin: at a memory of a quarter of the cycle, 0.56 bits unplaced over the first cycle against 0.73 for either rival against a pulse in threes. Six of eight named timelines do the same. A timeline's order of gaps looks chosen for how it sits against the beat, not for how it sounds alone.
A bass chord low enough to balance has already hidden its tenor
Played as loud as the written register, a progression two octaves down is 343 times rougher than the same progression an octave up — if every partial on the page is counted. Count only the partials that stand above what the rest of the chord masks and that register is the smoothest of the four, with nothing left that beats. The balance is not what does it: the extra thirteen decibels move no voice by more than two partials. The register had already buried the tenor at the written dynamic.
A bass that holds through a change marks the barline
At two chords a bar the chords change on the barline and on the half-bar alike, so a reading of where they change is at chance, and the metre ties the two. The bass has one more piece of evidence: a change inside the bar can be voiced over the note already sounding, and a change on the barline is voiced over its root. Hold the bass through one mid-bar change in six and the metre and bass together place the barline in 85 per cent of eight-bar passages; one in three, 97. The convention cannot be stronger than that, and the chord changes, asked first, only get in the way.