Concept

Memory decay — where it appears

The falling away of a representation with time, which decides what of a passage is available when a later part of it arrives. It sets what a form can be built out of, because a return can only be heard as a return if the first statement is still there.

Named by 15 essays across 6 fields — each of them below, with the objects they name alongside it.

The ranking is settled either side of one narrow band. Remembered repetition — each bar's best match to an earlier bar, discounted by exp(−Δt/τ) with Δt in seconds — for 6 schemes at 108 beats a minute, against the decay constant τ on a logarithmic axis. The order of the schemes changes only between 8 and 13 seconds; outside that band it is fixed, so an estimate of τ wrong by any amount that stays outside it leaves the ranking alone.

A return has to be remembered

A stripe four bars off the diagonal and a stripe twenty-four bars off it are the same ink and are not the same experience. Convert the lag axis to seconds, discount every comparison by how long ago it was, and the ranking of these six schemes by how repetitive they are changes — and the decay constant and the tempo turn out to enter the arithmetic as one number rather than two.

perception · Repetition
What a contour costs to remember. A melody of n notes over 8 degrees carries 3 bits a note. Its contour carries fewer, and fewer than the number of distinct contours suggests, because the contours are not equally likely: at 6 notes there are 243 of them but the entropy is 6.59 bits, an effective alphabet of 96. Each further note adds 1.28 bits of contour against three of melody, so the shape keeps a stable 37 per cent of what is there however long the tune.

The part of the tune that is kept

Contour survives transposition, retuning, a change of instrument and a doubling of every interval, and the usual explanation is that it is what a listener retains. That can be counted rather than assumed. A six-note melody over eight degrees carries eighteen bits; its contour carries 6.59 — not the 7.92 the number of distinct shapes suggests, because the shapes are wildly unequal — and the effective alphabet is ninety-six out of two hundred and forty-three. Each further note adds 1.28 bits of shape against three of melody, and at about nine notes a contour is specific enough to pick one tune out of a thousand.

perception · Melody
Crescendo, and what the impression does. A crescendo of 20 dB over 8 seconds, drawn as three loudnesses in sones. The pale line is what is physically sounding, the middle line is the short-term loudness of the moment and the heavy line is the long-term loudness, which is the passage's loudness as a listener would report it. The gap between the last two is the whole of the effect: at its widest the moment is 1.02 times the running impression, and the impression takes 2 seconds to come down against 99 milliseconds to go up.

Loud is relative, and it comes down slowly

The account of loudness had a model of a moment and the account of closure asked it for a model of a form. The published one exists and its content is a pair of numbers that are not the same: a listener's running impression of how loud the music is rises to meet a step in a fifth of a second and takes seven seconds to come back down. A twenty-decibel crescendo spread over eight seconds therefore buys almost no contrast at all, and the same twenty decibels taken as a step buys a factor of two.

form · Loudness
What the notes in between do to the anchor the interval is measured against. How finely a 7-semitone interval can be judged when its two notes are separated by other notes rather than by silence, under the two published accounts. Confirming material restates the key and refreshes the shared reference, so the correlation climbs from 0.5 toward a ceiling and the limen falls to 4.81 cents. Overwriting material competes for the same memory, so the correlation decays to 0.04 and the limen rises to 9.28. By 8 notes the two accounts differ by 4.5 cents, which is 47 per cent of the limen with no anchor at all — and no experiment here distinguishes them.

The notes in between

Every figure until now is about two notes with nothing between them, and a melody is notes with other notes between them. Two published accounts of what the intervening material does predict opposite signs — one says the key is restated and the shared reference is refreshed, the other says each note competes for the same memory and it decays. By eight notes they differ by four and a half cents, which is nearly half the limen the interval would have with no anchor at all.

intervals · Pitch-acuity
A listener has a quarter of an analyst's confidence and the same answer. The mean margin between a passage's best two key readings, against how many bars a listener's memory of the evidence takes to halve. The two-sided reading — the one that uses bars that have not happened yet — sits at 3.90; the forward pass with perfect recall at 1.79; a forward pass whose evidence halves every 6.6 bars at 1.00. The dots' size is how often that reading names the same key as the two-sided one: 100 per cent at perfect recall and 81 at a one-bar half-life. So forgetting costs a great deal of confidence and very little accuracy — the key is robust and the certainty is not.

The listener who forgets

Setting an analyst's reading of a key against a listener's measures what arrives late. Both passes assume perfect recall of their own half — which is as wrong going forward as knowing the future is going back. Put a decay on the forward pass and a listener with a memory of a few bars keeps a quarter of the confidence and nine tenths of the answers.

harmony · Key-relations
The same eight notes are four times as much to read. How many bits each note of a line carries, taken as minus the log of the probability of the interval that reached it, under the distribution of melodic steps measured over the tunes used throughout. A scale costs 1.76 bits a note and a wide leaps costs 7.02 — a factor of 4.0 at the same number of notes on the page. Every quantity computed until now counts notes, and the page cannot tell these apart: eight quavers are eight quavers of horizontal space whichever line they spell.

A reader does not read notes

Eleven earlier essays count notes, and the page cannot tell one line of eight quavers from another. A reader can: a scale of eight is one object where eight leaps are eight. Measured against the melodic interval distribution, the same eight notes are four times as much to read — and the eye–hand span, the best-measured quantity in the reading literature, is four notes of a tune and one of a leaping line.

scales · Notation
Where a cycle of 16 at 16, 8, 4 outruns the listener's memory. The residual uncertainty a listener is left with once the evidence has stopped accumulating, against how long one turn of a 16-step cycle takes. The listener's memory of a step halves after 3.5 seconds throughout; what changes is how many steps that is. At a cycle of 1.6 seconds it is 35 steps and every design reaches certainty, which is the regime a clave is played in. At 60 seconds it is 0.93 steps and none of them does: layers at 16, 8, 4 settles at 1.89 bits, son clave settles at 2.27 bits, the bossa-nova pattern settles at 2.38 bits, the best single line of 7 settles at 1.85 bits. That is the range a gong cycle occupies, and it is the design that wins there.

The cycle that outruns the memory

A timeline and a colotomy were compared at equal strokes and the comparison had no clock in it. A memory span is a number of seconds and a cycle is a number of steps, so the two only meet through a tempo — and at a clave's two seconds a listener's memory covers twenty-eight steps and forgets nothing, while at a gong cycle's forty it covers 1.4 and forgets almost everything. The single line is the better locator up to twenty-three seconds a cycle and the layered code is better after it, which is very close to where each is actually used.

rhythm · Cyclic rhythm
Which bars the key is decided by, and which bars it is believed on. Every bar of a 32-bar scheme removed in turn, with what its absence costs. The column is how far the passage's mean margin falls without that bar — how much of the model's certainty it supplies. The dot is how many bars are then read as a different key — how much of the answer it supplies. The 18 bars an analysis would point at — a section opening or closing, a dominant, the chord a dominant resolves to — average 0.079 bits of certainty and 0.50 bars moved; the 14 ordinary bars average 0.047 and 0.07. So the structural bars carry 1.7 times as much certainty as the ordinary ones, and 7.0 times of the answer. Those are different quantities, and the second is the one an analysis is about: an ordinary bar can carry a great deal of a passage's certainty and none of its reading.

The bars a key is made of

A discount that treats every bar alike is the wrong shape for a memory, so take each bar away in turn and see what it was worth. The bars an analysis points at carry seven times as much of the answer as the ordinary ones and slightly less of the certainty — and a memory built only of them reads the passage worse than a memory with no structure in it at all.

harmony · Key-relations
The same forms by the clock and by what is stored. Six forms, each drawn twice: its sections sized by their share of the bars, and sized by their share of what a listener has to store when a bar counts only if it is recognised from 1 bar of context. Returns are drawn pale with a dashed edge. twelve-bar blues: returns take 67 per cent of the clock and 13 per cent of the storage; thirty-two-bar AABA: returns take 25 per cent of the clock and 5 per cent of the storage; rondo, ABACA: returns take 40 per cent of the clock and 10 per cent of the storage; verse and chorus: returns take 50 per cent of the clock and 13 per cent of the storage; two eight-bar phrases: returns take 0 per cent of the clock and 0 per cent of the storage; a four-bar ostinato: returns take 88 per cent of the clock and 0 per cent of the storage.

A return is shorter than its first hearing

A rondo's refrain takes three fifths of the clock and a verse-and-chorus song is balanced to the bar. Count instead the bars a listener could not have predicted when they arrived, and the returns shrink to between a tenth and a quarter of what is kept — so a song equal by the clock is between three and seven times heavier in its first half. A coder that learns repeats one bar at a time says the halves are equal, and the two memories disagree by more than any proportion a listener could confuse.

form · Proportion
With a fading memory the slowest order of the gaps 1 1 2 2 2 2 2 is not the one a perfect memory finds. For each of the 3 cyclic orders of the gaps 1, 1, 2, 2, 2, 2, 2 in 12 steps, the bits of position still unknown once listening has settled, against how many steps a listener's memory of a step takes to halve, with a mismatch costing 6. 2 2 2 1 2 1 2: 12 → 4e-9, 6 → 4e-5, 4 → 0.007, 3 → 0.070, 2 → 0.540, 1.5 → 1.130, 1 → 1.786; 2 2 2 2 1 1 2: 12 → 2e-5, 6 → 0.035, 4 → 0.335, 3 → 0.795, 2 → 1.405, 1.5 → 1.694, 1 → 2.005; 2 2 1 2 2 1 2 (the standard bell pattern): 12 → 2e-5, 6 → 0.027, 4 → 0.231, 3 → 0.535, 2 → 1.028, 1.5 → 1.368, 1 → 1.819. With perfect memory the slowest to locate is the standard bell pattern; at a half-life of 6 steps the highest floor is the order 2 2 2 2 1 1 2, and at 1 it is the order 2 2 2 2 1 1 2.

The bell pattern is slowest only to a perfect memory

Among the orders of its own gaps, a named timeline is usually both the most even and the slowest to locate — for a listener who never forgets. Give the listener a memory that halves and the result comes apart. Of six timelines slowest among their orders with perfect memory, only the fume-fume stays slowest for every forgetting listener, and the standard bell pattern, which is the fume-fume with onsets and rests exchanged and settles at exactly the same floors, is second of its three orders for every memory of half its cycle or less. The census ranking survives better, and in fourteen of twenty-one censuses it was the arithmetic of a pattern that repeats.

rhythm · Euclidean rhythm
A count is least reliable at both ends and best in the middle. How precisely a listener knows the length of a section they are counting, against how many units long it is, at three kinds of timing judgement. A slip — one unit miscounted, at 2% a unit — accumulates as a random walk, so its relative cost FALLS as the section lengthens. A lapse — the count lost altogether, at 1% a unit — compounds, so the chance of still having the count falls geometrically and a long enough section is certain to lose it. A listener who has lost the count is back to timing, so the two failures mix into a floor. Against a Weber fraction of 7.5% the count is worth most at 21 units, where it is 1.7 times finer than timing, and falls back under a quarter better by 98 units; Against a Weber fraction of 15% the count is worth most at 10 units, where it is 2.4 times finer than timing, and falls back under a quarter better by 101 units; Against a Weber fraction of 38% the count is worth most at 4 units, where it is 3.7 times finer than timing, and falls back under a quarter better by 102 units. The length at which it stops being worth much is nearly the same in all three, because it is set by the lapse rate alone.

A count is not an estimate

Both established routes to a proportion are estimates — a duration timed, blurred by a Weber fraction, and a duration stored, biased by what was new. A listener who has induced a hypermetre has a third, and it is exact until it fails. It fails two ways that pull opposite: a slip miscounts one unit and its relative cost falls as the section lengthens, while a lapse loses the count entirely and its chance compounds. The mixture has a floor at about four units, where counting is 3.7 times finer than timing, and it is worth almost nothing past a hundred.

form · Proportion
A cycle already known, against a cycle just arrived at. How many bits of uncertainty about position a listener has, against how long the cycle takes, for a single timeline and for a layered colotomy — each drawn twice, once as a listener arriving and once as a listener who has been hearing it long enough to settle. At 1.6 seconds a cycle the timeline goes 0.94 bits arriving and 0.00 settled, and the colotomy 0.71 and 0.00; At 16 seconds a cycle the timeline goes 1.10 bits arriving and 0.08 settled, and the colotomy 0.93 and 0.23; At 60 seconds a cycle the timeline goes 2.47 bits arriving and 2.27 settled, and the colotomy 2.14 and 1.89. The gap between each pair is what the repetitions are worth, and it narrows as the cycle slows. The two designs are drawn at their own step counts rather than at equal strokes, so the levels here are not the earlier ones and the gaps are.

Repetition buys least where it is needed most

Every locating figure so far is a listener arriving — the uncertainty averaged over the first cycle heard. Cyclic music comes round dozens of times, and the same model already carries the answer for a listener who has settled: a floor of uncertainty that nothing had read. At two seconds a cycle the repetitions close the whole gap. At sixty they close eight per cent for a single timeline and twelve for a layered code. A slow cycle is worse on the first hearing and gains less from the second, and the two disadvantages compound.

rhythm · Cyclic rhythm
Checkpoints sharpen the middle of a form and leave its top vague. A piece of 480 seconds divided 7 times, with how many proportions between 1 : 1 and 3 : 1 a listener tells apart at each level: timed, counted in 2-second units, and counted with a second count of 16-second phrases that can mend a lapse in the first. 480 s: 2.1 timed, 2.2 counted, 3.9 with 0 per cent of lapses shared and 2.5 with 50 per cent of lapses shared; 240 s: 2.1 timed, 2.5 counted, 7.1 with 0 per cent of lapses shared and 3.1 with 50 per cent of lapses shared; 120 s: 2.1 timed, 3.1 counted, 13.2 with 0 per cent of lapses shared and 4.0 with 50 per cent of lapses shared; 60 s: 2.1 timed, 4.1 counted, 19.4 with 0 per cent of lapses shared and 5.4 with 50 per cent of lapses shared; 30 s: 2.1 timed, 5.4 counted, 18.1 with 0 per cent of lapses shared and 7.2 with 50 per cent of lapses shared; 15 s: 5.2 timed, 12.2 counted, 12.2 with 0 per cent of lapses shared and 12.2 with 50 per cent of lapses shared; 7.5 s: 5.2 timed, 10.3 counted, 10.3 with 0 per cent of lapses shared and 10.3 with 50 per cent of lapses shared. With the two counts failing independently, the level of 60 seconds goes from 4.1 to 19.4, and the whole piece only from 2.2 to 3.9.

Checkpoints sharpen the middle of a form, not its top

A listener who counts bars loses the count somewhere in a long section and is thrown back on timing the whole of it. A listener who also counts phrases can mend the lapse at the last phrase. If the two counts fail independently, the level a minute long goes from four distinguishable proportions to nineteen; the whole eight-minute piece goes only from two to four, because thirty phrases are long enough to lose a count as well. And if a fifth of lapses take both counts at once, three quarters of the gain is gone.

form · Proportion
Against a pulse, the bell pattern is the quickest of its orders to place. The bits of position a listener is still missing, averaged over the first cycle heard, for each cyclic order of the gaps 1 1 2 2 2 2 2 in 12 steps, heard alone, against a pulse every three steps and against a pulse every four, at perfect memory, half-life 3 steps, half-life 1.5 steps. 2 2 2 1 2 1 2: alone 1.08, 1.23, 1.73; against a pulse every 3 steps 0.69, 0.75, 0.96; against a pulse every 4 steps 0.63, 0.67, 0.85. 2 2 2 2 1 1 2: alone 1.22, 1.63, 2.02; against a pulse every 3 steps 0.60, 0.73, 0.92; against a pulse every 4 steps 0.83, 1.07, 1.36. 2 2 1 2 2 1 2 (the standard bell pattern): alone 1.25, 1.45, 1.82; against a pulse every 3 steps 0.54, 0.56, 0.67; against a pulse every 4 steps 0.55, 0.57, 0.68. Alone, the bell pattern is not the quickest order to place at any memory. Against either pulse it is the quickest at every memory.

Against a pulse the bell pattern is the easiest to place

Heard alone, the standard bell pattern is not the quickest order of its own gaps to place in its cycle, for a listener with any memory. Heard against a pulse every three steps or every four — which is how anyone hears it — it is the quickest, at every memory and at every alignment of pulse and bell, and by a wide margin: at a memory of a quarter of the cycle, 0.56 bits unplaced over the first cycle against 0.73 for either rival against a pulse in threes. Six of eight named timelines do the same. A timeline's order of gaps looks chosen for how it sits against the beat, not for how it sounds alone.

rhythm · Euclidean rhythm
A timed expectation would erase a slow cycle's cost, and a listener cannot time a slow cycle that well. Bits of position a listener with a 3.5-second memory is still missing over the first cycle of son clave, against how long the cycle takes, for a newcomer with no expectation, a listener timing the cycle with the Weber fraction a duration that long is judged with, and a listener timing it to ten per cent. a newcomer, no expectation: 2 s 0.94, 8 s 0.97, 24 s 1.41, 40 s 2.02, 60 s 2.47; timing as well as listeners do: 2 s 0.68 (w 0.150), 8 s 0.69 (w 0.150), 24 s 0.85 (w 0.150), 40 s 1.90 (w 0.375), 60 s 2.38 (w 0.375); timing the cycle to ten per cent: 2 s 0.47, 8 s 0.48, 24 s 0.55, 40 s 0.73, 60 s 1.02. At ten per cent even a sixty-second cycle is placed about as well as a newcomer places a two-second one. At the precision a listener actually has for durations of half a minute or more, the expectation is worth a tenth of a bit.

An expectation cannot rescue a cycle too slow to time

A listener who knows a piece arrives with an expectation of where in the cycle they are, and the size of that expectation was the number the last essay said nobody had. It can be given one: a listener who has been timing the cycle carries a spread of their Weber fraction times the cycle, which is the same number of steps at any tempo. Timed to ten per cent, a forty-second cycle would be placed better than a newcomer places a two-second one. But forty seconds is judged in the band where the Weber fraction is nearer forty per cent, and there the expectation is worth a tenth of a bit.

rhythm · Cyclic rhythm

Named alongside it

The objects these essays reach for when they reach for this one.

InformationMusical formDurationExpectationProportionRepetitionRotationTimelineWeber fractionCycleDescription lengthEntrainment

All concepts