Concept

Duration — where it appears

How long a note lasts, treated as a quantity in its own right rather than as a fraction of a bar. Performed durations fall into categories with boundaries, in the same way pitch does, and the categories are not exactly the notated ratios.

Named by 19 essays across 5 fields — each of them below, with the objects they name alongside it.

The price of a tonic. Every note of Dorian is given the same duration except its tonic, which is lengthened; the horizontal axis is the share of the total that goes to it. The key-finder answers with the parent key until 25.0 per cent of the time is spent on the modal tonic, and with D minor above it. At the left-hand edge every note has equal weight, which is the pitch-class set itself — and with every weight identical the correlation is not merely low but undefined, because a flat histogram has no variance to correlate with anything.

What a tonic costs in seconds

The standard key-finding algorithm cannot be run on a pitch-class set at all — a flat histogram has no variance and the correlation is undefined. Give it durations and it answers with the parent key for all seven modes identically, and it takes between 15.8 and 30.0 per cent of the total time spent on one note before it names that note instead.

perception · Modes
Three ways to arrive at the same final tempo. Tempo against position in the closing passage, ending at 35 per cent of the opening tempo, for curvature exponents 1, 2, 3. All three begin and end at the same tempo, so what separates them is the middle: at the halfway point they read 68 per cent for linear in score position, 75 per cent for constant deceleration, 80 per cent for q = 3. The straight line is the one nobody plays. Measured ritardandos fit the decelerating curves, which is the whole of Kronman and Sundberg's argument: a closing gesture has the shape of a body stopping rather than of a dial being turned, and the parameter that varies between performances is the final tempo rather than the shape.

An ending is a deceleration

Every performance slows down at the end and the slowing has a shape. Tempo read against score position is the velocity of a body stopping — a square root rather than a straight line — and the three candidate curves agree at both ends by construction, so the whole audible difference is in the middle, where they part by fifteen per cent of the passage's length.

form · Closure
The same steps, counted by the clock. The step distribution of Ode to Joy and Twinkle, twinkle counted two ways: once per interval, which is what every earlier figure did, and once weighted by how long the note it leaves is held. The two disagree because a tune's long notes are not distributed evenly over its interval sizes — in Ode to Joy the 2-semitone step is 55.2 per cent of the moves and 60.0 per cent of the time. Which of the two a claim about melodic motion means has never been stated here, and the answer matters most exactly where a tune slows down, which is at the ends of its phrases.

The note that has a length

Every melodic figure so far reads a table where each note is a pair — a pitch and a duration — and throws the second number away. A step between two minims and a step between two quavers have been one event in every histogram it has drawn. Weighting the same statistics by time moves the step distribution by up to seven points, changes forty-four of a hundred and one contour signs, and turns up an off-by-one in the one figure that did use the durations: it took the length of the note arrived at where the time between two onsets is the length of the note left.

rhythm · Melody
note length against the onsets. Every candidate metre's fit to a 16-step pattern with 4 onsets, plotted against how strongly note length is weighted. At a gain of zero the scoring is the onset-only one every earlier model used, and the winner is step 1 and step 4. At a gain of 0.05 the answer becomes step 4. 2 candidates are exactly flat — step 2 and step 3 have no onset on any strong position, so there is no credit for the cue to multiply and no weighting of it can move the line.

What the onsets left out

Eight essays induce a metre from a list of ones and zeros, and every failure they recorded was argued about as a failure of the rules. Two of the three are not. Note length is already in that list and the scoring throws it away: put it back and the son clave's two-way tie resolves to the notated downbeat. But the groove with its beat removed cannot be repaired by any cue at any strength, and the reason is arithmetic rather than empirical — the true phase has no onset on any of its strong positions, so there is no credit for a cue to multiply and its line is exactly flat.

form · Metre induction
How long a note has to be before its pitch is worth arguing about. The smallest audible frequency difference at 440 Hz, against how long the note lasts. The flat line is the steady-tone difference limen of 4.0 cents that every tuning argument on this site rests on. The falling line is the bound a finite duration imposes on its own frequency, 1/2T in cents, which no listener can beat. They cross at 486 milliseconds: below that the note is the limit and above it the listener is. A tenth of a second gives 19.6 cents and a quarter gives 7.9, against the commas drawn across the figure.

How long a note has to be

Every difference limen quoted so far is for a tone that lasts as long as the listener needs, and no note in music does. A tone of duration T occupies a band about 1/2T wide whatever the ear does with it, so at 440 hertz the quoted five-cent limen is the right number only for notes longer than 486 milliseconds. A tenth of a second gives 19.6 cents, which does not clear the syntonic comma. Most of the tuning arguments in this collection are about a quantity that only exists in long notes, and the essays that made them said so about the listener and not about the note.

perception · Pitch-acuity
A duration category has a tempo range of its own. Each simple ratio's short note is the beat divided by one more than the ratio, so at a high enough tempo it falls under the fastest interval that can be a beat at all — 100 milliseconds. Each bar here runs from the slowest tempo at which the ratio's long note still belongs to a beat to the fastest at which its short note is still a note: 1:1 ends at 300 bpm, 2:1 ends at 200 bpm, 3:1 ends at 150 bpm, 4:1 ends at 120 bpm. The line is this site's swing curve, and where it crosses a category boundary the category it is leaving has already ceased to exist.

The short note is sitting on the floor

Swing is modelled here as a short note of constant absolute length, which was measured from drummers and left as a fitted parameter. The tempo window's fast edge — the shortest interval a series of events can be a beat at — was measured from listeners tapping. Both are a hundred milliseconds, and if that is not a coincidence then the swing ratio has no free parameter in it at all: the short note is not held constant, it is resting on the floor.

rhythm · Microtiming
What the engraver used here gives a note, measured off the page. The horizontal distance VexFlow allots each duration, read back off a formatted system rather than quoted from a manual. It is not a power of the duration: it is a constant of 58 points plus 15 points a crotchet, and the constant is 93 per cent of the width the shortest note here gets. A note four times as long as another is about 2 times as wide, not four. The floor is the notehead, its stem, its accidental and the space a reader needs to see them as separate events — which is a claim about legibility and not about time at all.

The axis that is not a time axis

Eight earlier essays have measured the staff's vertical axis to a position. Its horizontal one has never been asked about, and the answer is that it is proportional to nothing: measured off the typesetter used here, a note gets 58 points before its duration is considered at all and 15 points a crotchet after — so the constant is 93 per cent of what the shortest note gets, and a note four times as long is not four times as wide.

scales · Notation
Articulation is worth 2.6 phons, and nobody counts it. A passage of notes at 80 decibels, 2 to the beat, at 7 tempi and 3 articulations, scored against the same level held continuously. The variable is the fraction of each inter-onset interval that is sounding — 0.95 is a legato, 0.4 a staccato — and the vertical axis is what that costs the passage's running loudness in phons. Nothing here is anybody playing harder or softer. At 40 to the beat the span from legato to staccato is 1.47 phons; at 200 it is 2.61, because a staccato note there lasts 60 milliseconds and no longer reaches its own loudness either.

A staccato is a dynamic mark

Every loudness figure in this collection is of a sound that has been going on long enough, and no note in music has. Run the running-loudness model on notes with lengths in them and an articulation turns out to command 1.5 phons at a slow tempo and 8.1 at a fast one — more than the 1.8 decibels a whole texture commands, on the same page, written down in the same ink, and counted by nobody.

perception · Loudness
A long note and a strong note disagree, and the winner is neither. The same 8 notes scored against every triad and seventh at every root, with the weighting run from the metrical one always used to a durational one never drawn. On the left each note counts for its metrical weight; on the right, for how long it is held. The long notes here are on beats 2, 4, 6, 8, which are the weak ones. The two cues point at different chords — C major7 on the left and D minor7 on the right — turning over at a mixture of 40 per cent. And at the crossing the winner is A minor7, which is neither cue's answer — a chord that shares three notes with each and is not the reading either rule asks for. Nothing about the notes changed. What changed is which of two cues a theorist would call obvious is being believed.

The long note and the strong note

The segmentation that produces every object connected here has carried a free parameter since the day it was written: whether a note counts for its metrical weight or for how long it is held. Only the first has ever been drawn. The two name different chords on sixteen per cent of passages where the cues agree about the notes and forty-three per cent where they do not — and where they disagree most sharply a mixture of them picks a third chord neither one asks for.

harmony · Progression
The reading the joint search was never offered. The best chord at each mixture of the two segmentation cues, and what the same weighting gives the same notes shuffled into a different order. Both fall along the axis, and most of the fall is the ruler rather than the music: a metrical weighting over a bar of eight spans a factor of eight and a three-to-one duration spans three, so the weighted note mass is 2.1 times more concentrated at the left of the figure than at the right, and a concentrated mass is easier for four notes to cover. What is not the ruler is the gap. It is widest at a mixture of 0.75, where the reading is D minor7 at 2.15 standard deviations above its own null, against 1.12 for C major7 at a mixture of nought. The joint search holds this axis at nought, so D minor7 is not among the hypotheses it considers.

A fourth decision, and two that were never made

The joint search resolves key, metre and segmentation together and holds the segmentation's cue mixture at zero. Adding the mixture is one loop, and reading the search in order to add it turns up something worse than a missing axis: on the passages it is drawn on, the key it reads is the same key at all forty-eight of its hypotheses and the metre scores every barline identically. The fourth axis then cannot be ranked at all until each reading is measured against its own null, because a mixture changes the ruler and not only the answer.

harmony · Progression
The term that was owed, and the corpus cannot hold it. For each of the three tunes everything here is measured on, how many of its notes have a duration that differs from the gap to the next onset. The answer is none, in 101 notes: these tunes are stored as a list of pitches and lengths with no rests in them, so a note's duration IS its inter-onset interval and conditioning one on the other leaves exactly zero bits. That is a fact about the representation rather than about music. The prediction was that the term would be small, and it could not have been known that the corpus would make it identically zero — which means the prediction cannot be tested here and the exceptions have to be priced directly.

A note lasts until the next one starts

Pricing where a note is against which note it is left duration as the term it had not, with a prediction that it would be small. Measured on the three tunes these readings are built on, it is exactly zero — and it is zero by construction, because those tunes are stored as pitches and lengths with no rests in them, so every duration is its own inter-onset interval. The prediction cannot be tested on the corpus that produced it. Priced directly, a rest costs 0.67 bits a note where a tenth of the notes have one, which is not well under half a bit.

scales · Notation
A tie is charged twice, and the second charge is the larger one. What a tie costs a reader, against the share of noteheads that are the second of a tied pair. The lower curve is the decision itself — is this notehead an event or a continuation? — at 0.52 bits a note where a tenth of them are tied. The upper curve adds what the extra noteheads cost on every other axis: a tied continuation has a pitch and a position and is read like any other notehead before the reader discovers it carries no event, at 4.79 bits each. The total is 1.05 bits a note, which is 2.0 times the decision alone and is a fifth of what a whole note of music costs. A tie is the most expensive mark on the staff per occurrence, and every published account of notational difficulty treats it as a minor one.

The notehead that is not a note

Every quantity so far is charged per notehead, and a tie is the one mark on the staff that puts a notehead on the page carrying no event. Its cost is not the decision that identifies it — that is half a bit where a tenth of the noteheads are continuations. It is the decision plus the whole reading of a notehead that turns out to have been unnecessary, which is 1.05 bits, twice the decision and a fifth of what a note of music costs. Set beside a dot and a longer note value, the tie is five times the price of either and is the only one of the three that can cross a barline.

scales · Notation
Six named proportions, as blurred as the durations that make them. Six proportions between two parts of a piece — 1 : 1, 4 : 3, 3 : 2, golden section, 2 : 1, 3 : 1 — placed on one axis by the logarithm of the ratio of the longer part to the shorter, and drawn as bars one criterion wide (d′ = 1) for a listener timing both parts with a Weber fraction of 7%, 15%, 35%. Bars that overlap are proportions that listener cannot tell apart. At 7%, 4 of 5 neighbouring pairs stay apart; at 15%, 3 of 5 neighbouring pairs stay apart; at 35%, 0 of 5 neighbouring pairs stay apart.

A proportion is only as fine as its two durations

Analyses of form measure proportions in bars and report them to three figures — a climax at 0.618, a section in the ratio 3 : 2. A listener has each part only as an estimate of how long it lasted, and a ratio of two estimates is blurred by both. Timed as well as anyone times a single second, eleven proportions fit between 1 : 1 and 3 : 1; timed from memory over minutes, two do. The golden section is told from 3 : 2 only below a Weber fraction of 5.4 per cent.

form · Proportion
The same forms by the clock and by what is stored. Six forms, each drawn twice: its sections sized by their share of the bars, and sized by their share of what a listener has to store when a bar counts only if it is recognised from 1 bar of context. Returns are drawn pale with a dashed edge. twelve-bar blues: returns take 67 per cent of the clock and 13 per cent of the storage; thirty-two-bar AABA: returns take 25 per cent of the clock and 5 per cent of the storage; rondo, ABACA: returns take 40 per cent of the clock and 10 per cent of the storage; verse and chorus: returns take 50 per cent of the clock and 13 per cent of the storage; two eight-bar phrases: returns take 0 per cent of the clock and 0 per cent of the storage; a four-bar ostinato: returns take 88 per cent of the clock and 0 per cent of the storage.

A return is shorter than its first hearing

A rondo's refrain takes three fifths of the clock and a verse-and-chorus song is balanced to the bar. Count instead the bars a listener could not have predicted when they arrived, and the returns shrink to between a tenth and a quarter of what is kept — so a song equal by the clock is between three and seven times heavier in its first half. A coder that learns repeats one bar at a time says the halves are equal, and the two memories disagree by more than any proportion a listener could confuse.

form · Proportion
A count is least reliable at both ends and best in the middle. How precisely a listener knows the length of a section they are counting, against how many units long it is, at three kinds of timing judgement. A slip — one unit miscounted, at 2% a unit — accumulates as a random walk, so its relative cost FALLS as the section lengthens. A lapse — the count lost altogether, at 1% a unit — compounds, so the chance of still having the count falls geometrically and a long enough section is certain to lose it. A listener who has lost the count is back to timing, so the two failures mix into a floor. Against a Weber fraction of 7.5% the count is worth most at 21 units, where it is 1.7 times finer than timing, and falls back under a quarter better by 98 units; Against a Weber fraction of 15% the count is worth most at 10 units, where it is 2.4 times finer than timing, and falls back under a quarter better by 101 units; Against a Weber fraction of 38% the count is worth most at 4 units, where it is 3.7 times finer than timing, and falls back under a quarter better by 102 units. The length at which it stops being worth much is nearly the same in all three, because it is set by the lapse rate alone.

A count is not an estimate

Both established routes to a proportion are estimates — a duration timed, blurred by a Weber fraction, and a duration stored, biased by what was new. A listener who has induced a hypermetre has a third, and it is exact until it fails. It fails two ways that pull opposite: a slip miscounts one unit and its relative cost falls as the section lengthens, while a lapse loses the count entirely and its chance compounds. The mixture has a floor at about four units, where counting is 3.7 times finer than timing, and it is worth almost nothing past a hundred.

form · Proportion
Timing blurs a whole form evenly; counting sharpens it downward. A piece of 480 seconds divided 7 times, each level half the length of the one above, with how many proportions between 1 : 1 and 3 : 1 a listener can tell apart at each. Timed, the answer is 2.1 at the top and 5.2 at the bottom, a spread of 2.4 — because a timing judgement's Weber fraction is a step function of duration and almost every level of a piece falls in one step of it. Counted, in units of 2 seconds, the answer runs 2.2 to 10.3, a spread of 5.5. At no level does timing separate 3 : 2 from the golden section.

A form is sharp at the bottom and vague at the top

A movement is divided into sections, each into phrases, each into bars, and every level is a ratio of two estimates. Timed, the hierarchy is almost uniformly blunt — 2.1 distinguishable proportions at the top and 5.2 at the bottom, because a Weber fraction is a step function of duration and six of a piece's seven levels fall in one step of it. Counted, the same hierarchy runs from 2.2 to 12.2 and sharpens monotonically downward. At no level of either does timing separate 3 : 2 from the golden section.

form · Proportion
The golden section and an equal division are one judgement. Where a boundary falls in a piece, as a share of its length, with the band a listener cannot tell from the golden section shaded. A stretch of minutes is judged with a Weber fraction of about 38%, so one criterion's worth of ratio spread around 0.618 covers everything between 0.492 and 0.730 — a quarter of the piece wide, and containing the halfway point. 1 : 1 and 4 : 3 and 3 : 2 and golden section and 2 : 1 are inside it. A claim that a climax falls at the golden section rather than at the middle is, at this resolution, not a claim about anything a listener could hear.

A golden section is a coin toss with six coins

An analysis that reports a climax at 0.618 of a piece has not tested one prediction; it has looked at a piece with several defensible boundaries and reported whichever landed nearest. The rate at which that happens under no hypothesis is one line of arithmetic, and the tolerance it needs is not a number chosen on the page — it is the blur a listener's own timing puts on the judgement. Over a stretch of minutes that blur covers everything from 0.492 to 0.730 of the piece, which contains the halfway point, and six candidate boundaries produce a hit eighty per cent of the time.

form · Proportion
The breath is the looser ceiling nearly everywhere. How long a trained singer can hold a phrase on one breath, across a compass and at four dynamics, against the 8-second ceiling the psychological present puts on the same phrase. The flow through the folds rises with pitch and with loudness, so the breath ceiling falls both ways: at 60 decibels it runs 32.6 seconds at the bottom of the compass to 21.2 at the top; at 70 decibels it runs 23.1 seconds at the bottom of the compass to 15.0 at the top; at 80 decibels it runs 16.4 seconds at the bottom of the compass to 10.6 at the top; at 90 decibels it runs 11.6 seconds at the bottom of the compass to 7.5 at the top. The shaded line is the listener's ceiling and it does not move. The breath binds only where the two lines cross — 1 of the 40 cells drawn, all of them loud and high. So the constraint everybody names when asked why a phrase is the length it is, is almost never the constraint that decides it.

The ceiling everybody names is the loose one

Ask why phrases are the length they are and the answer given is the breath. It is arithmetic — usable lung volume over the air a note costs per second — and it comes out between fifteen and twenty-three seconds at a comfortable dynamic and between seven and twelve at a loud one. The ceiling the present moment imposes, the two-to-eight seconds inside which a stretch is heard as one thing rather than as a series, is two to three times tighter at almost every note and dynamic. A singer in an adagio is not running out of breath at the phrase end. They are running out of present.

form · Phrase
A timed expectation would erase a slow cycle's cost, and a listener cannot time a slow cycle that well. Bits of position a listener with a 3.5-second memory is still missing over the first cycle of son clave, against how long the cycle takes, for a newcomer with no expectation, a listener timing the cycle with the Weber fraction a duration that long is judged with, and a listener timing it to ten per cent. a newcomer, no expectation: 2 s 0.94, 8 s 0.97, 24 s 1.41, 40 s 2.02, 60 s 2.47; timing as well as listeners do: 2 s 0.68 (w 0.150), 8 s 0.69 (w 0.150), 24 s 0.85 (w 0.150), 40 s 1.90 (w 0.375), 60 s 2.38 (w 0.375); timing the cycle to ten per cent: 2 s 0.47, 8 s 0.48, 24 s 0.55, 40 s 0.73, 60 s 1.02. At ten per cent even a sixty-second cycle is placed about as well as a newcomer places a two-second one. At the precision a listener actually has for durations of half a minute or more, the expectation is worth a tenth of a bit.

An expectation cannot rescue a cycle too slow to time

A listener who knows a piece arrives with an expectation of where in the cycle they are, and the size of that expectation was the number the last essay said nobody had. It can be given one: a listener who has been timing the cycle carries a spread of their Weber fraction times the cycle, which is the same number of steps at any tempo. Timed to ten per cent, a forty-second cycle would be placed better than a newcomer places a two-second one. But forty seconds is judged in the band where the Weber fraction is nearer forty per cent, and there the expectation is worth a tenth of a bit.

rhythm · Cyclic rhythm

Named alongside it

The objects these essays reach for when they reach for this one.

Musical formNotationPhraseProportionWeber fractionPerceptual presentTempoHypermetreInformationInter-onset intervalMemory decayMetrical weight

All concepts