Concept

Streaming — where it appears

The assignment of successive events to one perceived line or to several, which is decided by rate and separation rather than by the score. It bounds how far a fast melody may leap and how close two written voices may be written.

Named by 7 essays across 3 fields — each of them below, with the objects they name alongside it.

A 7-semitone sequence at 120 ms a tone. Tones drawn as pitch against time, one bar per tone. The events are the same in both readings of this pattern; what changes is whether a listener assigns them to one line that leaps back and forth or to two lines that each stay put. Nothing in the drawing decides which, and nothing in the sound does either.

The ear builds objects, and sometimes offers a choice

What arrives at an ear is one pressure signal. What a listener gets is a set of separate things — a violin, a voice, a car outside. The assignment is a construction, and the clearest evidence is that it can be flipped by changing nothing but the speed: one sequence of tones is a single line when slow and two lines when fast, with a wide region in between where the listener may choose.

perception · Auditory scene
Four voices, placed by the arithmetic. The rules in force are: no parallel octaves; no parallel fifths; no voice crossing; no gap over an octave above the tenor; the leading note is not doubled; the leading note resolves, outer voices; no augmented melodic interval. The four parts are drawn lowest to highest — bass, tenor, alto, soprano. I – vi – ii – V – I in C major, realised in four voices by the cheapest set of voicings obeying 7 rules, at 16 semitones of motion in all. Every gap between adjacent voices narrower than the fission boundary of 5.2 semitones is marked, and below that boundary two parts cannot be heard as two however hard a listener tries.

A voice is a stream, and the ear decides which

Seven earlier essays have assigned voices to notes. Whether a listener follows the assignment is a separate question with laboratory numbers attached, and the numbers are unkind to it: two parts closer than about five semitones cannot be heard as two at any speed, a third of the gaps in the cheapest four-part writing are inside that limit, and in a third of chord changes the ear's own rule for continuing a line does not recover the parts as written.

perception · Voice-leading
Ode to Joy as a path. Ode to Joy plotted as 30 notes against the 8 scale degrees it uses, one column per note. Its largest melodic interval is 2 semitones and it spans 7; the mean absolute step is 1.24 semitones. Beethoven, Ninth Symphony, finale, 1824 — the theme as first stated, eight bars.

A melody is a walk, not a set

Nine essays here are about which seven of the twelve a scale takes, and every one of them describes a set. A tune is not a set; it is a path across one, and the path is nearly all small steps. That is not a matter of taste. Above about eight notes a second the ear stops being able to hold a large interval and a small one in the same line, and at sixteen the choice disappears altogether — so a fast passage is scalar because a fast passage that leaps is two pieces of music.

scales · Melody
The gap that decides whether two clocks are two. The closest approach of the two streams in each ratio, at 100 to the minute in 4-time — a bar of 2.4 seconds. It is the bar divided by the product of the two numbers, which is an identity and is checked here against the measured minimum. 9:5 is the last ratio whose onsets are securely separate at this tempo; past it the two streams' events fall inside the window in which the ear cannot put two onsets in order, and what is heard is one irregular pattern rather than two clocks. The bound has a product in it, which is why 3:2 and 4:3 are everywhere and 11:7 is a notation.

The ratio that stops being two

Eight earlier essays have taken two clocks at a rational ratio to be a thing a listener can hold. There is a ratio past which it is not, and the bound is not where anyone would look for it: the cycle is exactly one bar long at every ratio, so the length of the pattern separates nothing. What separates them is the closest the two streams ever come, which is one part in their least common multiple — 400 milliseconds for three against two, and seventeen for thirteen against eleven, which is not two events at all.

rhythm · Polyrhythm
The same steps, counted by the clock. The step distribution of Ode to Joy and Twinkle, twinkle counted two ways: once per interval, which is what every earlier figure did, and once weighted by how long the note it leaves is held. The two disagree because a tune's long notes are not distributed evenly over its interval sizes — in Ode to Joy the 2-semitone step is 55.2 per cent of the moves and 60.0 per cent of the time. Which of the two a claim about melodic motion means has never been stated here, and the answer matters most exactly where a tune slows down, which is at the ends of its phrases.

The note that has a length

Every melodic figure so far reads a table where each note is a pair — a pitch and a duration — and throws the second number away. A step between two minims and a step between two quavers have been one event in every histogram it has drawn. Weighting the same statistics by time moves the step distribution by up to seven points, changes forty-four of a hundred and one contour signs, and turns up an off-by-one in the one figure that did use the durations: it took the length of the note arrived at where the time between two onsets is the length of the note left.

rhythm · Melody
What a competition decides when the two cues do not agree. Each spectrum with its two cue readings and the grouping the competition chooses. Harmonicity asks whether a partial is near enough a whole multiple to belong; common fate asks whether it decays at the same rate as the rest. Where they disagree there is no rule in this collection, so the published apparatus is used instead: every way of splitting the partials into one stream or two is scored for the partials each cue says it has wrongly grouped and wrongly separated, and the cheapest wins. an ideal string — harmonicity 100 per cent, common fate 20, and the competition says one stream; a piano string — harmonicity 90 per cent, common fate 20, and the competition says a cut after partial 2; a bell — harmonicity 88 per cent, common fate 13, and the competition says a cut after partial 1; a bar — harmonicity 33 per cent, common fate 67, and the competition says a cut after partial 4; a kettledrum — harmonicity 60 per cent, common fate 100, and the competition says one stream. The exchange rate between the two cues is the number nobody here can supply, so what is reported beside each is how many decades of it leave the answer unchanged.

The exchange rate nobody has

There are now two cues that disagree about how a spectrum divides, and every figure so far reports them separately because there is no principled way here to weigh one against the other. The published apparatus is a competition between grouping hypotheses with a cost per cue — and the useful thing it produces is not the winner but how much of the exchange rate the winner survives.

perception · Auditory scene
The cue that settles it. Every spectrum to hand, arbitrated by the earlier competition and then again with the onset cue added at equal weight. 3 of the 5 change their verdict, and all 3 change the same way — from splitting into two streams to staying as one: a piano string, a bell, a bar. Nothing changes the other way, because the onset cue on a struck source votes for fusion on every partial and can only ever push toward one stream. The bell is the case worth naming: its partials are wildly inharmonic and it is heard as one sound, which is a fact the harmonicity cue alone cannot produce.

The cue that settles it

Arbitrating between two grouping cues meant sweeping an exchange rate nobody could supply. The cue it had no term for at all is the one every account calls strongest, and its strength is computable: a struck string's partials start together to within a tenth of a millisecond against a threshold of twenty. Put that into the competition and three of five verdicts change, all the same way — and a bell becomes one sound.

perception · Auditory scene

Named alongside it

The objects these essays reach for when they reach for this one.

Auditory scene analysisTemporal coherenceFission boundaryFusionCommon fateContourMelodic intervalPartialBistabilityCompound melodyCross-rhythmDecay

All concepts