Every essay — page 20
Pitch and tuning Intervals and chords Scales and modes Harmony and voice leading Rhythm and metre Timbre and acoustics Perception and the listener Instruments and their design Form and structure Series Objects Sounds Search
Form and structure
The shape a piece has in time, computed rather than labelled. A distance matrix finds the sections before anybody names them, a phrase turns out to be a number of seconds rather than of bars, and an ending is five separate signals that can each arrive without the others.
The twelfth, and where it comes from
Melodies occupy about an octave and a fifth, and an earlier essay set out to explain that by the singer's register break and found that it does not: the chest mechanism alone spans two octaves and a semitone. The answer is in a parameter the essay on leaps fitted and then put down. A walk with no walls whose central tendency reproduces the post-skip reversal rate has a range that grows logarithmically — six semitones at eight notes, twelve at thirty, eighteen at a hundred and twenty — so across every length a tune plausibly has, the span is between an octave and a fifteenth.
Where a phrase ends
Run a boundary detector over the three tunes used throughout and it agrees with the notated phrasing on one of them perfectly and on another almost not at all. The reason is which cue each tune uses: Twinkle's phrases all end on a long note, so a duration-weighted detector finds five of five with no false alarms; Ode to Joy's run on in crotchets and its phrasing is in the intervals, where a duration detector finds one of three and a pitch detector finds all three and eight others. No fixed weighting serves both, and the published one is worse on each tune than the single cue that tune uses.
The repeat that is not in the notes
Eight earlier essays compare bars by writing each one as a bag of pitch classes and taking a cosine. Nothing chose that encoding — the first used it and the other seven inherited it. Encode the same six schemes four other ways and one of the three findings survives untouched, one survives with different numbers, and one turns out to have been a statement about the encoding all along: the boundary operator finds none of the section edges in three schemes as a bag of pitch classes and every one of them as tonic, subdominant and dominant.
What the onsets left out
Eight essays induce a metre from a list of ones and zeros, and every failure they recorded was argued about as a failure of the rules. Two of the three are not. Note length is already in that list and the scoring throws it away: put it back and the son clave's two-way tie resolves to the notated downbeat. But the groove with its beat removed cannot be repaired by any cue at any strength, and the reason is arithmetic rather than empirical — the true phase has no onset on any of its strong positions, so there is no credit for a cue to multiply and its line is exactly flat.
A boundary at a stated level
A boundary detector run over three tunes agreed with the notation on one and barely at all on another, and left two things owing: a version with a scale parameter, and a version run on performance timings. Both are paid here, and they pay differently — the scale removes a free parameter from the comparison and does not rescue the hard case, while two per cent of rubato does.
Loud is relative, and it comes down slowly
The account of loudness had a model of a moment and the account of closure asked it for a model of a form. The published one exists and its content is a pair of numbers that are not the same: a listener's running impression of how loud the music is rises to meet a step in a fifth of a second and takes seven seconds to come back down. A twenty-decibel crescendo spread over eight seconds therefore buys almost no contrast at all, and the same twenty decibels taken as a step buys a factor of two.
The level the tempo chooses
The boundary detector has a scale parameter and an earlier essay left it free, ending with the sentence that names this one: the scale is in notes and the psychological present is in seconds. The psychological present is two to eight seconds, a tempo converts one to the other, and the level a listener reads then stops being a parameter at all — which is a prediction with teeth, because the same tune at two tempos should be phrased differently at levels the arithmetic names in advance.
The dynamics are in the score already
Count the parts in each bar, realise them in their ranges, put every partial in its critical band, sum the loudnesses and run the result through the two smoothers built earlier. What comes out is a dynamic curve for a piece with no performance in it anywhere — and it says that doubling the number of parts inside a fixed register adds three decibels of power and about one of loudness, because the extra parts land in bands that were already occupied. Let the register widen with the parts and the same arithmetic gives eight phon, which is what a tutti actually is.
A final chord is not made loud by adding to it
An earlier essay on closure said the loudest cue an ending has needs a corpus rather than an arithmetic. The arithmetic was built one essay ago, so it does not. Run four ending textures through it and two things come out backwards: a final chord three parts thicker than the rest arrives *quieter* against the running impression than the passage it ends, and a texture that drops a part a bar does not get quieter at all until the bar where there is one part left.
Who plays what and how loud is one question
Two lines of argument, one about spectrum and one about loudness, each stopped at the same wall and each said so. One of them can choose who plays which note and has every player at the same level; the other can choose how loud each part is and has nobody assigned to anything. Put together they are a single problem with two kinds of variable, and solving it in stages picks a different answer from solving it at once — nine per cent rougher, at the same loudness, on an ordinary triad.
The parameter that did not decide the answer
An earlier essay on closure ended by naming the tempo as the thing every number in it was resting on, and said it was the kind of parameter that had caused trouble before by turning out to decide the answer. Turned across a tenfold range at a closing gesture of fixed length it moves the reading by four per cent, against a twenty-three per cent gap between the gestures it is distinguishing. The parameter beside it in the same figure — how many bars the gesture occupies — moves it by twenty, and nobody had named that one at all.
A detector whose resolution the performance sets
The boundary detector lost its free parameter when the psychological present became a number of notes at a stated tempo, and what that held still was named at the time: a performance slows into a phrase end, so the number of notes inside the present is not the same everywhere in a tune — it falls exactly where a boundary is. Making the width follow the performance recovers half of what removing the parameter cost, and honestly leaves the other half.
The reading was a step response
Sweeping the tempo found it did not decide the answer. This one sweeps the deceleration across a factor of three and finds a null to five figures, and then sweeps the length of the closing gesture across a factor of forty-eight and finds it moves the reading by eight per cent — but not as a function of seconds. Sorted by seconds the twelve runs scatter; sorted by how many bars the instruction covers they fall into three tight groups. One sentence explains the null and the not-null together.
A subito piano is a rate, not a level
All three earlier essays score one chord held still. An orchestration is a succession, and the running impression carries a chord into the one after it — so a written dynamic is an instruction to the listener's impression rather than to the instantaneous sound, and there are markings that cannot be produced at all. The correction runs to silence and the impression still sits above the target.
The dissonance arrives and the dynamic does not
A scoring decides two things at once and both of them have to be integrated by a listener before they exist. The loudness smoother's release is two seconds and the roughness window is thirty-seven milliseconds, and that ratio of fifty decides which of the two survives at the pace music is actually played. Nothing anybody performs is fast enough to blur a dissonance, and a great deal of it is fast enough to average a dynamic.
A rest is a diminuendo
Three parameters swept over factors of ten, three and forty-eight returned the same reading to four significant figures. The one manipulation left unswept moves it by twenty-four decibels: a silence. A listener's running impression decays at the loudness smoother's two-second release, so a general pause is a diminuendo nobody wrote, and at the length that makes a gap an ending it is worth more than any marking a composer has.
A soft chord has to fade in
Forward masking sits between the two integration times already in play — two hundred milliseconds against a thirty-seven millisecond roughness window and a two-second loudness release — and it was owed as the term that might eat the dissonance contrast. It does not. It widens it, by three per cent at a chorale's pace and fifty-nine at four chords a second, because it can only ever remove partials and it can only reach the chord a dynamic has already made quiet. What it does instead is stranger: one chord in the passage is entirely inaudible for its first twelve milliseconds and takes a quarter of a second to arrive whole.
A rest needs a dry room
The twenty-four decibels a general pause is worth assume the sound stops when the players do. Put a hall under it and a bar of silence keeps 60 per cent of its value in a shoebox concert hall and 22 per cent in a cathedral, while a three-and-a-half-second one keeps 86 and 45 — because a hall's tail has a length and a rest either outlives it or does not. A written bar of silence is worth half its dry value at 2.76 seconds of reverberation, which falls between the concert hall and the stone church.
A general pause is spent by the note after it
Whether a composer should write the pause before the diminuendo or after it looks like a question about how big the ensemble is. It is not. Forty decibels of ensemble are worth one decibel of silence, and the order is worth seven — because a running impression rises twenty times faster than it falls, so half a general pause is spent by ninety-seven milliseconds of sound.
An entrance is a change of colour
Eight essays on orchestration move the assignment and hold the ensemble still, and a score does the opposite: it brings players in and takes them out. Loudness is a sum over parts and roughness is a sum over pairs, so the player who joins adds one term to the first and one to the second for everybody already there. What the entrance is worth in phons falls by a factor of forty-eight across the range an ensemble spans and crosses the difference limen at five players; what it is worth in roughness rises by twelve, and by a further factor of ten for every ten decibels the passage is played at.
The release is on the wrong side
Whether the loudness model's two-second release makes an entrance inaudible has the answer no, for a reason the question did not anticipate. The smoother is asymmetric — ninety-nine milliseconds going up and two seconds coming down — so a rise is tracked twenty times faster than a fall, and an entrance is received promptly by every one of a listener's three readings. The colour of it arrives first, at twenty-five milliseconds against ninety, and the reading that moves with the ensemble is the one nobody would have picked.
A part that leaves is not a part that arrives
The same player, the same note, the same level, and the only difference is which way round it happens. A listener's loudness reading takes 1.43 seconds to receive half of a departure and 90 milliseconds to receive half of an arrival — a factor of sixteen with nothing asymmetric in the sound at all, since both readings integrate the same two states in the same order. The colour reading receives the two identically, because a window has no direction, so a departure is a change whose grain arrives at once and whose level takes most of two seconds.
The chord that has room for an entrance
Three essays have made the ensemble something a score can change and none of them has asked when. The ensemble already sounding puts a masked threshold over whatever register an entering part takes, and that threshold is set by the voicing rather than by the dynamic — so the five chords of one passage differ by 8.7 decibels in how much of an entering oboe survives them, and the quietest chord of the five is the worst place in the passage to bring somebody in. Swept over the entrant's own pitch, the choice of moment is worth as much as the choice of register.
Room is used up by whoever enters first
The chord with the most room for a part entering alone is a fact about that chord. It stops being a fact the moment two parts want it, because each part that comes in raises the mask over everybody after it. Given six parts waiting to enter a five-chord passage, choosing each part's moment the way one part's moment is chosen puts three of them into the same chord and lands in the bottom fifth of all 15,625 schedules. Placing them one at a time does no better. The schedule under which the least audible entrance is heard best is unique, and it brings the low and middle parts in while the texture is thin and holds the three highest back for the last three chords — because a high part keeps its room over a full texture and a middle part does not.