← Files AI Film Pipeline MasterARCHIVED FILE
skills/ai-film-pipeline-master/references/phase-05-song-generation/vocal-arrangement-harmonies.md
42.8 KB · Oct 7, 2026 · 00:35 UTC
# Vocal Arrangement, Choral Textures & Multi-Part Harmonies 26 cards. Professional vocal arrangement, parallel thirds/sixths, choral counterpoint, whisper doubling, Eastern vocal ornamentation, and multi-layered vocal stacking. --- ### Name the Room Before the Stack **Also called:** the harmony selector, room before stack, who is singing and for what **What it is:** The card at the head of the `harmony.` shelf. It does not pick a harmony. It asks the three questions the shelf's own pair graph asks over and over — **what room is this song for, how many voices must the listener hear, and how hard is the throat pushed** — and hands back a named shortlist for each answer. The axes were read off the shelf's `Avoid when:` boundaries and the conditions its `Yields to:` fields repeat; they were not invented. Every card named here is named by id. **Effect on the audience:** None. This card never reaches a listener. It exists so the voicing that does reach the listener was chosen against the occasion, the cast of singers and the force of the delivery, rather than stacked because stacking is what a chorus is assumed to want. **Used for and where it works best:** Answer the three questions **in order, out loud, one sentence each, before you open the shelf.** The first two answers together leave at most five cards; the third splits the single-voice cards, which the first two cannot separate. **One fact goes before all three: is the melody in a maqam?** If it is, `harmony.tarab_vocal_ornamentation` and `harmony.microtonal_vocal_mordent` are live, and read the warning under `Avoid when:` before trusting any stacked answer. Where two answers fight for one section — a verse for one voice and a chorus for a congregation — that is two sections and two runs of this card. Three of the shelf's exits land in dated engine settings rather than on another shelf; when the map sends you out of the shelf, check which kind of family you have landed in. **Best in:** formats: Song Lyric, Music Video, Film Soundtrack, Musical Theatre, Sacred Music, Heritage Documentary | genres: All **Avoid when:** **The melody lives in a maqam and your answer to Axis 2 is *fused* or *thickened*.** Not one card under those two values carries a boundary against a quarter-tone melody — `harmony.parallel_thirds_harmony`, `harmony.barbershop_close_harmony`, `harmony.eight_layer_vocal_stacking` and `harmony.vocal_breath_synchronization` are all offered as though every melody were tempered — and the only edges on the shelf that mention tuning run **away** from the maqam cards, never toward them. On that brief this map will hand you a Western stack without a word of warning. Also distrust **Axis 1**, because it is the axis nearest to `Effect on the audience`: its values are occasions, and a song that is two occasions at once — a sacred song people dance to — has no clean answer. `harmony.gospel_choral_call_response` is that song, and it sits under *formal* only because the cards it trades with do. And avoid it for a voice that speaks rather than sings: nothing on this shelf governs narration or dialogue. **Example:** `selector_run: first "a formal room — a liturgy" -> second "several voices fused into one body" -> third "full voice, not pushed" -> shortlist: harmony.monastic_pedal_tone.` --- **The axes** **Axis 1 — What room is this song for: the popular one, the formal one, or the dark one?** Asked first, because it is the boundary the shelf states most often — eight conditions, and every one names the commission, not a feeling: *"Dark, dissonant horror or tense psychological sequences"* (`harmony.parallel_thirds_harmony`), *"Upbeat commercial advertisements or feel-good family songs"* (`harmony.choral_cluster_dissonance`), *"Light, bouncy children's animation songs"* (`harmony.bass_baritone_vocal_anchor`), *"Bright upbeat romantic comedy songs"* (`harmony.guttural_throat_drone`), *"Upbeat commercial dance music"* (`harmony.monastic_pedal_tone`), *"Quiet sacred monastic hymns where individual ego must be submerged"* (`harmony.post_chorus_ad_lib_runs`), *"Nihilistic dark noir"* (the first half of `harmony.gospel_choral_call_response`), *"Factual documentary voiceover where reality must remain grounded"* (`harmony.detuned_vocal_chorus`). **Why this is not `Effect on the audience` in another coat:** a cartoon song, a dance track, a hymn and a horror cue are known from the brief before a note is sung; what the listener feels is known only afterwards. The first is a fact you can fork on. The second is what every card here claims. **Axis 2 — How many voices, and must the listener hear them as one or as several?** Eight conditions, and every one of them is a **measured move between the values of this axis**, not a mood: *"Fragile acoustic solo ballads where a single vulnerable voice must stand completely alone"* (`harmony.lead_vocal_doubling`, which yields to the stripped single voice), *"Intimate bedroom conversations between lovers where individual identity matters"* (`harmony.octave_doubling_blend`, which yields from a fused body to two separate people), *"Solo tracks with only one intended speaker"* (`harmony.vocal_formant_duet`, which yields back from two people to one voice thickened), *"Intimate acoustic folk pieces where natural minimalism is the goal"* (`harmony.eight_layer_vocal_stacking`), *"Simple commercial jingles where a single slogan must be heard clearly"* (`harmony.contrapuntal_vocal_counterpoint`), *"Intentionally simulating a chaotic, disorganized civilian crowd"* (`harmony.vocal_breath_synchronization`), *"Instrumental background wallpaper music"* (`harmony.naked_vocal_breakdown`), and *"solitary acoustic tragedy"* — the second half of the gospel card's line, which yields from a congregation to one wordless voice. **Axis 3 — How hard is the throat pushed: breath, full voice, or projection?** Five conditions on the shelf and one from outside it, all about force and nothing else: *"Shouted rock or metal vocals where a whisper layer causes phase flanging"* (`harmony.whisper_track_sizzle`), *"Aggressive modern trap or heavy metal"* (`harmony.barbershop_close_harmony`), *"Authoritative military commands where falsetto breaks sound weak"* (`harmony.falsetto_chest_voice_flip`), *"Classical bel canto opera or formal institutional anthems"* (`harmony.vocal_fry_onset`, whose pair field names the reason — *full formant projection*), *"Whispered conversational dialogue"* (`harmony.open_air_vocal_projection`), and from the hook shelf, `hook_psych.climax_octave_leap_catharsis` sending a reader to `harmony.octave_doubling_blend` when *"The singer cannot sustain the higher octave"* — size bought without push. It is asked last because it is the only axis that separates the three breath cards from their neighbours; the first two questions leave them sitting beside a belt and a shout. --- **The map** *A card whose own boundary is on another axis is placed on this one by its own `What it is:` and `Best in:`.* **Axis 1 — what room is this song for** | Answer | The shortlist, by id | | :--- | :--- | | **Popular — a chart, a party, a stage musical, a child or an advert** (fourteen) | `harmony.parallel_thirds_harmony` · `harmony.post_chorus_ad_lib_runs` · `harmony.lead_vocal_doubling` · `harmony.whisper_track_sizzle` · `harmony.eight_layer_vocal_stacking` · `harmony.naked_vocal_breakdown` · `harmony.barbershop_close_harmony` · `harmony.falsetto_chest_voice_flip` · `harmony.vocal_fry_onset` · `harmony.belting_vs_head_voice` · `harmony.vocal_formant_duet` · `harmony.contrapuntal_vocal_counterpoint` · `harmony.vocal_breath_synchronization` · `harmony.octave_doubling_blend` | | **Formal — sacred, courtly, memorial, or a classical art tradition** (eight) | `harmony.monastic_pedal_tone` · `harmony.antiphonal_stereo_choir` · `harmony.open_air_vocal_projection` · `harmony.bass_baritone_vocal_anchor` · `harmony.wordless_vocalise_anthem` · `harmony.tarab_vocal_ornamentation` · `harmony.microtonal_vocal_mordent` · `harmony.gospel_choral_call_response` | | **Dark — dread, the primordial, or the unreal** (three) | `harmony.choral_cluster_dissonance` · `harmony.guttural_throat_drone` · `harmony.detuned_vocal_chorus` | **Axis 2 — how many voices, heard as one or as several** | Answer | The shortlist, by id | | :--- | :--- | | **One singer, heard as one — the voice itself is the choice** (ten) | `harmony.falsetto_chest_voice_flip` · `harmony.belting_vs_head_voice` · `harmony.vocal_fry_onset` · `harmony.tarab_vocal_ornamentation` · `harmony.microtonal_vocal_mordent` · `harmony.open_air_vocal_projection` · `harmony.wordless_vocalise_anthem` · `harmony.post_chorus_ad_lib_runs` · `harmony.naked_vocal_breakdown` · `harmony.guttural_throat_drone` | | **One singer made bigger — layers under the lead that add no new person** (four) | `harmony.lead_vocal_doubling` · `harmony.whisper_track_sizzle` · `harmony.eight_layer_vocal_stacking` · `harmony.detuned_vocal_chorus` | | **Several voices fused into one body** (seven) | `harmony.octave_doubling_blend` · `harmony.parallel_thirds_harmony` · `harmony.barbershop_close_harmony` · `harmony.bass_baritone_vocal_anchor` · `harmony.monastic_pedal_tone` · `harmony.choral_cluster_dissonance` · `harmony.vocal_breath_synchronization` | | **Several voices heard as separate people or sides** (four) | `harmony.contrapuntal_vocal_counterpoint` · `harmony.vocal_formant_duet` · `harmony.antiphonal_stereo_choir` · `harmony.gospel_choral_call_response` | **Axis 3 — breath, full voice, or projection** | Answer | The shortlist, by id | | :--- | :--- | | **Breath — under full voice; the microphone hears air, creak or a crack** (three) | `harmony.whisper_track_sizzle` · `harmony.vocal_fry_onset` · `harmony.falsetto_chest_voice_flip` | | **Full voice — sung, not pushed** (sixteen) | `harmony.lead_vocal_doubling` · `harmony.parallel_thirds_harmony` · `harmony.contrapuntal_vocal_counterpoint` · `harmony.monastic_pedal_tone` · `harmony.octave_doubling_blend` · `harmony.tarab_vocal_ornamentation` · `harmony.barbershop_close_harmony` · `harmony.choral_cluster_dissonance` · `harmony.eight_layer_vocal_stacking` · `harmony.antiphonal_stereo_choir` · `harmony.microtonal_vocal_mordent` · `harmony.vocal_breath_synchronization` · `harmony.detuned_vocal_chorus` · `harmony.naked_vocal_breakdown` · `harmony.vocal_formant_duet` · `harmony.wordless_vocalise_anthem` | | **Projection — belt, shout, growl, or a voice meant to carry across a valley** (six) | `harmony.gospel_choral_call_response` · `harmony.post_chorus_ad_lib_runs` · `harmony.belting_vs_head_voice` · `harmony.bass_baritone_vocal_anchor` · `harmony.guttural_throat_drone` · `harmony.open_air_vocal_projection` | *A card appears once per axis. The answer is the intersection of the three shortlists, and it is normally one to four cards wide.* **Two intersections are empty and that is information, not a gap in the table:** no card thickens one voice for a rite, and no card sets separate voices against each other in the dark. A brief that lands in either is asking for a card the shelf does not hold. **Two disagreements sit inside this map, and it routes to both cards anyway.** `harmony.vocal_breath_synchronization` declares itself for *all* multi-track vocal arrangements, and `harmony.contrapuntal_vocal_counterpoint` prescribes lines with *"completely independent rhythms"*, which cannot share one attack. The map puts them under different values of Axis 2 — *fused* and *separate* — which is where the disagreement comes from: breath lock is a law of the fused value written as a law of every value. And `harmony.microtonal_vocal_mordent` makes the quarter-tone ornament part of *"every authentic vocal performance"* in its maqamat, while `harmony.tarab_vocal_ornamentation` scopes the same ornament to mawal introductions and climaxes. **Neither is settled here.** --- **The default, and the one against it** **The default is `harmony.lead_vocal_doubling` — and it is argued from the count of distinct conditions, not of edges.** Counting edges, `harmony.parallel_thirds_harmony` leads: three cards send a reader to it. But all three fire on **one** condition — the piece is bright — from `harmony.choral_cluster_dissonance` (*"Upbeat advert or feel-good family song"*), `harmony.bass_baritone_vocal_anchor` (*"Light bouncy children's animation song"*) and `harmony.guttural_throat_drone` (*"Bright upbeat romantic comedy song"*). That makes it the door out of the dark and formal rooms, not a fallback. `harmony.antiphonal_stereo_choir` is reached twice on one condition too — the brief is Western and tempered — and is likewise a door. **`harmony.lead_vocal_doubling` is reached twice for two different reasons, and both come from inside the shelf:** `harmony.whisper_track_sizzle` (*"Shouted rock or metal, where a whisper layer flanges"*) and `harmony.vocal_formant_duet` (*"Solo track with only one intended speaker"*). No other card the shelf itself falls back on is reached for more than one reason. It is also the smallest arrangement there is: it adds no person and no note. That is its strength and its cost — it is the sound of almost every chorus the listener has ever heard. **The stated runner-up is `harmony.wordless_vocalise_anthem`, which has the higher count and loses the argument.** It is reached for three distinct reasons — *"Nihilistic noir or solitary acoustic tragedy"* from the gospel card, *"Quiet contemplative end credits"* from `sync.epic_hybrid_crescendo_sync_finale`, *"Wordless or instrumental track: open vowels instead of a line"* from `hook_psych.semantic_repetition_refrain` — but two of the three come from other shelves, and all three are about **losing the lyric or the crowd**. It is where the package goes when a song cannot carry words. Made the default, it would make *drop the words* the standard way to voice a song. **The one against it is `harmony.microtonal_vocal_mordent`.** Nothing routes to it anywhere in the package. Every other card on the shelf makes harmony by **adding** — a second voice, a layer, a drone, a choir, a formant-shifted twin. This one adds nothing: it makes the colour inside one voice, on one syllable, in under a tenth of a second — harmony made in **time** rather than in stacked pitch. The default thickens the lead with a copy of itself; this card thickens it with a note the tempered stack cannot even name. Its only stated cost is strict Western tempered choral singing, which a maqam melody is not. **Its near cousin against the grain is `harmony.guttural_throat_drone`** — one throat producing two pitches, a harmony with no second singer — which is also reached by nothing. --- **The unasked** **Twelve of the shelf's working cards are the target of no `Yields to:` field anywhere in the package.** They can be chosen, but nothing ever sends you to them. By id: `harmony.contrapuntal_vocal_counterpoint` · `harmony.whisper_track_sizzle` · `harmony.tarab_vocal_ornamentation` · `harmony.barbershop_close_harmony` · `harmony.falsetto_chest_voice_flip` · `harmony.vocal_fry_onset` · `harmony.belting_vs_head_voice` · `harmony.eight_layer_vocal_stacking` · `harmony.microtonal_vocal_mordent` · `harmony.vocal_breath_synchronization` · `harmony.detuned_vocal_chorus` · `harmony.guttural_throat_drone` **And the shape of that list is the finding, not its length. Every throat technique on the shelf but one is on it** — the whisper, the fry, the flip, the belt, the tarab run, the quarter-tone mordent, the throat drone. The single-voice cards that *are* reached — `harmony.naked_vocal_breakdown`, `harmony.post_chorus_ad_lib_runs`, `harmony.wordless_vocalise_anthem` — are placements in the song's form, a stripped bar, a post-chorus fill, a wordless cue; the one throat technique reached, `harmony.open_air_vocal_projection`, is reached as a *register*, the formal answer to the fry. **The pair graph compares arrangements, standing textures that hold across a section, and a throat technique is an event on a syllable.** A flip is not an alternative to a mordent; each lands on one word. That is why Axis 3 is on this card and why it is asked last rather than not at all. **The second thing the list shows is a direction.** Both ornament cards are unasked, and both yield **outward** — to `harmony.antiphonal_stereo_choir` — the moment a brief is Western. Nothing yields **inward** when a brief is in a maqam, and the tempered stacking cards carry no boundary against one. **The shelf knows when to leave the maqam and never when to enter it.** The card it wants is the one that says *this melody carries a neutral interval — do you harmonise it in stacked thirds, or voice it with a drone, an octave, or a second singer ornamenting the same line?* Nothing on the shelf asks that question, and it is the first question of every maqam song. **The third:** exactly one card here knows where the song will be heard. `harmony.antiphonal_stereo_choir` stops at a mono phone speaker; `harmony.eight_layer_vocal_stacking`, which pans hard left and hard right, and `harmony.detuned_vocal_chorus`, which pans wide, say nothing about it. `harmony.room_before_the_stack` --- ### Lead Vocal Automatic Double Tracking (ADT) **Also called:** vocal doubling, tight double, artificial double tracking **What it is:** Layering a second vocal performance identical to the lead, panned center or slightly offset by 15-25ms with subtle pitch micro-detuning (±8 cents). **Effect on the audience:** Thickens the lead vocal, smooths out pitch imperfections, and gives the voice modern radio authority and presence. **Used for and where it works best:** Chorus vocal lines, rock anthems, pop hooks, and modern rap choruses. **Best in:** formats: Pop Single, Rock Anthem, Hip-Hop | genres: Pop, Rock, Electronic **Avoid when:** Fragile acoustic solo ballads where a single vulnerable voice must stand completely alone. **Example:** `arrangement: Lead Vocal Center + Double Vocal (offset +18ms, detuned -6 cents), sitting under the lead rather than beside it — how far under is a balance set in the edit.` **Source:** assumption — the package's working figure, not a published standard; set the offset and detune by ear in the mix. **Yields to:** `harmony.naked_vocal_breakdown` — Fragile ballad where one voice must stand alone. `harmony.lead_vocal_doubling` --- ### Parallel Thirds Vocal Harmony **Also called:** country thirds, pop third harmony, close interval harmony **What it is:** Adding a harmony vocal line that mirrors the lead melody exactly a diatonic third above (or sixth below). **Effect on the audience:** Provides immediate, sweet, comforting harmonic richness; the most universal and accessible harmony in world music. **Used for and where it works best:** Chorus payoffs, romantic duets, and anthemic refrains across pop, country, and folk. **Best in:** formats: Folk Ballad, Pop Chorus, Duet | genres: Folk, Pop, Country **Avoid when:** Dark, dissonant horror or tense psychological sequences. **Example:** `harmony_spec: Lead vocal sings E4 -> Harmony vocal tracks simultaneously on G4 (minor third above).` **Yields to:** `harmony.choral_cluster_dissonance` — Dark dissonant horror or psychological tension. `harmony.parallel_thirds_harmony` --- ### Contrapuntal Vocal Counterpoint **Also called:** independent vocal lines, polyphonic singing, Bach vocal counterpoint **What it is:** Arranging two or more vocal lines that move with completely independent rhythms, directions, and words while maintaining harmonic consonance. **Effect on the audience:** Grips the brain with musical complexity; creates the thrilling sensation of multiple minds or perspectives speaking simultaneously. **Used for and where it works best:** The climax of musical theatre songs, dramatic confrontations, and epic concept suites. **Best in:** formats: Musical Theatre, Progressive Rock, Concept Album | genres: Classical, Epic, Musical **Avoid when:** Simple commercial jingles where a single slogan must be heard clearly. **Example:** `arrangement: Singer A holds sustained high note while Singer B sings rapid staccato rhythmic counter-melody beneath.` **Yields to:** `lyric.jingle_micro_hook` — Jingle where one slogan must be heard clearly. `harmony.contrapuntal_vocal_counterpoint` --- ### Monastic Pedal Tone Drone **Also called:** pedal point vocal, Byzantine ison, church drone **What it is:** Bass singers holding an unbroken, continuous root note (pedal drone, e.g. low D2) beneath an undulating solo vocal improvisation. **Effect on the audience:** Anchors the music in eternal stillness and ancient sacred reverence; strips away all earthly restlessness. **Used for and where it works best:** Ancient church sequences, Sumerian temple liturgies, and contemplative desert scenes. **Best in:** formats: Sacred Music, Heritage Film, Ambient Score | genres: Faith / Biblical Epic, Heritage, Ambient **Avoid when:** Upbeat commercial dance music. **Example:** `choral_spec: Three deep bass vocalists hold continuous D2 throat drone while tenor improvises modal hymn above.` **Yields to:** `harmony.post_chorus_ad_lib_runs` — Upbeat commercial dance music. `harmony.monastic_pedal_tone` --- ### Whisper Track High-Frequency Sizzle **Also called:** whisper doubling, ghost vocal layer, breath sizzle **What it is:** Recording a fully whispered version of the vocal line and high-passing it hard, so that only the breath and the sibilant air survive, then placing that layer far beneath the lead — felt as texture, never heard as a second voice. Recording and filtering it is this phase's work; how far beneath it sits is a balance set in the edit. **Effect on the audience:** Adds an expensive, breathy, intimate "air" and tactile crispness to the lead voice without needing excessive treble EQ. **Used for and where it works best:** Pop ballads, intimate female/male vocals, and close-mic cinematic monologues. **Best in:** formats: Pop Ballad, Cinematic Narration, Intimate Single | genres: Pop, Noir, Drama **Avoid when:** Shouted rock or metal vocals where a whisper layer causes phase flanging. **Example:** `chain: Lead Vocal (Center) + Whisper Vocal (Center, high-pass @ 4kHz, de-essed heavily), buried under the lead — if it can be identified as a second voice, it is too loud.` **Yields to:** `harmony.lead_vocal_doubling` — Shouted rock or metal, where a whisper layer flanges. `harmony.whisper_track_sizzle` --- ### Octave Doubling (Male / Female Unison) **Also called:** octave unison, gender-octave blend, thick lead **What it is:** Having a male vocalist and a female vocalist sing the identical melody line exactly one octave apart in pitch. **Effect on the audience:** Creates a massive, universal, and timeless vocal timbre that feels larger than any single individual human. **Used for and where it works best:** Anthems, protest songs, mythological declarations, and iconic chorus hooks. **Best in:** formats: Epic Anthem, Folk Song, Musical | genres: All **Avoid when:** Intimate bedroom conversations between lovers where individual identity matters. **Example:** `arrangement: Male Baritone sings at A2; Female Soprano sings at A3 in perfect unison lock.` **Yields to:** `harmony.vocal_formant_duet` — Intimate scene where individual identity matters. `harmony.octave_doubling_blend` --- ### Oriental Arabesque Tarab Vocal Ornamentation **Also called:** Tarab runs, microtonal vocal trills, oriental melisma **What it is:** Virtuosic rapid vocal ornamentation featuring quarter-tone bends, mordents, and cascading melismatic runs on a single vowel. **Effect on the audience:** Induces the ecstatic aesthetic state of Tarab (طرب); commands visceral admiration for vocal soul and mastery. **Used for and where it works best:** Mawal introductions, emotional ballad climaxes, and Near Eastern heritage climaxes; whether it is scoped this narrowly or, as `harmony.microtonal_vocal_mordent` frames it, a feature of every performance, is attributed, not settled (see Source). **Best in:** formats: Traditional Arabic / Assyrian Song, Heritage Epic | genres: World, Folk, Heritage **Avoid when:** Western classical hymns where straight, non-vibrato singing is required. **Example:** `directing: "[Vocalist performs intricate melismatic run descending through Maqam Bayati with microtonal quarter-tone inflection on E-half-flat]".` **Source:** reference — searched 2026-09-23; no academic source located quantifies how often this ornament occurs across a performance. A.J. Racy's *Making Music in the Arab World: The Culture and Artistry of Tarab* (Cambridge University Press, 2003) and Johnny Farraj and Sami Abu Shumays's *Inside Arabic Music* (Oxford University Press, 2019) are the standard academic treatments of tarab ornamentation, but neither gave quotable text today either confirming or ruling out a narrow scoping. The disagreement with `harmony.microtonal_vocal_mordent` over how often the ornament is used stays open. **Yields to:** `harmony.antiphonal_stereo_choir` — Western classical hymn, straight non-vibrato tone. `harmony.tarab_vocal_ornamentation` --- ### Barbershop Close Four-Part Harmony **Also called:** 4-part vocal harmony, close harmony quartet, dense vocal stack **What it is:** Arranging four vocal parts (Lead, Tenor, Baritone, Bass) where all voices sit within an octave and a half of each other. **Effect on the audience:** Produces shimmering, ringing acoustic overtones (the "fifth voice" illusion); delivers pure acoustic pleasure. **Used for and where it works best:** Traditional vocal ensembles, nostalgic vintage scenes, and lush a cappella arrangements. **Best in:** formats: A Cappella, Traditional Quartet, Vintage Scene | genres: Jazz, Traditional, Vintage **Avoid when:** Aggressive modern trap or heavy metal. **Example:** `arrangement: Lead melody in center; Tenor sings 3rd above; Baritone fills chord tone below; Bass anchors root.` **Yields to:** `suno.guttural_battle_directing` — Aggressive trap or heavy metal delivery. `harmony.barbershop_close_harmony` --- ### Gospel Choral Call-and-Response Wall **Also called:** choir surge, gospel crescendo, anthemic choir answer **What it is:** A dynamic 20-to-40 voice choir answering every short solo vocal phrase with an explosive, harmonized choral shout. **Effect on the audience:** Electrifies the room with spiritual ecstasy and collective power; makes the listener feel part of a triumphant congregation. **Used for and where it works best:** Climax of uplifting anthems, spiritual awakenings, and grand festival finales. **Best in:** formats: Gospel Track, Inspirational Anthem, Epic Finale | genres: Gospel, Soul, Epic **Avoid when:** Nihilistic dark noir or solitary acoustic tragedy. **Example:** `lyrics: "[Soloist] Did you see the light? / [Full Choir EXPLODES] We saw the light! / [Soloist] In the darkest night? / [Choir] We saw the light!"` **Yields to:** `harmony.wordless_vocalise_anthem` — Nihilistic noir or solitary acoustic tragedy. `harmony.gospel_choral_call_response` --- ### Falsetto-to-Chest Voice Flip **Also called:** vocal break, yodel break, emotional crack flip **What it is:** Intentionally flipping from powerful chest voice into pure, breathy head voice/falsetto on a single syllable. **Effect on the audience:** Conveys profound vulnerability, heartbreak, and emotional surrender; instantly commands empathy. **Used for and where it works best:** Emotional ballad hooks (Radiohead style, ancient folk laments). **Best in:** formats: Indie Folk, Pop Ballad, Art Song | genres: Indie, Folk, Pop **Avoid when:** Authoritative military commands where falsetto breaks sound weak. **Example:** `directing: "Lead vocal belts high F# in chest voice, then flips suddenly into soft falsetto on the word 'alone'."` **Yields to:** `harmony.bass_baritone_vocal_anchor` — Authoritative military command, where a break reads weak. `harmony.falsetto_chest_voice_flip` --- ### Vocal Fry Onset & Terminal Decay **Also called:** glottal rattle, vocal fry intro, low-frequency creak **What it is:** Starting a vocal line with low-frequency, relaxed glottal rattle (fry) before engaging full resonant pitch on the vowel. **Effect on the audience:** Communicates effortless intimacy, world-weary exhaustion, and modern bedroom proximity. **Used for and where it works best:** Modern indie pop, intimate memoirs, and late-night blues verses. **Best in:** formats: Indie Pop, Acoustic Singer-Songwriter, Noir | genres: Indie, Pop, Blues **Avoid when:** Classical bel canto opera or formal institutional anthems. **Example:** `directing: "[starts with slight vocal fry creak] I haven't slept since the river rose..."` **Yields to:** `harmony.open_air_vocal_projection` — Bel canto or formal anthem needs full formant projection. `harmony.vocal_fry_onset` --- ### Post-Chorus Vocal Ad-Lib Runs **Also called:** vocal fills, ad-libs, ad-libitum riffs **What it is:** Improvisational vocal runs, cries, and soulful shouts performed over the instrumental hook following the chorus. **Effect on the audience:** Adds spontaneous human joy and virtuosity; prevents repetitive instrumental loops from feeling sterile. **Used for and where it works best:** The post-chorus instrumental drop in pop, R&B, soul, and Near Eastern dance tracks. **Best in:** formats: Pop Single, R&B, Dance Track | genres: Pop, R&B, Dance **Avoid when:** Quiet sacred monastic hymns where individual ego must be submerged. **Example:** `directing: "[Post-Chorus: Oud solo plays; Lead vocalist ad-libs soulful high melismatic cries in background with wide stereo reverb]".` **Yields to:** `harmony.monastic_pedal_tone` — Quiet sacred hymn, ego submerged. `harmony.post_chorus_ad_lib_runs` --- ### Choral Cluster Chord Dissonance **Also called:** tone cluster choir, Ligeti vocal cloud, horror vocal texture **What it is:** Having multiple choral voices sing notes spaced by semitones and microtones simultaneously, forming a dense, shimmering vocal cloud. **Effect on the audience:** Induces existential dread, cosmic vertigo, and psychological disorientation (e.g. 2001: A Space Odyssey monolith cue). **Used for and where it works best:** Encounters with ancient cosmic deities, entering ruined cursed temples, and catastrophic visions. **Best in:** formats: Horror Epic, Dark Fantasy, Sci-Fi Film | genres: Horror, Sci-Fi, Avant-Garde **Avoid when:** Upbeat commercial advertisements or feel-good family songs. **Example:** `choral_spec: "24-voice choir: Each singer chooses random microtone between C4 and D4, sustaining swelling dissonant cloud".` **Yields to:** `harmony.parallel_thirds_harmony` — Upbeat advert or feel-good family song. `harmony.choral_cluster_dissonance` --- ### Belting Vocal Tension vs Clean Head Voice **Also called:** vocal registration contrast, belt vs head voice **What it is:** Directing when a vocalist should push raw chest resonance up into dangerous high ranges (belting) vs floating into pure head voice. **Effect on the audience:** Belting conveys desperate, life-or-death defiance and passion; head voice conveys ethereal purity and divine detachment. **Used for and where it works best:** Structuring the dramatic arc of emotional power ballads. **Best in:** formats: Musical Theatre, Power Ballad, Rock Anthem | genres: Rock, Pop, Musical **Avoid when:** Leaving vocal registration up to chance in generative AI prompts. **Example:** `verse: Ethereal head voice (soft and floating) -> final_chorus: Full chest belting with audible vocal strain and grit.` `harmony.belting_vs_head_voice` --- ### Eight-Layer Stereo Vocal Stacking **Also called:** vocal wall of sound, multi-tracking choir, Queen vocal stack **What it is:** Recording 8 layers of the identical harmony part (2 center, 2 hard left, 2 hard right, 2 wide mid-sides) to create a colossal acoustic wall. **Effect on the audience:** Turns a single human singer into an immense, stadium-sized choir; envelops the listener in pure sonic luxury. **Used for and where it works best:** Queen-style rock choruses, epic cinematic anthems, and modern hyper-pop drops. **Best in:** formats: Stadium Rock, Epic Anthem, Hyper-Pop | genres: Rock, Pop, Epic **Avoid when:** Intimate acoustic folk pieces where natural minimalism is the goal. **Example:** `stack_spec: 2 Lead Center + 2 Harmonies Left (100% & 70%) + 2 Harmonies Right (100% & 70%) + 2 Octaves Under.` **Yields to:** `lyria.intimate_acoustic_proximity` — Intimate acoustic folk, natural minimalism. `harmony.eight_layer_vocal_stacking` --- ### Resonant Bass-Baritone Vocal Anchor **Also called:** sub-vocal anchor, chest bass foundation, Russian bass octave **What it is:** Adding a deep bass-baritone vocalist singing the fundamental root note an octave below the main melody in the center channel. **Effect on the audience:** Gives the music physical weight, grounding, and imperial authority; rattles the chest cavity of the listener. **Used for and where it works best:** Sovereign decrees, military marches, and ancient imperial temple liturgies. **Best in:** formats: Epic Anthem, Sacred Choral, Historical Drama | genres: Heritage (civilisation-focused), Epic, Classical **Avoid when:** Light, bouncy children's animation songs. **Example:** `arrangement: Lead melody sung by tenor @ C3; Bass-baritone doubles @ C2 with heavy chest resonance.` **Yields to:** `harmony.parallel_thirds_harmony` — Light bouncy children's animation song. `harmony.bass_baritone_vocal_anchor` --- ### Antiphonal Stereo Choir Placement **Also called:** split choir, Venetian polychoral style, stereo call-and-response **What it is:** Panning Choir A 100% Left and Choir B 100% Right, alternating musical phrases across the stereo field before uniting in the center. **Effect on the audience:** Immerses the listener inside a three-dimensional cathedral; sound physically moves back and forth across headphones. **Used for and where it works best:** Sacred choral epics, ancient temple dialogues, and theatrical spectacles. **Best in:** formats: Theatrical Feature, High-End Audio Drama, XR | genres: Classical, Sacred, Epic **Avoid when:** Mono single-speaker mobile phone releases without stereo headphone listening. **Example:** `panning: Choir Left sings question -> Choir Right sings answer -> Combined Choir unites in massive stereo center.` **Yields to:** `harmony.gospel_choral_call_response` — Mono phone release: call-and-response without stereo split. `harmony.antiphonal_stereo_choir` --- ### Microtonal Vocal Quarter-Tone Mordent **Also called:** Eastern vocal trill, quarter-tone bend, Tarab ornament **What it is:** Directing vocalists to perform a rapid mordent (flick) up to a neutral quarter-tone and back to the parent note in under 100ms. **Effect on the audience:** Authenticates authentic Near Eastern, Arabic, and Assyrian vocal mastery; creates spine-tingling soulfulness. **Used for and where it works best:** Vocal performance in Maqam Bayati, Rast, Saba, or Hijaz; how often within a performance -- a constant feature, or concentrated at mawal introductions and climaxes as `harmony.tarab_vocal_ornamentation` frames it -- is attributed, not settled (see Source). Verify the exact mode, regional style and target note before directing it. **Best in:** formats: World Music, Heritage Feature, Folk Single | genres: Heritage (civilisation-focused), Folk, World **Avoid when:** Strict Western baroque choral pieces requiring pure Pythagorean or equal temperament. **Example:** `ornament_directive: "Vocalist applies rapid quarter-tone upper mordent on syllable 'Al-' before landing on long held note."` **Source:** reference — searched 2026-09-23; no academic source located quantifies how often this ornament occurs across a performance. A.J. Racy's *Making Music in the Arab World: The Culture and Artistry of Tarab* (Cambridge University Press, 2003) and Johnny Farraj and Sami Abu Shumays's *Inside Arabic Music* (Oxford University Press, 2019) are the standard academic treatments of tarab ornamentation, but neither gave quotable text today either confirming or ruling out an "every performance" framing. The disagreement with `harmony.tarab_vocal_ornamentation` over how often the ornament is used stays open. **Yields to:** `harmony.antiphonal_stereo_choir` — Strict Western baroque choral, equal temperament. `harmony.microtonal_vocal_mordent` --- ### Vocal Breath Synchronization Across Harmonies **Also called:** choral breath lock, ensemble inhalation **What it is:** Ensuring all backing harmony singers inhale at the exact same millisecond as the lead vocalist before attacking the chord. **Effect on the audience:** Produces a massive, unified punch; eliminates messy, staggered breath noises that sound amateurish. **Used for and where it works best:** Multi-track vocal arrangements in pop, choral music, and musical theatre where the parts attack the chord together. Lines written with independent rhythms (`harmony.contrapuntal_vocal_counterpoint`) share no attack, so each breathes on its own phrase. **Best in:** formats: Pop Chorus, Choral Suite, Anthemic Rock | genres: All **Avoid when:** Intentionally simulating a chaotic, disorganized civilian crowd. **Example:** `editing: Align all 6 vocal stem waveforms so that inhalation transients begin on frame 14 and chord hits cleanly on frame 24.` **Yields to:** `soundscape.bazaar_marketplace` — Deliberately chaotic disorganised civilian crowd. `harmony.vocal_breath_synchronization` --- ### De-Tuned Vocal Chorus for Dream Sequences **Also called:** chorus effect vocal, detuned choir, psychedelic vocal wash **What it is:** Modulating the pitch of backing vocal tracks with a slow LFO (±12 cents) and wide stereo panning to simulate hallucinatory beauty. **Effect on the audience:** Induces a dream-like, floating, mystical altered state of consciousness; un-anchors the listener from reality. **Used for and where it works best:** Dream sequences, mythological visions, drug hallucinations, and celestial encounters. **Best in:** formats: Psychological Film, Dream Sequence, Art Song | genres: Fantasy, Psychological, Sci-Fi **Avoid when:** Factual documentary voiceover where reality must remain grounded. **Example:** `vocal_chain: Backing vocals routed to stereo chorus (rate: 0.8Hz, depth: 35%, 100% wet) panned 90% wide.` `harmony.detuned_vocal_chorus` --- ### Guttural Throat Singing Drone (Kargyraa / Hoomii) **Also called:** overtone singing, sub-harmonic throat drone, ancient growl **What it is:** Producing two pitches simultaneously using false vocal folds (vestibular folds) to generate deep subterranean sub-harmonics. **Effect on the audience:** Evokes prehistoric primordial antiquity, shamanic rituals, and terrifying elemental power. **Used for and where it works best:** Bronze Age rituals, nomad warrior camps, and terrifying mythic encounters. **Best in:** formats: Epic Feature, Mythological Game, Shamanic Soundtrack | genres: Mythology, World, Fantasy **Avoid when:** Bright upbeat romantic comedy songs. **Example:** `style: "Sub-harmonic throat singing drone holding low C1 note, with whistle-like overtones dancing above at 1.8kHz".` **Yields to:** `harmony.parallel_thirds_harmony` — Bright upbeat romantic comedy song. `harmony.guttural_throat_drone` --- ### Call-to-Prayer / Adhan Acoustic Vocal Projection **Also called:** open-air vocal projection, minaret acoustics, mountain call **What it is:** Maximal un-amplified vocal projection utilizing the singer's formant (2.8kHz-3.2kHz) designed to carry across vast open desert valleys. **Effect on the audience:** Pierces through all background sound with breathtaking, soaring majesty and spiritual gravity. **Used for and where it works best:** Desert dawn awakenings, calls to arms, and sacred invocations. **Best in:** formats: Heritage Documentary, Epic Film | genres: Heritage (civilisation-focused), Religion, Epic **Avoid when:** Whispered conversational dialogue. **Example:** `vocal_directive: "High resonant male tenor projection in Maqam Hijaz, long sustained melismatic arches decaying over 3-second open air reverb".` **Yields to:** `suno.whispered_vocal_directing` — Whispered conversational delivery. `harmony.open_air_vocal_projection` --- ### The Naked Vocal Breakdown (The Solo Drop) **Also called:** a cappella drop, solo vocal breakdown, stripped bar **What it is:** Suddenly cutting all instruments, drums, and bass for exactly 1 or 2 bars, leaving the vocal melody completely naked before the final drop. **Effect on the audience:** Freezes the entire room; focuses 100% of human attention on the words before the massive explosive chorus hits. **Used for and where it works best:** The bridge-to-final-chorus transition in pop, rock, and epic anthems. **Best in:** formats: Pop Anthem, Rock Ballad, EDM Fusion | genres: Pop, Rock, Electronic **Avoid when:** Instrumental background wallpaper music. **Example:** `arrangement: Full band cuts out -> Lead singer delivers line alone: 'We will rebuild the gates' -> [DRUM EXPLOSION] into Chorus 3.` **Yields to:** `hook_psych.sudden_silence_punch_drop` — Instrumental track with no vocal to strip to. `harmony.naked_vocal_breakdown` --- ### Vocal Formant Tracking for Artificial Duets **Also called:** gender formant shift, AI duet generator, timbre pair **What it is:** Shifting vocal formants by -2.5 semitones on Track A (male chest) and +2.5 semitones on Track B (female throat) to create believable duets from a single AI vocal model. **Effect on the audience:** Creates two distinctly colored human characters interacting with perfect musical phrasing. **Used for and where it works best:** Scripting musical duets between kings and queens, or lovers separated by war. **Best in:** formats: Musical Duet, Narrative Opera, Dramatic Single | genres: Musical, Drama, Romance **Avoid when:** Solo tracks with only one intended speaker. **Example:** `processing: Verse 1 (Formant -2 semitones, Baritone) -> Verse 2 (Formant +2 semitones, Mezzo-Soprano) -> Chorus (Harmonized together).` **Yields to:** `harmony.lead_vocal_doubling` — Solo track with only one intended speaker. `harmony.vocal_formant_duet` --- ### The Vocalise (Wordless Emotional Anthem) **Also called:** wordless vocal, soaring vocalise, cinematic vowel crying **What it is:** A singer performing a soaring, heartbreaking melody using only open vowels (/ah/, /oh/) with no written lyrical words. **Effect on the audience:** Transcends all human language and logic; strikes directly into pure, unmediated emotional grief and beauty. **Used for and where it works best:** The emotional climax of tragedy films (e.g. Gladiator's "Now We Are Free" style), memorials, and end credits. **Best in:** formats: Film Soundtrack, Memorial Feature, Epic Trailer | genres: Tragedy, Epic, Classical **Avoid when:** Informational corporate presentations where facts must be spoken. **Example:** `cue: Solo female dramatic soprano singing soaring wordless vocalise in Maqam Bayati over bowed cello and harp.` **Yields to:** `sync.tech_innovation_pulsing_arpeggio` — Informational corporate presentation, facts spoken. `harmony.wordless_vocalise_anthem`
SHA-256: 990413173a415afe3c5e9f2ceaabeff04da049b5b7c2995045122a41361a7c07