Music and Voices

Status: Approved initial tool direction; prompt format remains Candidate

Suno is always the first media-generation stage. After the episode concept, learning goal, and story/song brief are approved, Suno generates both music and narration. No Anima storyboard or Wan shot production begins until a specific Suno version is selected, imported into Palmier, transcribed with Whisper, and converted into a reviewed timing map.

The working pipeline is strongest when a piece is designed for two distinct voices, uses explicit spoken-word tags, and describes the emotional movement of each section.

The selected Suno audio on the Palmier timeline is the episode’s timing and emotional spine. Whisper provides the timecoded transcript used to place shot boundaries. The storyboard interprets that audio; the audio is not reshaped to justify an already-generated storyboard.

Voice functions

Voice Dramatic job
Lumi Ground the idea, model listening, carry reassurance and wonder
Pip Create momentum, ask the active question, embody surprise and play
Together Deliver the hook and the shared discovery

The functions can swap for story reasons, but the voices should never feel interchangeable.

Song brief structure

song:
  episode: S01E__
  letter: ""
  duration_target: ""
  emotional_arc: curious -> uncertain -> delighted -> warm
  spoken_word:
    opening: "[Lumi, softly spoken]"
    pivot: "[Pip, excited spoken interjection]"
  sections:
    - intro
    - call_and_response_verse
    - chorus
    - discovery_bridge
    - final_chorus
    - gentle_button
  hook: ""
  anchor_words: []
  generation_record:
    tool: Suno
    output_mode: integrated_music_and_narration
    prompt_version: ""
    output_link: ""
    status: Draft
  audio_lock:
    selected_version: ""
    duration: ""
    palmier_project: ""
    palmier_audio_track: ""
    whisper_transcript: ""
    whisper_timing_map: ""
    locked_on: YYYY-MM-DD

For episodes where control is stronger with separate generations, preserve distinct music_output and narration_output records and document how they are combined in Palmier.

Rules

  • Give each voice an intention and feeling, not just a name.
  • Use spoken tags where story clarity matters.
  • Let the chorus express the discovery, not merely repeat the letter.
  • Write visual beats and song structure together.
  • Preserve generation prompts and selected output links.
  • Import the chosen Suno version into Palmier before Anima storyboarding.
  • Run Whisper on the exact imported song and preserve the timecoded transcript.
  • Review Whisper’s words against the approved lyrics while retaining verified timing; sung words can require text correction.
  • Lock the reviewed transcript and shot timing map before Anima storyboarding.
  • Treat any later audio change as a dependency change that may invalidate the storyboard, Wan shots, and Palmier timeline.