Editing to Sound

Sound, Score and Direction Vocabulary

Overview

Editing picture to sound and acquiring the vocabulary of camera direction. Students generate a music track to a chosen mood, cut the picture to its beat, and chain a text model into an audio model to produce narration and score.

  • 01 By the end, you can Choose or build the track first and let it set the trailer section’s duration and pace.
  • 02 By the end, you can Direct an audio model to generate music that matches an intended mood.
  • 03 By the end, you can Cut the picture so changes land on the beat, counting BPM by hand to line the words up.
  • 04 By the end, you can Combine both, keeping AI-made sound on its own track for a clear mix, and assign every task to a named crew role.
  1. Matching Music To Scenes
    Applied craft
    • Demo One Scene Two Scores Play the same clip under two different music tracks so students hear how the sound changes the scene.
    • Game Guess The Music Drops Watch three silent trailers and mark where you think the music starts, stops, and hits, then check against the real sound.
    • Role-play Emotion And Genre Scene Groups draw an emotion and genre, get a music clip, and act out a short scene that matches it.
  2. Prompting Suno And ElevenLabs
    Tool teaching
    • Demo Change One Word The teacher prompts Suno on the screen and changes just one word each time, like the mood, the tempo, or the instrument, so the class hears how each word changes the track.
    • Worksheet Prompt Recipe Card Students fill in slots for mood, genre, instruments, and beats per minute to draft an audio prompt on paper before they type it into Suno.
    • Build Make Your Trailer Track Each team generates a music track in Suno and a matching sound effect or narration line in ElevenLabs for their scene, then keeps the best version.
  3. Counting Beats And Cutting To Them
    Applied craft
    • Quiz Name The Move Watch clips shot by shot and have the class name the camera move and the pace of each shot.
    • Demo Cut On The Beat The teacher edits a clip live and talks through each cut, then teams talk through the choices on their own rough cut.
  4. Crew Roles And Task List
    Deliverable prep
    • Role-play Crew Role Draft Students pick or draw director, editor, sound engineer, or producer and read a card out loud that lists what that job does on this trailer.
    • Worksheet Job Card Task List Each role fills in a card listing their own named tasks for the trailer and writes down who they hand each finished task to next.
    • Discussion Daily Stand Up Each crew holds a two-minute standing meeting where every person says what they finished, what they will do next, and what is blocking them.

Applied · Film

# Knowledge & Skills

ACA-Applied
K Knowledge · what you know
The music track sets the length and speed of the scene before any clip is made, so film is shot longer than needed and then cut down to match the track.
Pace words say how fast a shot moves, for example very slow (glacial), creeping, slow and steady, one sudden movement, speeding up, and violent and sudden.
AI-made sound needs its own separate track, apart from your own recording, because mixing them together makes the sound unclear.
S Skills · what you can do
Make a piece of music that matches a mood, then cut it together with the picture. point w2d3-s1-p1; evidence w2d3-e3
Split the crew into director, editor, sound engineer and producer, and give each person named tasks. point w2d3-s3-p1; evidence w2d3-e1
Layer AI-made sound over your own recording and keep the mix clear.
T Techniques · named subskills
Chained Generation One model writes the audio prompt, a second model executes it; revise by editing the prompt, not regenerating blind.
Track Sets the Cut Duration and pace are fixed by the audio track before any visual clip is generated.

AI

Aligned with AI4K12 · Five Big Ideas in AI View framework →
AI4K12 · Five Big Ideas in AI

Sync Lyrics by BPM

To line up spoken words or song lyrics with the music, count the track's beats per minute by hand, give that number and the words to the AI model, then fix the timing by hand.

AI4K12 · Five Big Ideas in AI

Chain Text Into Sound

Chained generation means one AI model (a text model) writes the sound instructions, and a second AI model (an audio model) turns those instructions into sound, and you make changes by editing the instructions rather than starting over.

AI4K12 · Five Big Ideas in AI

Ten Reliable Camera Words

There are ten camera words the AI model reliably understands, including move in and out (dolly), turn side to side (pan), tilt up and down, move on a crane, and a still shot (locked off).

Aligned with UNESCO · AI Competency Framework for Students View framework →
AI Techniques and Applications
4.3.3

Creating AI Tools

Developing advanced technical creativity to design, build, and innovate with AI tools and methods.

Aligned with Derived from course activity
Derived from course activity

Structured Audio-Direction Prompting

Writing a role-based, constraint-bound prompt (role, duration, mood) to direct AI audio/voice generation. Evidence: "You are a sound engineer. make music for a horror movie trailer. It must be 60 [seconds]..." (editing-to-sound).

# Competency Goals

ACA-AI.Goals
Aligned with UNESCO · AI Competency Framework for Students View framework →
UNESCO · AI Competency Framework for Students
CG4.3.3.2

Creativity in Applying AI

Innovatively combining and adapting AI methods to address novel problems and create new solutions.

UNESCO · AI Competency Framework for Students

Write Camera Directions

Write camera directions in words the AI model can follow: where the camera is, how it moves, what it focuses on, the shot type, and what the background is doing.

UNESCO · AI Competency Framework for Students

Measure BPM by Hand

Count a track's beats per minute (how many beats happen in one minute) by hand, then give that number to the model so the sound lines up.

Deliverable

# Editing to Sound

Projects.Briefs

Make a section of the trailer with music, where the picture is cut to match an AI-generated music track. A piece of music made to match a mood and cut together with the picture, plus a task list where every job has one named person responsible.

Transfers to: Directing a generative tool in the exact language it understands.

See past cohorts ship this in the Showcase →

# Assessed toward these aims

Projects.Rubrics

This deliverable is where students demonstrate

  1. C Produce narration and score, and acquire the vocabulary of camera direction
    Look out for
    • Music matches mood: The music matches the intended mood, and the picture is cut so changes land on the beat.
    • Camera direction: Camera moves are written in words the AI model can read: where the camera is, how it moves, what it focuses on, the shot type, and the background.
    • Sound timing: Beats per minute are counted by hand and given to the AI model, so spoken words or lyrics line up. Any remaining timing gaps are fixed by hand.
    • Clear mix: AI-made sound sits on its own track, separate from the team's own sound, so the mix stays clear.
    • Task ownership: Every task still to be done is given to one named crew member.