Modern Creator
Gavin Herman · YouTube

How To Sound Design Any Video In 15 Minutes

A ten-year editor breaks his three-pass audio workflow — vocals, sound effects, music — into checklists anyone can run without becoming an audio engineer.

Posted
1 months ago
Duration
Format
Tutorial
educational
Views
49.5K
3.3K likes
Big Idea

The argument in one line.

Professional-sounding video comes from three sequential audio passes — cleaning the voice, designing sound effects with a five-step process, and fitting music around both — not from better cameras, color grading, or being a trained audio engineer.

Who This Is For

Read if. Skip if.

READ IF YOU ARE…
  • You already edit video regularly and can cut and color grade competently, but your finished videos still feel amateur for a reason you can't pinpoint.
  • You want a repeatable audio workflow you can run in Premiere, DaVinci Resolve, or Final Cut without formal audio training.
  • You currently skip sound effects and ambience almost entirely and want a structured way to start layering them in.
SKIP IF…
  • You're looking for a stock-music or sound-library comparison rather than a technique breakdown — the video treats one sponsor's library (Epidemic Sound) as the default source throughout.
  • You need help with spoken-word or podcast mixing specifically — this is built around narrative/talking-head video sound (vocals plus SFX plus music), not audio-only production.
TL;DR

The full version, fast.

Great-sounding video comes from three sequential passes, not talent or better gear. First, clean the voice with the same five-effect chain every time: noise reduction, EQ, compression, de-essing, and a limiter. Second, design sound effects with a five-step system called SHAPE — source a sound, hone its pitch and length, amplify it by layering frequencies, place it exactly on the visual's peak moment, and fill the environment with subtle diegetic ambience. Third, choose music by emotional intent rather than genre, then EQ out the vocal-range frequencies so it doesn't compete with the voice. Do all three passes and the same footage plays ten times more polished.

Free for members

Chat with this breakdown — free.

Sign in and you get 23 free chat messages on us — ask for the hook, quote a framework, find the exact transcript moment, generate a markdown action plan. Bring your own key when you want unlimited.

Create a free account →
Chapters

Where the time goes.

00:0000:32

01 · Why your edit still feels off

Cold open diagnosing the real problem with edits that feel unfinished: it's the audio, not the picture.

00:3200:59

02 · PASS 1: Creating clean vocals

Introduces the first of three passes and previews the five-effect vocal chain.

00:5901:24

03 · Remove background noise

Denoise effect plus a noise gate to remove background hiss and silence gaps between words.

01:2401:59

04 · Using EQ for clarity and cleanup

Cutting frequencies outside the 80 Hz–12 kHz vocal range to remove rumble and hiss.

01:5902:12

05 · Shaping vocals with EQ

A second EQ pass with a vocal-enhancer preset adds warmth, then gets dialed back for subtlety.

02:1203:01

06 · Level audio loudness with compression

Setting compressor threshold at the voice's quietest point, 2:1 ratio, then makeup gain to even out volume swings.

03:0103:28

07 · Reduce harsh S sounds and limit peaks

De-esser softens sibilance; a limiter caps the loudest peaks as a final safety net.

03:2803:40

08 · PASS 2: Immersive sound effects

Introduces the SHAPE system (Source, Hone, Amplify, Place, Environment) for sound effects.

03:4005:01

09 · My go-to source for sfx and music

Sourcing sounds via Epidemic Sound's AI-searchable plugin instead of manually digging through folders.

05:0105:56

10 · Manipulating sounds to work in every situation

Pitch-shifting and time-stretching a raw sound to generate variations from one source file.

05:5606:15

11 · Turn one sound into three

Cutting a whoosh at its inflection point makes an impact; reversing that cut makes a riser.

06:1507:00

12 · The power of frequency layering

Layering low- and high-pitched copies of the same sound so it fills the frequency spectrum instead of sounding thin.

07:0007:34

13 · Timing, timing, timing.

Aligning a sound effect's loudest waveform peak with the exact frame of the visual's peak motion.

07:3408:42

14 · Build an environment with sound

Layering subtle diegetic ambience (birds, wind, traffic) beneath dialogue to make a location feel real.

08:4209:56

15 · PASS 3: Choose the right music

Choosing a track by the emotional story being told rather than by genre or aimless browsing.

09:5610:29

16 · Make music fit around the voice

EQing out vocal-range frequencies from the music track instead of just lowering its volume under dialogue.

10:2911:06

17 · Seamlessly transition between tracks

Reversing the outgoing track's last beat and placing it at the start of the incoming track to create a natural riser.

11:0612:11

18 · Adjust tracks to any length and add reverb

Using an AI remixer to fit track length, then adding reverb to a nested clip ending for a cinematic fade.

12:1112:52

19 · Use submix routing for cohesive music

Routing all music tracks to one submix for a single master volume and shared EQ across every track.

12:5213:44

20 · Bonus trick #1

Layering additional risers, swells, and suckbacks between song transitions, and ending songs on a reverbed impact hit.

13:4414:58

21 · Bonus trick #2

Keyframing stereo pan so a sound effect physically travels across the speakers with the on-screen subject's movement.

Atomic Insights

Lines worth screenshotting.

  • A video can have clean cuts, dialed color, and perfect pacing and still feel wrong to a viewer — because the actual problem is unfinished audio, not the picture.
  • A repeatable five-effect vocal chain — noise reduction, EQ, compression, de-essing, limiter — applied in the same order every time replaces the need for audio engineering expertise.
  • Human speech sits roughly between 80 Hz and 12 kHz, so cutting everything outside that range on an EQ removes rumble and hiss without touching the voice itself.
  • Compression isn't about making a voice louder — it's about shrinking the gap between its loudest and quietest moments so it never gets buried or blown out.
  • A single raw sound effect can become three different effects: cut it at the inflection point for an impact, reverse that cut for a riser, or leave it whole for the original whoosh.
  • A sound effect played alone often sounds thin because it only occupies one part of the frequency spectrum — layering a lower-pitched copy and a higher-pitched copy underneath and above it makes it feel like it fills the room.
  • The loudest point of a sound effect's waveform should land on the exact frame of the visual's peak motion — a ten-second alignment step editors describe as the difference between amateur and professional sound.
  • Diegetic ambience (wind, birds, traffic, distant voices) that viewers never consciously notice is exactly what they notice the absence of — it's what makes a location feel real instead of staged.
  • Instead of just lowering music volume under dialogue, cutting the mid-frequencies where the human voice sits out of the music track lets the track keep its full-band vibe while the voice still cuts through cleanly.
  • Reversing the last big beat of an outgoing song and placing it at the start of the incoming song creates a natural riser that makes a track transition feel seamless instead of like a hard cut.
  • Routing every music track through a single submix gives one master volume control and lets one EQ move (carving out vocal frequencies) apply to every track at once, instead of repeating it per track.
  • Panning a sound effect's stereo position to match a subject's on-screen movement makes the audio physically travel with the visual, reinforcing motion the eye is already tracking.
Takeaway

Three sequential passes turn a clean edit into an immersive one.

THE WORKFLOW

Professional-feeling video sound comes from running vocals, sound effects, and music through their own repeatable checklists in that order, not from better gear or natural talent.

01Why your edit still feels off
  • A video can have perfect cuts, color, and pacing and still feel wrong to a viewer, because the actual problem is unfinished audio, not the picture.
02PASS 1: Creating clean vocals
  • Cleaning dialogue is the first of three sequential passes, and skipping straight to music or sound effects without fixing the voice first undermines everything layered on top of it.
03Remove background noise
  • A subtle noise-reduction effect removes ambient hiss without making a voice sound processed, and a gate then mutes anything below a set decibel threshold so silence between words stays truly silent.
04Using EQ for clarity and cleanup
  • Human speech sits roughly between 80 Hz and 12 kHz, so cutting everything below and above that range on an EQ removes rumble and hiss the ear doesn't need.
05Shaping vocals with EQ
  • A second, shaping EQ pass can add warmth or presence to a voice's natural tone, then gets dialed back in intensity so it reads as tone rather than obvious processing.
06Level audio loudness with compression
  • Compression evens out a speaker's natural volume swings: set the threshold near the quietest point the voice hits, use roughly a 2:1 ratio, shorten attack and release, then add makeup gain to restore overall level.
  • The goal of compression isn't to make audio louder, it's to shrink the gap between the loudest and quietest moments so the voice never gets buried or blown out.
07Reduce harsh S sounds and limit peaks
  • A de-esser tames sharp 's' and 't' sounds that broader EQ moves can't fix without dulling the whole voice, and a limiter caps the absolute loudest peaks as a final safety net.
08PASS 2: Immersive sound effects
  • Sound effects are what make a viewer feel physically inside a scene rather than just watching it, the difference between a video that's edited and one that's designed.
09My go-to source for sfx and music
  • An AI-searchable sound library lets you type a description like 'aggressive impact' or 'soft whoosh' and get matching samples instantly, replacing what used to be twenty minutes of digging through folders.
10Manipulating sounds to work in every situation
  • A five-step system, Source, Hone, Amplify, Place, Environment, governs every sound effect, turning ad hoc sound-dropping into a repeatable checklist.
  • Pitch-shifting and time-stretching a single raw sound effect can generate multiple distinct-feeling variations from one source file.
11Turn one sound into three
  • Cutting a whoosh at its inflection point turns it into an impact, and reversing that same cut produces a riser, three usable effects mined from one source clip.
12The power of frequency layering
  • A single sound effect played alone often sounds thin because it only occupies one part of the frequency spectrum; layering a lower-pitched copy and a higher-pitched copy underneath and above it makes it feel like it fills the room.
13Timing, timing, timing.
  • The loudest point of a sound effect's waveform should line up exactly with the peak of the visual motion it's paired with, a ten-second alignment step that separates amateur sound from professional sound.
14Build an environment with sound
  • Diegetic ambience that would realistically exist in a scene but isn't shown on camera is what makes a location feel alive rather than staged; viewers rarely notice it consciously but they notice its absence.
  • Environmental sound should sit low in the mix, underneath dialogue and primary effects, aiming for a felt sense of place rather than something consciously heard.
15PASS 3: Choose the right music
  • Music choice should start with the emotional story being told, not with browsing a library aimlessly, searching by the feeling a moment needs rather than by genre.
16Make music fit around the voice
  • Cutting the mid-frequencies where the human voice sits out of a music track with an EQ lets the track keep its full-band vibe while the voice still cuts through cleanly, a more professional result than a simple volume dip.
17Seamlessly transition between tracks
  • Reversing the last big beat of an outgoing track and placing it at the start of the incoming track creates a natural riser, then a quick volume crossfade makes the splice inaudible.
18Adjust tracks to any length and add reverb
  • An AI music remixer can stretch or shrink a track to match a video's exact length in seconds when the mood fits but the timing doesn't.
  • Adding reverb with increased decay and wetness to the tail of a music clip gives endings a cinematic, trailing-off quality, and nesting just that ending in its own sequence lets the reverb apply without affecting the rest of the track.
19Use submix routing for cohesive music
  • Routing every music track through a single submix gives one master volume fader and lets one EQ move apply to every track at once instead of repeating it individually.
20Bonus trick #1
  • Beyond a reversed-beat riser, additional risers, swells, or suckbacks can further smooth song transitions, and the same effects reversed can ease a scene out of music into silence to emphasize a quiet moment.
  • Ending a song on a dropped impact or hit with reverb applied creates a more dramatic, deliberate conclusion than simply letting a track fade out.
21Bonus trick #2
  • Panning a sound effect's stereo position to match a subject's on-screen movement makes the audio physically travel with the visual, reinforcing motion the eye is already tracking.
Glossary

Terms worth knowing.

Noise gate
An effect that mutes audio below a set volume threshold, commonly used to silence background noise in the gaps between spoken words.
De-esser
An audio effect that specifically reduces harsh sibilant sounds, like sharp 's' and 't' sounds, in a vocal recording.
Diegetic sound
Sound that would realistically exist within the scene being shown, such as wind or traffic, even though it isn't the main dialogue or a deliberate sound effect.
Submix
A single audio bus that several tracks are routed through, giving them one shared volume control and one shared set of effects instead of adjusting each track individually.
Riser
A rising sound effect used to build tension or smooth a transition into a new scene or piece of music, often created by reversing part of another sound.
Compressor threshold and ratio
Compressor settings that control the volume level a signal must cross before it gets reduced (threshold) and how much reduction is applied once it does (ratio).
High-pass / low-pass filter
Filters that remove everything below a chosen frequency (high-pass) or above it (low-pass), used to cut unwanted low rumble or high-end hiss.
Resources

Things they pointed at.

Quotables

Lines you could clip.

00:06
It's the audio. It's always the audio.
tight, punchy standalone thesis line with no setup neededTikTok hook↗ Tweet quote
05:56
A whoosh isn't just a whoosh.
playful wordplay that teases a concrete techniqueIG reel cold open↗ Tweet quote
03:29
Sound design is the other half of every great video.
recurring thesis line reused throughout the video as both narration and literal demo audionewsletter pull-quote↗ Tweet quote
The Script

Word for word.

Read-along

Don't just watch it. Burn it in.

See every word as it's spoken — crank it to 2× and still catch all of it. The same dual-channel trick behind Amazon's Kindle + Audible.

You spent hours editing this video. The cuts are clean. The color is dialed.
The pacing feels right. But something's off, and you can't figure out why. It's the audio.
It's always the audio. Sound design is the other half of every great video. It immerses viewers in your world.
It makes them feel some type of way. And without it, even the best edit falls flat. I've been editing for ten years now and I'm going to walk you through the three audio passes I run on every single video.
Starting with our foundation, the human voice. If your vocals sound bad, the video will be less engaging.
Simple as that. Luckily, we don't have to be audio experts for this. It's actually as simple as five effects in the same order every time.
I'm going to show this in Adobe Premiere, but every effect I'm covering exists in DaVinci, Final Cut, whatever you use. So here's the vocal chain I use on every video. First, we want to remove background noise from the vocals.
The simplest fix here is a denoise or noise reduction effect, which I set to a subtle amount and it works pretty well.
To completely remove all audio below a certain loudness level, you can use a gate and set a decibel threshold to remove any noise below that specific audio level.
The cuts are clean.
The cuts are clean. Next, we add some EQ, which allows us to adjust the volume of specific frequency ranges independently.
Generally, the human voice sits between eighty and twelve thousand hertz, so we can add an equalizer effect and reduce low frequencies below 80 and reduce high frequencies above 12 to around 16 k. And by doing this, the equalizer adds clarity to the voice and removes unnecessary noise.
The cuts are clean. The color is dialed.
The cuts are clean. The color is dialed. I then shape the voice with another equalizer.
Here, I'm using the parametric EQ starting with the vocal enhancer preset, which gives my deeper voice a bit more warmth. The cuts are clean. The color is dialed.
The cuts are clean. The color is dialed. And then I dial it back to sound a little more subtle.
Next, we level the audio with a compressor because when you record someone speaking, their volume fluctuates up and down. But compression makes it more consistent. I use a tube modeled compressor simply finding the lowest decibel level your voice hits.
So here it's about negative 20, negative 21, and set your threshold there. Nothing happens yet because the ratio is one to one, but set the ratio to two to one, drop the attack and release, and then add some output gain to bring the overall levels back up.
And this flattens the loud spikes so that your voice stays a little more consistent and doesn't get too loud or too quiet. The cuts are clean. The color is dialed.
The cuts are clean. The color is dialed. After this, I like to use a de esser, which makes the s's sound less harsh in your voice.
So before, s sounds are super sharp. After, s sounds are less sharp. Finally, we add a limiter, which limits how loud the audio can get so that you don't completely destroy your viewers hearing.
So with clean vocals locked in now, we've built a strong connection between the narrator and the viewer. But to take this to the next level, the viewer needs to feel like they are inside the world that you are showing them.
And that's where pass two comes in, sound effects. Every sound effect that makes it into my videos goes through the same five steps that I call shape. First, we source our sounds.
People often ask me where I get my music and sound effects. And if you've watched any of my videos, it's probably using Epidemic Sound, who I've also partnered with on this video. Now not only do you get access to a huge sound library, but the quality is what keeps me coming back.
Every sound is crisp, and the music, well, they've got some bangers. Plus, it's all restriction free, so you don't have to worry about copyright claims.
But what will revolutionize your workflow is their plugin because it works directly inside Premiere or Resolve. Before this, you would spend twenty minutes searching through folders for the right sound. Now you just search from inside the plugin and it's an AI enhanced search.
So I can type aggressive impact or soft whoosh and get samples that are spot on.
I mean, that's perfect. And if you don't want to spend time sourcing sounds at all, they have a new feature called studio that automatically adds a soundtrack to your video for you.
It's still in beta, but it's improving fast, and it gives you a solid starting point for your sound designers. Editing this video. The cuts are clean.
So if any of this sounds helpful, there's a link in the description below to check out Epidemic Sound. And if you use code Gavin at checkout, you can get 50% off your first two months, which is live until August 20. So now that we have our raw sounds, we hone them.
Meaning, we manipulate them into exactly what the moment needs. So here are a few tricks that I use constantly. First, we've got the pitch shifter.
This allows you to drop a sound's pitch to make it deeper or push it up to make it brighter and snappier. Another way to adjust pitch is adjusting the speed of the sound.
So if we slow a sound down, it will pitch it down. But not only that, it also makes the sound last longer and sound slower.
Of course, we can also speed up a sound to pitch it up and make it sharper and faster. High pass and low pass filters are handy for adjusting frequencies. High pass will cut everything below a certain frequency, and low pass cuts everything above.
And get creative with what you already have. A whoosh isn't just a whoosh. If you cut it on the inflection point, you've got an impact.
Reverse that and you've got a riser.
So from one sound, you get three completely different uses.
So honing transforms individual sounds, but amplifying makes those sounds feel bigger. And the trick is layering sounds on top of each other.
But not only that, we also want to layer across different frequencies because this helps fill the space. So for example, take a whoosh by itself. It sounds thin because it's only hitting one part of the frequency spectrum.
But if I take that same whoosh and drop a copy down low or a lower frequency and push another copy up high for a higher frequency, then suddenly the sound fills the room. And you can do this with one sound using a pitch shifter or you can layer different sounds of varying frequencies.
So we want to stack our layers in a way that feels three-dimensional instead of flat. Place is where your sound needs to be placed in order to be most effective.
So my process here is really simple. Let's say you have a camera shake. The loudest moment of your whoosh should hit the exact frame where the shake is most violent.
So drop the sound in roughly where it belongs, then find the peak of the visual motion, find the peak of the audio waveform, and drag to line them up. It takes ten seconds and it's the difference between pro and amateur sound.
The environment of a scene is full of sounds you don't see visually, but something would feel missing if they were gone. And this is what separates a scene that feels real from one that feels staged. So for an outdoor shot like this, we could add birds chirping, wind rustling through the trees, and maybe a stream of water.
And that's already much better. In a city, should be kind of the opposite. Layer in traffic, some distant voices, and an ambient drone that shows that the place feels alive.
And these sounds that exist in your environment are diegetic sounds. Sounds that would actually exist in the world that you're showing. And having these is something that viewers generally don't notice when they are there, but if they're missing, then the audience knows.
Now when you're adding these, add them subtly, make them low volume underneath your main sound effects and dialogue because usually you're not trying to make them heard, you're just trying to make the space feel a bit more real. And like we talked about before, if we have a voice speaking, we can also cut out certain frequencies to make the sound effects a little bit more subtle.
Starting with our foundation, the human voice. Starting with our foundation, the human voice.
So that's shape. We source, we hone, we amplify, we place, and we fix the environment. Now let's get on to pass three, music.
I think about music in three parts. First is choosing the right track, and this starts with the story that you're trying to tell. Because music carries emotion before any words are said, so if a track fights your story, then your edits just won't land.
So for example, if you're telling a fond childhood memory, maybe look for nostalgic or dreamy vibes. A search term you could use for this is ambient synth pads.
Wait. Did you get that? And that just fits really well.
For your typical documentary or educational content, the pros reach for marimbas or pizzicato strings. And there's a reason for that.
Those instruments sit in a frequency range that doesn't compete with voiceovers. And they're also emotionally neutral, which keeps the viewer focused on what's being said. So before you even start looking for music, ask what the vibe is that you need to convey, ask what the viewer should feel, and then search around that vibe instead of browsing aimlessly.
Once you have tracks that fit the vibe, you need to make them fit your video. And here are three techniques that I use to do that. So remember how the human voice sits in the mid frequencies?
Well, most editors including the old me would just lower the volume of the music when someone's talking. They're half of every great video. But a better move in addition to lowering the volume is to add an equalizer to the music track and cut the mid frequencies where the voice sits.
Sound design is the other half of every great video. It immerses viewers in your world. And now the music keeps its vibe, but the voice cuts through clean.
And it sounds way more professional than just lowering the volume. The second technique is to transition between tracks. Now putting two songs back to back almost never sounded good for me before I learned this.
So here's that trick. To make it feel seamless, find the last big beat of your track, cut it, and reverse it. And then drag that reversed beat to the start of your track, and you get a natural riser that leads perfectly into the new song.
But something's off. Then key frame a quick volume fade between the two, and the cut becomes invisible. The color is dialed.
The pacing feels right, but something's off. Just like that. The third technique is to adjust your track length.
So when a track is too long or too short for your visual, the fast way is to just use the AI music remixer to do it for you. It's not perfect, but it does the job for you and it takes like five seconds. And finally, the polish.
The final touches that make everything feel cohesive. So one thing I really like using is reverb for dramatic endings in your world. So if you want a track to fade out with that cinematic echo, then you can add a studio reverb to the end of the clip and crank up the decay and the wetness.
Viewers in your world.
It immerses viewers in your world.
And if you want this effect on a single clip without affecting the whole track, can just cut the end of the song, nest it in its own sequence, extend the nested clip with an adjustment layer, and then apply the reverb there.
And crank up the decay and wet. Viewers in your world.
The last thing that I sometimes do that allows you to control all of your different music tracks at once is to use submix routing. So once all your music is laid in, you can route every music track to a single submix track, and this gives you one master volume control for all your music. Plus you can apply any effects you want to all of those different music tracks at once.
So we could add an equalizer that carves out the vocal space across all of the tracks, and it's way faster than EQ ing each individual track. And the mix sounds more cohesive because every track is being treated the same way. So the final mix.
Once you've done all three passes, vocals, sound effects, and music, your video is 10 times better than it was before. But there are two bonus tricks I want to show you. So the first is to layer sound effects between song transitions.
Although we already covered creating natural risers from the song itself, we can also use other risers, swells, or suckbacks to further ease into a song. Sound design is the other half of every great video.
It immerses viewers, and we can use the same sounds in reverse to get out of a song into silence. Sound design is the other half of every great video. It immerses viewers in your world.
And this is especially powerful if you want to emphasize a quiet moment. And to wrap a song up impactfully, we can drop an impact, slam, or a hit on the final beat, and make sure that reverb is applied, and we get a really dramatic effect here.
The second trick is panning, and this one's really cool. So when your visual movement goes from one side of the screen to the other, you can use panning to make the sound move with it. Now in Premiere, if you set the pan to negative 100, that means the sound will be on the left side.
It'll come out of the left side of my speakers, the left side of your headset, and when it's at zero, then it's in the middle. It comes out of both sides. And then at plus 100, it comes out on the right.
So if we set a keyframe at negative 100 when the subject is on the left, and then a keyframe at 100 when it's on the right, now the audio physically travels across the viewers ears the same way the visual travels across their eyes.
So that's it. Sound design is one of the things that most editors completely neglect, which means putting in the work here is the quickest way to make your videos stand out. And again, Epidemic Sound makes 90% of what we covered today possible.
So if you want to check it out, you can use the link in the description below. And don't forget to use code Gavin to get 50% off your first two months before August 20. And if you wanna take your editing one step further, I made a video on what separates videos that people can't stop watching and videos that people click off of in seconds.
And you can check that out right here.
The Hook

The bait, then the rug-pull.

The video opens on a diagnosis, not a promise: your edit can be flawless and still feel wrong, because the actual problem is almost always unfinished audio. From there a ten-year editor walks through the exact three passes — vocals, sound effects, music — he runs on every video he makes.

Frameworks

Named ideas worth stealing.

00:13list

3 Audio Passes

  1. Vocals
  2. Sound Effects
  3. Music

The top-level structure of the entire video: three sequential passes run on every project, in this order.

Steal forany video-editing workflow checklist
03:11list

Vocal Chain

  1. Noise Reduction
  2. EQ
  3. Compression
  4. De-essing
  5. Limiter

A five-effect chain applied to raw dialogue in the same order on every single video.

Steal forcleaning any talking-head voiceover or podcast vocal
03:40acronym

SHAPE (sound effect workflow)

  1. Source
  2. Hone
  3. Amplify
  4. Place
  5. Environment

A five-step system for turning a raw stock sound effect into a polished, well-placed one.

Steal forany project that uses stock sound effects
08:42list

Music in Three Parts

  1. Choose
  2. Fit
  3. Polish

A framework for picking a music track by emotion, fitting it to the edit, then polishing it with reverb and routing.

Steal forscoring any narrative or talking-head video
CTA Breakdown

How they asked for the click.

VERBAL ASK
14:40link
So if you want to check it out, you can use the link in the description below. And don't forget to use code Gavin to get 50% off your first two months before August 20.

The sponsor (Epidemic Sound) is woven into the content itself — its plugin is demoed live as the 'source' step of the SHAPE framework — then reinforced with an explicit mid-video and end-video link-and-code CTA, so the pitch doubles as a product demo rather than a bolted-on ad read.

FROM THE DESCRIPTION
PRIMARY CTAWhere the creator wants you to go next.
OTHER LINKSAlso linked in the description.
Storyboard

Visual structure at a glance.

cold open
hookcold open00:00
vocal chain
valuevocal chain03:11
SHAPE: source
valueSHAPE: source03:56
frequency layering
valuefrequency layering06:15
choose music
valuechoose music08:42
sponsor CTA
ctasponsor CTA14:25
Frame Gallery

Visual moments.

Watch next

More from this channel + related breakdowns.

15:15
Andy Diep · Tutorial

I Trained Claude to Place B-Roll in DaVinci Resolve

A content director runs 336 unsorted vlog clips through a Claude Code + DaVinci Resolve Studio pipeline that classifies A-roll from B-roll, proposes cutaway placements against four editorial rules, and drops the picks onto a real timeline — then shows exactly where it still needs a human.

August 5th