Mixes almost always fail the same way: music laid at full level, speech fighting to be heard, and no hours left to draw volume keyframes by hand. Ask for ducking and Sparki rides the music level against the voice — down while someone speaks, up between lines, gone cleanly at the end.
The volume relationship between voice and music, handled for you.
Sparki finds the speech — recorded narration, a generated voiceover, someone talking on camera — and lowers the music beneath it for exactly as long as the sentence runs. No keyframes to draw, no decibel guessing: the level follows the voice, and a phrase that runs long keeps its headroom the whole way.

Between sentences the bed breathes again, so the edit keeps its energy instead of flattening. When the speech stops for good, the music comes up for the outro or fades out where you say. The result sounds like someone rode the fader all the way through — someone did, just not you.

Ducking is one part of the mix. Sparki also keeps effects, ambience and original audio in their lane — speech on top, music beneath, everything else out of the way — and normalises the export so the balance survives a phone speaker, not just the machine that made it.

Six mixes that go wrong without it.
Explainers, ads and faceless-channel videos where a narration track plays over continuous music. Sparki dips the bed under every narration section and restores it between them, so the read stays forward without losing the momentum the music provides. Any video where the words are the content and the music is the mood.
Gaming highlights and stream recuts where the commentator talks over loud in-game sound. Sparki holds the commentary on top and pulls the game audio down while the voice runs, letting the play's own sound back in during the gaps. Clips where both the reaction and the action need to be heard.
Cooking, fitness and DIY videos where spoken steps sit on top of a track chosen for energy. Sparki keeps the bed low through each instruction and lifts it in the quiet stretches between steps, so the pace never dies mid-recipe. Step-by-step content filmed to music.
Cutdown interviews and testimonial edits with music underneath the conversation. Sparki rides the bed under long answers without pumping, and stays out of the way when the room tone carries a pause. Testimonials and case-study videos with underscore.
Podcast clips, duets and dialogue edits where several voices share one music bed. Sparki follows whoever is speaking and keeps the bed clear of both voices, so handoffs between speakers never leave the music louder than the next person. Conversation formats cut for social.
The final seconds of promos and episodes, where the call to action ends and the music should close the video. Sparki lets the bed swell once the speech finishes and fades it on the end card, instead of chopping the track off mid-phrase. Endings that feel finished rather than truncated.
Four people whose voice keeps losing to the music.
Cuts podcast moments for social with the show's theme track underneath. Sparki ducks the theme under every spoken line and brings it back between clips, so the hook is audible in a feed where most viewers never turn the sound up twice.
Pulls highlights from VODs where the commentator drowns in their own game audio. Sparki drops the capture under the reaction, restores it for the play and exports a clip where both the call and the clutch survive.
Films to music because the kitchen audio is unusable. Sparki keeps the track beneath every instruction and lifts it while the oven does the work, which is what separates a followable recipe from a merely pretty one.
Delivers dozens of reels a month where a spoken offer sits over licensed music. Sparki applies the same voice-forward balance to every client cut, so nothing ships with the offer buried and no afternoon is lost drawing volume curves.
Upload, say what stays on top, export the balanced mix.
Send the file or paste a link. Videos with narration, on-camera speech, a generated voiceover or a baked-in track all work — Sparki sorts the layers before touching levels.
Name the layer the viewer must not miss — the voiceover, the commentary, the instructions. Sparki lowers the music under that layer's speech and brings it back in the gaps, with fades where the track starts or ends.
Play the mix, ask for a deeper dip under a specific section or more music in the quiet parts, then export with the balance normalised for platform playback.
Ducking is one mix control. Music selection, sound effects and captions share the same upload.
Upload the video and name the layer that matters most — Sparki ducks the music under it, restores it between lines and exports a mix that holds up everywhere.
Duck your music automatically