Voice-over music bed
Prompt: Keep speech in front while the music breathes between lines.
A playable ducking demo source pair.
automatic audio ducking
Add a voice track and a music bed, then continue to the Sonilo audio-ducking API workflow for automated speech-first mixing.
Duration
Match foreground and background assets
Prompt
Optional voice priority, mix notes, and output format
Model version
Sonilo audio ducking
Input format
Voice audio/video plus music audio
Upload or reference voice and music assets, then use automatic ducking so speech remains clear over the background.
Examples
Prompt: Keep speech in front while the music breathes between lines.
A playable ducking demo source pair.
Prompt: Lower the bed under narration, then restore energy during the reveal.
A commercial clip context for speech-first music balancing.
Prompt: Keep the background audible but out of the way during explanation.
A video context for automatic ducking around spoken guidance.
Compare
The source clip keeps its original camera audio for reference.
The same clip muted, so timing and visual beats are easy to judge.
A finished video mix where music is balanced against speech and picture.
Use cases
FAQ
Automatic audio ducking lowers background music when speech is present, then raises it between phrases so narration stays clear.
No. Audio ducking mixes existing voice and music assets. Use text-to-music or video-to-music to generate a new soundtrack.
The Sonilo API supports foreground speech or dialogue inputs and background music inputs for an async ducking workflow.
Use POST /v1/audio-ducking and poll the returned task until the mixed output is ready.
Use Sonilo audio ducking when you already have voice and music and need an automatic mix.