Voice-over music bed
Prompt: Keep speech in front while the music breathes between lines.
A playable ducking demo source pair.
automatic audio ducking
Mix an existing voice track and music bed through the Sonilo audio-ducking API for automated speech-first mixing.
Duration
Match foreground and background assets
Prompt
Not required for audio ducking
Model version
Sonilo audio ducking
Input format
Voice audio/video plus music audio
API workflow for mixing existing voice and music tracks. An API key and eligible API access are required.
POST /v1/audio-ducking accepts a voice track and a music track. Submit files or public URLs through the API, then poll GET /v1/tasks/{task_id} for the mixed output.
Need a music track first?
Generate music from textGenerate music from videoExamples
Prompt: Keep speech in front while the music breathes between lines.
A playable ducking demo source pair.
Prompt: Lower the bed under narration, then restore energy during the reveal.
A commercial clip context for speech-first music balancing.
Prompt: Keep the background audible but out of the way during explanation.
A video context for automatic ducking around spoken guidance.
Compare
The source clip keeps its original camera audio for reference.
The same clip muted, so timing and visual beats are easy to judge.
A finished video mix where music is balanced against speech and picture.
Use cases
FAQ
Automatic audio ducking lowers background music when speech is present, then raises it between phrases so narration stays clear.
No. Audio ducking mixes existing voice and music assets. Use text-to-music or video-to-music to generate a new soundtrack.
The Sonilo API supports foreground speech or dialogue inputs and background music inputs for an async ducking workflow.
Use POST /v1/audio-ducking and poll the returned task until the mixed output is ready.
Use Sonilo audio ducking when you already have voice and music and need an automatic mix.