Industry
AI Music for E-Learning and Course Videos
- Written by
- Sonilo Team
- Published

Course creators and L&D teams need background music that stays quiet under narration instead of competing with it, and Sonilo generates that music and automatically ducks it under your voice track with a single API call.
Quick answer
Sonilo generates subtle, low-key background music for e-learning and course videos from a short text prompt describing the mood you want ("subtle, focused, low-key ambient") at the exact duration you need, or directly from your lesson recording by analyzing its pacing. Then, instead of you manually riding volume faders in an editor, Sonilo's audio-ducking API automatically mixes that music under your narrator or instructor track and lowers it whenever speech is present, so the music supports the lesson without ever masking the instruction. Every track is fully licensed via Shutterstock, so it's commercial-safe for paid courses sold on Udemy, Teachable, or an internal corporate LMS.
Why it fits e-learning video production
| Need | Sonilo capability | Why it matters |
|---|---|---|
| Music that never masks narration | POST /v1/audio-ducking automatically mixes a narrator or instructor voice track with background music, lowering the music around speech | Learners can follow every word of the lesson without you manually automating volume in a video or audio editor |
| Low-key, unobtrusive mood instead of generic stock music | Text-to-music generation from a prompt describing mood and energy (e.g. "subtle, focused, low-key ambient") at an exact requested duration | Course music should support attention, not pull focus away from the instructor, and a precise prompt keeps it consistently calm |
| Commercial-safe licensing for paid courses | All music is fully licensed via Shutterstock | Removes the copyright and licensing risk that comes with unverified royalty-free tracks when you're selling courses on Udemy, Teachable, or a corporate LMS |
| Consistent background music across an entire course library | REST API at api.sonilo.com with Bearer auth | L&D teams and LMS vendors can generate and duck music programmatically for dozens or hundreds of lesson videos instead of scoring each one by hand |
| Music that matches each lesson's pacing | Video-to-music mode analyzes the lesson recording itself to generate matching music | Useful when you'd rather generate music directly from the finished recording than write a text prompt for every video |
How it works
- Generate your background music: describe the mood you want (for example, "subtle, focused, low-key ambient") and request the exact duration you need, or let Sonilo analyze your recorded lesson video directly and generate music that matches its pacing.
- Upload your narration or instructor voice track alongside the generated music track.
- Call POST /v1/audio-ducking to automatically mix the two tracks, lowering the music whenever the narrator is speaking and bringing it back up in the gaps.
- Download the mixed result and drop it straight into your lesson video, course upload, or LMS asset library.
FAQ
Will the music drown out my narration?
No. Sonilo's audio-ducking endpoint (POST /v1/audio-ducking) automatically mixes your narrator or instructor track with the background music and lowers the music whenever speech is present, so instruction stays clearly audible without any manual volume automation in your editor.
Is the music safe to use on paid course platforms like Udemy or Teachable?
Yes. All Sonilo music is fully licensed via Shutterstock, so it's commercial-safe for paid courses sold through Udemy, Teachable, or an internal corporate LMS, an important distinction from unlicensed royalty-free tracks that can carry real legal and business exposure.
Can we automate music for our entire course library instead of doing it lesson by lesson?
Yes. Sonilo's REST API at api.sonilo.com, authenticated with a Bearer token, lets e-learning platforms and LMS vendors generate and duck background music programmatically across an entire library of lesson videos, keeping the sound consistent from course to course instead of scoring each video by hand.
Do I need to write a detailed mood description, or can I just use my video?
Either works. You can describe the mood and energy you want in a short text prompt, such as "subtle, focused, low-key ambient," and request an exact duration between 5 and 360 seconds, or let Sonilo generate music directly from your video by analyzing the lesson recording's pacing.
Related Sonilo pages
- Audio ducking API docs: https://platform.sonilo.com/docs/api/audio-ducking
- Text-to-music API docs: https://platform.sonilo.com/docs/api/text-to-music
- Pricing: https://sonilo.com/pricing


