Sonilo Raises $11M for Generative Audio
Sonilo has raised $11M to build AI that generates music and sound effects matched to video.

Generative media is moving from creating isolated assets toward systems that understand how different parts of a production fit together.
What happened
Sonilo raised $11 million to scale its generative-audio models.
The startup builds systems that analyse video and generate music and sound effects matched to the movement, timing and emotional changes in a scene.
The funding will support additional model training, creator distribution and the company’s music-licensing infrastructure.
Why it matters
Video creators currently have to source, edit and synchronise music and sound effects separately.
A model that understands both the visual scene and the audio timeline could compress much of that workflow into one generation step.
That creates a more specialised opportunity than generic text-to-music generation.
The bigger picture
AI media tools are becoming increasingly multimodal.
The next generation will not simply generate an image, video or song in isolation. They will understand how different media layers interact.
Sonilo is betting that synchronised audio generation becomes a core part of AI-native video production.
