Skip to main content
  1. Home
  2. Computing
  3. News

Adobe wants Firefly to handle the entire soundtrack for your videos

Generate Music, Generate Speech, and Generate Sound Effects are now generally available, covering most of the audio work around a typical video

Add as a preferred source on Google
Text, Sphere, Business Card
Adobe

Adobe Firefly has spent the past few years learning to make images and video. Now Adobe wants it handling the soundtrack too.

Generate Music, Generate Speech, and Generate Sound Effects are now generally available in Firefly, giving creators tools for producing music, narration, and custom effects without leaving Adobe’s AI workspace. Adobe says Generate Music can create fully licensed, royalty-free tracks, while Generate Speech supports voices in more than 20 languages.

Recommended Videos

That gives Firefly a much bigger role in video production. Instead of generating a clip and then bouncing between other tools to finish it, creators can now keep more of the audio work in the same place.

What the three tools actually do

Generate Music can build instrumental tracks from a description, with controls for things like genre and tempo. Firefly can also analyze an uploaded video and create music that better fits the footage.

Generate Speech handles narration by turning scripts into voiceovers with adjustable delivery and pronunciation. Adobe says it supports more than 20 languages, which could make alternate versions of the same video easier to produce.

Generate Sound Effects covers the smaller details around those two. Creators can describe an effect with text or use reference audio as guidance, including a vocal performance meant to shape the result.

Taken together, those tools cover a large chunk of the audio work that would usually send creators elsewhere.

Why Adobe wants this inside Firefly

Every extra tool adds another handoff. Music may come from one service, narration from another, and effects from somewhere else entirely.

Adobe is trying to cut down those detours by keeping more of the workflow inside Firefly. For creators already using it for images or video, that could mean fewer exports and less time jumping between apps before a project is ready to publish.

Convenience will only carry this so far, though. Creators who already trust dedicated audio tools still need a reason to replace them.

What creators should watch next

The real test is whether Firefly’s generated audio is consistently good enough to earn that bigger role. A soundtrack or voiceover can meet the prompt and still sound generic or awkward once it’s attached to a polished video.

Adobe has made the direction clear. Firefly is moving deeper into the production process, and audio gives creators another reason to stay there after the first image or video has been generated.

Paulo Vargas
Paulo Vargas is an English major turned reporter turned technical writer, with a career that has always circled back to…
Windows 11 is about to turn on a setting that could hurt gaming performance
Microsoft will start enabling Memory Integrity on more Windows 11 PCs in October
A gaming PC with RGB synced lights running Apex Legends.

Microsoft is preparing to enable Memory Integrity on more Windows 11 PCs starting in October. The security feature is meant to protect systems from malicious code, but there is one reason gamers may want to keep an eye on it.

Microsoft has previously acknowledged that Memory Integrity can affect gaming performance on some Windows 11 systems. Back in 2022, the company even published instructions explaining how gamers could temporarily disable Memory Integrity and Virtual Machine Platform if they were causing problems.

Read more
AI chatbots will agree and misinform if you just put a little pressure, warns research
ChatGPT 3.5 proved the most vulnerable when false claims were repeated over and over
Claude AI on an iPhone.

AI chatbots have a well-documented habit of hallucinating information and sometimes agreeing with users even when they are wrong. A new study suggests that simply refusing to take no for an answer can make the problem worse.

Researchers from the University of Arizona tested seven AI models, including GPT-3.5, GPT-4o, GPT-4o-mini, Claude 3.5 Sonnet, Gemini 1.5 Pro, Llama 3 70B, and DeepSeek-R1. Instead of judging them from a single response, researchers kept conversations going while repeatedly feeding the models information they knew was false.

Read more
Google brings its best AI music model Lyria 3.5 to the Gemini app
Developers get full API access through Google AI Studio.
Google-gemini-lyria-3.5

Google has added Lyria 3.5, its most advanced music generation model yet, to the Gemini app. Previously available through Google's AI filmmaking tool Flow, the model is now rolling out to all Gemini users, making it easier to generate polished songs, instrumentals, and soundtracks from simple text prompts or even photos.

Lyria 3.5 makes AI-generated music sound more natural

Read more