Adobe Firefly's Generate Music, Speech, and Sound Effects Tools Now Generally Available
Adobe has moved Firefly's Generate Music, Generate Speech, and Generate Sound Effects tools out of beta and into general availability on desktop and mobile web. The tools produce commercially licensed audio without requiring a separate subscription, and Firefly has also added Gemini Omni Flash, Runway Aleph 2.0, and Kling 2.0 to its third-party model roster.
Adobe Firefly Audio Tools Exit Beta
Adobe announced today that three audio generation tools inside Firefly — Generate Music, Generate Speech, and Generate Sound Effects — are now generally available on the platform's desktop and mobile web apps. The tools had previously been in limited testing and are now open to all Firefly users without requiring a separate subscription for commercially safe output, according to Adobe.
What each tool does
Generate Music creates original tracks tuned to a video's length and mood. According to Adobe, the tool can analyze a video's content and produce editable prompts suggesting a fitting musical direction. Users can adjust tempo, energy, and duration before generating, and each request returns four track options. Adobe states all output is fully licensed, eliminating copyright or takedown risk for commercial use.
Generate Speech converts scripts into voiceovers using either Adobe's own Firefly speech model or ElevenLabs' voice technology. Users can insert expressivity and pronunciation tags into specific parts of a script to control delivery.
Generate Sound Effects produces custom sound effects matched to on-screen action, timing, and energy. Users can prompt with text or record a reference sound with their own voice to guide the output.
The research behind the launch
Adobe partnered with Berklee College of Music on a survey of video creators, musicians, and marketers ahead of the release. Adobe claims 100% of respondents use music in their videos, and 32.7% already use AI-generated music. Legal and copyright risk was cited by 43.2% of respondents as the top barrier to using music, while 38.9% pointed to licensing costs — figures Adobe is using to justify positioning Generate Music as a fully licensed alternative to stock or third-party licensing services.
Firefly's expanding model roster
Alongside the audio tools' general release, Adobe added Google's Gemini Omni Flash to Firefly's list of available AI models. According to Adobe, Gemini Omni Flash allows users to prompt with video, audio, and image inputs alongside text, supporting iterative back-and-forth editing of video content. The addition follows Adobe's integration of Runway Aleph 2.0 and Kling 2.0 just days earlier, expanding a third-party model lineup that already includes tools from Google, Kling AI, Luma AI, OpenAI, and Runway alongside Adobe's own Firefly models.
Pricing for the audio tools was not disclosed beyond Adobe's statement that no separate subscription is required for commercially safe generation, suggesting the features are bundled into existing Firefly and Creative Cloud plans.
What this means
Adobe is betting that licensing risk, not creative capability, is the biggest obstacle keeping video creators from using AI-generated audio — the Berklee-backed survey data is designed to support that argument. By bundling music, voiceover, and sound effects generation directly into Firefly at no extra cost, Adobe is competing less with music-generation specialists like Suno and more with the friction of licensing stock audio altogether. The simultaneous expansion of third-party video models (Gemini Omni Flash, Runway Aleph 2.0, Kling 2.0) shows Adobe positioning Firefly as an aggregation layer rather than a single-model product — a strategy that trades model ownership for breadth, letting Adobe capture workflow lock-in even as the underlying generation technology comes from competitors.
Related Articles
Mistral Launches Agentic Search, Claims 3x Accuracy Gain on Financial Document Retrieval
Mistral has released Agentic Search, a retrieval layer that lets AI models navigate, read, and verify information across complex documents instead of relying on single-pass chunk retrieval. The company claims accuracy improvements from 26.7% to 86% on FinanceBench and up to 39.6% lower p90 latency.
Black Forest Labs Launches FLUX Upscale, a Dedicated Tool for 4K Video Upscaling
Black Forest Labs has released FLUX Upscale, a dedicated tool that takes existing video and regenerates it at higher resolution, up to native 4K. The tool is now available as its own FLUX endpoint.
Meta AI Launches Mac App With System-Wide Dictation and Screen Context Awareness
Meta released a new Mac app for Meta AI with system-wide dictation and the ability to answer questions based on what's visible on screen, using its Muse Spark model. The launch also bundles business-focused features letting merchants connect Instagram, Facebook, ad accounts, and Google Workspace to the assistant.
OpenAI Launches 'Private Safety Processing' to Detect Misuse Without Storing Enterprise Data
OpenAI has built a system called Private Safety Processing that detects misuse patterns across multiple interactions without storing customer inputs or outputs. The company says it only receives narrow safety signals—type and severity of activity—while data stays encrypted on customer infrastructure.
Comments
Loading...