product update

Adobe Firefly's Generate Music, Speech, and Sound Effects Tools Now Generally Available

TL;DR

Adobe has moved Firefly's Generate Music, Generate Speech, and Generate Sound Effects tools out of beta and into general availability on desktop and mobile web. The tools produce commercially licensed audio without requiring a separate subscription, and Firefly has also added Gemini Omni Flash, Runway Aleph 2.0, and Kling 2.0 to its third-party model roster.

3 min read
0

Adobe Firefly Audio Tools Exit Beta

Adobe announced today that three audio generation tools inside Firefly — Generate Music, Generate Speech, and Generate Sound Effects — are now generally available on the platform's desktop and mobile web apps. The tools had previously been in limited testing and are now open to all Firefly users without requiring a separate subscription for commercially safe output, according to Adobe.

What each tool does

Generate Music creates original tracks tuned to a video's length and mood. According to Adobe, the tool can analyze a video's content and produce editable prompts suggesting a fitting musical direction. Users can adjust tempo, energy, and duration before generating, and each request returns four track options. Adobe states all output is fully licensed, eliminating copyright or takedown risk for commercial use.

Generate Speech converts scripts into voiceovers using either Adobe's own Firefly speech model or ElevenLabs' voice technology. Users can insert expressivity and pronunciation tags into specific parts of a script to control delivery.

Generate Sound Effects produces custom sound effects matched to on-screen action, timing, and energy. Users can prompt with text or record a reference sound with their own voice to guide the output.

The research behind the launch

Adobe partnered with Berklee College of Music on a survey of video creators, musicians, and marketers ahead of the release. Adobe claims 100% of respondents use music in their videos, and 32.7% already use AI-generated music. Legal and copyright risk was cited by 43.2% of respondents as the top barrier to using music, while 38.9% pointed to licensing costs — figures Adobe is using to justify positioning Generate Music as a fully licensed alternative to stock or third-party licensing services.

Firefly's expanding model roster

Alongside the audio tools' general release, Adobe added Google's Gemini Omni Flash to Firefly's list of available AI models. According to Adobe, Gemini Omni Flash allows users to prompt with video, audio, and image inputs alongside text, supporting iterative back-and-forth editing of video content. The addition follows Adobe's integration of Runway Aleph 2.0 and Kling 2.0 just days earlier, expanding a third-party model lineup that already includes tools from Google, Kling AI, Luma AI, OpenAI, and Runway alongside Adobe's own Firefly models.

Pricing for the audio tools was not disclosed beyond Adobe's statement that no separate subscription is required for commercially safe generation, suggesting the features are bundled into existing Firefly and Creative Cloud plans.

What this means

Adobe is betting that licensing risk, not creative capability, is the biggest obstacle keeping video creators from using AI-generated audio — the Berklee-backed survey data is designed to support that argument. By bundling music, voiceover, and sound effects generation directly into Firefly at no extra cost, Adobe is competing less with music-generation specialists like Suno and more with the friction of licensing stock audio altogether. The simultaneous expansion of third-party video models (Gemini Omni Flash, Runway Aleph 2.0, Kling 2.0) shows Adobe positioning Firefly as an aggregation layer rather than a single-model product — a strategy that trades model ownership for breadth, letting Adobe capture workflow lock-in even as the underlying generation technology comes from competitors.

Related Articles

product update

Suno launches Speech public beta: AI voiceovers with background music, up to about eight minutes

Suno has launched Speech in public beta on web and mobile. It generates spoken voice from a script or text prompt, with optional AI-generated background music, for clips up to roughly eight minutes. Suno claims it is the first audio model to generate voice and music together as one track.

product update

Google limits free Gemini app users to Flash-Lite from Oct. 9; AI Pro adds Deep Think

Starting October 9, Gemini app users without a Google AI subscription will be limited to Flash-Lite, losing access to Flash and Pro. AI Plus ($4.99/month) subscribers lose Pro, while AI Pro ($19.99/month) gains Deep Think, according to an updated Google support document.

product update

Anthropic adds Mods to Claude Code, a plugin system that hooks into tool calls, prompts and UI rendering

Anthropic released Mods for Claude Code, a plugin system built on JavaScript and TypeScript functions that hook into events such as tool calls, user prompts and UI rendering. Mods are not sandboxed and run with the user's permissions. They work in the CLI, the desktop app and, partly, the VS Code extension.

product update

Meta launches Muse Gadgets, open-source firmware and Linux SDK for building hardware that connects to its Muse agent

Meta introduced Muse Gadgets on Friday, an open-source project with firmware and a Linux SDK for building hardware that connects to its Muse personal AI agent. Meta also built a first-party device, Muse Home Link, and says it is giving away 5,000 units to Muse subscribers.

Comments

Loading...