model release

Meta Releases Muse Spark 1.3, a Free Multimodal Reasoning Model with 1M-Token Context

TL;DR

Meta has released Muse Spark 1.3, a multimodal reasoning model with a 1M-token context window, listed as free on OpenRouter. The model targets long-running agentic, multi-agent, and coding workflows, though audio input support remains incomplete.

2 min read
0

Meta Releases Muse Spark 1.3

Meta has released Muse Spark 1.3, a multimodal reasoning model designed for long-running agentic, multi-agent, and coding workflows. The model is now listed on OpenRouter with a 1 million token context window and is currently available at no cost.

What Muse Spark 1.3 Does

According to Meta, Muse Spark 1.3 is built to maintain state across extended tasks, reconcile conflicting inputs, and pause to request clarification or confirmation from users when needed. The model emphasizes concise execution over verbose output, positioning it for use in autonomous and semi-autonomous agent pipelines rather than single-turn chat.

The 1M-token context window puts Muse Spark 1.3 in the same range as other long-context models designed to hold entire codebases, multi-document research sets, or extended multi-agent transcripts in a single session.

Known Limitations

Meta's own listing flags a specific gap: audio understanding in Muse Spark 1.3 is "currently not fully supported," and response quality on requests that include audio content may be degraded. This suggests the model's multimodal capabilities are more mature for text and likely image/document inputs than for audio, and teams building audio-dependent pipelines should treat this as a known constraint rather than a fully shipped feature.

Pricing and Availability

Muse Spark 1.3 is currently listed as free for both input and output tokens on OpenRouter, with no per-token cost disclosed beyond the zero-price listing. OpenRouter reports 100% uptime over the trailing three days and 99.69% availability over the last 24 hours across its routing infrastructure, though these figures reflect platform-level routing performance rather than model quality or accuracy.

No benchmark scores, parameter count, or training data cutoff date have been disclosed by Meta for this release. OpenRouter's listing shows a release date of September 2, 2026; readers should treat this date as sourced directly from the OpenRouter model card, as Meta has not published an independent announcement with corroborating details at the time of writing.

What This Means

A free, 1M-token multimodal reasoning model lowers the barrier for developers experimenting with long-running agent workflows, particularly multi-agent systems that need to retain context across many steps without expensive re-summarization. The lack of published benchmarks makes it difficult to assess how Muse Spark 1.3 compares to other reasoning models like GPT or Gemini variants on standard coding and reasoning suites — buyers evaluating it for production agentic pipelines should run their own evals rather than relying on Meta's framing. The explicit audio limitation is a useful signal: this is a model still being hardened for full multimodal input, and teams should scope their initial use to text, code, and likely image-based tasks until Meta confirms audio support is production-ready.

Related Articles

model release

Meta Releases Muse Spark 1.3 Contributor, a Low-Cost Multimodal Reasoning Model With 1M Context Window

Meta has released Muse Spark 1.3 Contributor, described as the cost-efficient contributor tier of its multimodal reasoning model line. The model offers a 1 million token context window at $0.10 per 1M input tokens and $0.20 per 1M output tokens, targeting experimentation and early-stage agentic workflows.

model release

Google Lists Gemini 3.8 Flash on OpenRouter With 1M-Token Context, September 2026 Release Date

Google's Gemini 3.8 Flash has surfaced on OpenRouter with a 1-million-token context window and discounted pricing of $0.75 per 1M input tokens and $3.75 per 1M output tokens. Google has not issued a separate public announcement, and the listed release date of September 2, 2026 is unusually far out, leaving key details unconfirmed.

model release

Inception Launches Mercury 2.5 Preview, a Diffusion LLM Claiming 1,107 Tokens/Sec

Inception released Mercury 2.5 Preview, a diffusion-based language model that generates tokens in parallel rather than sequentially, claiming throughput of 1,107 tokens per second on standard GPUs. The model is available on OpenRouter with a 260K context window and an 80% launch discount through September 8, 2026.

model release

Meta Launches Muse Voice Transcribe, Real-Time Speech Model Handling 20+ Speakers Across 70+ Languages

Meta Superintelligence Lab released Muse Voice Transcribe, its first real-time audio perception model, claiming state-of-the-art streaming speech-to-text with native speaker diarization. The model handles 20+ speakers and code-switching across languages, priced at $3 per 1,000 audio minutes.

Comments

Loading...