AionLabs Launches Aion 3.5, a Multi-Model Storytelling System Built on GLM
AionLabs has released Aion 3.5, a collaborative multi-model system for roleplaying and storytelling built on the GLM model family. It offers a 262K token context window at $3 per 1M input tokens and $6 per 1M output tokens.
Aion 3.5 — Quick Specs
AionLabs has released Aion 3.5, the latest version of its purpose-built roleplaying and storytelling AI system, now available through OpenRouter. The model is built on the GLM family of models and extends the company's approach of using multiple specialized models in a single collaborative generation pipeline.
What's new
According to AionLabs, Aion 3.5 uses a "collaborative generation process" in which several specialized models each contribute to a single response. The company claims this produces stronger narrative structure and more compelling tension and conflict than single-model generation — a persistent weak point for general-purpose LLMs used in long-form fiction and character roleplay.
Aion 3.5 supports a 262,000-token context window, doubling the 131K context of its predecessor, Aion-3.0. Pricing is set at $3 per 1M input tokens and $6 per 1M output tokens, with cached input priced at $0.75 per 1M tokens — identical to Aion-3.0's rate card despite the larger context window.
On OpenRouter's infrastructure, the model has shown a median latency of 0.54 seconds and throughput of 53 tokens per second, with 100% uptime and availability over the past three days of tracked data. These figures reflect performance through AionLabs' own provider endpoint rather than independently verified benchmarks.
Alongside Aion 3.5, AionLabs also released Aion 3.5 Mini, a smaller, lower-cost variant using the same GLM-based collaborative architecture and matching 262K context window, priced at $0.70 per 1M input tokens and $1.40 per 1M output tokens.
Product lineage
Aion 3.5 is the sixth major release in AionLabs' Aion series. The company has iterated across different base model families: Aion-1.0 and Aion-1.0-Mini were built on DeepSeek-R1 with Tree of Thoughts and Mixture of Experts techniques for reasoning tasks; Aion-2.0 was built on DeepSeek V3.2; and Aion-3.0 shifted to the GLM family, the same lineage Aion 3.5 continues. A separate model, Aion-RP 1.0, is a fine-tuned Llama-3.1-8B base model the company says ranks highest in the character evaluation portion of the RPBench-Auto benchmark.
No independent third-party benchmark scores for Aion 3.5 itself were disclosed at release. AionLabs has not published a technical report detailing the specific GLM checkpoint used as the base model or training data cutoff.
What this means
Aion 3.5 targets a narrow but active niche: AI systems optimized specifically for roleplay and narrative fiction rather than general assistant tasks. The multi-model "collaborative generation" approach — routing pieces of a response through different specialized models — is an architectural bet that ensemble methods can outperform a single fine-tuned model on creative writing quality, though AionLabs has not published data isolating that claim from simply using a larger context window or better base model.
The pricing sits above budget-tier models but well below frontier general-purpose models like GPT-5 or Claude Opus, positioning Aion 3.5 as a specialized tool rather than a general competitor. Buyers evaluating it for production roleplay or storytelling applications should treat the uptime and latency figures as infrastructure-level metrics, not evidence of output quality, and look for independent benchmark comparisons before switching from established general-purpose models.
Related Articles
Aion Labs Launches Aion 3.5 Mini, a $0.70/M-Token Roleplaying Model with 262K Context
Aion Labs has released Aion 3.5 Mini, a lower-cost version of its multi-model roleplaying system Aion 3.5. Built on the GLM model family, it offers a 262K token context window at $0.70 per 1M input tokens and $1.40 per 1M output tokens.
Anonymous Stealth Model "Space Bunny Alpha" Debuts on OpenRouter With 1M-Token Context, Free During Preview
A previously unknown AI provider has released Space Bunny Alpha, a stealth model on OpenRouter offering a 1M-token context window, adjustable reasoning effort, and multimodal input support. The model is free during its preview period, though its developer remains unnamed.
Upstage Releases Solar Mini 4: 35B MoE Model with 524K Context at $0.05/$0.20 per Million Tokens
Upstage has released Solar Mini 4, a compact mixture-of-experts model with 35B total parameters, 3B active parameters, and a 524K token context window. The model targets agentic workloads and is priced at $0.05 per 1M input tokens and $0.20 per 1M output tokens, a promotional 50% discount off standard rates.
Anthropic Ships Claude Opus 5.5, OpenAI Launches GPT-6 Sol and Luna — All Cheaper Than Predecessors
Anthropic released Claude Opus 5.5 at $4/$20 per million input/output tokens, undercutting Opus 5's $5/$25 pricing while claiming better agentic coding scores. OpenAI countered with GPT-6 Sol ($2/$10) and GPT-6 Luna ($0.10/$0.50), both up to 50% cheaper than GPT-5.6's promotional rates.
Comments
Loading...