model release

AionLabs Releases Aion-3.0: Multi-Model Roleplaying System with 131K Context at $3/$6 per 1M Tokens

TL;DR

AionLabs has released Aion-3.0, a multi-model system designed for roleplaying and storytelling that uses collaborative generation from specialized models. The system offers a 131K context window and is priced at $3 per 1M input tokens and $6 per 1M output tokens.

2 min read
0

Aion-3.0 — Quick Specs

Context window131K tokens
Input$3/1M tokens
Output$6/1M tokens

AionLabs Releases Aion-3.0: Multi-Model Roleplaying System with 131K Context

AionLabs has released Aion-3.0, a multi-model system designed for roleplaying and storytelling applications with a 131,000-token context window, priced at $3 per million input tokens and $6 per million output tokens.

Technical Architecture

According to AionLabs, Aion-3.0 is built on the GLM family of models and uses a collaborative generation process where multiple specialized models each contribute to a single response. The company claims this approach produces "stronger narrative structure and more compelling tension and conflict" compared to single-model systems.

The model is listed as text-only (text in, text out) and was released July 7, 2025. It is currently hosted exclusively through OpenRouter, which forwards requests directly to the provider without routing decisions.

Pricing and Performance

  • Input: $3 per 1M tokens
  • Output: $6 per 1M tokens
  • Context window: 131,000 tokens

OpenRouter's data indicates that customers using prompt caching can reduce effective costs by 60-80% below list prices for workloads with repeated context.

Market Position

The multi-model collaborative approach distinguishes Aion-3.0 from standard single-model systems. However, the company has not released benchmark scores on standard evaluation tasks like MMLU or HumanEval, making direct comparisons with general-purpose models difficult.

The pricing sits in the mid-range tier: more expensive than budget models like Gemini 1.5 Flash (input $0.075/1M, output $0.30/1M) but significantly cheaper than premium models like Claude 3.5 Sonnet (input $3/1M, output $15/1M for similar context lengths).

What This Means

Aion-3.0 targets a specific niche—narrative generation and roleplaying—rather than competing as a general-purpose model. The collaborative multi-model architecture is unusual in the current landscape, where most systems rely on a single large model. Whether this approach delivers meaningfully better creative outputs remains to be validated through independent testing. The 131K context window and mid-tier pricing make it accessible for applications requiring long-form narrative coherence, though the lack of published benchmarks limits comparisons with alternatives.

Related Articles

model release

Inference.net Launches Schematron V2 Turbo, a 3B-Parameter Model for High-Volume HTML-to-JSON Extraction

Inference.net has released Schematron V2 Turbo, a 3-billion-parameter model built specifically for high-volume HTML-to-JSON extraction. The model supports a 128K context window and is priced at $0.03 per 1M input tokens and $0.15 per 1M output tokens.

model release

Inference.net Releases Schematron V2 Small, a 3B-Parameter Model for HTML-to-JSON Extraction

Inference.net has released Schematron V2 Small, a 3B-parameter model specialized in converting HTML pages into structured JSON output. The model supports a 128K context window and requires extraction schemas to be passed via response_format rather than standard prompts.

model release

Unverified 'GPT Astra' Model Appears on OpenRouter With 1.05M Token Context, No OpenAI Confirmation

OpenRouter is listing a model called 'OpenAI GPT Astra Latest' with a 1.05 million token context window and $10/$50 per-million-token pricing. OpenAI has made no public announcement, and the listing's own description says it is an auto-redirecting alias rather than a fixed model.

model release

OpenRouter Lists 'GPT Sol Latest' — An Alias Pointer to OpenAI's Newest Sol-Family Model, Not a Standalone Release

OpenRouter has added a listing called '~openai/gpt-sol-latest,' described as an alias that always points to the newest model in an undisclosed 'GPT Sol' family from OpenAI. The listing shows a 1050K token context window and pricing of $2.00 per million input tokens and $10.00 per million output tokens, but OpenAI has not publicly confirmed a model line by this name.

Comments

Loading...