model release

OpenRouter Releases Elephant Alpha: 100B-Parameter Model with 256K Context Window and Free Pricing

TL;DR

OpenRouter has released Elephant Alpha, a 100B-parameter text model with a 256K context window and 32K output token limit. The model is available at no cost through OpenRouter's platform, supporting function calling, structured output, and prompt caching.

2 min read
1

OpenRouter Releases Elephant Alpha: 100B-Parameter Model with 256K Context Window and Free Pricing

OpenRouter has released Elephant Alpha, a 100B-parameter text model designed for "intelligence efficiency" with a 256K context window and support for up to 32K output tokens. The model is available at $0 per million tokens for both input and output through OpenRouter's routing platform.

Technical Specifications

Elephant Alpha features:

  • 100 billion parameters
  • 262,144 (256K) token context window
  • 32,768 (32K) maximum output tokens
  • Function calling support
  • Structured output capabilities
  • Prompt caching
  • Released April 13, 2025

According to OpenRouter, the model focuses on "delivering strong reasoning performance while minimizing token usage," though specific benchmark scores have not been disclosed.

Target Use Cases

OpenRouter positions Elephant Alpha for three primary applications:

  • Code completion and debugging
  • Rapid document processing
  • Lightweight agent interactions

The model is available through OpenRouter's unified API, which routes requests across multiple providers with automatic fallbacks. OpenRouter notes that prompts and completions may be logged by the provider and used for model improvement.

Pricing and Access

The model is currently available at zero cost through OpenRouter's platform, with no charges for input or output tokens. This pricing is managed through OpenRouter's routing system, which normalizes requests and responses across providers.

The model supports OpenAI-compatible API calls and can be accessed through the OpenAI SDK as well as various third-party SDKs and frameworks.

What This Means

Elephant Alpha enters a crowded field of large language models with a distinctive positioning around "intelligence efficiency" and a notably large context window at 256K tokens. The free pricing through OpenRouter makes it accessible for experimentation, though the lack of published benchmarks makes it difficult to assess performance claims against established models. The 32K output token limit is substantially higher than many competing models, which could be useful for document generation tasks. However, the data logging policy and absence of performance metrics warrant careful evaluation for production deployments.

Related Articles

model release

Moonshot AI releases Kimi K3, China's largest model at 2.8 trillion parameters

Beijing-based Moonshot AI released Kimi K3, China's largest AI model at 2.8 trillion parameters. The company claims the model consistently outperforms OpenAI's GPT 5.5 and Anthropic's Claude Opus 4.8 on benchmarks including coding and general agents, though it still trails the leading-edge GPT 5.6 Sol and Claude Fable 5 in overall performance.

model release

Moonshot AI Releases Kimi K3: Open-Weight Multimodal Reasoning Model with 1M Context Window

Moonshot AI has released Kimi K3, an open-weight multimodal reasoning model with a 1-million token context window. The model is priced at $3 per 1M input tokens and $15 per 1M output tokens, available through OpenRouter.

model release

Moonshot AI's Kimi K3 to launch with 2-3 trillion parameters, targets Anthropic Claude Opus 4.8 performance

Moonshot AI will release Kimi K3 in the coming days with a parameter count between 2 trillion and 3 trillion, according to Financial Times sources. The open-weight model is expected to perform at par with or surpass Anthropic's Claude Opus 4.8, making it the largest open-weight AI model from China.

model release

Kwaipilot Releases KAT-Coder-Air V2.5 with 256K Context Window at $0.15/$0.60 Per Million Tokens

Kwaipilot has released KAT-Coder-Air V2.5, a coding-specialized model with a 256K token context window. The model is priced at $0.15 per million input tokens and $0.60 per million output tokens, positioning it as a mid-tier coding assistant option.

Comments

Loading...