OpenRouter Releases Elephant Alpha: 100B-Parameter Model with 256K Context Window and Free Pricing
OpenRouter has released Elephant Alpha, a 100B-parameter text model with a 256K context window and 32K output token limit. The model is available at no cost through OpenRouter's platform, supporting function calling, structured output, and prompt caching.
OpenRouter Releases Elephant Alpha: 100B-Parameter Model with 256K Context Window and Free Pricing
OpenRouter has released Elephant Alpha, a 100B-parameter text model designed for "intelligence efficiency" with a 256K context window and support for up to 32K output tokens. The model is available at $0 per million tokens for both input and output through OpenRouter's routing platform.
Technical Specifications
Elephant Alpha features:
- 100 billion parameters
- 262,144 (256K) token context window
- 32,768 (32K) maximum output tokens
- Function calling support
- Structured output capabilities
- Prompt caching
- Released April 13, 2025
According to OpenRouter, the model focuses on "delivering strong reasoning performance while minimizing token usage," though specific benchmark scores have not been disclosed.
Target Use Cases
OpenRouter positions Elephant Alpha for three primary applications:
- Code completion and debugging
- Rapid document processing
- Lightweight agent interactions
The model is available through OpenRouter's unified API, which routes requests across multiple providers with automatic fallbacks. OpenRouter notes that prompts and completions may be logged by the provider and used for model improvement.
Pricing and Access
The model is currently available at zero cost through OpenRouter's platform, with no charges for input or output tokens. This pricing is managed through OpenRouter's routing system, which normalizes requests and responses across providers.
The model supports OpenAI-compatible API calls and can be accessed through the OpenAI SDK as well as various third-party SDKs and frameworks.
What This Means
Elephant Alpha enters a crowded field of large language models with a distinctive positioning around "intelligence efficiency" and a notably large context window at 256K tokens. The free pricing through OpenRouter makes it accessible for experimentation, though the lack of published benchmarks makes it difficult to assess performance claims against established models. The 32K output token limit is substantially higher than many competing models, which could be useful for document generation tasks. However, the data logging policy and absence of performance metrics warrant careful evaluation for production deployments.
Related Articles
Moonshot AI releases Kimi K3, China's largest model at 2.8 trillion parameters
Beijing-based Moonshot AI released Kimi K3, China's largest AI model at 2.8 trillion parameters. The company claims the model consistently outperforms OpenAI's GPT 5.5 and Anthropic's Claude Opus 4.8 on benchmarks including coding and general agents, though it still trails the leading-edge GPT 5.6 Sol and Claude Fable 5 in overall performance.
Moonshot AI Releases Kimi K3: Open-Weight Multimodal Reasoning Model with 1M Context Window
Moonshot AI has released Kimi K3, an open-weight multimodal reasoning model with a 1-million token context window. The model is priced at $3 per 1M input tokens and $15 per 1M output tokens, available through OpenRouter.
Moonshot AI's Kimi K3 to launch with 2-3 trillion parameters, targets Anthropic Claude Opus 4.8 performance
Moonshot AI will release Kimi K3 in the coming days with a parameter count between 2 trillion and 3 trillion, according to Financial Times sources. The open-weight model is expected to perform at par with or surpass Anthropic's Claude Opus 4.8, making it the largest open-weight AI model from China.
Kwaipilot Releases KAT-Coder-Air V2.5 with 256K Context Window at $0.15/$0.60 Per Million Tokens
Kwaipilot has released KAT-Coder-Air V2.5, a coding-specialized model with a 256K token context window. The model is priced at $0.15 per million input tokens and $0.60 per million output tokens, positioning it as a mid-tier coding assistant option.
Comments
Loading...