LLM
7 articles tagged with LLM
Upstage Releases Solar Pro 4 With 524K Token Context Window at $0.03/M Input Pricing
Upstage has released Solar Pro 4, a large language model with a 524K token context window aimed at agentic workflows, document processing, and coding. The model is priced at $0.03 per million input tokens and $0.12 per million output tokens, and is available now via OpenRouter.
Moonshot AI's Kimi K3 to launch with 2-3 trillion parameters, targets Anthropic Claude Opus 4.8 performance
Moonshot AI will release Kimi K3 in the coming days with a parameter count between 2 trillion and 3 trillion, according to Financial Times sources. The open-weight model is expected to perform at par with or surpass Anthropic's Claude Opus 4.8, making it the largest open-weight AI model from China.
OpenAI Releases GPT-5.6 Terra Pro with Enhanced Reasoning Mode at $2.50/$15 Per Million Tokens
OpenAI has released GPT-5.6 Terra Pro, a variant of GPT-5.6 Terra configured with enhanced reasoning capabilities for complex tasks. The model features a 1 million token context window and is priced at $2.50 per million input tokens and $15 per million output tokens.
Base44 launches Base1 LLM trained on tens of millions of user interactions
Base44, the vibe coding platform acquired by Wix for $80 million in 2025, has released Base1, a custom LLM trained on tens of millions of user interactions. The company claims the model will eventually outperform frontier models through optimization for latency, cost, and efficiency specific to app creation workflows.
IBM's Granite 4.1: 8B Dense Model Matches 32B MoE Performance on 15T Tokens
IBM released Granite 4.1, a family of dense decoder-only LLMs (3B, 8B, 30B parameters) trained on approximately 15 trillion tokens using a five-phase pre-training pipeline. The 8B instruct model matches or surpasses the previous Granite 4.0-H-Small (32B-A9B MoE) despite using fewer parameters and a simpler dense architecture. All models support up to 512K context windows and are released under Apache 2.0 license.
DeepSeek releases V4 model preview with agent optimization, pricing undisclosed
DeepSeek released a preview of its V4 large language model on April 24, 2026, available in 'pro' and 'flash' versions. The Hangzhou-based company claims the open-source model achieves strong performance on agent-based tasks and has been optimized for tools like Anthropic's Claude Code and OpenClaw.
OpenRouter Releases Elephant Alpha: 100B-Parameter Model with 256K Context Window and Free Pricing
OpenRouter has released Elephant Alpha, a 100B-parameter text model with a 256K context window and 32K output token limit. The model is available at no cost through OpenRouter's platform, supporting function calling, structured output, and prompt caching.