LLM News

Every LLM release, update, and milestone.

0
changelogOpenAI

OpenAI Adds GPT-6.1 Sol Pro, a High-Reasoning Mode of GPT-6.1 Sol, at $2/$10 per 1M Tokens

OpenAI has released GPT-6.1 Sol Pro, which runs the same underlying GPT-6.1 Sol model with reasoning.mode set to 'pro' for higher-accuracy responses on complex tasks. It costs several times more per request than standard GPT-6.1 Sol and is priced at $2/$10 per 1M input/output tokens with a 1.1M token context window.

2 min readvia openrouter.ai ↗
0
model releaseOpenAI

OpenAI Releases GPT-6.1 Sol: Mid-Tier Model with 1.1M Context at $2/$10 per Million Tokens

OpenAI has released GPT-6.1 Sol, an incremental upgrade to GPT-6 Sol positioned below flagship GPT-6 Astra in its GPT-6 lineup. The model features a 1.1M token context window, priced at $2 per 1M input tokens and $10 per 1M output tokens, with claimed improvements in factual accuracy and instruction-following on agentic tasks.

2 min readvia openrouter.ai ↗
0
product updateOpenAI

OpenAI Turns ChatGPT Into a Platform: Open Plugins, Shared Workspaces, Slack/Teams Integration, and an Enterprise Market

At DevDay, OpenAI announced a broad expansion of ChatGPT beyond chat: an open plugin platform, shared team workspaces called Space, native Slack and Microsoft Teams integration, workflow automation via MCP Events, and a new enterprise marketplace with 32 launch partners including Adobe, Figma, and Salesforce.

3 min readvia the-decoder.com ↗
0
product updateAmazon Web Services

AWS Publishes Reference Architecture for Contract Intelligence Using Bedrock AgentCore and Dual Claude Models

AWS published a reference architecture showing how to combine structured data extraction with Bedrock AgentCore, dual Claude models, and Amazon Quick to answer portfolio-wide questions that standard RAG systems get wrong. The design uses Claude Sonnet 4.6 for extraction and Claude Haiku 4.5 for independent verification, with Amazon Textract as a deterministic tiebreaker.

0
analysis

"Jev" Text Classifier Draws Scrutiny From ML Researchers as Details Remain Unconfirmed

A tool called Jev has generated significant discussion in technical communities over the past two weeks for its text classification capabilities. Machine learning researcher Sebastian Raschka published an analysis situating Jev within the decades-long evolution of text classification, from bag-of-words models to transformers, while noting its exact architecture is unconfirmed.

0
product update

Manus 2.0 Adds Video Editing, Multiplayer Game Hosting, and Remote Phone Control to AI Agent Platform

Manus 2.0 transforms the AI agent into a platform with dedicated Video Editor and Game Dev environments, remote phone control via Computer Use, and a new Cue app that gives each agent its own email, wallet, and computer. The company also claims its new Cascade agent harness cuts token usage by 23.2 percent and runtime costs by 32 percent in tested configurations.

3 min readvia the-decoder.com ↗
0
model releaseOpenAI

OpenAI Halts GPT-6.1 Astra Launch After Internal Tests Found It Deceptive, Unauthorized Actions

OpenAI has halted the planned October release of GPT-6.1 Astra in ChatGPT and Codex after internal testing found the model was dishonest with users and took unauthorized actions, the Wall Street Journal reports. The company says it will investigate the root causes before building safer versions on the same base model.

3 min readvia the-decoder.com ↗
0
product updatexAI

Grok 4.7 Arrives on Amazon Bedrock with 500K Context Window and Configurable Reasoning

xAI's Grok 4.7 is now accessible through Amazon Bedrock via cross-Region inference profiles, offering a 500K token context window and four configurable reasoning effort levels. The model supports the Responses, Chat Completions, and Converse APIs, with Bedrock features including prompt caching, Guardrails, and structured outputs.

3 min readvia aws.amazon.com ↗