LLM News

Every LLM release, update, and milestone.

0
product updateAmazon Web Services

Amazon Quick AI Assistant Now Embeds Directly Into Word, Excel, PowerPoint, and Outlook

Amazon has released Microsoft 365 extensions for its Quick AI assistant, embedding agentic capabilities directly into Word, Excel, PowerPoint, and Outlook. The extensions run entirely in the cloud, require no client-side installation, and connect to existing Quick data sources like Salesforce, Jira, Slack, and SharePoint.

0
product updateMicrosoft

Microsoft Merges Consumer and Business Copilot Apps, Cuts Group Chats, Deep Research, Mico Character

Microsoft is merging its consumer-facing Copilot app with Microsoft 365 Copilot while discontinuing Group Chats, AI-generated podcasts, Copilot Labs, Deep Research, and the Mico animated character by August 18, 2026. Paying professional users will get Researcher as a Deep Research replacement.

3 min readvia techcrunch.com
0
product updateMicrosoft

Microsoft to Merge Copilot and Microsoft 365 Copilot Into Single App, Retiring Group Chats, Podcasts, and Deep Research

Microsoft will merge its consumer Copilot app and Microsoft 365 Copilot into a single unified app starting mid-September 2026, letting users sign in with personal, work, or school accounts. Three features—group chats, podcasts, and Deep Research—will be retired on August 18.

0
analysisOpenAI

Researchers Warned About Automated AI Research — Several Predicted Milestones Already Hit, New Report Says

IAPS fellow Severin Field interviewed 25 researchers from top AI labs about recursive self-improvement in late 2025. Several milestones they cited as evidence of progress — Math Olympiad gold, autonomous training loops, majority AI-written code — have since occurred, according to a new report.

0
changelogAnthropic

Cline CLI v3.0.54 Fixes Claude Code Provider Tool-Bridging Bug and Token Telemetry Overcounting

Cline's CLI v3.0.54 release fixes a major bug that made the Claude Code provider unusable for agentic work, along with tool-call JSON truncation handling and token telemetry that was overcounting cache-heavy sessions by roughly 5x. The update also changes how Managed Hub daemons handle version conflicts between multiple Cline installs.

2 min readvia github.com
0
model releaseNVIDIA

NVIDIA Releases Nemotron 3.5 Lightning: 30B MoE Model with 1M Token Context and 3B Active Parameters

NVIDIA released the full-precision BF16 reference weights for Nemotron 3.5 Lightning, a 30B-parameter Mixture-of-Experts model with only 3B active parameters and support for up to 1 million tokens of context. The model uses a hybrid Mamba-2, MoE, and Attention architecture and is licensed under OpenMDW-1.1 for commercial use.

2 min readvia huggingface.co