Nvidia

10 articles tagged with Nvidia

July 15, 2026
model release

Mira Murati's Thinking Machines releases Inkling, 975B-parameter open-weight model trained on 45T tokens

Thinking Machines Lab released Inkling, a 975-billion-parameter mixture-of-experts model that uses 41 billion active parameters per task. The open-weight model was trained on 45 trillion tokens across text, image, audio, and video, marking the first public release from Mira Murati's AI startup.

June 12, 2026
model releaseApple

Apple releases AFM 3 lineup: 20B-parameter on-device model and cloud AI running on Google's Nvidia infrastructure

Apple announced five third-generation foundation models at WWDC26, headlined by AFM 3 Core Advanced—a 20-billion-parameter sparse model that runs on-device by activating only 1-4 billion parameters at a time. For the first time, Apple extended Private Cloud Compute to third-party infrastructure, with AFM 3 Cloud Pro running on Nvidia GPUs in Google Cloud.

June 9, 2026
product updateApple

Apple's AFM Cloud Pro Model Runs on Nvidia GPUs in Google Cloud, Execs Confirm

Apple executives revealed at WWDC that its most advanced AI model, AFM Cloud Pro, runs on Nvidia GPUs deployed in Google's cloud infrastructure while maintaining Apple's privacy guarantees. The company disclosed it uses Google's Gemini frontier models to refine its own custom models built for Apple Silicon.

product updateApple

Apple deploys 1.2T-parameter Gemini model on Nvidia Blackwell GPUs for rebuilt Siri

Apple announced at WWDC 2026 that the rebuilt Siri runs on a custom 1.2-trillion-parameter model based on Google's Gemini technology, hosted on Google Cloud servers powered by Nvidia Blackwell B200 GPUs. The company unveiled a three-tier privacy architecture and five new Apple Foundation Models to handle queries across device, private cloud, and Google Cloud infrastructure.

June 8, 2026
product updateApple

Apple deploys Google-trained models in iOS 27 Siri via Private Cloud Compute on Nvidia GPUs

Apple's senior vice president Craig Federighi disclosed that iOS 27's Siri AI uses a family of third-generation Apple Foundation Models trained with outputs from Google's Gemini frontier models. The most capable model, AFM Cloud Pro, runs on Nvidia GPUs in Google's cloud infrastructure while maintaining Apple's Private Cloud Compute privacy architecture.

June 5, 2026
model releaseNVIDIA+1

Nvidia releases Nemotron 3 Ultra: 550B-parameter MoE model with 1M context window for agentic workflows

Nvidia has released Nemotron 3 Ultra, a 550-billion parameter mixture-of-experts model with 55 billion active parameters and support for up to 1 million token context windows. The model uses a hybrid Transformer-Mamba architecture and is designed specifically for long-running agentic workflows including agent orchestration, coding agents, and complex enterprise tasks.

June 4, 2026
product updateApple

Apple to Use Nvidia Blackwell B200 GPUs in Google Cloud for Gemini-Powered Siri

Apple will process some Siri queries using Nvidia's Blackwell B200 data center GPUs deployed in Google Cloud, according to The Information. The company plans to use Nvidia's confidential compute feature to encrypt data during processing on the chips.

May 21, 2026
analysisOpenAI

OpenAI reasoning model solves 80-year math problem as Anthropic hits $10.9B quarterly revenue

In a two-hour span Wednesday, OpenAI announced its reasoning model autonomously solved an 80-year-old geometry problem while Anthropic reported it's on track for $10.9 billion in Q2 revenue with $559 million in operating profit—two years ahead of internal projections. The developments came alongside Nvidia's $81.6 billion quarter, Anthropic's $1.25 billion monthly SpaceX compute deal, and a White House AI executive order signing.

April 24, 2026
product update

Alibaba's Qwen AI integrates with BYD, Volkswagen and 8 other Chinese automakers for voice-controlled services

Alibaba announced Friday that its Qwen AI model will be integrated into vehicles from 10 Chinese automakers including BYD, Geely, Li Auto, and SAIC Volkswagen. The system runs on Nvidia's automotive chip platform and allows drivers to order food delivery, book hotels, and make payments through voice commands, even with limited network connectivity.

February 27, 2026
fundingOpenAI

OpenAI raises $110B in largest private funding round ever, valued at $730B

OpenAI has raised $110 billion in what is now the largest private funding round in history, according to the company. The round values OpenAI at $730 billion and includes a $50 billion investment from Amazon alongside $30 billion each from Nvidia and SoftBank.