Perplexity Launches Hybrid Compute for Mac, Splitting AI Tasks Between Cloud and Local Models
Perplexity's Mac app now supports Hybrid Compute, which starts tasks in the cloud and shifts sensitive steps to a local model running on-device. The feature requires Apple silicon with at least 24GB of unified memory and uses an open-sourced on-device PII classifier to mask private data before any cloud processing.
Perplexity has launched Hybrid Compute, a new feature in its Mac app that splits AI tasks between cloud models and a local model running on-device. The rollout follows an initial announcement of the work in June and arrives after Apple featured Perplexity's Personal Computer app as a productivity use case for the new M6 Mac mini.
How it works
According to Perplexity, its "Computer" assistant now starts each task in the cloud but can shift to a local model for steps that touch private files or sensitive data. A task triggered from an iPhone, for example, can hand off to the Mac to access local files and complete sensitive processing without that data leaving the device.
Privacy protection is handled by an on-device PII classifier that scans each task before anything is sent to the cloud. Names, addresses, and account numbers are replaced with placeholder values, then restored once the cloud model returns its answer. Perplexity says it developed the classifier with its Secure Intelligence Institute and has open-sourced it, though the company has not published independent benchmark data verifying its accuracy at detecting or masking personal information.
Hardware requirements
Hybrid Compute only runs on Apple silicon Macs with macOS 15 or later. Perplexity sets 24GB of unified memory as the minimum requirement and recommends 32GB for best results — meaning Macs with 8GB or 16GB of RAM cannot run the feature at all. This effectively limits Hybrid Compute to higher-end MacBook Pro, Mac Studio, and upper-tier Mac mini configurations.
Setup requires a one-click download of what Perplexity calls "PPLX Qwen 3.8 27B," a local model built on Alibaba's Qwen 3.8 architecture at 27 billion parameters. Perplexity says the process requires no manual runtime configuration or third-party tools like Ollama. Tasks handled entirely by the local model consume no cloud credits and require no API key, according to the company.
What this means
Hybrid Compute is a product feature, not a new foundation model — the underlying local component is a repackaged version of an existing Qwen checkpoint, not a Perplexity-trained model from scratch. Its real significance is architectural: Perplexity is betting that a split-execution model, where sensitive data never leaves the device while general reasoning still runs in the cloud, becomes a meaningful differentiator as AI assistants gain deeper access to personal files.
The 24GB memory floor is a hard constraint that excludes a large share of the existing Mac installed base, including every base-model MacBook Air and many entry Mac mini configurations sold in the past two years. That narrows Hybrid Compute's near-term audience to higher-end hardware buyers — likely by design, since Apple's own promotion of Perplexity alongside the M6 Mac mini suggests the two companies are coordinating on positioning higher-memory Macs as AI-capable machines. Whether the open-sourced PII classifier holds up to independent scrutiny will determine if this privacy pitch is substantive or largely marketing.
Related Articles
Perplexity Says It Runs End-to-End Engineering Systems on OpenAI's GPT-6 Astra
Perplexity says it has shifted core engineering workflows, including code changes and production monitoring, onto OpenAI's GPT-6 Astra model. The claim comes from an OpenAI-published case study with no independent benchmark data released.
Perplexity Deploys OpenAI's Astra Model for Autonomous Code and Systems Management
Perplexity is using an OpenAI model referred to as Astra to handle software changes, communications, and production monitoring with less frequent human check-ins. OpenAI published the case study; specific model specs and benchmarks have not been disclosed.
ElevenLabs Launches Music v2.5, Adds API Access and Free Tier for AI-Generated Songs
ElevenLabs has released Music v2.5, an updated version of its ElevenMusic generator, now available through both the app and API. The company says blind testing with nearly 48,000 comparison pairs showed listeners preferred v2.5 over the prior version, particularly for R&B, Hip-Hop, and orchestral genres.
Augment Code Claims 4.5x Developer Output Increase From Internal 'Software Factory' of AI Agents
Augment Code says its internal 'software factory'—a network of specialized agents built on its Cosmos platform—drove a 4.5x increase in size-adjusted developer output and cut median PR merge time from 11.2 to 3.1 hours over nine months. The company frames this as evidence that once AI writes nearly all new code, the bottleneck shifts to review, verification, and incident response.
Comments
Loading...