Perplexity unveils local AI platform for NVIDIA DGX Spark
Reported by MarkTechPost: Nvidia and Perplexity partner to bring AI to local hardware - Jon Peddie Research. Analysis and context written by TickrWire.
Perplexity introduced Portable Computer, a bundled local‑first AI agent system that runs on NVIDIA DGX Spark and eliminates per‑token fees for on‑device processing.

- Portable Computer bundles local models, orchestrator and sandbox into a single install on NVIDIA DGX Spark
- Local inference incurs no per‑token cost, while optional cloud escalation adds a modest fee per step
- Benchmarks show the system outperforms open‑source Pi and Hermes harnesses on several tasks
- The product requires high‑end hardware and is currently Linux‑only, with Windows support pending
Perplexity announced the launch of Portable Computer, a self‑contained AI agent platform that runs entirely on NVIDIA’s DGX Spark hardware. The offering combines the company’s agent harness, planner, tool router and post‑trained language models into a single installable package, allowing every task to start on the local device. When a step requires live web access or advanced reasoning, the system pauses, prompts the user, and forwards that single step to one of more than fifteen cloud models. This hybrid approach preserves data privacy while still tapping external expertise when needed.
The core of Portable Computer is a local inference engine that supports either the Qwen 3.8 27‑billion‑parameter model or Perplexity’s own PPLX 27B variant, which is fine‑tuned for the platform’s orchestrator. Both models run at 3‑bit quantization, with the Qwen model requiring a 17.4 GB download and about 32 GB of RAM, while the upcoming Nemotron 3.5 Lightning model, a 30‑billion‑parameter mixture‑of‑experts, will be delivered at 4‑bit precision and needs roughly 36 GB of RAM. The software package also includes an OS‑enforced sandbox that isolates tool execution, limits filesystem and network access, and disables tool calls if the sandbox cannot be established. Connectors for Gmail, Outlook, Slack and GitHub are built in, and users can bring their own models or inference servers.
Perplexity’s move reflects a broader industry shift toward “local‑first” AI, where enterprises seek to keep data processing on‑premises to reduce latency, control costs, and meet regulatory requirements. By eliminating per‑token charges for locally handled steps, the platform makes long verification loops and large‑scale migrations economically viable on owned hardware. The requirement for a GB10‑class box or an RTX GPU with at least 24 GB of VRAM positions the product for data‑center or high‑end workstation deployments rather than consumer laptops.
In benchmark testing, Portable Computer demonstrated notable gains over open‑source alternatives. On the 53‑task Local Knowledge Work Bench, the Qwen‑based configuration achieved an 82.6 % success rate, surpassing the open‑source Pi harness (77.6 %) and Hermes (74.0 %) on the same model. The PPLX 27B variant lifted performance to 85.4 %. On the BrowseComp suite, the system reached 66.7 % versus 50.2 % for Pi and 43.9 % for Hermes, while using 51 % less wall‑time and 70 % fewer tokens. Visual document understanding on ParseBench‑100 yielded 65.1 % compared with 34.6 % and 13.9 % for the competitors. A hybrid run on Terminal Bench 2.1 showed a fully local score of 59.6 % at effectively zero marginal cost, which rose to 73.0 % when a cloud adviser was invoked at roughly $0.415 per rollout, narrowing the gap to frontier models like Claude Opus 5.
Despite the strong performance numbers, the solution has clear constraints. It currently supports only Linux for Pro‑level subscribers, with Windows slated for a later release and macOS not planned. Only a single DGX Spark node is supported at launch, and clustering capabilities are listed as future work. The hardware prerequisites, GB10 superchip, 128 GB memory, and at least 1 TB storage, make the offering inaccessible to smaller teams. Moreover, the reliance on a sandbox means that any tool that cannot be sandboxed is simply disabled, potentially limiting functionality in edge cases.
Looking ahead, Perplexity plans to open‑source the 53‑task Local Knowledge Work Bench, which could foster community contributions and broader adoption. The roadmap includes support for additional hardware platforms, clustering, and the upcoming Nemotron 3.5 Lightning model. Observers will watch for how the hybrid escalation model balances cost and performance, and whether the zero per‑token pricing model spurs wider migration of enterprise workloads to on‑premise AI.
Provides a ready‑to‑run local AI stack with built‑in sandboxing, reducing setup time for enterprise applications
Enables cost‑effective, privacy‑preserving AI workloads without per‑token cloud fees
Signals Perplexity’s move into high‑margin enterprise AI infrastructure
Offers a glimpse of how hybrid local‑cloud AI systems can balance performance and cost
- OS‑enforced sandbox
- A security layer that isolates tool processes, restricting file system and network access
- DGX Spark
- NVIDIA’s high‑performance AI server platform featuring the GB10 superchip
- Qwen 3.8 27B
- A 27‑billion‑parameter language model with a 260 k token context window
AI bias estimate: The source emphasizes performance gains and cost benefits without discussing potential vendor lock‑in or long‑term support concerns (Automated estimate, not a definitive judgement.)
- Nvidia and Perplexity partner to bring AI to local hardware - Jon Peddie Research ↗
- Perplexity AI launches Portable Computer on-device AI agent - SiliconANGLE ↗
- Perplexity and NVIDIA team up to release a local AI agent - How-To Geek ↗
- Perplexity, Nvidia partner to run AI directly on desktop instead of major cloud providers - CNBC ↗
- Perplexity partners with Nvidia to launch Portable Computer, a fully local AI agent with zero token costs - Venturebeat ↗
- Perplexity Ships Portable Computer on NVIDIA DGX Spark: Local Harness, OS-Enforced Sandbox, and Zero Per-Token Cost for Local Steps ↗
- Portable Computer is Perplexity's new local AI agent - why it's a game changer ↗
AI ToolsIBM Says Granite Speech 5.0 Transcribes 3.5 Hours of Speech in One Second
AI ToolsMeta's paid AI agent Hatch launches soon, with a new model called Watermelon due in October
AI ToolsAccel-backed Keenable is indexing the web for AI agents
AI ToolsNvidia says its Groq 3 LPX is four times faster than Cerebras, but the math is more complicated
AI Tools‘The world seems to be ready’: An interview with OpenAI head of product Thibault Sottiaux
FundingIndia’s Ringg gets backing from Peak XV as it pushes voice AI past the phone call
Indian voice AI startup Ringg has raised $10 million in a Series A extension led by Peak XV Partners, bringing its total funding in the round to $15.5 million.
RoboticsRobotics startup Generalist reaches $3B valuation, sources say
Robotics startup Generalist secured a nearly $200 million funding extension led by 8VC, lifting its valuation to $3 billion just months after a major Series B round.
BusinessOpenAI loses a top data center exec as stream of high-profile departures continues
OpenAI’s head of data centers, Chris Malone, has left the company as part of a broader executive exodus, raising questions about leadership stability ahead of a planned IPO.
AI ResearchAI Method Reveals What Genomic Models Learn From DNA and Exposes Hidden Experimental Bias
Researchers at the Stowers Institute introduced PISA, a pairwise influence by sequence attribution method that visualizes, at single‑base resolution, what deep‑learning models learn from DNA and can strip experimental bias from MNase‑seq data.
FundingStability AI, maker of image generator Stable Diffusion, raises $76 million in fresh funding
Stability AI announced a $76 million Series B round, bringing its total funding to $232 million. Investors include Universal Music Group, Sony Music, Warner Music, Electronic Arts, AMD Ventures and Pacific Alliance Ventures.
SecurityRussia used ChatGPT to run a covert influence campaign pushing pro-Kremlin narratives across the West
OpenAI banned 36 ChatGPT accounts linked to a Russian influence campaign that used AI to generate pro-Kremlin content across Western platforms.