Perplexity is now running its entire system on NVIDIA’s DGX Spark computer, integrating the agent harness, orchestrator, and post-trained models onto this hardware. The system allows agents to run locally while accessing over 15 cloud models, but only after the orchestrator requests permission before routing a task to them, prioritizing user control, Portable Computer says. Portable Computer currently utilizes PPLX 27B or Qwen 27B for on-device AI processing, with Nemotron 3.5 Lightning scheduled to be added soon, expanding the capabilities of local AI workflows.
Local Agent Stack Deploys on NVIDIA DGX Spark
Portable Computer is now deploying its complete local-first agent stack on NVIDIA DGX Spark systems, a move indicating confidence in the hardware’s capacity for advanced artificial intelligence workflows. This architecture prioritizes local processing for speed and privacy, engaging cloud models only after explicit approval. Each task initiated within Portable Computer executes locally on the DGX Spark; however, the orchestrator functions as a gatekeeper, ensuring user consent before routing any task requiring extensive reasoning to one of over 15 integrated cloud models.
This approach balances the benefits of on-device AI with the power of remote computation, giving users control over data handling and processing location. A demonstration of this process involved reviewing pull requests on GitHub, where the local orchestrator grouped, tagged, and prepared a summary for posting to a Slack channel, all while maintaining data within the DGX Spark environment.
Portable Computer currently utilizes Qwen 3.8 27B and PPLX 27B for on-device AI processing, with Nemotron 3.5 Lightning coming soon, according to the company. The system’s sandbox infrastructure isolates code and tool execution, providing controlled access to local files and connected applications like Gmail, Outlook, and Slack, the company says. Work completed by the local models incurs no per-token charge, making large-scale operations such as repository migrations and batch file summarization economically viable on owned hardware.
The installation process sets up the local agent harness, orchestrator, and post-trained models with sandboxed execution. Available to Pro and Max subscribers, the system requires a DGX Spark with GB10, 128 GB of memory, and at least 1 TB of storage to function optimally.
See today’s quantum computing news on Quantum Zeitgeist for the latest breakthroughs in qubits, hardware, algorithms, and industry deals.
