Perplexity is doing what it does best. Releasing something that feels like others could soon copy.
But this time, Nvidia is joining in as a partner.
On Tuesday, Perplexity announced Portable Computer, a version of its AI agent that can run open models locally on your own desktop hardware to save a ton of money in tokens, get faster performance, and boost data privacy.
That's the kind of stuff more and more people have been trying to do because of soaring token bills due to AI agents and because of wanting to run agents on their most sensitive data, which they'd prefer to run locally and not send to the latest frontier models in the cloud from Anthropic, OpenAI, Google, and others.
Where Nvidia comes in is that Portable Computer will run on Nvidia's DGX Spark machine, which has cult favorite status among AI builders. It's a tiny box the size of Mac mini, but with the power of a Mac Studio for running AI. Spark has 128GB of unified memory and runs up to a petaflop of compute. Nvidia says it runs models up to 200B parameters and can fine-tune models up to 70B.
Perplexity's Personal Computer ran on a Mac mini at launch, so a big part of the announcement is Perplexity's AI agent now running on PC hardware powered by Nvidia GPUs. That means it's likely to come to other DGX Spark competitors like the ones from Dell, Lenovo, Asus, MSI, and others. Right now, the software only runs on Linux. But Perplexity is also working on a version that will run on Windows, as well as a version that will run on Nvidia's more powerful version of DGX Spark called DGX Station, which is the size of a full computer tower like a Mac Pro.
The other big part of the announcement, of course, is being able to run Perplexity's agent on open models that run locally on your machine. At launch, Portable Computer will run a post-trained version of Alibaba's Qwen 3.8 27B called PPLX 27B, with a version of Nvidia's Nemotron 3.5 Lightning coming to the device soon, according to Perplexity. You can also use a local inference server like Ollama or LM Studio to run any open model you'd like, but that starts to get a little more complicated.
The problem that Perplexity and Nvidia are trying to solve is that more and more people are trying their hand at AI agents and would like to save the token costs and get the privacy you can get by running these models locally, but it's a very involved process that requires time and technical expertise.
"Now that we're seeing people actually run frontier intelligence on their desk, the remaining bottleneck becomes that it's quite cumbersome to set up," said Nate Kupp, VP of infrastructure at Perplexity, in a briefing with the media. "We really focused on making this a straightforward experience where you can get up and running very quickly."
While DGX Spark and competitors cost $4,000-$5,000, they can save you so much in token costs that AI builders don't even blink at that price. Heavy AI agent users can spend up to twice that a month in token costs. And if Portable Computer catches on, and other AI companies offer something similar, then it could increase competition and drive down the price tag.
Our Deeper View #
Perplexity's best attribute is arguably how fast it can execute. It's small and nimble and, again and again, the team has shown the ability to launch things quickly. Its AI agent was an idea that was hatched in about a month and launched around the same time OpenClaw went viral in early 2026. Perplexity's general purpose agent can now help you code your own software like Claude Code or OpenAI's Codex, but it can also help you carry out knowledge work tasks the same way Claude Cowork and ChatGPT Work can. The ability to run locally on open models is a super power. But even highly technical people have complained about how involved the setup is to get Nvidia's DGX Spark up and running, so this would be a win for both Nvidia and AI enthusiasts if Perplexity can make that process more streamlined.