Why AI agents are moving to your desktop Perplexity is doing what it does best. Releasing something that feels like others could soon copy. But this time, Nvidia is joining in as a partner. On Tuesday, Perplexity announced Portable Computer, a version of its AI agent that can run open models locally on your own desktop hardware to do three things: Save significant money on tokens Get faster performance Boost data privacy That’s the kind of stuff more and more people have been trying to do because of soaring token bills due to AI agents and because of wanting to run agents on their most sensitive data, which they’d prefer to run locally and not send to the latest frontier models in the cloud from Anthropic, OpenAI, Google, and others. Where Nvidia comes in is that Portable Computer will run on Nvidia’s DGX Spark machine, which has cult favorite status among AI builders. It’s a tiny box the size of Mac mini, but with the power of a Mac Studio for running AI. Spark has 128GB of unified memory and runs up to a petaflop of compute. Nvidia says it runs models up to 200B parameters and can fine-tune models up to 70B.Perplexity’s Personal Computer ran on a Mac mini at launch, so a big part of the announcement is Perplexity’s AI agent now running on PC hardware powered by Nvidia GPUs. That means it’s likely to come to other DGX Spark competitors like the ones from Dell, Lenovo, Asus, MSI, and others. Right now, the software only runs on Linux. But Perplexity is also working on a version that will run on Windows, as well as a version that will run on Nvidia’s more powerful version of DGX Spark called DGX Station, which is the size of a full computer tower like a Mac Pro.The other big part of the announcement, of course, is being able to run Perplexity’s agent on open models that run locally on your machine. At launch, Portable Computer will run a post-trained version of Alibaba’s Qwen 3.8 27B called PPLX 27B, with a version of Nvidia’s Nemotron 3.5 Lightning coming to the device soon, according to Perplexity. You can also use a local inference server like Ollama or LM Studio to run any open model you’d like, but that starts to get a little more complicated.The problem that Perplexity and Nvidia are trying to solve is that more and more people are trying their hand at AI agents and would like to save the token costs and get the privacy you can get by running these models locally, but it’s a very involved process that requires time and technical expertise.”Now that we’re seeing people actually run frontier intelligence on their desk, the remaining bottleneck becomes that it’s quite cumbersome to set up,” said Nate Kupp, VP of infrastructure at Perplexity, in a briefing with the media. “We really focused on making this a straightforward experience where you can get up and running very quickly.”While DGX Spark and competitors cost $4,000-$5,000, they can save you so much in token costs that AI builders don’t even blink at that price. Heavy AI agent users can spend up to twice that a month in token costs. And if Portable Computer catches on, and other AI companies offer something similar, then it could increase competition and drive down the price tag. Perplexity’s best attribute is arguably how fast it can execute. It’s small and nimble and, again and again, the team has shown the ability to launch things quickly. Its AI agent was an idea that was hatched in about a month and launched around the same time OpenClaw went viral in early 2026. Perplexity’s general purpose agent can now help you code your own software like Claude Code or OpenAI’s Codex, but it can also help you carry out knowledge work tasks the same way Claude Cowork and ChatGPT Work can. The ability to run locally on open models is a super power. But even highly technical people have complained about how involved the setup is to get Nvidia’s DGX Spark up and running, so this would be a win for both Nvidia and AI enthusiasts if Perplexity can make that process more streamlined. ![]() |
-
Archives
- August 2026
- July 2026
- June 2026
- May 2026
- April 2026
- March 2026
- February 2026
- January 2026
- December 2025
- November 2025
- October 2025
- September 2025
- August 2025
- July 2025
- June 2025
- May 2025
- April 2025
- March 2025
- February 2025
- January 2025
- December 2024
- November 2024
- October 2024
- September 2024
- August 2024
- July 2024
- June 2024
- May 2024
- April 2024
- March 2024
- February 2024
- January 2024
- December 2023
- November 2023
- October 2023
- September 2023
- August 2023
- July 2023
- June 2023
- May 2023
- April 2023
- March 2023
- February 2023
- January 2023
- December 2022
- November 2022
- October 2022
- September 2022
- August 2022
- July 2022
- June 2022
- May 2022
- April 2022
- March 2022
- February 2022
- January 2022
- September 2021
- August 2021
- July 2021
- June 2021
- May 2021
- April 2021
- February 2021
- January 2021
- December 2020
- November 2020
- October 2020
- September 2020
- August 2020
- July 2020
- June 2020
- May 2020
- April 2020
- March 2020
- February 2020
- January 2020
- December 2019
- November 2019
- October 2019
- September 2019
- August 2019
- July 2019
- June 2019
- May 2019
- April 2019
- March 2019
- February 2019
- January 2019
- December 2018
- November 2018
- October 2018
- September 2018
- July 2018
- June 2018
- May 2018
- April 2018
- March 2018
- February 2018
- January 2018
- December 2017
- November 2017
- October 2017
- September 2017
- August 2017
- July 2017
- June 2017
- May 2017
- April 2017
- March 2017
- February 2017
- January 2017
- December 2016
- November 2016
- October 2016
- September 2016
- August 2016
- July 2016
- June 2016
- May 2016
- April 2016
- March 2016
- February 2016
- January 2016
- December 2015
- November 2015
- October 2015
- September 2015
- August 2015
- July 2015
- June 2015
- March 2015
- January 2015
-
Meta

Perplexity’s best attribute is arguably how fast it can execute. It’s small and nimble and, again and again, the team has shown the ability to launch things quickly. Its AI agent was an idea that was hatched in about a month and launched around the same time OpenClaw went viral in early 2026. Perplexity’s general purpose agent can now help you code your own software like Claude Code or OpenAI’s Codex, but it can also help you carry out knowledge work tasks the same way Claude Cowork and ChatGPT Work can. The ability to run locally on open models is a super power. But even highly technical people have complained about how involved the setup is to get Nvidia’s DGX Spark up and running, so this would be a win for both Nvidia and AI enthusiasts if Perplexity can make that process more streamlined. 