
Perplexity and Nvidia Launch a Zero-Token-Cost Local AI Agent Device
The AMW Read
Concretizes yesterday's Nvidia investment/licensing talks into a shipped zero-marginal-cost local agent device, updating the AI Agents player map without yet resolving a cross-segment or structural debate.
Perplexity and Nvidia Launch a Zero-Token-Cost Local AI Agent Device
Perplexity has partnered with Nvidia to launch Portable Computer, a device that runs a fully local AI agent with no per-token inference costs. Unlike cloud-based agent products that meter usage through API calls, the device performs inference entirely on local Nvidia hardware, eliminating the recurring cloud and API fees that typically accompany agentic AI use.
The launch lands one day after reports that Nvidia was in talks to invest in Perplexity at a $30 billion-plus valuation and possibly sign a technology licensing deal, and it gives concrete form to an argument Perplexity's CEO made last month β that AI competition is shifting from raw model size toward operational and orchestration efficiency. A hardware product that removes the per-token cost line turns that thesis into a shipped item rather than commentary, and it deepens Nvidia's role in Perplexity's stack beyond simply supplying chips. Perplexity, tracked in the AI Market Watch index as an AI Agents company since its 2022 founding, is one of a limited set of players trying to move agent economics off metered cloud inference entirely.
For builders, a zero-token-cost local agent removes a cost structure that has constrained how aggressively agentic products can be used, which could reshape usage patterns for latency- or cost-sensitive workflows. For investors, the timing suggests Nvidia is using investment and licensing conversations to embed itself into the application layer of agentic AI, not just the chips underneath it β worth watching if the reported $30B+ investment talks close.



