Perplexity is releasing Portable Computer today, an on-device version of its agentic AI platform that eliminates cloud dependency and token costs entirely. The system runs locally on hardware users already control, starting with Nvidia's DGX Spark desktop supercomputer and Linux machines equipped with RTX GPUs.

This launch represents a fundamental shift in how AI agents operate. Traditional cloud-based AI services charge per token consumed. Running inference locally removes that friction. Users pay once for hardware, then run unlimited queries without accumulating API bills. For enterprises processing thousands of daily requests, this economics change matters considerably.

The partnership with Nvidia signals strategic alignment around edge AI deployment. DGX Spark positions itself as an accessible entry point to local inference, targeting professionals and smaller organizations that lack massive data center infrastructure. RTX GPU support broadens compatibility across existing workstations, making adoption easier for teams already running Nvidia cards.

Portable Computer preserves the agentic capabilities of Perplexity's cloud version. Agents autonomously break down complex tasks, search information, and execute actions across connected systems. Running this locally improves latency, eliminates network roundtrips, and keeps sensitive data on premise rather than routing it through external servers. For regulated industries handling proprietary information, this addresses a genuine compliance barrier.

The technology does carry tradeoffs. Local inference demands sufficient GPU VRAM and computational power. Not every machine qualifies. Users lose instant access to the largest frontier models unless they have enterprise-grade hardware. Perplexity presumably bundles optimized models sized for consumer and prosumer GPUs, sacrificing some capability for practical deployability.

This move follows broader industry momentum toward local AI. Companies like Ollama, oobabooga, and others proved substantial demand exists for open-source models running on personal devices. Large language model providers including Meta with Llama and Mistral with its open weights models enabled this shift. Perplexity competes in this space by wrapping agency and search capabilities around local inference rather than just offering bare models.

The zero token cost framing matters for positioning. Cloud providers like OpenAI, Anthropic, and Google charge per million tokens. A typical enterprise workflow burns millions monthly. Portable Computer offers escape velocity from that cost structure. Early adopters pay capital costs upfront but eliminate operational spending tied to query volume.

Perplexity's Computer agent already operates in enterprise markets. Adding a local variant creates product segmentation. Cloud-based Computer serves teams wanting managed infrastructure and latest models. Portable Computer targets organizations prioritizing cost control, data privacy, and latency reduction. Both feed Perplexity's wider strategy to position itself beyond simple search competitor into the AI agent infrastructure layer.

Integration with Nvidia's ecosystem strengthens the partnership's reach. DGX Spark targets the professional market Perplexity already pursues. RTX GPU support includes GeForce and RTX cards used by millions globally. This distribution potential exceeds what Perplexity achieves selling cloud compute alone.

The practical impact depends on model quality in the local variant. If Perplexity bundles competitive models with genuine agentic reasoning capabilities, organizations genuinely can replace cloud-dependent workflows. If the local versions underperform the cloud offering, adoption stalls. Early results will determine whether this represents genuine infrastructure shift or marketing maneuver.

Portable Computer launches today, available through Perplexity's hub platform.