Local inference · 26 Aug 2026
Perplexity and Nvidia ship a local AI agent platform that bills nothing per token
Portable Computer runs Qwen, Perplexity's own PPLX model and Nvidia's Nemotron on a single RTX GPU with at least 24GB of memory, escalating to cloud models only when needed.