@stretchcloud: Local inference is becoming a product category, not just a benchmark table. Perplexity launched Portable Computer on Wi…
Summary
Perplexity launched Portable Computer on Windows RTX PCs, enabling local AI inference with an on-device agent harness, shifting the focus from inference cost to harness quality in the competitive landscape.
View Cached Full Text
Cached at: 09/15/26, 01:52 PM
Local inference is becoming a product category, not just a benchmark table.
Perplexity launched Portable Computer on Windows RTX PCs today. The orchestrator LLM, subagent LLM, and the full agent harness all run on-device. No cloud dependency for most tasks. Zero token costs when running locally.
The hardware requirement: an NVIDIA RTX GPU with at least 24GB of VRAM. RTX 3090 and newer. The default local model is Qwen 3.8 27B, post-trained specifically for Perplexity Computer’s harness. Frontier cloud models are called on-demand when reasoning demands them.
This is not another chat wrapper running locally. It is a full agentic stack with task planning, subagent delegation, file access, and connected apps running on the hardware you own.
The competitive landscape: Ollama does local model serving. LM Studio and http://Jan.ai add desktop UIs. AnythingLLM adds RAG. None of them ship a production-grade agent harness with orchestrator and subagent tiers as a first-class primitive.
NVIDIA has over 100 million RTX GPUs in market. The share with 24GB or more is a fraction today, but each new RTX generation pushes 24GB to mainstream price points.
My read: the bottleneck shifts from inference cost to harness quality. Perplexity is positioning as the harness, not the model. That is a real strategic bet.
Jan - Open-Source ChatGPT Replacement
Source: https://www.jan.ai/
Meet Jan
Personal Intelligence that answers only to you



An open model ecosystem
Jan Agent
The core agent, distributed separately — run it on your own VM or container.
Tokamak
Router, fusion model, and governance/audit — the self-hosted backend Jan agents connect to.
Over 4 million downloads
Jan is built in public
We believe AI should be open, and grow through the people who build and use it

All the tools you need to make Jan yours
1
Models
Choose from open models or plug in your favorite online models.
ChatGPTOpenAI
ClaudeAnthropic
GeminiGoogle
LlamaMeta
MistralMistral AI
QwenAlibaba
DeepSeekDeepSeek
GemmaGoogle
KimiMoonshot AI
2
MemoryComing Soon
Your context carries over, so you don’t repeat yourself. Jan remembers your context and preferences.
![]()
Things Jan keeps in mind
- •Minimalist UI tasted
- •Currently on a portfolio refresh
- •Wants brief, to-the-point answers
- •Frequent Figma/prototyping questions
- •Dark-mode sharer
- •Curious about type trends (Mostly harmless)
Ask Jan anything
+6.6Mdownloads, Free & Open source

Perplexity (@perplexity_ai): Portable Computer is now available on Windows PCs with @NVIDIA RTX GPUs.
Run the harness, agents, and models locally on your PC.
Work with local files and connected apps without sending tasks to the cloud. Use frontier cloud models when needed.
Similar Articles
Perplexity Portable Computer Is Now Available on Windows, Powered by NVIDIA RTX
Perplexity introduces Portable Computer, a local AI agent for Windows PCs powered by NVIDIA RTX, enabling on-device task handling with optional cloud integration.
Perplexity partners with Nvidia to launch Portable Computer, a fully local AI agent with zero token costs (13 minute read)
Perplexity has partnered with Nvidia to launch Portable Computer, a fully local AI agent that runs on user-owned hardware like DGX Spark and RTX GPUs, eliminating cloud token costs and keeping data private.
@omarsar0: Can't wait to have always-on agents running locally 24x7. Local agents are something Perplexity seems to be chasing the…
NVIDIA announces Perplexity's Portable Computer, a local-first agent stack for NVIDIA DGX Spark that provides one-click local inference and optimized agentic experiences for always-on local agents.
@LinusEkenstam: Local model in Perplexity Just in time for the Apple event next week, Perplexity once again shows the pathway forward f…
Perplexity introduces hybrid compute for its Mac app, enabling local model inference for sensitive data while offloading to the cloud, marking a trend towards transparent local AI usage.
The Data Center Moves to Your Machine (4 minute read)
Perplexity unveiled a hybrid local-cloud inference system at Computex 2026 that intelligently routes queries between on-device and cloud models, building on its earlier Personal Computer agent.