Perplexity Introduces Hybrid Local-Server AI Orchestrator at Computex
Perplexity AI has introduced what it calls the first hybrid local-server inference orchestrator, announced in a company blog post. The system, demonstrated at Computex 2026, automatically determines which AI workloads should execute on a user's device and which should run on cloud-based frontier models.
The new system, named Personal Computer, uses local inference for sensitive data such as health or financial records while routing more complex tasks to remote models. It operates autonomously, deciding on a task-by-task basis where each computation should occur. This approach aims to balance accuracy, privacy, and efficiency without requiring user input.
Perplexity unveiled the technology with Intel during the Computex keynote. The orchestrator runs on Intel’s Core Ultra Series 3 chips and is compatible with other hardware, including Nvidia’s RTX Spark. By distributing workloads between local devices and cloud infrastructure, Perplexity aims to reduce reliance on centralized data centers while keeping confidential information on users’ machines.
The company stated that Personal Computer with local inference will be available in July. The release extends Perplexity’s existing orchestration framework, previously used to manage AI models, to now include compute location as part of the decision process.
We hope you enjoyed this article
Consider subscribing to one of our newsletters like Silicon Brief or Daily AI Brief.
Also, consider following us on social media:
More from Data Centers
Sep 15 Cisco Expands Splunk AI for Private and Isolated Environments Sep 15 Delos Data Raises Over $100 Million for AI Infrastructure Sep 15 MediaTek Introduces Dimensity 9600 Pro Smartphone Chip Sep 15 EUCLYD Raises More Than €200 Million for AI Inference Infrastructure Sep 15 Lambda Signs White House Ratepayer Protection PledgeSilicon Brief
Weekly coverage of AI hardware developments including chips, GPUs, cloud platforms, and data center technology.
Whitepaper
Tensordyne Napier: What If One Rack Could Do the Work of Nine?
Tensordyne
This Tensordyne whitepaper presents Napier, an inference-focused AI processor and rack-scale system based on the company’s TDN Math logarithmic number system. It examines infrastructure requirements for large mixture-of-experts and agentic models, compares major inference architecture approaches, and details the TDN AIP processor, TDN72 pod, TDN Link fabric, and Napier Ultra configuration. The paper reports simulation-based performance, cost, and accuracy-validation results, including Tensordyne’s projected comparison of one Napier rack with a nine-rack Nvidia Rubin plus Groq deployment; the chip is reported as taped out and in fabrication.
Read moreYou may also like
Perplexity Open Sources Lily Inference Engine for Apple Silicon
ASUS Expands AI Infrastructure From Cloud to Edge
Acer Introduces Veriton RI110 AI Mini Workstation
ScitiX Introduces Enterprise AI Inference Platform
SCX.ai and DDN Partner on Australian Sovereign AI Inference Cloud
Daily AI Brief: the AI news that matters, in your inbox.