Perplexity uses GPT-6 Astra for end-to-end system management
Perplexity has adopted OpenAI's GPT-6 Astra to handle communications, software changes, and production monitoring, allowing the team to check in less frequently than with earlier models.
Perplexity has adopted OpenAI's GPT-6 Astra to handle communications, software changes, and production monitoring, allowing the team to check in less frequently than with earlier models.
At the 37th Hot Chips conference, OpenAI presented benchmark details for its custom inference chip, Jalapeño, claiming superior performance per watt compared to NVIDIA's GB200 and GB300 systems. The chip delivers 1.5–1.9× more work per watt at peak throughput and 1.7–3.6× lower end-to-end latency, with deployment into OpenAI's infrastructure scheduled by year-end.
Perplexity has released a hybrid compute feature for Mac that splits agentic tasks between cloud frontier models and local models, using an on-device privacy gate to handle sensitive data. This allows users to leverage powerful cloud reasoning while keeping private files like deal documents or client records strictly on their hardware.
Perplexity has released Portable Computer, a local-first build of its agentic platform that runs the agent harness, orchestrator, planner, tool router, and post-trained models directly on NVIDIA DGX Spark hardware. This packaged system allows every task to begin on-device, carrying no per-token charge for work handled by local models, while escalating only specific steps to 15+ cloud models upon user approval.