Introducing Fast Search in the Perplexity Search API.
Fast Search runs on Photon, our new Rust-based retrieval and ranking service that we built with a small team of engineers and hundreds of agents.
It returns 95% of search results in 230 ms or less.
Show more
Portable Computer for Windows is now available on
@AMD Ryzen AI Max Series processors.
It makes it easy to run local AI agents that work with your connected apps and local files. Kick off tasks or schedule recurring work that runs entirely on your device.
Show more
New research: We post-trained a Computer model to learn from its own errors using hint-guided self-distillation.
In a live A/B test, a later trained checkpoint reduced tool-call failures by 21.2% relative to an earlier checkpoint.
Show more
GPT-6 Sol is now available in Perplexity and Computer.
It's now the default Light option in Computer's effort selector.
Claude Opus 5.5 is now available as the Standard effort level in Perplexity Computer.
On our WANDR benchmark, it scored 0.610 at $4.13 per task, slightly outperforming Claude Fable 5.1 while costing 67.6% less per task.
Show more
Perplexity Computer can now create videos using MiniMax H3 and ByteDance Seedance 2.5.
Ask Computer for a campaign clip, product demo, or social asset, and it produces finished video next to copy and creative in same thread.
Available now for Pro and Max subscribers.
Show more
We’re rolling out effort controls in Computer’s model selector.
Effort presets combine the orchestrator model and reasoning depth to control how deeply and efficiently Computer works through a task.
Available now on web. Coming to mobile and desktop soon.
Show more
We’re publishing research on how we built CobbleDB, our key-value database that serves web content for Perplexity search.
Two engineers and a team of hundreds of proactive, always-on AI agents built the core infrastructure in two months.
Show more
Perplexity Computer will now come pre-installed on HP ZBook Ultra G3a. .
Use Computer to run complex multi-step work from a simple interface. Computer agents are grounded in accurate deep research and connected to hundreds of tools, now including Autodesk.
Show more
Portable Computer is now available on Windows PCs with
@NVIDIA RTX GPUs.
Run the harness, agents, and models locally on your PC.
Work with local files and connected apps without sending tasks to the cloud. Use frontier cloud models when needed.
Show more
The corpus combines the top 5,000 production retrieval results per query, deduplicated with MinHash-LSH.
Each document is a plausible match for at least one query, including difficult distractors that match the topic but miss a required date, entity, or version.
Show more
We're introducing Q2D-Web (Query2Doc-Web), a benchmark and public leaderboard for evaluating retrieval in agentic RAG systems.
Q2D-Web tests how embedding models perform on large-scale web search using agent-reformulated search queries.
Read more:
Show more
GPT-6 Astra is now available in Perplexity Computer for Pro and Max subscribers.
Our serving infrastructure lowers latency and improves throughput across both online and batch embedding workloads.
Combining Ivy, Tulip, and ROSE results in faster search at a reduced cost compared to off-the-shelf solutions.
Show more
Ivy is the HTTP gateway.
It handles CPU-side request prep: parsing, tokenization, templating, and splitting large batches before sending them to Tulip over gRPC.
It allows us to tune request formatting and tokenization without touching the heavier inference servers.
Show more
Every answer in Perplexity starts with embedding and ranking models picking the most relevant results for the query.
Today we published research on how we built SoTA serving infrastructure behind those models.
Read the research:
Show more
We evaluated GPT-6 Astra on WANDR. It scored 0.682 at $11.98 per task, the highest score of any model we tested.
GPT-6-Astra scored 13.5% higher than Fable 5.1 at 6.1% lower cost, and 27.0% higher than Opus 5 at 3.3% higher cost.
Show more
Get started with Portable Computer:
Portable Computer is now available on Linux for
@NVIDIA RTX GPUs with 24GB of VRAM or higher.
Today we’re launching Portable Computer on
@NVIDIA DGX Spark.
Portable Computer is a fully local version of Perplexity Computer, where the entire runtime: orchestrator LLM, subagent LLM, agent harness all run on your local hardware. No cloud dependency.
Show more
Today we’re open-sourcing Lily, the local inference engine we built for hybrid compute in Perplexity Computer.
Lily is specialized for Qwen3.6-35B-A3B on Apple silicon, built so on-device compute doesn’t bottleneck Computer tasks.
Read more:
Show more