A new pareto-frontier for Deep Research price-performance | Parallel
Introducing Parallel Search Turbo: the fastest and lowest-cost search for AI. Learn more.Learn more.
HumanMachine
September 9, 2025
\# A new pareto-frontier for Deep Research price-performance
Expanded results that demonstrate Parallel's complete price-performance advantage in Deep Research.
Tags:Benchmarks
Reading time: 4 min
We previously released benchmarkspreviously released benchmarks for Parallel Deep Research that demonstrated superior accuracy and win rates against leading AI models. Today, we're publishing expanded results that showcase our complete price-performance advantage - delivering the highest accuracy across every price point.
\## **Parallel leads in accuracy at every price point**
We evaluated Parallel against all available deep research APIs on two industry-standard benchmarks. Our processors consistently deliver the highest accuracy at each price tier.
\### **BrowseComp Benchmark**
OpenAI's BrowseComp tests deep research capabilities through 1,266 complex questions requiring multi-hop reasoning, creative search strategies, and synthesis across scattered sources.
BrowseComp
Accuracy (%)
Ultra8x58% / 2400CPM
Ultra4x56% / 1200CPM
Ultra2x51% / 600CPM
Ultra45% / 300CPM
GPT-538% / 488CPM
Pro34% / 100CPM
Exa14% / 402CPM
Anthropic7% / 5194CPM
Perplexity6% / 709CPM
CPM: USD per 1000 requests. Cost is shown on a Log scale.
Parallel
Others
Benchmark comparison across Cost (CPM) and Accuracy (%). CPM: USD per 1000 requests. Cost is shown on a Log scale.
+−Methodology
\### About the benchmark
This benchmarkbenchmark, created by OpenAI, contains 1,266 questions requiring multi-hop reasoning, creative search formulation, and synthesis of contextual clues across time periods. Results are reported on a random sample of 100 questions from this benchmark.
\### Methodology
- - Dates: All measurements were made between 08/11/2025 and 08/29/2025.
- - Configurations: For all competitors, we report the highest numbers we were able to achieve across multiple configurations of their APIs. The exact configurations are below.
- - GPT-5: high reasoning, high search context, default verbosity
- - Exa: Exa Research Pro
- - Anthropic: Claude Opus 4.1
- - Perplexity: Sonar Deep Research reasoning effort high
\### About the benchmark
\### Methodology
\### New Browsecomp
| Series | Model | Cost (CPM) | Accuracy (%) |
| --------- | ---------- | ---------- | ------------- |
| Parallel | Pro | 100 | 34 |
| Parallel | Ultra | 300 | 45 |
| Parallel | Ultra2x | 600 | 51 |
| Parallel | Ultra4x | 1200 | 56 |
| Parallel | Ultra8x | 2400 | 58 |
| Others | GPT-5 | 488 | 38 |
| Others | Anthropic | 5194 | 7 |
| Others | Exa | 402 | 14 |
| Others | Perplexity | 709 | 6 |
CPM: USD per 1000 requests. Cost is shown on a Log scale.
\### About the benchmark
\### Methodology
Our results demonstrate clear price-performance leadership, with our Ultra processor achieving 45% accuracy at $300 CPM at up to 17X lower cost compared to alternatives. Our newly available high-compute processors push accuracy even further for critical research tasks, with Ultra8x reaching 58%.
**DeepResearch Bench**
DeepResearch Bench evaluates the quality of long-form deep research reports across 22 fields including Business & Finance, Science & Technology, and Software Development. The benchmark consists of 100 PhD-level tasks and assesses the multistep web exploration, targeted retrieval, and higher-order synthesis capabilities of deep research agents.
DeepResearch Bench
Win Rate vs Reference (%)
Ultra8x96% / 2400CPM
Ultra4x92% / 1200CPM
Ultra2x86% / 600CPM
Ultra82% / 300CPM
GPT-566% / 628CPM
O3 Pro30% / 4331CPM
O326% / 605CPM
Perplexity6% / 538CPM
300,-10ULTRA82% / 300CPMULTRA2X86% / 600CPMULTRA4X92% / 1200CPMULTRA8X96% / 2400CPMGPT-566% / 628CPMO3 PRO30% / 4331CPMO326% / 605CPMPERPLEXITY6% / 538CPM
COST (CPM)
WIN RATE VS REFERENCE (%)
CPM: USD per 1000 requests. Cost is shown on a Log scale.
Parallel
Others
Benchmark comparison across Cost (CPM) and Win Rate vs Reference (%). CPM: USD per 1000 requests. Cost is shown on a Log scale.
+−Methodology
\### About the benchmark
This benchmarkbenchmark contains 100 expert-level research tasks designed by domain specialists across 22 fields, primarily Science & Technology, Business & Finance, and Software Development. It evaluates AI systems' ability to produce rigorous, long-form research reports on complex topics requiring cross-disciplinary synthesis. Results are reported from the subset of 50 English-language tasks in the benchmark.
\### Methodology
- - Dates: All measurements were made between 08/11/2025 and 08/29/2025.
- - Win Rate: Calculated by comparing RACERACE scores in direct head-to-head evaluations against reference reports.
- - Configurations: For all competitors, we report results for the highest numbers we were able to achieve across multiple configurations of their APIs. The exact GPT-5 configuration is high reasoning, high search context, and high verbosity.
- - Excluded API Results: Exa Research Pro (0% win rate), Claude Opus 4.1 (0% win rate).
\### About the benchmark
\### Methodology
\### RACER
| Series | Model | Cost (CPM) | Win Rate vs Reference (%) |
| -------- | ---------- | ---------- | ------------------------- |
| Parallel | Ultra | 300 | 82 |
| Parallel | Ultra2x | 600 | 86 |
| Parallel | Ultra4x | 1200 | 92 |
| Parallel | Ultra8x | 2400 | 96 |
| Others | GPT-5 | 628 | 66 |
| Others | O3 Pro | 4331 | 30 |
| Others | O3 | 605 | 26 |
| Others | Perplexity | 538 | 6 |
CPM: USD per 1000 requests. Cost is shown on a Log scale.
\### About the benchmark
\### Methodology
Parallel Ultra achieves an 82% win rate against reference reports at $300 CPM, compared to GPT-5's 66% win rate at $628 CPM - delivering superior quality at half the cost. Our highest compute processor, Ultra8x, reaches a 96% win rate, representing a significant improvement from our previously published 82% benchmark.
We also measured win rate against GPT-5 directly by comparing the RACE scores of Parallel processors vs GPT-5. The results demonstrate that Ultra8x achieves an 88% win rate against GPT-5.
Head-to-head comparison with GPT-5
0%10%20%30%40%50%60%70%80%90%Ultra8xUltra4xUltra2xUltraModel Win Rate vs GPT-5 (%)
Performance comparison proving Parallel delivers the best enterprise deep research API for ChatGPT and AI agents with 48% accuracy vs competitors' 14% max across Model and Win Rate vs GPT-5 (%). Multi-hop research benchmark shows Parallel's structured AI agent deep research outperforms GPT-4, Claude, Exa, and Perplexity. Enterprise-ready structured deep research API with MCP server integration.
\### DeepResearch Bench against GPT-5
| Category | Win Rate (%) |
| -------- | ------------ |
| Ultra8x | 88 |
| Ultra4x | 84 |
| Ultra2x | 80 |
| Ultra | 74 |
\## **Beyond benchmarks: Flexible outputs, fully verifiable**
These benchmark results translate directly to production value. Parallel Deep Research delivers the same high accuracy in whichever format you need - human-readable reports for strategic analysis or structured JSON for machine consumption and database ingestion.
Every output, regardless of format, includes our comprehensive Basis framework:
- - **Citations**: Direct links to source materials
- - **Reasoning**: Explanations for each finding
- - **Confidence**: Calibrated scores (low/medium/high) for intelligent routing
- - **Excerpts**: Relevant text snippets from cited sources
This complete verification layer means the accuracy demonstrated in our benchmarks comes with the audibility and transparency required for production workflows where every detail matters.
\## **Built for scale: 1000x more research, predictably priced**
Our price-performance advantage unlocks new possibilities. At these price points, you can run 1000x the number of queries compared to token-based alternatives - transforming deep research from an occasional tool to core infrastructure.
Consider the possibilities:
- - **Build research databases**: Run thousands of queries, store structured results, and query them downstream
- - **Continuous intelligence**: Monitor competitors, marketsmarkets, and trends with daily deep research updates
- - **Pipeline integration**: Use research outputs as inputs for downstream analysis, decision-making, or automation
- - **Parallel processing**: Research hundreds of entities simultaneously for large-scale enrichment
Our per-query pricing model ensures complete cost predictability. Unlike token-based systems where a single complex query can unexpectedly consume your budget, every Parallel query costs exactly what you expect. This predictability enables confident scaling - whether you're running 10 queries or 10,000.
\## **Start building with Deep Research**
Parallel Deep Research is available today through our Task API. Choose the processor that matches your accuracy and budget requirements, from Pro for simpler deep research to Ultra8x for the most demanding deep research tasks.
Get started in our Developer PlatformDeveloper Platform or explore our documentationdocumentation.
\## **Notes on Methodology**
_Benchmark Dates_: Benchmarks were run from Aug 11 to Aug 29.
_DeepResearchBench Evaluation_ **: ** We evaluated all available DeepResearch API solutions on the 50 English-language tasks in the benchmark, measuring both RACE and FACT scores for generated reports. Given that RACE is a relative scoring metric benchmarked against reference materials, we calculated win-rates by comparing each vendor's performance to the human reference reports included in the dataset. A candidate report achieves a "win" when its RACE score exceeds that of the corresponding human reference report.
_BrowseComp Evaluation_ **: ** For the BrowseComp benchmark, we tested our processors alongside other APIs on a random 100-question subset of the original 1,266-question dataset. All systems were evaluated using the same standard LLM evaluator with consistent evaluation criteria, comparing agent responses against verified ground truth answers.
_Cost Calculation_: Token-based pricing is normalized to cost per thousand queries (CPM) based on actual usage in benchmarks.
\## Ready to get started?
Sign up for free. No credit card required.
Try ParallelTry Parallel Contact salesContact sales
Are you an agent? Read this to onboard ParallelAre you an agent? Read this to onboard Parallel
By Parallel
September 9, 2025
\## Related Posts80
Jul 30, 2026\ - Building an always-on background agent to proactively support customers
Author: By Khushi Shelat
Jul 21, 2026\ - Introducing the Parallel Responses API
Author: By Parallel
Jul 20, 2026\ - Building a vendor intelligence system with Parallel
Author: By Sahith Jagarlamudi
Jul 16, 2026\ - Parallel and Google Cloud Announce Partnership for Agentic Web Search on Gemini Enterprise Agent Platform
Author: By Parallel
Jul 15, 2026\ - $5 in free Parallel credits, every month
Author: By Parallel
\ \ Jul 13, 2026\ \ - [Introducing Parallel Search Turbo](/content/blog/parallel-search-turbo/index.html) \ \ Author: By Parallel](/content/blog/parallel-search-turbo/index.html)
Jul 12, 2026\ - Building a realtime voice agent with GPT-Realtime-2.1 and Parallel Search Turbo
Author: By George Pickett
Jul 10, 2026\ - How Nooks cut web search costs 70.5% by switching to Parallel
Author: By Parallel
Jul 8, 2026\ - How Build created live geofenced alerts powered by Parallel for institutional real estate
Author: By Parallel
Jun 9, 2026\ - OpenClaw now has free, LLM-optimized web search by default powered by Parallel
Author: By Parallel
Jun 5, 2026\ - Introducing real-time Entity Search
Author: By Parallel
Jun 3, 2026\ - How we enrich & triage inbound leads using the Parallel Task API
Author: By Khushi Shelat
May 20, 2026\ - How AirOps creates citation-worthy content at scale, powered by Parallel
Author: By Parallel
May 18, 2026\ - Introducing Index by Parallel
Author: By Parallel
May 7, 2026\ - Parallel Monitor API: New processor tiers, snapshots and event streams, and Basis on every event
Author: By Parallel
May 4, 2026\ - How we built parallelmpp.dev
Author: By Son Do
Apr 29, 2026\ - How Actively's Per Account Agents use Parallel to turn the entire web into a proactive sales intelligence layer
Author: By Parallel
Apr 28, 2026\ - Parallel Raises at $2 Billion Valuation to Scale Web Infrastructure for Agents
Author: By Parallel
Apr 24, 2026\ - Building a free CLI agent with Pi, Ollama, Gemma 4, and Parallel
Author: By Matt Harris
Apr 23, 2026\ - Parallel Search is now free for agents via MCP
Author: By Parallel
Apr 21, 2026\ - Upgrades to the Parallel Search & Extract APIs
Author: By Parallel
Apr 20, 2026\ - How Finch is scaling plaintiff law with AI agents that research like associates
Author: By Parallel
Apr 8, 2026\ - Genpact and Parallel Web Systems Partner to Drive Tangible Efficiency from AI Systems
Author: By Parallel
Apr 8, 2026\ - How Genpact helps top US insurers cut contents claims processing times in half with Parallel
Author: By Parallel
Apr 7, 2026\ - A new deep research frontier on DeepSearchQA with the Task API Harness
Author: By Parallel
Mar 30, 2026\ - How Modal saves tens of thousands annually by building in-house GTM pipelines with Parallel
Author: By Parallel
Mar 25, 2026\ - How Opendoor uses Parallel as the enterprise grade web research layer powering its AI-native real estate operations
Author: By Parallel
Mar 19, 2026\ - Introducing stateful web research agents with multi-turn conversations
Author: By Parallel
Mar 18, 2026\ - Parallel is live on Tempo, now available natively to agents with the Machine Payments Protocol
Author: By Parallel
Mar 17, 2026\ - How Parallel helped Kepler build AI that finance professionals can actually trust
Author: By Parallel
Mar 10, 2026\ - Introducing the Parallel CLI
Author: By Parallel
Mar 4, 2026\ - How Profound helps brands win AI Search with high-quality web research and content creation powered by Parallel
Author: By Parallel
Mar 2, 2026\ - How Harvey is expanding legal AI internationally with Parallel
Author: By Parallel
Feb 23, 2026\ - How Tabstack by Mozilla enables agents to navigate the web with Parallel’s best-in-class web search
Author: By Parallel
Feb 4, 2026\ - Parallel Web Tools and Agents now available across Vercel AI Gateway, AI SDK, and Marketplace
Author: By Parallel
Jan 28, 2026\ - Authenticated page access for the Parallel Task API
Author: By Parallel
Jan 21, 2026\ - Introducing structured outputs for the Monitor API
Author: By Parallel
Jan 15, 2026\ - Introducing research models with Basis for the Parallel Chat API
Author: By Parallel
Jan 8, 2026\ - Build a real-time fact checker with Parallel and Cerebras
Author: By Parallel
Dec 17, 2025\ - Parallel Task API achieves state-of-the-art accuracy on DeepSearchQA
Author: By Parallel
Dec 16, 2025\ - Introducing Granular Basis for the Task API
Author: By Parallel
Dec 11, 2025\ - How Amp’s coding agents build better software with Parallel Search
Author: By Parallel
Dec 10, 2025\ - Latency improvements on the Parallel Task API
Author: By Parallel
Nov 20, 2025\ - Introducing Parallel Extract
Author: By Parallel
Nov 18, 2025\ - Introducing Parallel FindAll
Author: By Parallel
Nov 13, 2025\ - Introducing Parallel Monitor
Author: By Parallel
Nov 12, 2025\ - Parallel raises $100M Series A to build web infrastructure for agents
Author: By Parallel
Nov 11, 2025\ - How Macroscope reduced code review false positives with Parallel
Author: By Parallel
Nov 6, 2025\ - Introducing Parallel Search
Author: By Parallel
Nov 3, 2025\ - Parallel processors set new price-performance standard on SealQA benchmark
Author: By Parallel
Oct 30, 2025\ - Introducing LLMTEXT, an open source toolkit for the llms.txt standard
Author: By Parallel
Oct 23, 2025\ - How Starbridge powers public sector GTM with state-of-the-art web research
Author: By Parallel
Oct 22, 2025\ - Building a market research platform with Parallel Deep Research
Author: By Parallel
Oct 17, 2025\ - How Lindy brings state-of-the-art web research to automation flows
Author: By Parallel
Oct 16, 2025\ - Introducing the Parallel Task MCP Server
Author: By Parallel
Oct 9, 2025\ - Introducing the Core2x Processor for improved compute control on the Task API
Author: By Parallel
Oct 8, 2025\ - How Day AI merges private and public data for business intelligence
Author: By Parallel
Oct 7, 2025\ - Full Basis framework for all Task API Processors
Author: By Parallel
Oct 6, 2025\ - Building a real-time streaming task manager with Parallel
Author: By Parallel
Sep 30, 2025\ - How Gumloop built a new AI automation framework with web intelligence as a core node
Author: By Parallel
Sep 16, 2025\ - Introducing the TypeScript SDK
Author: By Parallel
Sep 12, 2025\ - Building a serverless competitive intelligence platform with MCP + Task API
Author: By Parallel
Sep 11, 2025\ - Introducing Parallel Deep Research reports
Author: By Parallel
Sep 5, 2025\ - Building a Full-Stack Search Agent with Parallel and Cerebras
Author: By Parallel
Aug 21, 2025\ - Webhooks for the Parallel Task API
Author: By Parallel
Aug 14, 2025\ - Introducing Parallel: Web Search Infrastructure for AIs
Author: By Parallel
Aug 7, 2025\ - Introducing SSE for Task Runs
Author: By Parallel
Aug 5, 2025\ - A new line of advanced Processors: Ultra2x, Ultra4x, and Ultra8x
Author: By Parallel
Aug 4, 2025\ - Introducing Auto Mode for the Parallel Task API
Author: By Parallel
Jul 31, 2025\ - A state-of-the-art search API purpose-built for agents
Author: By Parallel
Jul 31, 2025\ - Parallel Search MCP Server in Devin
Author: By Parallel
Jul 28, 2025\ - Introducing Tool Calling via MCP Servers
Author: By Parallel
Jul 14, 2025\ - Introducing the Parallel Search MCP Server
Author: By Parallel
Jul 8, 2025\ - Introducing Source Policy
Author: By Parallel
Jul 2, 2025\ - The Parallel Task Group API
Author: By Parallel
Jun 17, 2025\ - State of the Art Deep Research APIs
Author: By Parallel
Jun 10, 2025\ - Parallel Search API is now available in alpha
Author: By Parallel
May 29, 2025\ - Introducing the Parallel Chat API
Author: By Parallel
May 16, 2025\ - Introducing Basis with Calibrated Confidences
Author: By Parallel
Apr 24, 2025\ - Introducing the Parallel Task API
Author: By Parallel
Ultra