Perplexity Sonar Reasoning

Text & chatUnavailable
by PerplexityModel ID: sonar-reasoning

Perplexity's reasoning model with chain-of-thought and integrated web search.

Status
Unavailable
Context
127,000 tokens
Max. output
8,192 tokens
Input โ†’ output
Text โ†’ Text
Developer
Perplexity
Updated
September 23, 2026

Perplexity Sonar Reasoning is currently unavailable

You can still read the details on this page. Pick one of the available alternatives below to run a comparable model right away.

Go to alternatives
01

Comparable models

All in this category
02

Playground

Try Perplexity Sonar Reasoning

Input & output

Currently unavailable

Currently unavailable.

The playground is disabled. You can find comparable models in the same category: Browse alternatives

Try Perplexity Sonar Reasoning
Output
The answer appears here.

This run

No price โ€“ currently unavailable.

New here?

10 free credits (US$0.10) when you sign up with Google

Usable 24 hours after sign-up, up to 5 runs per day and at most 2 credits per run. Other sign-in methods start without credits.

03

About Perplexity Sonar Reasoning

TL;DRAs of September 23, 2026

Perplexity Sonar Reasoning is a model by Perplexity in the Text & chat category. Perplexity Sonar Reasoning is currently not available on Railwail. The context window holds 127,000 tokens, and one response can be up to 8,192 tokens long.

Background

About Perplexity AI

Founded 2022 ยท San Francisco, California, USA

Perplexity AI was founded in August 2022 by Aravind Srinivas (CEO, former OpenAI and DeepMind researcher), Denis Yarats (former Quora), Johnny Ho (former Quora) and Andy Konwinski (Databricks co-founder). Headquartered in San Francisco, Perplexity built one of the first AI-native search products โ€” an answer engine combining live web retrieval with LLM synthesis and inline citations. The company has raised over $500M from investors including IVP, Nvidia, NEA, Jeff Bezos and SoftBank, with a 2024 valuation above $9B. Perplexity offers its Sonar API for developers to access search-grounded LLM responses. Sonar Reasoning, released January 2025, is built on the open-source DeepSeek-R1 reasoning model wrapped in Perplexity's live web-search retrieval stack with the reasoning trace exposed in the API response.

Visit Perplexity AI

Architecture

Reasoning Mixture-of-Experts Transformer with live web retrieval

Sonar Reasoning is Perplexity's hosted deployment of DeepSeek-R1 (671B Sparse Mixture-of-Experts, 37B active per token, DeepSeekMoE architecture with Multi-head Latent Attention) wrapped in Perplexity's retrieval pipeline. The base model uses 256 fine-grained experts plus shared experts, top-8 routing, 128K context, and was trained by DeepSeek with the GRPO reinforcement-learning recipe described in the R1 technical report. Perplexity hosts the model on US-based infrastructure, applies its own safety post-training on the Perplexity domain, and exposes the reasoning chain-of-thought as a separate field in the API response alongside the final answer. Each request triggers a live web-search retrieval that injects ranked passages into the model context with source URLs returned alongside the answer. Released January 2025 via the Sonar API.

Parameters
671B total, 37B active per token (DeepSeek-R1 base)
Context
127,000 tokens

Capabilities

  • Combines live web search with explicit reasoning traces
  • Returns ranked citations with source URLs alongside answers
  • Strong on math, logic and research-style questions
  • Reasoning chain-of-thought exposed in the API
  • Knowledge always current via live search โ€” bypasses model cutoff limits
  • Cheaper than OpenAI o-series for comparable reasoning quality at release
  • Hosted on Perplexity's US infrastructure
  • Best for: research-grade Q&A, market intelligence, citation-required workflows, fact-grounded reasoning.

Training & license

Inherits DeepSeek-R1's training pipeline (DeepSeek-V3-Base ~14.8T tokens of multilingual web, code and math, followed by the R1 reasoning RL stage with GRPO). Perplexity adds retrieval-grounded supervised fine-tuning on its domain. Base model knowledge cutoff approximately July 2024 โ€” but live search overrides for current questions.

License: Perplexity API Terms of Service. The base DeepSeek-R1 weights are MIT-licensed and can be self-hosted separately, but the live-search and reasoning-citation pipeline are Perplexity-specific and hosted-only.

Safety testing: Perplexity applies safety post-training and operates the model under its hosted API safety policies. The base DeepSeek-R1 has its own published model card; Perplexity's hosting modifies refusal behaviour for the Sonar deployment.

Known limitations

  • Reasoning latency is high (typically 5-20s) due to thinking tokens plus search
  • Quality depends on what the search retrieves โ€” bad sources lead to bad answers
  • Per-call cost includes reasoning tokens plus search โ€” can be expensive on long answers
  • No vision input
  • Reasoning trace can be verbose and not directly user-facing
  • Hosted-only โ€” Perplexity does not offer self-hosting for the Sonar stack
04

Pricing

Currently unavailable. There is no price for this model at the moment, so it cannot be run.

05

API

Call Perplexity Sonar Reasoning with your Railwail API key. Use this model ID in the request:

No verified API example

The public API passes a different input format than this model needs. Use the playground above.

06

Specifications

Model ID
sonar-reasoning
Developer
Perplexity
Category
Text & chat
Input
Text
Output
Text
Context window
127,000 tokens
Max. output
8,192 tokens
Model size
671B total, 37B active per token (DeepSeek-R1 base)
License
Perplexity API Terms of Service. The base DeepSeek-R1 weights are MIT-licensed and can be self-hosted separately, but the live-search and reasoning-citation pipeline are Perplexity-specific and hosted-only.
Catalog entry updated
September 23, 2026

Tags

  • perplexity
  • web-search
  • reasoning
07

Use cases

What it is used for

  • Research-grade question answering with citations
  • Market and competitive intelligence
  • Real-time fact-grounded reasoning
  • Journalism and legal research support
  • Academic literature review assistance
  • Sales and customer research workflows
08

Frequently asked questions

What is Perplexity Sonar Reasoning?

Perplexity Sonar Reasoning is a model by Perplexity in the Text & chat category. It is listed on Railwail but cannot be run at the moment.

How much does Perplexity Sonar Reasoning cost on Railwail?

Perplexity Sonar Reasoning cannot be run on Railwail at the moment, so there is no current price. Available alternatives with prices are listed further down this page.

What is the context window of Perplexity Sonar Reasoning?

The context window of Perplexity Sonar Reasoning holds 127,000 tokens. One response can be up to 8,192 tokens long.

How fast is Perplexity Sonar Reasoning?

There are not enough measured runs of Perplexity Sonar Reasoning on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.

Is Perplexity Sonar Reasoning better than Claude Fable 5.1?

That depends on the task. Perplexity Sonar Reasoning (Perplexity) and Claude Fable 5.1 (Anthropic) are both models in the Text & chat category. The comparison page shows their prices and specifications side by side.

Compare Perplexity Sonar Reasoning and Claude Fable 5.1

Can I use Perplexity Sonar Reasoning right now?

Currently unavailable. The page stays online; available alternatives from the same category are listed further down.

All models through one API

One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = US$0.01.