Perplexity Sonar

Text & chatUnavailable
by PerplexityModel ID: sonar

Perplexity's fastest and cheapest web-grounded chat model. Live-source citations included.

Status
Unavailable
Context
127,000 tokens
Max. output
8,192 tokens
Input โ†’ output
Text โ†’ Text
Developer
Perplexity
Updated
September 23, 2026

Perplexity Sonar is currently unavailable

You can still read the details on this page. Pick one of the available alternatives below to run a comparable model right away.

Go to alternatives
01

Comparable models

All in this category
02

Playground

Try Perplexity Sonar

Input & output

Currently unavailable

Currently unavailable.

The playground is disabled. You can find comparable models in the same category: Browse alternatives

Try Perplexity Sonar
Output
The answer appears here.

This run

No price โ€“ currently unavailable.

New here?

10 free credits (US$0.10) when you sign up with Google

Usable 24 hours after sign-up, up to 5 runs per day and at most 2 credits per run. Other sign-in methods start without credits.

03

About Perplexity Sonar

TL;DRAs of September 23, 2026

Perplexity Sonar is a model by Perplexity in the Text & chat category. Perplexity Sonar is currently not available on Railwail. The context window holds 127,000 tokens, and one response can be up to 8,192 tokens long.

Background

About Perplexity AI

Founded 2022 ยท San Francisco, USA

Perplexity AI was founded in August 2022 by Aravind Srinivas (CEO, former OpenAI researcher and DeepMind/Google Brain alumnus), Denis Yarats (CTO, former Meta AI researcher), Andy Konwinski (co-founder of Databricks) and Johnny Ho. The company's product is an 'answer engine' that combines real-time web search with LLM-based reasoning to return cited answers, positioning itself as a search alternative to Google. Perplexity launched its consumer chat product in late 2022 and quickly raised over $500M across multiple rounds from investors including IVP, NEA, NVIDIA, Jeff Bezos and Susan Wojcicki, reaching a $9B valuation in late 2024. The company released its own Sonar model family in 2024-2025, fine-tuned for grounded, citation-bearing answers from live web retrieval. Sonar is positioned as a fast, cost-efficient default model in the Perplexity API and consumer product, with Sonar Pro and Sonar Reasoning as higher-capability variants. Perplexity also ships Perplexity Pages (long-form content), Perplexity Spaces (collaboration) and partnership integrations with Apple Intelligence and several US news organisations.

Visit Perplexity AI

Architecture

Search-grounded LLM (Llama-derived, tuned for retrieval-augmented answering)

Perplexity Sonar is the default, fast and low-cost model in Perplexity's API and consumer product, released as a public API tier in January 2025. Sonar is built on a Llama-3.x-derived base, fine-tuned by Perplexity for retrieval-augmented generation, citation insertion and concise web-grounded answers. The training process is not fully disclosed but follows standard supervised fine-tuning and preference optimisation on Perplexity's curated logs of high-quality, well-cited answers, plus synthetic question-answer pairs grounded in real web pages. At inference time, every Sonar query is augmented with a live web search step run by Perplexity's in-house search index, which retrieves the most relevant pages, ranks them, and supplies them as context to the model. The model is trained to cite sources inline using numbered references, to refuse to answer when the corpus does not support the claim, and to compose answers in a concise, structured format. Sonar supports a 127K token context window and the standard OpenAI-compatible chat completions API, making it a drop-in replacement for OpenAI in many search-enabled agent stacks. The cost is roughly $1 per 1,000 search-augmented requests at launch, making it one of the cheapest fully-grounded options on the market.

Parameters
Undisclosed (Llama-3.x-derived base, mid-sized)
Context
127,000 tokens

Capabilities

  • Live web search built into every request
  • Inline citations with numbered source references
  • Concise, structured answers tuned for search use cases
  • 127K context window
  • OpenAI-compatible chat completions API
  • Low latency and low cost (around $1 per 1k search-augmented requests at launch)
  • Returns sources, images and related questions in the response
  • Drop-in replacement for retrieval-augmented agents
  • Refuses to answer when source corpus is insufficient
  • Available in Perplexity API and consumer product
  • Best for: real-time research, grounded Q&A, news summarisation, RAG without managing your own retrieval.

Training & license

Built on a Llama-3.x-derived base, fine-tuned on Perplexity's curated logs of high-quality cited answers plus synthetic retrieval-grounded QA pairs. At inference time augmented with Perplexity's proprietary live web search index.

License: Proprietary commercial license via the Perplexity API. Weights not released. Standard Perplexity Terms of Service apply.

Safety testing: Limited public safety evaluations. Refuses to answer when retrieved sources are insufficient. Inherits some safety properties from the Llama 3 base.

Known limitations

  • Quality depends heavily on Perplexity's retrieval ranking
  • May cite low-quality sources if highly ranked
  • Closed weights; no on-prem option
  • Citations occasionally do not support the exact claim
  • Less capable than Sonar Pro on complex multi-step reasoning
04

Pricing

Currently unavailable. There is no price for this model at the moment, so it cannot be run.

05

API

Call Perplexity Sonar with your Railwail API key. Use this model ID in the request:

No verified API example

The public API passes a different input format than this model needs. Use the playground above.

06

Specifications

Model ID
sonar
Developer
Perplexity
Category
Text & chat
Input
Text
Output
Text
Context window
127,000 tokens
Max. output
8,192 tokens
Model size
Undisclosed (Llama-3.x-derived base, mid-sized)
License
Proprietary commercial license via the Perplexity API. Weights not released. Standard Perplexity Terms of Service apply.
Catalog entry updated
September 23, 2026

Tags

  • perplexity
  • web-search
  • citations
  • cost-efficient
07

Use cases

What it is used for

  • Real-time research assistants
  • News summarisation
  • Cited Q&A for product UX
  • Lightweight RAG replacement
  • Live web fact-checking
  • Cost-sensitive search-augmented chat
08

Frequently asked questions

What is Perplexity Sonar?

Perplexity Sonar is a model by Perplexity in the Text & chat category. It is listed on Railwail but cannot be run at the moment.

How much does Perplexity Sonar cost on Railwail?

Perplexity Sonar cannot be run on Railwail at the moment, so there is no current price. Available alternatives with prices are listed further down this page.

What is the context window of Perplexity Sonar?

The context window of Perplexity Sonar holds 127,000 tokens. One response can be up to 8,192 tokens long.

How fast is Perplexity Sonar?

There are not enough measured runs of Perplexity Sonar on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.

Is Perplexity Sonar better than Claude Fable 5.1?

That depends on the task. Perplexity Sonar (Perplexity) and Claude Fable 5.1 (Anthropic) are both models in the Text & chat category. The comparison page shows their prices and specifications side by side.

Compare Perplexity Sonar and Claude Fable 5.1

Can I use Perplexity Sonar right now?

Currently unavailable. The page stays online; available alternatives from the same category are listed further down.

All models through one API

One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = US$0.01.