Separate Perplexity model-token charges from request, search and tool charges before estimating a workload. The live Sonar table is one part of a product-specific billing decision.

How Perplexity pricing works

Perplexity’s price reference separates Router, Agent, Search, Sonar and Embeddings billing. Sonar can include model-token and request-related categories, while Agent tool usage is separate from model-token usage. Read the product heading and unit before comparing rates. Official Perplexity documentation.

The official Sonar documentation states that Sonar support continues until September 27, 2026 and directs developers to the Agent API migration guide. Official Perplexity documentation.

Build an estimate from a complete accepted task. A researched answer includes evidence gathering as well as final text, and a multi-step application can make several provider calls before producing its artifact. Record which product performs each step. A token-only estimate may be useful for one component, but should not be labeled the total cost of an entire research workflow.

Keep unknown quantities visible. If the selected route reports search or reasoning usage only after execution, start with a bounded evaluation and inspect those results before extrapolating. Do not make a missing calculator field disappear by treating it as zero. A useful estimate states what is included and which charge categories still require actual usage evidence.

Original diagram of documented Perplexity charge categories and conditional rates.

Full price list

Verified model prices
ModelPrice typeUSDUnitTierRegionSource
SonarInput$1per 1M tokensStandardSee sourceOfficial source ↗
SonarOutput$1per 1M tokensStandardSee sourceOfficial source ↗
SonarPer Request$0.012per requesthigh search contextSee sourceOfficial source ↗
SonarPer Request$0.005per requestlow search contextSee sourceOfficial source ↗
SonarPer Request$0.008per requestmedium search contextSee sourceOfficial source ↗
Sonar Deep ResearchInput$2per 1M tokensStandardSee sourceOfficial source ↗
Sonar Deep ResearchOutput$8per 1M tokensStandardSee sourceOfficial source ↗
Sonar ProInput$3per 1M tokensStandardSee sourceOfficial source ↗
Sonar ProOutput$15per 1M tokensStandardSee sourceOfficial source ↗
Sonar ProPer Request$0.014per requesthigh search contextSee sourceOfficial source ↗
Sonar ProPer Request$0.006per requestlow search contextSee sourceOfficial source ↗
Sonar ProPer Request$0.01per requestmedium search contextSee sourceOfficial source ↗
Sonar Reasoning ProInput$2per 1M tokensStandardSee sourceOfficial source ↗
Sonar Reasoning ProOutput$8per 1M tokensStandardSee sourceOfficial source ↗
Sonar Reasoning ProPer Request$0.014per requesthigh search contextSee sourceOfficial source ↗
Sonar Reasoning ProPer Request$0.006per requestlow search contextSee sourceOfficial source ↗
Sonar Reasoning ProPer Request$0.01per requestmedium search contextSee sourceOfficial source ↗

Last verified · Source ↗

This live catalog tracks the documented Sonar models and supported charge categories. Conditions such as search-context request fees remain separate rows. The table does not substitute for the official Router or Agent catalogs. Consult Perplexity model and product selection before choosing the rate that applies to your endpoint.

Compare like units. A per-request category should not be added to a per-token rate as though the values shared a denominator. Preserve the number of applicable requests and their condition alongside token usage. If the application changes search behavior or product, recalculate from the new contract rather than carrying forward a monthly total from the previous configuration.

Estimate your cost

Interactive tool

Estimate your API costs

Your text and estimates stay in this browser. No API requests are sent to model providers.

Loading verified model records…

Use the AI API cost calculator for the supported input-output component and a repeatable set of workload assumptions. Add the applicable official request, search or tool charges separately where the calculator does not expose them. Select the exact supported model and condition; a Sonar entry is not a proxy for every response produced by the Perplexity platform.

After a representative call, inspect the provider’s returned usage and cost details where available. Compare those categories with the account record. If the result differs from the estimate, check product selection, retries, reasoning, search behavior and output length before adjusting the forecast. Preserve the original assumptions so the investigation explains what changed.

Three worked scenarios

For a chatbot that answers current documentation questions, define whether every turn needs fresh search or whether some turns only refer to the conversation. In the calculator, model the messages you actually send and the intended answer shape. Keep any search-related component separate. Test follow-up questions that refer to a cited source, and verify that the application preserves enough evidence context to answer them correctly.

For summarization, distinguish a supplied-document summary from a researched brief. The former should be checked against the supplied material; the latter may require additional source discovery and comparison. Use a fixture containing an intentionally missing fact so the application demonstrates how it handles gaps. Count any preprocessing, search or correction calls in the complete task rather than comparing only the final paragraph.

For a coding assistant, decide whether the role is finding current API documentation or generating and repairing code. Search-backed documentation lookup can help locate evidence, but the returned explanation still needs a compatible version and a checkable example. Estimate the source lookup and the coding step under their respective routes. Measure an accepted change that passes your checks, including correction work when the initial answer is insufficient.

How to reduce cost on Perplexity

Use the narrowest product that produces the artifact you need. The official quickstart distinguishes ranked Search results from generated Agent answers. A workflow that already performs its own synthesis can evaluate whether it needs another generated answer at the retrieval stage. Official Perplexity documentation.

For Sonar, use its documented query and filter controls to express the actual research question. Select appropriate source restrictions and freshness requirements, then inspect what they exclude. An overly narrow filter can create an incomplete answer rather than a useful saving. Keep a fixture with the expected primary source so you can see whether a configuration still retrieves the evidence the task requires. Official Perplexity documentation.

Keep answers focused on the requested artifact. A documentation lookup can request the relevant behavior and a source reference rather than a broad history of the technology. Evaluate whether additional research depth changes the accepted answer. More extensive output is useful when the task needs it; it should not be the default substitute for a precise question.

When migrating to Agent, review the documented search-tool and preset mapping rather than copying Sonar controls unchanged. Compare the new configuration against the same evidence and answer criteria. If a preset performs more work than your application needs, evaluate a more explicit configuration using the official guide and actual usage. Do not infer a cache or batch discount that the selected product has not documented. Official Perplexity documentation.

Billing, credits and limits notes

Perplexity’s API project uses purchased credits and offers automatic reload preferences. Inspect those preferences in the intended project before attaching a scheduled workload. An enabled reload can change what happens when the balance falls, so it should be a deliberate operating choice. Official Perplexity documentation.

Read Perplexity free-access and promotion conditions separately from usage-tier admission. A higher usage tier and a usable current balance are different account facts. Keep an application stop condition for the evaluation and review actual consumption before scaling public or background traffic.

Keep the endpoint and selected configuration beside the estimate so another reviewer can reproduce its assumptions.

Price change history

  1. Perplexity · Sonar Deep Research — MetadataRecord updated; consult the linked source for details. → Record updated; consult the linked source for details.Source ↗
  2. Perplexity · Sonar Reasoning Pro — MetadataRecord updated; consult the linked source for details. → Record updated; consult the linked source for details.Source ↗
  3. Perplexity · Sonar Pro — MetadataRecord updated; consult the linked source for details. → Record updated; consult the linked source for details.Source ↗
  4. Perplexity · Sonar — MetadataRecord updated; consult the linked source for details. → Record updated; consult the linked source for details.Source ↗
  5. Perplexity · Sonar — ValueNot previously recorded → Tier: Tier 5 · Metric: RPM · Value: 4000 · Notes: Sonar APISource ↗
  6. Perplexity · Sonar Pro — ValueNot previously recorded → Tier: Tier 5 · Metric: RPM · Value: 4000 · Notes: Sonar APISource ↗
  7. Perplexity · Sonar Reasoning Pro — ValueNot previously recorded → Tier: Tier 5 · Metric: RPM · Value: 4000 · Notes: Sonar APISource ↗
  8. Perplexity · Sonar Deep Research — ValueNot previously recorded → Tier: Tier 5 · Metric: RPM · Value: 100 · Notes: Sonar APISource ↗
  9. Perplexity · Sonar — ValueNot previously recorded → Tier: Tier 4 · Metric: RPM · Value: 4000 · Notes: Sonar APISource ↗
  10. Perplexity · Sonar Pro — ValueNot previously recorded → Tier: Tier 4 · Metric: RPM · Value: 4000 · Notes: Sonar APISource ↗

Subscribe to the changelog RSS feed

Retain the product and request configuration with each saved estimate. A Sonar-to-Agent migration is not merely a replacement price row: search behavior, response handling and charge categories can change. Use observed price history for public rate changes and the account’s actual usage record for private billed work. Return to the Perplexity API overview for the current integration paths.

Frequently asked questions

Is Perplexity billed only for input and output tokens?
The applicable categories depend on the product and configuration. Inspect request, search and tool charges where the official price reference defines them.
Does the calculator cover every Agent tool charge?
No. Use it for supported model-token assumptions and preserve additional categories separately until actual usage or the official tool rate supplies the needed evidence.
Can I substitute a Sonar price for an Agent preset?
No. The preset can select a different model and tools. Use the actual product configuration and reported usage.
Does a higher tier mean a lower token rate?
Do not infer a discount from an admission tier. Check the rate applicable to the model and product in the official pricing reference.
How should I estimate research work?
Count the complete accepted task, including source discovery, model calls and correction work. Preserve category-level assumptions and verify them with representative usage.
What changes when I enable automatic reload?
It changes the account’s funding behavior. Review the actual project preference and keep the application’s operating allowance explicit.

Sources

Last verified · Source ↗