<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Perplexity API — The AI Toolchain</title>
    <link>https://aitoolchain.io/tools/perplexity-api</link>
    <description>New releases and features in Perplexity API, tracked by The AI Toolchain.</description>
    <language>en</language>
    <lastBuildDate>Fri, 28 Aug 2026 20:21:49 GMT</lastBuildDate>
    <atom:link href="https://aitoolchain.io/tools/perplexity-api/rss.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Perplexity API changelog-20260828-02a9ef78</title>
      <link>https://docs.perplexity.ai/changelog</link>
      <guid isPermaLink="true">https://docs.perplexity.ai/changelog</guid>
      <pubDate>Fri, 28 Aug 2026 20:21:49 GMT</pubDate>
      <description>Perplexity API adds GLM 5.3 model and automatic prompt cache keys for Agent API presets
• An explicit `prompt_cache_key` in a request still overrides the preset-derived default, preserving full manual control over cache partitioning.</description>
    </item>
    <item>
      <title>Perplexity API snapshot-20260828</title>
      <link>https://docs.perplexity.ai/changelog</link>
      <guid isPermaLink="true">https://docs.perplexity.ai/changelog</guid>
      <pubDate>Fri, 28 Aug 2026 18:25:29 GMT</pubDate>
      <description>Perplexity Agent API and Router API add support for GLM 5.3 model.
• Adds `perplexity/glm-5.3` to the Agent API and Router API at $1.40/M uncached-input tokens, $0.26/M cached-input tokens, and $4.40/M output tokens.</description>
    </item>
    <item>
      <title>Perplexity API changelog-20260828-9fb6fe49</title>
      <link>https://docs.perplexity.ai/changelog</link>
      <guid isPermaLink="true">https://docs.perplexity.ai/changelog</guid>
      <pubDate>Fri, 28 Aug 2026 13:16:10 GMT</pubDate>
      <description>Perplexity Agent API presets now share stable prompt cache keys automatically, cutting costs ~5% with no request changes.
• Agent API presets automatically use stable prompt cache keys so independent requests sharing the same preset reuse the cached system prompt and tool definitions — no request changes required.
• An explicit `prompt_cache_key` on a request still overrides the preset-level cache default when per-request control is needed.</description>
    </item>
    <item>
      <title>Perplexity API snapshot-20260822</title>
      <link>https://docs.perplexity.ai/changelog</link>
      <guid isPermaLink="true">https://docs.perplexity.ai/changelog</guid>
      <pubDate>Sat, 22 Aug 2026 09:40:39 GMT</pubDate>
      <description>Perplexity Agent API presets now use stable prompt cache keys automatically, reducing costs by ~5% with no request changes required.
• Agent API presets automatically assign stable prompt cache keys, enabling independent requests sharing the same preset to reuse the cached system prompt and tool definitions prefix — no request changes required.
• An explicit `prompt_cache_key` still overrides the preset default when per-request cache control is needed.</description>
    </item>
    <item>
      <title>Perplexity API 1.0.0</title>
      <link>https://docs.perplexity.ai/openapi.json</link>
      <guid isPermaLink="true">https://docs.perplexity.ai/openapi.json</guid>
      <pubDate>Thu, 20 Aug 2026 10:01:19 GMT</pubDate>
      <description>Perplexity API now publishes an API — 15 endpoints across 3 areas: V1, Search, V2
• V1 (13 endpoints) — create, read
• Search (1 endpoint) — create
• V2 (1 endpoint) — read</description>
    </item>
    <item>
      <title>Perplexity API snapshot-20260820</title>
      <link>https://docs.perplexity.ai/changelog</link>
      <guid isPermaLink="true">https://docs.perplexity.ai/changelog</guid>
      <pubDate>Thu, 20 Aug 2026 09:49:57 GMT</pubDate>
      <description>Perplexity adds Router API, remote MCP server, finance_search tool, AWS Marketplace billing, and a wave of new Agent API models.
• New Router API provides unified access to open-weight models via a single endpoint using your existing Perplexity API key, with OpenAI Chat Completions and Anthropic Messages compatibility (base-URL swap), automatic health-based routing and failover, and per-token pricing with no per-request fees.
• New `GET /v1/models` endpoint lists all available Agent API models in OpenAI-compatible format with no authentication required, enabling dynamic model selection in integrations.
• Remote MCP Server now hosted by Perplexity at `https://api.perplexity.ai/mcp` — connect any MCP client supporting Streamable HTTP using your API key as a bearer token, with no local install; usage billed to your API key at standard API pricing.
• New `finance_search` tool added to the Agent API, returning structured financial and market data — quotes (near-real-time prices, OHLCV, pre/after-hours), income statement, balance sheet, cash flow, earnings transcripts, analyst estimates, and ETF constituents for public companies.
• MCP Server v1.0.0 backs `perplexity_ask`, `perplexity_reason`, and `perplexity_research` tools with Agent API presets (`fast`, `medium`, and `high` respectively); long-running research now streams progress to MCP clients, and cancelling an MCP request cancels the underlying run.
• Agent API presets now include inline citations: `fast` preset cites with numbered markers such as `[1]`; `low`, `medium`, and `high` presets cite with source-typed markers such as `[web:1]`.
• Agent API `fast` preset updated to use `openai/gpt-5.6-luna` with minimal reasoning effort and priority processing; frozen configurations must update model, reasoning effort, and set `service_tier` to `priority`.
• Agent API `low` preset updated to use `openai/gpt-5.6-luna` with minimal reasoning effort and a 32,768-token maximum output; frozen configurations must update these values manually.
• Agent API and Router API now support `google/gemini-3.7-flash` at $0.375/M input tokens, $0.0375/M cached-input tokens, and $1.875/M output tokens.
• Agent API now supports `xai/grok-4.6`, xAI&apos;s latest flagship reasoning and agentic model.
• Agent API and Router API now support `perplexity/nemotron-3-ultra-550b-a55b` at $0.25/M input or cached-input tokens and $2.50/M output tokens.
• Agent API and Router API now support `perplexity/nemotron-3.5-lightning-30b-a3b` at $0.0115/M input tokens, $0.00115/M cached-input tokens, and $0.17/M output tokens.
• Agent API and Router API now support `perplexity/deepseek-v4-flash-0731`, a fast open reasoning model with a 1M-token context window.
• Agent API now supports `openai/gpt-5.6-sol` Fast mode via `service_tier: &apos;priority&apos;` at 2× standard token pricing; `openai/gpt-5.6-luna` cut to $0.20/M input and $1.20/M output; `openai/gpt-5.6-terra` cut to $2/M input and $12/M output.
• Agent API now supports `anthropic/claude-opus-5`, `openai/gpt-5.6-sol`, `openai/gpt-5.6-terra`, `openai/gpt-5.6-luna`, `google/gemini-3.6-flash`, `google/gemini-3.5-flash-lite`, `xai/grok-4.5`, and `perplexity/kimi-k3`.
• Agent API now supports `anthropic/claude-sonnet-5`, `perplexity/glm-5.2`, `perplexity/kimi-k2.7-code`, and `nvidia/nemotron-3-super-120b-a12b`.
• Agent API now supports `anthropic/claude-opus-4-8`, `google/gemini-3.5-flash`, `google/gemini-3.1-flash-lite`, `xai/grok-4.3`, `xai/grok-4.20-non-reasoning`, and `xai/grok-4.20-multi-agent`.
• API key management upgraded to a one-time reveal model: full token values are returned only at creation and cannot be retrieved again from the console or any endpoint; `token_name` should be set at creation for ongoing identification.
• New native n8n integration ships a Perplexity node covering Chat Completions, Agent, Search, and Embeddings APIs with dynamic model loading from the API.
• New OpenClaw integration adds Perplexity Search API as a native web search provider, returning structured results (`title`, `url`, `snippet`) inside terminal workflows.
• Perplexity API credits now purchasable through the AWS Marketplace SaaS listing for consolidated billing under your AWS account.
Breaking changes:
• `google/gemini-3.1-flash-lite-preview` has been retired; requests for this model ID now return a &apos;model not supported&apos; error — use `google/gemini-3.1-flash-lite` instead.
• The `strip_thinking` and `reasoning_effort` parameters have been removed from MCP Server tool schemas (`perplexity_ask`, `perplexity_reason`, `perplexity_research`); clients sending them are ignored gracefully.
• The Agent API `fast` preset now uses `openai/gpt-5.6-luna` with priority processing (2× standard token prices); frozen configurations that pinned the previous model and `service_tier` must manually update model, reasoning effort, and set `service_tier` to `priority`.
• The Agent API `low` preset now uses `openai/gpt-5.6-luna` with a 32,768-token maximum output; frozen configurations must update model and max-output values to match.</description>
    </item>
  </channel>
</rss>
