apiPulse.app

Perplexity

AI search and answer platform for developers.

Perplexity API Changelog Monitoring

Perplexity is an AI search and answer platform for developers, offering APIs for search, research, and cited responses. This Perplexity API Changelog Monitoring page helps teams track product changes quickly so they can keep integrations, internal tools, and customer-facing features aligned with the latest platform behavior.

Strong Perplexity API Changelog Monitoring helps teams catch model updates, API changes, deprecations, pricing adjustments, and new developer features before they create downstream issues. It also supports better release planning by giving engineering and product teams a shared view of what changed and what may require action.

The benefit of Perplexity API Changelog Monitoring is simple: your team gets a repeatable way to review updates, evaluate integration risk, and respond before upstream platform changes affect production systems. Perplexity API changelog and developer updates

Perplexity API down?

Perplexity API down? For live availability checks, refer to Perplexity's official status or support channels. This page is focused on Perplexity API Changelog Monitoring, helping you follow releases, feature announcements, and breaking changes that may affect your integration over time.

Recent changes

Showing the last 10 changes from this feed.

08-28-2026

August 2026

GLM 5.3 The Agent API and Router API now support perplexity/glm-5.3 at $1.40 per million uncached-input tokens, $0.26 per million cached-input tokens, and $4.40 per million output tokens. See the Agent API Models reference or the Router model catalog.

08-27-2026

August 2026

DeepSeek V4 Flash 0731 The Agent API and Router API now support perplexity/deepseek-v4-flash-0731, a fast, efficient open reasoning model with a 1M-token context window. See pricing in the Agent API Models reference or the Router model catalog.

08-19-2026

August 2026

Fast preset updated The Agent API fast preset now uses openai/gpt-5.6-luna with minimal reasoning effort and priority processing. Dynamic fast preset requests pick up the change automatically. If you use a frozen configuration, update the model and reasoning effort and set service_tier to priority. Priority processing uses 2× the model's standard token prices.

08-15-2026

August 2026

Gemini 3.7 Flash The Agent API and Router API now support google/gemini-3.7-flash at launch pricing of $0.375 per million input tokens, $0.0375 per million cached-input tokens, and $1.875 per million output tokens. See the Agent API Models reference.

08-14-2026

August 2026

Gemini 3.7 Flash The Agent API and Gateway API now support google/gemini-3.7-flash at launch pricing of $0.375 per million input tokens, $0.0375 per million cached-input tokens, and $1.875 per million output tokens. See the Agent API Models reference.

08-13-2026

August 2026

Grok 4.6 The Agent API now supports xai/grok-4.6, xAI's latest flagship reasoning and agentic model. See pricing in the Agent API Models reference.

08-12-2026

August 2026

NVIDIA Nemotron 3.5 Lightning The Agent API and Gateway API now support perplexity/nemotron-3.5-lightning-30b-a3b, a fast, efficient open-weight reasoning model, at $0.0115 per million input tokens, $0.00115 per million cached-input tokens, and $0.17 per million output tokens. See the Agent API Models reference or the Gateway model catalog.

08-10-2026

July 2026

GPT-5.6 price cuts and Sol Fast mode GPT-5.6 Luna now costs $0.20 per million input tokens and $1.20 per million output tokens. GPT-5.6 Terra now costs $2 per million input tokens and $12 per million output tokens. GPT-5.6 Sol now supports Fast mode at 2× standard token pricing; send service_tier: "priority" to use it.

08-05-2026

August 2026

DeepSeek V4 Flash 0731 The Agent API and Gateway API now support perplexity/deepseek-v4-flash-0731, a fast, efficient open reasoning model with a 1M-token context window. See pricing in the Agent API Models reference or the Gateway model catalog.

07-31-2026

July 2026

Low preset updated The Agent API low preset now uses openai/gpt-5.6-luna with minimal reasoning effort and a 32,768-token maximum output. If you use a frozen configuration, update these values to match the current preset. Dynamic low preset requests pick up the change automatically.