October 2026
Update3 days agoUnverifiedAdded Oct 10, 2026
Claude Haiku 5.5 The Agent API now supports anthropic/claude-haiku-5-5 . Requests with more than 100k input tokens use higher rates. See the Agent API Models reference .
Topics: AI agents, LLMs, Developer tools
More Perplexity API releases
Every Perplexity API releaseSeptember 2026
Fast Search Set search_type: "fast" on the Search API or the Agent API web_search tool to use a lower-latency search path. Fast Search costs $1.00 per 1,000 Search API requests or web_search invocations, with model tokens billed separately for Agent API. Set search_type: "web" for standard web search. See Search API Fast Search or Agent API Web Search .
August 2026
GLM 5.3 The Agent API and Router API now support perplexity/glm-5.3 at $1.40 per million uncached-input tokens, $0.26 per million cached-input tokens, and $4.40 per million output tokens. See the Agent API Models reference or the Router model catalog .
July 2026
GPT-5.6 price cuts and Sol Fast mode GPT-5.6 Luna now costs $0.20 per million input tokens and $1.20 per million output tokens. GPT-5.6 Terra now costs $2 per million input tokens and $12 per million output tokens. GPT-5.6 Sol now supports Fast mode at 2× standard token pricing; send service_tier: "priority" to use it.
Also shipped on Oct 7, 2026
Perplexity in October 2026The microVM sandbox type is now Generally Available (GA) with GKE Sandbox in clusters that run version...
The microVM sandbox type is now Generally Available (GA) with GKE Sandbox in clusters that run version 1.37.0-gke.4713000 and later. MicroVM sandboxes provide hardware virtualization and isolation for untrusted workloads, AI agent runtimes, and multi-tenant environments.
SSH for Cloud Run services and instances is in Preview
SSH for Cloud Run services and instances is in Preview . Use this feature to establish a secure, interactive shell connection to your running instances.
Set credit usage limits and alerts for your workspace
Workspace admins and owners on paid plans can now set usage limits and alerts that keep shared credits under control. Set a credit threshold for the whole workspace, a project, a member, or, on Business and Enterprise plans, a group or an access token, and choose what happens when usage reaches it: an email or in-app alert, or a block that stops building or...
Faster inline text edits
Simple text changes you make with Edit text inline in the preview toolbar now apply in a few seconds. Text that your app builds from data or translates still takes a little longer.
Gemini 3.7 Flash is deprecated for your app's AI features
Google is retiring Gemini 3.7 Flash ( google/gemini-3.7-flash ), so it is now marked deprecated for your app's AI features and stops working on January 28, 2027. Apps that use it keep working until then, and Lovable no longer picks it for new AI features.
The Cost Analysis tab of the Billing page now includes Serverless Inference Live Usage, which lists the...
The Cost Analysis tab of the Billing page now includes Serverless Inference Live Usage, which lists the estimated cost and token counts of your Serverless Inference requests by model and model access key, for both teams and organizations. …