Claude Haiku 5.5 model in the Cortex REST API (Preview)
Preview2 days agoUnverifiedAdded Oct 9, 2026
The claude-haiku-5-5 model is now available in Public Preview in the Cortex REST API.
Topics: LLMs, Developer tools
More Snowflake Cortex AI releases
Every Snowflake Cortex AI releaseSnowflake Decision with AI_COMPLETE (Private Preview)
With this release, we introduce the private preview of Snowflake Decision, a model available through AI_COMPLETE that evaluates text and structured data against user-defined questions.
Data lineage for storage integrations, external tables, model monitors, and Cortex Search services
Data lineage now includes the following relationships:
Cortex Agent code execution tool (General availability)
The Cortex Agent code execution tool is now generally available. The code execution tool is a built-in tool that enables an agent to execute code during a conversation, so your agents can run scripts to process data, perform calculations, and produce visualizations.
Snowflake Native Apps: Observability for Cortex Agents
Observability for Cortex Agents in a Snowflake Native App now includes the following capabilities:
DCM Projects DEFINE SEMANTIC VIEW (General availability)
With this release, DEFINE SEMANTIC VIEW in DCM Projects is generally available. You can declaratively manage semantic view definitions, including tables, relationships, facts, dimensions, metrics, AI instructions, and verified queries.
Response caching for Cortex AI Functions
Snowflake Cortex AI Functions now automatically caches and reuses the results of AI_COMPLETE calls. When a query makes an identical AI_COMPLETE call more than once, Snowflake can return the stored result instead of running inference again, which lowers cost and speeds up query execution.
Cortex AI Function Evaluation for measuring quality (Public Preview)
Snowflake Cortex AI Function Evaluation is now available in public preview, enabling customers to measure the output quality, cost, and token usage of custom AI Functions and Cortex AI calls against labeled datasets.
Cortex AI Function Optimization for more efficient AI implementations (Public Preview)
Snowflake Cortex AI Function Optimization is now available in public preview, enabling customers to automatically search across prompts and models for improved implementations of custom AI Functions.
Also shipped on Oct 7, 2026
Snowflake in October 2026The microVM sandbox type is now Generally Available (GA) with GKE Sandbox in clusters that run version...
The microVM sandbox type is now Generally Available (GA) with GKE Sandbox in clusters that run version 1.37.0-gke.4713000 and later. MicroVM sandboxes provide hardware virtualization and isolation for untrusted workloads, AI agent runtimes, and multi-tenant environments.
SSH for Cloud Run services and instances is in Preview
SSH for Cloud Run services and instances is in Preview . Use this feature to establish a secure, interactive shell connection to your running instances.
Set credit usage limits and alerts for your workspace
Workspace admins and owners on paid plans can now set usage limits and alerts that keep shared credits under control. Set a credit threshold for the whole workspace, a project, a member, or, on Business and Enterprise plans, a group or an access token, and choose what happens when usage reaches it: an email or in-app alert, or a block that stops building or...
Faster inline text edits
Simple text changes you make with Edit text inline in the preview toolbar now apply in a few seconds. Text that your app builds from data or translates still takes a little longer.
Gemini 3.7 Flash is deprecated for your app's AI features
Google is retiring Gemini 3.7 Flash ( google/gemini-3.7-flash ), so it is now marked deprecated for your app's AI features and stops working on January 28, 2027. Apps that use it keep working until then, and Lovable no longer picks it for new AI features.
The Cost Analysis tab of the Billing page now includes Serverless Inference Live Usage, which lists the...
The Cost Analysis tab of the Billing page now includes Serverless Inference Live Usage, which lists the estimated cost and token counts of your Serverless Inference requests by model and model access key, for both teams and organizations. …