Skip to main content
Model ReleasesProvider CoveragePricing

Model and API Updates

  • DeepSeek V4 Flash 0731 is now a separate model from the original DeepSeek V4 Flash, with updated agent benchmarks, a 1M context window, 384K maximum output, reasoning controls, tool use, structured output, server-side web search, and native Responses API support.
  • Phaseo now routes DeepSeek’s current deepseek-v4-flash slug to V4 Flash 0731 through the native /responses endpoint. Existing third-party deployments remain attached to the original V4 Flash model, and V4 Pro remains on Chat Completions until DeepSeek enables Responses support for it.
  • Official DeepSeek pricing remains 0.0028permillioncachedinputtokens,0.0028 per million cached input tokens, 0.14 per million uncached input tokens, and $0.28 per million output tokens. DeepSeek’s announced peak pricing remains pending until an effective date is published.
  • Retired direct-provider deepseek-chat and deepseek-reasoner routes and their historical pricing were end-dated at 24 July 2026, 15:59 UTC. Other providers serving the corresponding models are unaffected.
FeaturesModel ReleasesProvider CoveragePricingSDKs

Model Releases

Pricing and Provider Coverage

Features and Platform

Model ReleasesProvider CoveragePricing

July 23, 2026

New Models


July 21, 2026

New Models

Gemini 3.6 Flash and Gemini 3.5 Flash-Lite model release artwork
  • Google: Gemini 3.6 Flash
  • Google: Gemini 3.5 Flash-Lite
  • Mind Lab: Macaron V1 Venti
  • Sakana AI: Fugu Cyber
  • Google’s release announcement reports that Gemini 3.6 Flash uses 17% fewer output tokens than Gemini 3.5 Flash on the Artificial Analysis Index, with published comparisons on DeepSWE (49% vs. 37%), MLE-Bench (63.9% vs. 49.7%), OSWorld-Verified (83.0% vs. 78.4%), and GDPval-AA v2 (1421 vs. 1349).
  • The same announcement reports Gemini 3.5 Flash-Lite at 350 output tokens per second on the Artificial Analysis Index, outperforming Gemini 3.1 Flash-Lite on Terminal-Bench 2.1 (54% vs. 31%), GDM-MRCR v2 (72.2% vs. 60.1%), and GDPval-AA v2 (1140 vs. 642), plus Gemini 3 Flash on SWE-Bench Pro (54.2% vs. 49.6%) and OSWorld-Verified (74.0% vs. 65.1%).
  • Both models are now catalogued with Google AI Studio and Google Vertex provider coverage; see Google’s Vertex AI Gemini model documentation for the provider-side model reference.

Features
  • Model pages are now easier to discover in search.
FeaturesModel Releases

Features

  • Improved search discovery across Phaseo and its documentation.

Model Releases

Features

Features

  • Authenticated MCP is now available.
  • Benchmark rankings now account for ties more consistently.
FeaturesModel Releases

Features

  • Passkeys are now available to all users for passwordless sign-in and account management with device biometrics, PINs, or security keys.
  • Chat model selection is faster and works correctly when starting a new chat.

Model Releases

Kimi K3 model release artwork
FeaturesModel Releases

Features

  • Dropdown menus and tab selectors now use Base UI render composition consistently across chat, settings, usage, rankings, and data screens (PR #938).
  • The OpenAI GPT-5.6 migration guide was added for Sol, Terra, and Luna readiness planning.
  • Model links now store human-readable titles alongside link kinds and preserve multiple links of the same kind.
  • Amazon Bedrock provider coverage was staged for SpaceXAI Grok 4.3 with Bedrock Mantle model id xai.grok-4.3, standard pricing, and current us-west-2 in-region availability. The provider-model capability is currently disabled pending setup. AWS currently lists Grok 4.3 as the only SpaceXAI model on Bedrock, and does not list Geo or Global inference for it.
  • Amazon Bedrock OpenAI frontier model coverage was checked; AWS still documents GPT-5.5 and GPT-5.4 as regional/Geo in-region offerings, with Global cross-region pricing marked as coming soon.

Model Releases

GPT 5.6 model release artwork
FeaturesModel Releases

Features

  • Chat now surfaces clearer timing and tool metadata, audio attachments, model-picker focus handling, tag management, and account credit display (PR #922).
  • Header search now lazy-loads a compact cached index only when opened, reducing normal dashboard payloads (PR #920).
  • Chat streaming, tool-call continuation, and model catalogue cache refresh paths were hardened for more reliable server-tool turns and model pages (PR #919, PR #918).
  • API key reveal flows, internal navigation, and new-model Discord notifications were tightened for cleaner settings, admin, and release workflows (PR #812, PR #925, PR #927).
  • SiliconFlow provider coverage was added for Tencent Hy3, including the free route and current pricing.
  • xAI catalogue and provider labels were renamed to SpaceXAI across model pages, provider records, aliases, logos, and generated SDK/OpenAPI surfaces.
  • Aion 1.0 and Aion 1.0 Mini were marked deprecated and retired on July 4, 2026 after Aion Labs removed them from the live model endpoint.

Model Releases

Aion 3.0 model release artwork
FeaturesDocsBrandingMigration
Phaseo migration overview

Features

  • AI Stats is now Phaseo across the product, documentation, SDK examples, and public site. Read the migration announcement in AI Stats is becoming Phaseo.
  • New developer-facing examples use Phaseo naming and PHASEO_* configuration while existing deployments can keep their current environment-variable names.
  • The blog and announcement surfaces were refreshed for the rebrand, including the new migration announcement layout and launch imagery.
  • Local, preview, and production builds now use distinct favicons so active environments are easier to tell apart while testing.
Model ReleasesProvider CoveragePricing

Model Releases

Provider Coverage

  • Amazon Bedrock Mantle was split into its own provider catalogue with model coverage and pricing for Anthropic, Google, MiniMax, MoonshotAI, NVIDIA, OpenAI, Qwen, SpaceXAI, and Z.AI models.
  • AtlasCloud and Novita provider coverage was added for Tencent Hy3, including free-route pricing where available.
  • Groq Qwen 3.6 gateway routing was disabled until provider pricing is available.

Pricing

  • DeepSeek V4 Flash and DeepSeek V4 Pro pricing now supports time-windowed provider rates.
Features

Features

  • Server-tool inputs, monitor history links, workspace credit cache invalidation, and pricing edge cases were hardened across gateway and web surfaces (PR #900).
  • Chat gateway target resolution now fails closed in production when the configured gateway URL is missing, while preserving local-development overrides (PR #899).
  • Anthropic and Bedrock request shaping now uses adaptive thinking controls for Claude Sonnet 5 and newer Opus 4.x models (PR #901).
  • Server-tool handling was hardened for datetime, web fetch, web search, image generation, and fusion workflows.
  • Monitor links now apply stricter URL safety checks.
  • GMICloud MiniMax M3 cached-read pricing and Google Vertex Gemini 3.1 Flash-Lite Image pricing validation were tightened.
Features

Features

  • Chat shortcuts, prompt queue controls, composer tool configuration, restored tags, bulk editing, and richer response timing metadata shipped together for the chat workspace (PR #896).
Features

Features

  • Individual model pages were overhauled with clearer lifecycle notices, restored routing details, and better handoffs from model pages into gateway workflows (PR #888).
  • Retired and replaced model pages now make status, alternatives, and historical context easier to scan without burying the current model link (PR #888).
  • Chat tags, tool-call markers, API target controls, and mobile viewport handling were restored or tightened across the chat experience (PR #894, PR #895, PR #891).
FeaturesModel Releases

Features

  • Core interactive UI primitives moved onto Base UI, improving accessibility and behaviour consistency across filters, menus, dialogs, and model controls (PR #854).
  • Chat and settings flows received follow-up polish after the Base UI migration, including steadier model pickers and cleaner menu spacing (PR #853).

Model Releases

Laguna XS 2.1 model release artwork
Model Releases
FeaturesModel Releases

Features

  • Pricing calculators now support time-windowed pricing and side-by-side model comparisons, making burst, realtime, and route-specific pricing easier to reason about (PR #864).
  • The pricing calculator layout was simplified so model, provider, and usage assumptions are easier to adjust before comparing costs (PR #864).

Model Releases

Claude Sonnet 5 model release artwork
Model Releases

Model Releases

Model Releases
Model Releases

Model Releases

GLM 5.2 model release artwork
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases

Model Releases

Gemma 4 12B model release artwork
Model Releases
Model Releases
Model Releases
FeaturesModel ReleasesProvider CoveragePricing

Features

  • Pricing selectors now present Qwen 3.7 Max variants more clearly, reducing ambiguity between preview, dated, and current model routes (PR #418).
  • Anthropic-on-AWS provider branding was refreshed so Bedrock and Claude Platform routes are easier to distinguish in provider lists (PR #415).

Model Releases

Claude Opus 4.8 model release artwork
Model Releases
Model Releases
Model Releases

Model Releases

Qwen 3.7 Max model release artwork
Model Releases

Model Releases

Command A+ model release artwork
FeaturesModel ReleasesSDKDocs

SDK Releases

  • The TypeScript Agent SDK launch added a dedicated package for building agentic applications on top of Phaseo Gateway (PR #495). See @phaseo/agent-sdk.

Model Releases

Gemini 3.5 Flash model release artwork
Model Releases

Model Releases

Composer 2.5 model release artwork
FeaturesModel ReleasesProvider CoverageSDKPricingDocs

Features

  • Provider activation states and region filters were tightened so unavailable routes are easier to identify before a request is made (PR #495).
  • Free-router behaviour was made more predictable across pricing, provider selection, and gateway routing surfaces (PR #495).

SDK Releases

  • Agent SDK documentation and examples shipped with the first public workflow guidance for durable loops, parallel tools, and agent orchestration (PR #495).

Model Releases

Model Releases

Model Releases

Grok Build 0.1 model release artwork
FeaturesModel ReleasesProvider CoveragePricing

Features

  • Provider availability and pricing summaries were expanded so model pages show clearer route, tier, and cost differences (PR #474).
  • Realtime minute pricing became visible in the product, making audio and live-session models easier to compare against token-priced models (PR #477).

Model Releases

Gemini 3.1 Flash-Lite model release artwork
FeaturesModel Releases

Features

  • Video generation and async job lifecycle coverage expanded, including clearer status handling for long-running provider work (PR #468).
  • Gateway surfaces now expose more of the reservation, completion, and failure states needed to debug async media requests (PR #468).

Model Releases

Grok Imagine Image Quality model release artwork
Model Releases

Model Releases

chat-latest model release artwork
FeaturesSDK

SDK Releases

  • TypeScript SDK 2.0.4: generated OpenAPI and model-surface updates for the latest gateway schema.
  • Python SDK 2.0.4: generated OpenAPI and model-surface updates for the latest gateway schema.
  • Go SDK 2.0.4: generated OpenAPI and model-surface updates for the latest gateway schema.
  • C# SDK 2.0.4: generated OpenAPI and model-surface updates for the latest gateway schema.
  • Java SDK 2.0.4: generated OpenAPI and model-surface updates for the latest gateway schema.
  • PHP SDK 2.0.4: generated OpenAPI and model-surface updates for the latest gateway schema.
  • Ruby SDK 2.0.4: generated OpenAPI and model-surface updates for the latest gateway schema.
Model Releases

Model Releases

Grok 4.3 model release artwork
FeaturesModel ReleasesSDK

SDK Releases

  • SDK 2.0.2 shipped for TypeScript, Python, Go, C#, Java, PHP, and Ruby.
  • This release separated runtime request model IDs from generated callable-helper constants, so newly released models can be used before the next SDK publish.
  • Generated helpers now track the current callable-on-gateway snapshot, reducing noisy SDK churn from catalog-only model updates.

Model Releases

Granite 4.1 30B model release artwork
FeaturesPricingModel Releases

Features

  • Low-balance alerts and billing safeguards were expanded so teams get clearer warning before usage runs into account limits (PR #371).
  • Pricing edge cases now fail more predictably when provider billing rules are incomplete or route metadata changes unexpectedly (PR #374).

Model Releases

DeepSeek V4 Pro Lightning model release artwork
Model Releases
Model Releases
FeaturesSDKModel Releases

Features

  • SDK 2.0.0 shipped with regenerated clients and broader parity against the current gateway contract.

Model Releases

Gemini Embedding 2 model release artwork
FeaturesModel Releases

Features

  • The announcements hub, workspace management surfaces, and related membership flows shipped.

Model Releases

Gemini Deep Research Max Preview (04-2026) model release artwork
Model Releases
FeaturesModel ReleasesSDKDocs

Features

  • SDK 1.2.0 synced model availability, generated types, and docs with the latest catalog changes.
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Music 2.6 model release artwork
Model Releases
Model Releases

Model Releases

GLM 5.1 model release artwork
FeaturesModel ReleasesSDKPricing

Features

  • Mobile model pages, pricing layouts, and free-route visibility were improved, and the Pay As You Go top-up fee dropped to 5%.
Model Releases

Model Releases

Wan 2.7 T2V model release artwork
Model Releases
Model Releases
Model Releases

Model Releases

Voxtral TTS model release artwork
FeaturesModel ReleasesSDK

Features

  • New language SDKs shipped, TypeScript and Python SDKs were overhauled, and chat rooms went multimodal.
  • Foundry launched, LLM Council became the first Foundry project, and SDK 1.1.0 aligned model availability and lifecycle warnings with the API.
Model Releases
Model Releases
Model Releases

Model Releases

GLM 5 Turbo model release artwork
FeaturesModel Releases

Features

  • The models display page was overhauled for clearer availability, metadata hierarchy, and support context.

Model Releases

Model Releases
FeaturesModel ReleasesProvider CoverageDocs

Features

  • A broader brand refresh shipped across the website and docs.
  • Alibaba Cloud provider coverage landed across Phaseo, gateway routes, and documentation.

Model Releases

Aion 2.5 model release artwork
FeaturesModel ReleasesDocs

Features

  • GPT 5.4 and GPT 5.4 Pro landed across discovery, gateway, and docs.

Model Releases

GPT 5.4 model release artwork
FeaturesModel ReleasesDocs

Features

  • Gemini 3.1 Flash Lite Preview landed across product surfaces, gateway routing, and docs.

Model Releases

Gemini 3.1 Flash Lite Preview model release artwork
FeaturesDocsModel Releases

Features

  • Responses API assistant phase support landed across request handling, response payloads, and OpenAPI docs.

Model Releases

Qwen 3.5 0.8B model release artwork
FeaturesModel ReleasesDocs

Features

  • Gemini 3.1 Flash Image Preview support landed across discovery, gateway compatibility, and docs.

Model Releases

Gemini 3.1 Flash Image Preview (Nano Banana 2) model release artwork
Model Releases
FeaturesModel ReleasesPricingDocs

Features

  • Gemini 3.1 Pro Preview coverage landed across discovery, pricing, gateway support, and docs.

Model Releases

Gemini 3.1 Pro Preview model release artwork
FeaturesModel ReleasesDocs

Features

  • Claude Sonnet 4.6 coverage landed across discovery, gateway compatibility, and docs.

Model Releases

Lyria 3 model release artwork
Model Releases
Model Releases

Model Releases

GLM 5 model release artwork
Model Releases

Model Releases

Composer 1.5 model release artwork
Model Releases
Model Releases
FeaturesModel ReleasesSDK

Features

  • The gateway was overhauled again, including Anthropic-compatible /messages support.
  • Database architecture, website UX, playground tooling, and SDK direction were all refreshed for broader expansion.

Model Releases

Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases

Model Releases

GLM 4.7 Flash model release artwork
Model Releases

Model Releases

Music 2.5 model release artwork
Model Releases

Model Releases

GLM Image model release artwork
Model Releases
Model Releases

Model Releases

Scribe V2 model release artwork
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases

Model Releases

Model Releases

Model Releases

Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
FeaturesModel ReleasesSDKPricingDocs

Features

  • OpenAI-compatible /responses support, aligned SDKs, and gateway quickstart or pricing page updates shipped.
  • Docs expanded ahead of the SDK launch, and nightly CI started automatically refreshing model statuses.

Model Releases

Model Releases
Model Releases
Model Releases

Model Releases

Features

Features

  • Dark mode shipped.
Model Releases
FeaturesModel ReleasesProvider CoverageSDKBenchmarks

Features

  • Phaseo relaunched with a refreshed brand, a broader gateway, and much wider data coverage.
  • Model availability, benchmark pages, provider pages, and the latest-updates surfaces were overhauled.
  • The unified gateway entered alpha, and the first Python and TypeScript or JavaScript SDKs launched alongside it.

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases

Model Releases

Model Releases

Model Releases

Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases

Model Releases

Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases
Model Releases
Model Releases
Model Releases

Model Releases

Model Releases
Model Releases

Model Releases

Model Releases

Model Releases

Model Releases

Model Releases

Model Releases

Model Releases

Last modified on July 31, 2026