Skip to content

Changelog

Major changes to AI Foundation Services — new models, feature updates, and deprecations.

Aligns the documentation with Service Description v1.23 (31 August 2026), which is the binding source for everything below.

  • New preview model in the Test catalog: GLM 5.3 Flash — natively multimodal, free on any active plan. See Test / Preview Models.
  • New models on the Essential, Professional and Agentic plans: GPT-5.6 Terra and GPT-5.6 Luna. See Plans & Pricing.
  • Gemini 3.1 Pro is now priced by context length.
  • Cached input token pricing on T-Cloud-hosted models. See Prefix Caching.
  • The availability target is 99.5%, not the 99.9% previously published.
  • GLM 5.2 token prices reduced, and it moved to the standard model tier.
  • GPT-Image-2 pricing corrected.
  • Prefix caching is on by default on T-Cloud-hosted models, not opt-in.
  • Server location and data processing for the Gemini 3 family now read Worldwide instead of Europe. See Enterprise Trust.
  • The Anthropic models are labelled GCP / Azure.
  • The Service Description shipped with the docs is now v1.23.
  • Enterprise Trust category 1 now lists the models the Service Description names.
  • Llama 3.3 70B, Claude 4.5 Sonnet and Claude 4.5 Opus reached their retirement date. See Retired Models for replacements.
  • GLM 5.2 is listed on Coding Models again.
  • OpenCode: corrected the authentication URL in the plugin instructions.
  • /category/model-serving now redirects to the guides instead of returning a 404.
  • New preview model in the Test catalog: Qwen 3.8 27B FP8 — natively multimodal, free on any active plan. See Test / Preview Models.
  • New models on the Essential, Professional and Agentic plans: GPT-5.4, GPT-5.4 mini, GPT-5.5 and GPT-Image-2 (OpenAI); Gemini 3.1 Pro and Gemini 3.5 Flash (Google). Claude 4.8 Opus (Anthropic) on Professional and Agentic. See Plans & Pricing.
  • Mistral Small 4 and Gemma 4 are now generally available, having been preview-only. See Plans & Pricing.
  • New preview model in the Test catalog: GLM 5.2 — free on any active plan. See Test / Preview Models.
  • Models across the GPT-4.1, GPT-4o, o-series, Qwen3, Gemini 2.5, Gemini 3 and Claude Sonnet / Opus families are scheduled for retirement on 2026-08-01, along with Mistral Small 3 and GPT Image 1. See Scheduled for Retirement for the table and replacements.

Archived snapshot: v1.1.0.

  • New top-level Models section with a comparison page covering every hosted model — context window, pricing, cloud, modalities, plan availability, and OpenAI-compatible endpoint reference. Image-capable models also surface on the dedicated Vision page via a capability filter.
  • Changelog page (you are reading it).
  • Two new preview models in the Test catalog: Mistral Small 4 (Mistral-Small-4-119B-2603-Preview) and Gemma 4 (gemma-4-31B-it-FP8-preview). See rate limits for the preview-model policy.
  • Fine-Tuning API documentation. Requests to /guides/fine-tuning now redirect to the optional services. Superseded: fine-tuning is offered as an optional service on request — see § 3.4.2 of the Service Description. The self-service API documentation remains withdrawn.

Archived snapshot: v1.0.0 — the initial released state of the docs prior to the changelog being introduced.