Skip to content

Coding Models

Models tuned for code completion, refactoring, and review. Stronger on programming benchmarks than general LLMs of similar parameter count, while remaining usable through the standard Chat Completions endpoint.

Provider
Cloud
Input
Output
Streaming output
Tool calling
Plans
Showing 2 of 2 models
Token-by-token output via Server-Sent Events. Suitable for low-latency, real-time UI. Function / tool calling (OpenAI-compatible). The model can return structured tool invocations. Price per million prompt tokens served from the prefix cache: 15% of the normal input price. T-Cloud-hosted models only — cached pricing for external providers will be added later.
Qwen 3.8 27B FP8 Alibaba Cloud Alibaba Hosted on T-Cloud Public — Telekom's sovereign infrastructure in Germany. Text, Image Text 256K Test
GPT-5 Codex OpenAI OpenAI Hosted on Microsoft Azure (EU regions). Text, Image Text 400K €1.10 €8.70 n/a EssentialProfessionalAgentic