Coding Models
Models tuned for code completion, refactoring, and review. Stronger on programming benchmarks than general LLMs of similar parameter count, while remaining usable through the standard Chat Completions endpoint.
Showing 2 of 2 models
| Token-by-token output via Server-Sent Events. Suitable for low-latency, real-time UI. | Function / tool calling (OpenAI-compatible). The model can return structured tool invocations. | Price per million prompt tokens served from the prefix cache: 15% of the normal input price. T-Cloud-hosted models only — cached pricing for external providers will be added later. | ||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Qwen 3.8 27B FP8 | | Hosted on T-Cloud Public — Telekom's sovereign infrastructure in Germany. | Text, Image | Text | ✓ | ✓ | 256K | — | — | — | Test | |
| GPT-5 Codex | | Hosted on Microsoft Azure (EU regions). | Text, Image | Text | ✓ | ✓ | 400K | €1.10 | €8.70 | n/a | EssentialProfessionalAgentic | |
| No models match the current filters. | ||||||||||||