Qwen3.8-27B qwen/qwen3.8-27b
Served by 16 providers. Cheapest: Chutes at $0.73 blended /1M tokens. The most expensive listing costs 2.2× as much for the same model.
| Provider | Type | Input /1M | Output /1M | Cached input | Blended | Context | Notes | Source |
|---|---|---|---|---|---|---|---|---|
| Chutes Qwen/Qwen3.8-27B-TEE |
Inference hosts | $0.24 | $2.20 | $0.024 | $0.73 | 262k | fp8 TEE (confidential compute); upstream weights Qwen/Qwen3.8-27B-FP8 |
✓ src src2 2026-09-18 |
| AkashML Qwen/Qwen3.8-27B |
Inference hosts | $0.25 | $2.20 | $0.05 | $0.738 | — | Limited Trial. |
✓ src src2 2026-09-18 |
| DeepInfra Qwen/Qwen3.8-27B |
Inference hosts | $0.20 | $2.50 | — | $0.775 | 262k | Price columns are DeepInfra's list price; its models API flags a 25% promotional discount on this model (no end date given), so billed price is lower. |
✓ src src2 2026-09-18 |
| OpenRouter qwen/qwen3.8-27b |
LLM routers | $0.214 | $2.55 | $0.15 | $0.798 | 262k | api modality text+image+video->text |
✓ src 2026-09-18 |
| CoreWeave Qwen/Qwen3.8-27B |
Inference hosts | $0.40 | $3 | $0.15 | $1.05 | 262k | Context from docs model table (rounded, e.g. "262k"). |
✓ src src2 2026-09-18 |
| Featherless Qwen/Qwen3.8-27B |
Inference hosts | $0.40 | $3 | $0.15 | $1.05 | 262k | Developer-plan per-token rate; model_class qwen3_5-27b; also usable flat-rate on $25/mo Chat plan (32K ctx, human chat only) |
✓ src src2 src3 2026-09-18 |
| Novita AI qwen/qwen3.8-27b |
Inference hosts | $0.42 | $3 | $0.085 | $1.06 | 1M |
✓ src src2 2026-09-18 |
|
| Cerebras Inference qwen-3.8-27b |
Inference hosts | $0.99 | $1.49 | — | $1.11 | 131k | Free trial: 64K context / 32K output; ~1850 tok/s; prompt caching supported (price not listed) |
✓ src 2026-09-18 |
| OVHcloud AI Endpoints Qwen3.8-27B |
Inference hosts | $0.459 | $3.10 | — | $1.12 | — | fp8 list price EUR 0.4/M in, EUR 2.7/M out; context shown on catalog as "262K" (rounded); converted at ECB EURUSD 1.1481 2026-09-17 (latest ECB reference rate) |
✓ src 2026-09-18 |
| Alibaba Cloud Model Studio (Qwen API, International) qwen3.8-27b |
Model labs | $0.50 | $3 | — | $1.12 | — | Singapore/International price. Flat 0-1M. Context-cache discount available. |
✓ src 2026-09-18 |
| Atlas Cloud qwen/qwen3.8-27b |
Inference hosts | $0.50 | $3 | $0.05 | $1.12 | 1.05M | fp8 |
✓ src 2026-09-18 |
| Cloudflare Workers AI @cf/qwen/qwen3.8-27b |
Inference hosts | $0.45 | $3.20 | — | $1.14 | — | Billed in neurons ($0.011/1k neurons); USD per-M-token equivalent as published by Cloudflare |
✓ src 2026-09-18 |
| GMI Cloud Inference Engine Qwen/Qwen3.8-27B |
Inference hosts | $0.45 | $3.20 | — | $1.14 | — | from GMI console public pricing API (console is client-rendered; docs say prices live in console Model Hub) |
✓ src src2 2026-09-18 |
| Venice.ai API qwen-3-8-27b |
Inference hosts | $0.45 | $3.20 | — | $1.14 | 262k | fp8 |
✓ src src2 2026-09-18 |
| Scaleway qwen3.8-27b |
Inference hosts | $0.689 | $3.79 | $0.138 | $1.46 | — | batch −50% list price EUR 0.6/M in, EUR 3.3/M out, EUR 0.12/M cached; Paris region; prices before tax; Batches API -50%; converted at ECB EURUSD 1.1481 2026-09-17 (latest ECB reference rate) |
✓ src 2026-09-18 |
| Groq qwen/qwen3.8-27b |
Inference hosts | $0.80 | $4 | — | $1.60 | 131k | batch −50% Preview model; ~450 tok/s |
✓ src 2026-09-18 |
✓ from the provider's own pricing page · ◐ partly confirmed there · 2° third-party source. Router prices may exclude the router's own fees; see each router's page.