Qwen3.8 27B API pricing
Of the 6 providers checked, DeepInfra lists the lowest blended rate for Qwen3.8 27B: $0.15 input and $1.875 output per 1M tokens, on its pricing page as of 27 September 2026.
The compact Qwen3.8 model: 27B parameters plus a vision encoder for images and video. Context is 262K natively and extends to 1M. Licensed Apache 2.0.
- Creator
- Alibaba (Qwen)
- Weights
- Open
- License
- Apache 2.0
- Size
- 27B total
- Context
- 256K tokens
- Input
- text, image, video
- Model card
- Hugging Face
Facts from the model card, checked 25 Sep 2026.
Across 6 providers, the most expensive charges 2.1× the cheapest on a blended basis. 1 provider is on a limited-time promotion.
Prices by provider
Per 1M tokens, on-demand, sorted cheapest first by blended cost (3:1 input to output). Off-peak and long-context rates are listed on each provider's page.
| # | Provider | Input / 1M | Cached / 1M | Output / 1M | Checked | Notes |
|---|---|---|---|---|---|---|
| 1 | DeepInfra | $0.15 | $0.038 | $1.875 | 27 Sep 2026 | Promotional price, 25% off list $0.20 / $2.50 |
| 2 | LLM Tech | $0.25 | $0.04 | $2.09 | 25 Sep 2026 | Cached input is automatic |
| 3 | CoreWeave | $0.40 | $0.15 | $3.00 | 27 Sep 2026 | |
| 4 | Novita AI | $0.42 | $0.085 | $3.00 | 27 Sep 2026 | |
| 5 | Cloudflare Workers AI | $0.45 | – | $3.20 | 27 Sep 2026 | @cf/qwen/qwen3.8-27b |
| 6 | Berget AI | €0.40 ≈ $0.4547 | €0.40 | €3.00 ≈ $3.4101 | 27 Sep 2026 | Preview; listed as Qwen3.8-27B-FP8. Cached input is billed at the input price |
Each price was read off the provider's own pricing page on the date shown; the date links to that page. Providers that publish no per-token price for this model are not listed. EUR prices are ranked at the ECB reference rate of 24 September 2026. The same data is open under CC-BY at /api/query/prices.
Compare every inference provider on free tiers, hosting and EU availability in the inference APIs table.
Is your product missing?