≫ Home / LLM API pricing / Qwen3.8 27B

Qwen3.8 27B API pricing

Of the 6 providers checked, DeepInfra lists the lowest blended rate for Qwen3.8 27B: $0.15 input and $1.875 output per 1M tokens, on its pricing page as of 27 September 2026.

The compact Qwen3.8 model: 27B parameters plus a vision encoder for images and video. Context is 262K natively and extends to 1M. Licensed Apache 2.0.

Creator
Alibaba (Qwen)
Weights
Open
License
Apache 2.0
Size
27B total
Context
256K tokens
Input
text, image, video
Model card
Hugging Face

Facts from the model card, checked 25 Sep 2026.

Across 6 providers, the most expensive charges 2.1× the cheapest on a blended basis. 1 provider is on a limited-time promotion.

Prices by provider

Per 1M tokens, on-demand, sorted cheapest first by blended cost (3:1 input to output). Off-peak and long-context rates are listed on each provider's page.

# Provider Input / 1M Cached / 1M Output / 1M Checked Notes
1 DeepInfra $0.15 $0.038 $1.875 27 Sep 2026 Promotional price, 25% off list $0.20 / $2.50
2 LLM Tech $0.25 $0.04 $2.09 25 Sep 2026 Cached input is automatic
3 CoreWeave $0.40 $0.15 $3.00 27 Sep 2026
4 Novita AI $0.42 $0.085 $3.00 27 Sep 2026
5 Cloudflare Workers AI $0.45 – $3.20 27 Sep 2026 @cf/qwen/qwen3.8-27b
6 Berget AI €0.40 ≈ $0.4547 €0.40 €3.00 ≈ $3.4101 27 Sep 2026 Preview; listed as Qwen3.8-27B-FP8. Cached input is billed at the input price

Each price was read off the provider's own pricing page on the date shown; the date links to that page. Providers that publish no per-token price for this model are not listed. EUR prices are ranked at the ECB reference rate of 24 September 2026. The same data is open under CC-BY at /api/query/prices.

Compare every inference provider on free tiers, hosting and EU availability in the inference APIs table.

Is your product missing?

Add it here →