Qwen3.8-27B
Official · BF16 · safetensors
Runtime certified: Transformers vLLM SGLang
| Precision / quantization | Format | Bits | Publisher trust | Deployment state |
|---|---|---|---|---|
| BF16 | safetensors | 16 | Official | Runtime certified |
Qwen · General · Qwen3.8
Qwen3.8 27B is a current General release in the Qwen family, certified from its immutable official source revision.
Qwen3.8 27B is an official self-hostable Qwen release in the Qwen General track. It uses Qwen3_5ForConditionalGeneration, carries 27B, supports a verified context window of 262,144 tokens, and exposes only source-certified checkpoints and runtime combinations. Its release, license and deployment facts are tied to the official repository revision rather than inferred from the family name.
Qwen3.8 27B uses Qwen3_5ForConditionalGeneration with 27B. The stored architecture profile records 64 layers, hidden width 5120, 24 attention heads and a 262,144 token context ceiling where the official configuration publishes those fields.
Only exact, source-certified variants attached to this release are shown.
Official · BF16 · safetensors
Runtime certified: Transformers vLLM SGLang
| Precision / quantization | Format | Bits | Publisher trust | Deployment state |
|---|---|---|---|---|
| BF16 | safetensors | 16 | Official | Runtime certified |
Official · FP8 · safetensors
Runtime certified: Transformers vLLM SGLang
| Precision / quantization | Format | Bits | Publisher trust | Deployment state |
|---|---|---|---|---|
| FP8 | safetensors | 8 | Official | Runtime certified |
Use only the listed transformers, vllm, sglang runtime configuration with the exact checkpoint precision. The recommendation engine checks weight memory, KV cache, runtime overhead, safety headroom, GPU architecture and topology before ranking rentable and reference hardware.
Recommended best value
160 GB total VRAM · nvlink
75.43 GB headroom · Available
USD 1,830 / monthLowest-cost viable
96 GB total VRAM · pcie
11.43 GB headroom · Available
USD 1,000 / monthHighest performance
160 GB total VRAM · nvlink
75.43 GB headroom · Available
USD 1,830 / monthThe full 27B resident weight set must fit across the selected GPUs. Context and concurrency add KV-cache memory; MoE active parameters do not replace resident-weight memory.
Qwen3.8 27B is published under Apache License 2.0. The release-scoped official license link and verification date are retained with the catalogue record.
Commercial Use Allowed. The official license permits commercial use; its obligations still apply.
Qwen3.8 2.4T-A95B, Qwen3 30B-A3B Instruct 2507, Qwen3 235B-A22B
The exact estimate depends on checkpoint precision, context and concurrency; the page recommendation uses the certified memory engine and compatible runtime configurations.
Certified upstream runtime evidence is available for transformers, vllm, sglang.
The stored commercial classification is Commercial Use Allowed. Consult the official Apache License 2.0 terms for the final obligations.
Last verified: 2026-08-26