Skip to main content

Qwen · General · Qwen3.8

Qwen3.8 27B: Specs, GPU requirements, hosting & deployment

Qwen3.8 27B is a current General release in the Qwen family, certified from its immutable official source revision.

Technical deployment reference illustration for Qwen3.8 27B
Reference deployment flow; exact architecture facts are listed separately.
Total parameters27.0B
Active parametersDense model
Context window262,144 tokens
Release statusCurrent · Current
ArchitectureQwen3_5ForConditionalGeneration
ModalitiesText, Image, Video
LicenseApache License 2.0
Deployment stateDeployable with certified configuration
Commercial stateCommercial Use Allowed
Release date2026-08-05

Architecture, strengths and use cases

Qwen3.8 27B is an official self-hostable Qwen release in the Qwen General track. It uses Qwen3_5ForConditionalGeneration, carries 27B, supports a verified context window of 262,144 tokens, and exposes only source-certified checkpoints and runtime combinations. Its release, license and deployment facts are tied to the official repository revision rather than inferred from the family name.

Qwen3.8 27B uses Qwen3_5ForConditionalGeneration with 27B. The stored architecture profile records 64 layers, hidden width 5120, 24 attention heads and a 262,144 token context ceiling where the official configuration publishes those fields.

Verified Qwen3_5ForConditionalGeneration architecture
262,144 token certified context
Official BF16, FP8, Unknown checkpoint coverage
Private LLM inference
Agentic workflows
Long-context document processing
Vision-language analysis

Certified checkpoints and quantizations

Only exact, source-certified variants attached to this release are shown.

Qwen3.8-27B

Qwen/Qwen3.8-27B

Official · BF16 · safetensors

Runtime certified: Transformers vLLM SGLang

Precision / quantizationFormatBitsPublisher trustDeployment state
BF16safetensors16OfficialRuntime certified

Qwen3.8-27B-FP8

Qwen/Qwen3.8-27B-FP8

Official · FP8 · safetensors

Runtime certified: Transformers vLLM SGLang

Precision / quantizationFormatBitsPublisher trustDeployment state
FP8safetensors8OfficialRuntime certified

Recommended hardware

Use only the listed transformers, vllm, sglang runtime configuration with the exact checkpoint precision. The recommendation engine checks weight memory, KV cache, runtime overhead, safety headroom, GPU architecture and topology before ranking rentable and reference hardware.

Estimated weight memory51.8 GB
KV cache2 GB
Runtime overhead14.4 GB
Total with safety margin84.57 GB

Recommended best value

2× NVIDIA A100 SXM 80GB

160 GB total VRAM · nvlink

75.43 GB headroom · Available

USD 1,830 / month

Lowest-cost viable

4× NVIDIA L4 24GB

96 GB total VRAM · pcie

11.43 GB headroom · Available

USD 1,000 / month

Highest performance

2× NVIDIA A100 SXM 80GB

160 GB total VRAM · nvlink

75.43 GB headroom · Available

USD 1,830 / month

License and self-hosting considerations

The full 27B resident weight set must fit across the selected GPUs. Context and concurrency add KV-cache memory; MoE active parameters do not replace resident-weight memory.

Qwen3.8 27B is published under Apache License 2.0. The release-scoped official license link and verification date are retained with the catalogue record.

Commercial Use Allowed. The official license permits commercial use; its obligations still apply.

Alternatives and sibling variants

Qwen3.8 2.4T-A95B, Qwen3 30B-A3B Instruct 2507, Qwen3 235B-A22B

Release timeline

qwen/qwen/qwen3-8-27b @ 3e7c75625b03cf6f2d7256b8dd37aae1edd2bdfed8b6489d80f23c9a09db6165

Official source

Frequently asked questions

How much VRAM does Qwen3.8 27B require?

The exact estimate depends on checkpoint precision, context and concurrency; the page recommendation uses the certified memory engine and compatible runtime configurations.

Which runtimes support Qwen3.8 27B?

Certified upstream runtime evidence is available for transformers, vllm, sglang.

Can Qwen3.8 27B be used commercially?

The stored commercial classification is Commercial Use Allowed. Consult the official Apache License 2.0 terms for the final obligations.

Official references

Last verified: 2026-08-26