Qwen / Qwen3-VL 32B Instruct

Transformation options for Qwen3-VL 32B Instruct

Validated per release from registry evidence, pricing components and runtime compatibility.

Technical transformation state

Transformation supported

Verified public open-weight release with authoritative checkpoint evidence.

Architecture

Qwen3VLForConditionalGeneration

32.00B parameters · 262144 context

Supported transformations

quantization

Validated per release

runtime packaging

Validated per release

managed hosting

Validated per release

Distillation

Validated per release

gpu targeting

Validated per release

Evaluation and benchmarking

Validated per release

download

Validated per release

Published packages

Engineering ready

Qwen3-VL 32B Instruct FP8 Production Package - Downloadable production build

Qwen · Qwen3-VL 32B Instruct

Package profile: Downloadable production build

Transformation operations: FP8 Quantization, Evaluation and benchmarking

Target GPU: Hardware selected in the builder

Runtime packaging: transformers

Configured delivery: Download package

Delivery options: Download package / Host with FusionPointAI / Download + host

FP8transformers Download packageHost with FusionPointAIDownload + host

USD 28,400.00
One-time engineering

Engineering ready

Qwen3-VL 32B Instruct FP8 Production Package - GPU-targeted hosted production

Qwen · Qwen3-VL 32B Instruct

Package profile: GPU-targeted hosted production

Transformation operations: FP8 Quantization, Runtime packaging, Evaluation and benchmarking

Target GPU: NVIDIA H200 SXM 141GB

GPU count: 1

Runtime packaging: transformers

Configured delivery: Download + host

Delivery options: Download package / Host with FusionPointAI / Download + host

FP8GPU targeted1 × NVIDIA H200 SXM 141GBtransformersRuntime optimized Download packageHost with FusionPointAIDownload + host

USD 34,400.00
One-time engineering

Engineering ready

Qwen3-VL 32B Instruct FP8 Production Package - GPU-targeted hosted production

Qwen · Qwen3-VL 32B Instruct

Package profile: GPU-targeted hosted production

Transformation operations: FP8 Quantization, Runtime packaging, Evaluation and benchmarking

Target GPU: NVIDIA H100 SXM 80GB

GPU count: 1

Runtime packaging: transformers

Configured delivery: Download + host

Delivery options: Download package / Host with FusionPointAI / Download + host

FP8GPU targeted1 × NVIDIA H100 SXM 80GBtransformersRuntime optimized Download packageHost with FusionPointAIDownload + host

USD 34,400.00
One-time engineering

Engineering ready

Qwen3-VL 32B Instruct FP8 Production Package - GPU-targeted hosted production

Qwen · Qwen3-VL 32B Instruct

Package profile: GPU-targeted hosted production

Transformation operations: FP8 Quantization, Runtime packaging, Evaluation and benchmarking

Target GPU: NVIDIA H100 PCIe 80GB

GPU count: 1

Runtime packaging: transformers

Configured delivery: Download + host

Delivery options: Download package / Host with FusionPointAI / Download + host

FP8GPU targeted1 × NVIDIA H100 PCIe 80GBtransformersRuntime optimized Download packageHost with FusionPointAIDownload + host

USD 34,400.00
One-time engineering

Engineering ready

Qwen3-VL 32B Instruct FP8 Production Package - GPU-targeted hosted production

Qwen · Qwen3-VL 32B Instruct

Package profile: GPU-targeted hosted production

Transformation operations: FP8 Quantization, Runtime packaging, Evaluation and benchmarking

Target GPU: NVIDIA A100 SXM 80GB

GPU count: 1

Runtime packaging: transformers

Configured delivery: Download + host

Delivery options: Download package / Host with FusionPointAI / Download + host

FP8GPU targeted1 × NVIDIA A100 SXM 80GBtransformersRuntime optimized Download packageHost with FusionPointAIDownload + host

USD 34,400.00
One-time engineering

Engineering ready

Qwen3-VL 32B Instruct FP8 Production Package - GPU-targeted hosted production

Qwen · Qwen3-VL 32B Instruct

Package profile: GPU-targeted hosted production

Transformation operations: FP8 Quantization, Runtime packaging, Evaluation and benchmarking

Target GPU: NVIDIA l40s

GPU count: 1

Runtime packaging: transformers

Configured delivery: Download + host

Delivery options: Download package / Host with FusionPointAI / Download + host

FP8GPU targeted1 × NVIDIA l40stransformersRuntime optimized Download packageHost with FusionPointAIDownload + host

USD 34,400.00
One-time engineering

Engineering ready

Qwen3-VL 32B Instruct FP8 Production Package - GPU-targeted hosted production

Qwen · Qwen3-VL 32B Instruct

Package profile: GPU-targeted hosted production

Transformation operations: FP8 Quantization, Runtime packaging, Evaluation and benchmarking

Target GPU: NVIDIA L4

GPU count: 2

Runtime packaging: transformers

Configured delivery: Download + host

Delivery options: Download package / Host with FusionPointAI / Download + host

FP8GPU targeted2 × NVIDIA L4transformersRuntime optimized Download packageHost with FusionPointAIDownload + host

USD 34,400.00
One-time engineering

Design validated

Qwen3-VL 32B Instruct DISTILLATION Production Package - Quality-retention distillation

Qwen · Qwen3-VL 32B Instruct

Package profile: Quality-retention distillation

Transformation operations: Distillation, Evaluation and benchmarking

Target GPU: Hardware selected in the builder

Runtime packaging: transformers

Configured delivery: Download package

Delivery options: Download package / Host with FusionPointAI / Download + host

DISTILLATIONQuality retentiontransformers Download packageHost with FusionPointAIDownload + host

USD 47,875.00
One-time engineering

Design validated

Qwen3-VL 32B Instruct DISTILLATION Production Package - Throughput-optimized distillation

Qwen · Qwen3-VL 32B Instruct

Package profile: Throughput-optimized distillation

Transformation operations: Distillation, Evaluation and benchmarking

Target GPU: Hardware selected in the builder

Runtime packaging: transformers

Configured delivery: Download package

Delivery options: Download package / Host with FusionPointAI / Download + host

DISTILLATIONThroughput optimizedtransformers Download packageHost with FusionPointAIDownload + host

USD 47,875.00
One-time engineering