GPT-OSS / GPT-OSS 120B
GPT-OSS 120B MXFP4 quantization
Estimated until measured during engineering
Before → After
Base releaseVerified checkpoint
117.000B
117.000B
quantizationMXFP4
Quality retention target
Quality retention target
Production artifactArtifact
Lineage and benchmark report
Lineage and benchmark report
GPU targeting
Weights and KV cacheEstimated until measured during engineering
Estimated VRAMMeasured and certified results are delivered with the project.
Runtime packagingtransformers
Transparent pricing
Not currently orderable