Skip to main content

GLM · General · glm-4-7

GLM-4.7-NVFP4: Specs, GPU requirements, hosting & deployment

GLM-4.7-NVFP4 is a recent General release in the GLM family, certified from its immutable official source revision.

Reference only Compare this model
Technical deployment reference illustration for GLM-4.7-NVFP4
Reference deployment flow; exact architecture facts are listed separately.
Total parametersNot published by manufacturer
Active parametersDense model
Context windowN/A / Not directly comparable
Release statusRecent · Recent
Architecture
ModalitiesText
Licensemit
Deployment stateReference only
Commercial stateCommercial Use Allowed
Release date

Architecture, strengths and use cases

GLM-4.7-NVFP4 is an official self-hostable Z.ai release in the GLM General track. It uses Not published by manufacturer, carries Not published by manufacturer, supports a verified context window of 0 tokens, and exposes only source-certified checkpoints and runtime combinations. Its release, license and deployment facts are tied to the official repository revision rather than inferred from the family name.

GLM-4.7-NVFP4 uses Not published by manufacturer with Not published by manufacturer. The stored architecture profile records layers, hidden width , attention heads and a 0 token context ceiling where the official configuration publishes those fields.

Verified Not published by manufacturer architecture
Official checkpoint and pipeline evidence
Official NVFP4 checkpoint coverage
Private LLM inference
Agentic workflows
Long-context document processing

Certified checkpoints and quantizations

Only exact, source-certified variants attached to this release are shown.

GLM-4.7-NVFP4

nvidia/GLM-4.7-NVFP4

Official · NVFP4 · safetensors

Precision / quantizationFormatBitsPublisher trustDeployment state

Recommended hardware

Use only the listed certified runtime configuration with the exact checkpoint precision. The recommendation engine checks weight memory, KV cache, runtime overhead, safety headroom, GPU architecture and topology before ranking rentable and reference hardware.

Reference only

The official license permits commercial use; its obligations still apply.

License and self-hosting considerations

The full 0B resident weight set must fit across the selected GPUs. Context and concurrency add KV-cache memory; MoE active parameters do not replace resident-weight memory.

GLM-4.7-NVFP4 is published under mit. The release-scoped official license link and verification date are retained with the catalogue record.

Commercial Use Allowed. The official license permits commercial use; its obligations still apply.

Alternatives and sibling variants

GLM-5.3, GLM-5.2-NVFP4, GLM-5-NVFP4, GLM-5.1-NVFP4, GLM-5.2-FP8, GLM-5.1-FP8, GLM-5-FP8, GLM-4.7-FP8, GLM-4.5-FP8, GLM-5.2, GLM-5.1

Release timeline

zai/glm/glm-47-nvfp4 @ 3a7e6eea05b153333007bc6dc850143b86b448a37662f5798635134d2b9587fd

Official source

Frequently asked questions

How much VRAM does GLM-4.7-NVFP4 require?

The exact estimate depends on checkpoint precision, context and concurrency; the page recommendation uses the certified memory engine and compatible runtime configurations.

Which runtimes support GLM-4.7-NVFP4?

No runtime is presented as supported until upstream evidence is verified.

Can GLM-4.7-NVFP4 be used commercially?

The stored commercial classification is Commercial Use Allowed. Consult the official mit terms for the final obligations.

Official references

Last verified: 2026-09-03