LocalOps LogoLocalOps
Back to Calculator

GLM-5.1

Z.ai next-gen flagship for agentic engineering. 744B MoE, 40B active. MIT licensed. #1 open-weight model on SWE-Bench Pro as of April 2026. Trained on Huawei Ascend chips.

Specifications

Source
ArchitectureTEXT
Parameters744B
Familyglm
VRAM (Q4)372.0G
MoE: 40B active.
codingagentsreasoning

Run in the Cloud

This model requires enterprise-grade VRAM. Rent GPUs on RunPod and start generating.

Deploy on RunPod

Instant Cloud GPUs

Running out of VRAM? Rent a high-end H100 or RTX 4090 on RunPod and deploy in seconds.

Deploy Now

Quantization Estimates

FormatVRAM NeedTier
FP161488.0 GBFull Precision
Q8_0744.0 GBHigh
Q6_K632.4 GBExcellent
Q5_K_M520.8 GBGreat
Q4_K_M372.0 GBSweet Spot
Q2_K223.2 GBEmergency

Share this Model

Send these specs directly to your community.

Post