LocalOps LogoLocalOps
Back to Calculator

Llama 3.1 405B

Frontier-class open model. Requires datacenter hardware.

Specifications

Source
ArchitectureTEXT
Parameters405B
Familyllama
VRAM (Q4)202.5G
flagshipmetadatacenter

Run in the Cloud

This model requires enterprise-grade VRAM. Rent GPUs on RunPod and start generating.

Deploy on RunPod

Instant Cloud GPUs

Running out of VRAM? Rent a high-end H100 or RTX 4090 on RunPod and deploy in seconds.

Deploy Now

Quantization Estimates

FormatVRAM NeedTier
FP16810.0 GBFull Precision
Q8_0405.0 GBHigh
Q6_K344.3 GBExcellent
Q5_K_M283.5 GBGreat
Q4_K_M202.5 GBSweet Spot
Q2_K121.5 GBEmergency

Share this Model

Send these specs directly to your community.

Post