Compatibility check
Can I run Llama 3.2 1B Instruct on H100 80GB SXM?
Yes — it runs.best quant:
Decode speed
3684 tok/sest
Usable context
33K
Memory
80 GB
Fits in memory3684 tok/sest · 33K ctx
Fits in memory3140 tok/sest · 33K ctx
Fits in memory2102 tok/sest · 33K ctx