Can a GB10 run a 120B model?
Can a GB10 run a 120B model?
Rentable machine
dgx-spark
Can a GB10 run a 120B model?
Maybe, but the parameter count by itself is not enough to answer it.
A heavily quantized or MoE model can be realistic where a 120B dense model may be too large or too slow for the result you want.
AxForge's GB10 page shows large MoE models running on the hardware class, but for a specific 120B model we would want the exact model name and precision first.
If you send us the Hugging Face model link, we can size the weights and likely KV-cache requirement before recommending the machine.
Sign in to reply.
Posting guidelines
Accounts that break these can lose forum access — paid plans included. Read the full guidelines →
Related topics