|
|
|
|
|
by canucker2016
297 days ago
|
|
I never claimed the 200B model was FP16. If the 200B model was at FP16, marketing could've turned around and claimed the DGX Spark could handle a 400B model (with an 8-bit quant) or a 800B model at some 4-bit quant. Why would marketing leave such low-hanging fruit on the tree? They wouldn't. |
|