FP16 MASS
QUANT LEVELS
|ERROR|
Simplified on purpose. Round-to-nearest symmetric int on a seeded
synthetic 64×64 tensor. Real quality depends on architecture, calibration,
method (GPTQ, AWQ, NF4), kernels, activation and KV precision, and the task.
Bit count alone does not predict usability. Memory counts weights plus group
scales only. Downstream numbers are simulated, not a real LLM.