bestmodel.run / model — NVIDIA-Nemotron-3.5-Lightning-30B-A3B-W4A16 (agent view) category chat · useful-quants/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-W4A16 · 38 runs in the model record answer: 226.4 tok/s (model median) Honesty ladder: measured > reported > extrapolated > formula > no data yet. Empty fields stay empty — never filled with an estimate. rig | quant | median result | basis | ttft | peak vram | max context CMP 170HX 64GB | 4-bit | 226.4 tok/s | measured | ttft 1,697ms | vram - GB | ctx 67,592