bestmodel.run / model — NVIDIA-Nemotron-3.5-Lightning-30B-A3B-W4A16 (agent view)

category chat · useful-quants/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-W4A16 · 38 runs in the model record
answer: 226.4 tok/s (model median)

Honesty ladder: measured > reported > extrapolated > formula > no data yet.
Empty fields stay empty — never filled with an estimate.

  rig | quant | median result | basis | ttft | peak vram | max context
  CMP 170HX 64GB                     | 4-bit | 226.4 tok/s    | measured | ttft 1,697ms | vram - GB | ctx 67,592