Where these numbers come from
TurboLLM users who opt into anonymous benchmark sharing contribute the measurements behind these pages — model, quant, hardware and measured speed, nothing else. The data refreshes weekly, so coverage grows as more people run more models. Synthetic auto-tune sweeps are kept separate from measurements taken during real use, and every figure is published with its sample size, however small.
One command detects your GPU, provisions a matching engine, and benchmarks each model on your card as it loads. See Install & first run, the model hub, or Quantization explained.