Local LLM Benchmark Database
Every number below was measured on hardware we own, with the exact settings documented on the methodology page. This page is re-run and updated as models and runtimes change — the changelog at the bottom lists every run.
First benchmark runs are in progress. Results are published here once they pass validation and human review — see how we test.