oyd@lab:~oyd bench --list --verbose

Benchmarks

How fast local language models run on our hardware: which hardware, which model, which settings and what result. The list grows with every new test.

stdout

> 0 results

> First benchmarks are on the way

> We will publish only what we measure ourselves on our own hardware: the model, the settings and the result. For now, the hardware is ready.

> lsdev --lab ↓

oyd@lab:~lsdev --lab --all

Our hardware

We measure on our own hardware: from a single consumer graphics card to several NVIDIA DGX Spark units and Apple Silicon machines.

qtydevicememstatusnote
4×NVIDIA RTX 309024 GB in use One runs as an eGPU server. We have a 4-slot NVLink bridge.
3×NVIDIA DGX Spark128 GB in use Linked over ConnectX-7. We measure how models scale from 1 to 4 units.
1×Mac Studio M5 Ultra— arriving
1×NVIDIA RTX 2080 Ti11 GB in use We plan a memory upgrade to 22 GB and a two-card 44 GB rig.
1×NVIDIA RTX 306012 GB in use
2×AMD RX 4808 GB in use
2×Mac (Apple Silicon)— in use Two Macs linked over Thunderbolt RDMA.
2×X99 (Intel)— in use
1×X299 (Intel)— arriving

oyd@lab:~cat METHOD.md

How we measure

  1. [1] We measure on our own hardware, with the models and settings listed next to each result.
  2. [2] We publish only what we measured ourselves. We do not repeat vendor numbers.
  3. [3] Where we tuned a setup, we show the result before and after.