llamaperf
Oct 3, 2026
Throughput
73.4 t/s gen · 1010.0 t/s pp
Quant
Q3-G128-LSQ (MLX)
System RAM
256 GB

Summary

User asks whether oMLX benchmarks can be spoofed, citing a third-party oMLX benchmark entry for deepseek-v4.1-flash-q3-g128-lsq-mlx on an M5 Ultra with 256GB RAM, posted two days earlier, reporting about 73.4 tok/s generation and 1,010 tok/s prompt processing. The run is not the user's own; the figures come from a linked oMLX benchmark page and screenshot.