llamaperf

M2 8GB

APPLE · 8GB unified memory · 1 report

See what fits on this GPU →
This page is thin (1 of 3 reports needed for indexing). Help fill it in.
Tone: positive
throughput:
5.5 t/s gen

Custom Swift/Metal engine. Also tested on M5 MacBook Pro: 31-35 tok/s. OpenAI-compatible server with streaming and tool-call support.