Inco Splash
An inference engine for running open-weight LLMs locally.
1 community report
This engine doesn't yet have an editorial profile on llamaperf. The community reports below show how it's been used in practice across different hardware.
Top GPUs running Inco Splash
| GPU | VRAM | Reports | Fastest t/s |
|---|---|---|---|
| M5 Max 36GBapple | 36GB | 1 | 144.0 |