llamaperf

Inco Splash

An inference engine for running open-weight LLMs locally.

1 community report

This engine doesn't yet have an editorial profile on llamaperf. The community reports below show how it's been used in practice across different hardware.

Top GPUs running Inco Splash

GPUVRAMReportsFastest t/s
M5 Max 36GBapple36GB1144.0

Top models on Inco Splash