llamaperf

text-generation-webui

An inference engine for running open-weight LLMs locally.

3 community reports

This engine doesn't yet have an editorial profile on llamaperf. The community reports below show how it's been used in practice across different hardware.

Top GPUs running text-generation-webui

GPUVRAMReportsFastest t/s
RTX 3090nvidia24GB212.5
RTX 4060 Ti 16GBnvidia16GB132.5

Top models on text-generation-webui