llamaperf

lemonade-sdk/lemonade

Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkMu8Zk

View on GitHub

Stars
5,835
Gained, 7 days
+36
Gained, 30 days
n/a
Reddit posts
3

C++ · Apache-2.0 · last push 2026-10-07 · created 2025-05-15

GitHub stars, last 90 days

2026-09-26: 5,7842026-10-07: 5,835

Does it work with local models?

Reports from people who ran it with a model on their own hardware. Tool calling depends on the model as much as on the server, so say which model and quant you used.

No reports yet.

Add your report

Sign in to say whether it worked for you.

Reddit posts linking it

In the official MCP registry

No registry entry points at this repository, so how it runs is not listed here. Check its README.