lemonade-sdk/lemonade
Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkMu8Zk
About the project
A local AI server that runs optimized LLMs, whisper, TTS, and image generation on the user's GPU and NPU, exposing OpenAI-, Anthropic-, and Ollama-compatible APIs. Available as a service or an embeddable binary for apps.
Useful for
- Run a local LLM server with an OpenAI-compatible API for your apps
- Download and run a GGUF model via CLI using lemonade pull and lemonade run
- Connect local models to third-party apps through OpenAI, Anthropic, or Ollama APIs
README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.
Why it’s trending
- 1 star so far today, about 4 expected by the end of the day.
- GitHub Trending C++ today: #20, +10 stars.
- Recent forks include notable developers: @kou (633 followers).
Stars per day
Bars are daily stars, the line is the usual pace. Red marks spike days.
Numbers
- Total stars
- 5,809
- Today
- 1 · ≈ 4 by evening
- Forks
- 513
- Issues and pull requests
- 3,646
- Watchers
- 34
- Language
- C++
- License
- Apache-2.0
- Latest release
- v2026.40.0 · September 30, 2026
- Created
- May 15, 2025
- Last push
- October 1, 2026
Star trust
Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.
These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.
Spotted in
- October 1, 2026GitHub Trending C++ today: #20, +10 stars
Similar by description
-
0.9
AtomicBot-ai/Atomic-Chat
Desktop app and local inference engine for running open-weight LLMs (Llama, Gemma, Qwen, Mistral, etc.) on your own machine, exposing an OpenAI-compatible API at localhost:1337/v1.
-
2.7
ollama/ollama
Run open language models (Kimi, GLM, DeepSeek, Qwen, Gemma and more) locally via CLI and REST API with integrations into editors and assistants.