BrowseFull catalogOutcomesSolve a specific problemRolesStack by teamTrustFilter by risk tier
← Back to the Claude Observatory

lemonade

Skill Development Solid
Works inClaude Code
Solid Scanned — metadata only

Runs inference on local hardware (GPU/NPU) without cloud dependencies, reducing latency, cost, and data residency risk for AI workloads.

Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkMu8Zk

5,396 starsApache-2.0 (commercial OK)FreeQuick setup

Runs inference on local hardware (GPU/NPU) without cloud dependencies, reducing latency, cost, and data residency risk for AI workloads.

Teams needing low-latency LLM inference on-premises or edge devices without external API calls.

Claude Code Claude Cowork Claude Chat

https://github.com/lemonade-sdk/lemonade

By lemonade-sdk

How to Get It

Option 1: Claude Desktop App (Code Mode)Click the + button next to the prompt box → PluginsAdd plugin. Search and click Install. Skills work in Claude Code only.
Option 2: Paste into Claude CodeCopy the command below and paste it into your conversation. Claude will install it.
Command
claude plugins install lemonade-sdk/lemonade

Tip: Paste this into a Claude Code conversation. Verify command matches your Claude Code version.

Auto-generated from the tool's public listing — not hands-on verified. Cross-check against the source repo's README before running.

First thing to try

After installing, paste this into Claude:

Help me run LLMs locally for document analysis without cloud uploads
CostFree

Trust Signals Auto-scanned

Stars5,396Contributors125Last updated2026-08-18LicenseApache-2.0 (OK for commercial use)Known CVEsNone foundSources: GitHub Advisory Database + OSV.dev · Scanned 2026-08-18 · scanner v1

Community Pulse Active

Discussed on Hacker News, Reddit

3 mentions across 2 sources

Reviewer notes

Auto-scanned review. These are observations, not a security certification.

catalog_hygiene stale-eval refresh: Scored from trust signals (evidence-eval-v1): 5,396 GitHub stars; 125 contributors; last commit 0d ago; license Apache-2.0.

2026-05-10: Lemonade is a local inference runner targeting GPU and NPU hardware — useful for teams with data residency constraints, air-gapped environments, or edge deployments where calling out to cloud APIs isn't an option. The 3.6k stars and 73 contributors suggest real traction, but production adoption is thin (one confirmed mention), so treat it as promising-but-unproven for anything critical. Worth benchmarking against Ollama or llama.cpp if you're already in that space, as those have wider production track records and larger model support matrices.

Things to check

  • Scanned, not hands-on tested — this entry was auto-scanned from public metadata (GitHub metrics, license, security flags). No reviewer has run it, and no tool-specific limitations have been documented yet.

How to evaluate tools before deploying →

Data shown here comes from public APIs and automated scanning. Reviewer notes reflect one person's experience. This is not a security certification or legal recommendation. Always evaluate tools according to your own organization's policies.

Evaluation

Ease of Use
4/5
Versatility
5/5
Reliability
5/5
Security
3/5
Overall score4.35 / 5.00 SolidEvaluatedAug 2026
catalog_hygiene stale-eval refresh: Scored from trust signals (evidence-eval-v1): 5,396 GitHub stars; 125 contributors; last commit 0d ago; license Apache-2.0.

← Back to the Claude Observatory

Rolling Claude out in your org? Let's talk.

Start a conversation →