claudeers.

Projects tagged #llm-inference

// llm-inference (2)

🔓

Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkM…

// frameworkslemonade-sdk/⟨C++⟩★ 5,790◷ Apache-2.0[ claude ]
🔓

Run a 105 GB AI model on a Mac that can't hold it. Slotstream streams Qwen3.8-Flash-Next (125B mixture of experts) from your SSD and caches the busiest exper…

// othercarloslfu/⟨Swift⟩★ 402◷ MIT[ claude ]