The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separa…
Projects tagged #apple-silicon
// apple-silicon (6)
// ragraullenchai/⟨Python⟩★ 3,871◷ Apache-2.0[ claude ]
Run Claude Code 100% on-device with local AI on Apple Silicon. MLX-native Anthropic-API server, 65 tok/s Qwen 3.5 122B, Llama 3.3 70B, Gemma 4 31B. Private,…
// mcp-serversnicedreamzapp/⟨Python⟩★ 3,345◷ MIT[ claude ]
OpenAI and Anthropic compatible server for Apple Silicon. Run LLMs and vision-language models (Llama, Qwen-VL, LLaVA) with continuous batching, MCP tool call…
// ragwaybarrios/⟨Python⟩★ 1,608◷ Apache-2.0[ claude ]
MLX Studio - Home of JANG_Q - Image Gen/Edit + Chat/Code All in one - + OpenClaw (Anthropic API)
// automationjjang-ai/★ 979[ claude ]
Run a 105 GB AI model on a Mac that can't hold it. Slotstream streams Qwen3.8-Flash-Next (125B mixture of experts) from your SSD and caches the busiest exper…
// othercarloslfu/⟨Swift⟩★ 402◷ MIT[ claude ]
Native macOS video editor that coding agents can drive: everything the UI does, Claude Code and Codex can do through the bashcut CLI or MCP server. Layered t…
// mcp-serversdongnguyenvie/⟨Swift⟩★ 30◷ MIT[ claude ]