Rapid-MLX

The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.

Recent spike, 1.7× faster than LLM Infrastructure this week

7d +1.0% +39 stars in 7d
30d Not tracked yet — first tracked 2026-09-02
90d Not tracked yet — first tracked 2026-09-02
Sep 2, 2026: 3,647 starsSep 3, 2026: 3,647 starsSep 5, 2026: 3,664 starsSep 6, 2026: 3,669 starsSep 7, 2026: 3,680 starsSep 8, 2026: 3,700 starsSep 9, 2026: 3,709 starsSep 10, 2026: 3,720 starsSep 11, 2026: 3,727 starsSep 12, 2026: 3,729 starsSep 13, 2026: 3,736 starsSep 14, 2026: 3,743 starsSep 15, 2026: 3,749 starsSep 16, 2026: 3,759 starsSep 17, 2026: 3,767 starsSep 18, 2026: 3,775 starsSep 19, 2026: 3,791 starsSep 20, 2026: 3,796 starsSep 21, 2026: 3,802 starsSep 22, 2026: 3,810 starsSep 23, 2026: 3,821 starsSep 24, 2026: 3,829 starsSep 25, 2026: 3,841 starsSep 26, 2026: 3,845 starsSep 27, 2026: 3,849 starsSep 28, 2026: 3,856 starsSep 29, 2026: 3,862 starsSep 30, 2026: 3,866 starsOct 1, 2026: 3,868 stars 3,6473,7583,868 Sep 2Oct 1

3,647 → 3,868 stars since Sep 2

3,645 Stars 424 Forks 114 Open issues
apple-siliconclaude-codecursordeepseekfastapihacktoberfestinferencellmlocal-llmm1m2m3macosmlxollama-alternativeopenai-apipythonqwentool-calling