Back to Tools Directory

Rapid-MLX

Python

The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.

Categories & Topics

apple-siliconclaude-codecursordeepseekfastapihacktoberfestinferencellmlocal-llmm1m2m3macosmlxollama-alternativeopenai-apipythonqwentool-calling

Repository Stats

3.4k
GitHub Stars
389
Forks
3.4k
Watchers
72
Open Issues
License
Apache License 2.0
Last Updated
Today
August 7, 2026

Related Tools