• https://github.com/raullenchai/Rapid-MLX
  • Fastest local AI engine for Apple Silicon. Claims 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling.
  • 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.

Connections