MCP & Tool Calling

Rapid-MLX

raullenchai/rapid-mlx

The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.

★ 3.5KStars
⑂ 401Forks
28Open issues
PythonLanguage
NOASSERTIONLicense
Q@project.QualityScoreEditorial score

Overview

The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.

Requirements, installation and quick start

Installation requirements are not stated in the repository metadata.

Usage

Usage is not stated in the repository metadata.

Model compatibility and use cases

Model compatibility is not stated in the repository metadata.

License and risk notes

NOASSERTION

ToolAI metadata-only listing: repository metadata is public; editorial content and manual verification are still pending.

MemPalace

mempalace/mempalace

★ 58KPython

litellm

berriai/litellm

★ 56.7KPython

QwenPaw

agentscope-ai/qwenpaw

★ 33.9KPython