Skip to content

Add Rapid-MLX to Inference engines - #11

Open
raullenchai wants to merge 1 commit into
0xSojalSec:mainfrom
raullenchai:add-rapid-mlx
Open

raullenchai wants to merge 1 commit into
0xSojalSec:mainfrom
raullenchai:add-rapid-mlx

Conversation

@raullenchai

Copy link
Copy Markdown

Adds Rapid-MLX to the Inference engines section, next to mlx-lm.

Rapid-MLX is an OpenAI-compatible LLM inference server for Apple Silicon, built on MLX (Apache-2.0). It serves 220+ MLX-native models behind the standard /v1/chat/completions API — so local models drop straight into tools that expect OpenAI semantics (Claude Code, Cursor, Codex, Aider, Continue). One-line install, one command to serve.

Placed alongside mlx-lm since both target Apple Silicon via MLX; Rapid-MLX adds the OpenAI-compatible serving layer on top.

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant