Skip to content

docs: add Rapid-MLX as a custom AI endpoint - #715

Open
raullenchai wants to merge 2 commits into
LibreChat-AI:mainfrom
raullenchai:add-rapid-mlx-endpoint
Open

raullenchai wants to merge 2 commits into
LibreChat-AI:mainfrom
raullenchai:add-rapid-mlx-endpoint

Conversation

@raullenchai

@raullenchai raullenchai commented Jul 28, 2026 •

Copy link
Copy Markdown

Adds a Rapid-MLX custom endpoint example to the LibreChat docs, alongside other local OpenAI-compatible servers. The example pins port 8000, uses a v0.15.7 model alias, and explains the Docker host address.

Rapid-MLX is an Apache-2.0 LLM inference server for Apple Silicon built on MLX. Install with brew install rapid-mlx or pip install rapid-mlx; its OpenAI-compatible base URL is http://localhost:8000/v1 when started on port 8000.

Rapid-MLX is an Apple-Silicon-native inference server (pure MLX) with an
OpenAI-compatible API, like the existing Apple MLX / vLLM / Ollama entries.
Adds the endpoint page + registers it in meta.json.
@vercel

vercel Bot commented Jul 28, 2026

Copy link
Copy Markdown

@raullenchai is attempting to deploy a commit to the LibreChat's projects Team on Vercel.

A member of the Team first needs to authorize it.

This branch has not been deployed

No deployments
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant