Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Includes MLX Core macOS app with chat, agent mode, and tool calling.
- ✓Actively maintained (<30d)
- ✓Healthy fork ratio
- ✓Clear description
- ✓Topics declared
- !Licence file present but not machine-readable
git clone https://github.com/ddalcu/mlx-serveTools overview
What people ask about mlx-serve
What is ddalcu/mlx-serve?
+
ddalcu/mlx-serve is tools for the Claude AI ecosystem. Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Includes MLX Core macOS app with chat, agent mode, and tool calling. It has 1.2k GitHub stars and its last recorded update is dated 2026-09-11.
How do I install mlx-serve?
+
You can install mlx-serve by cloning the repository (https://github.com/ddalcu/mlx-serve) or following the README instructions on GitHub. ClaudeWave also provides quick install blocks on this page.
Is ddalcu/mlx-serve safe to use?
+
Our security agent has analyzed ddalcu/mlx-serve and assigned a Trust Score of 82/100 (tier: Trusted). See the full breakdown of passed checks and flags on this page.
Who maintains ddalcu/mlx-serve?
+
ddalcu/mlx-serve is maintained by ddalcu. The last recorded GitHub activity is dated 2026-09-11, with 54 open issues.
Are there alternatives to mlx-serve?
+
Yes. On ClaudeWave you can browse similar tools at /categories/tools, sorted by popularity or recent activity.
Deploy mlx-serve to your cloud
Ship this repo to production in minutes. Each platform spins up its own environment with editable env vars.
Maintain this repo? Add a badge to your README
Drop the badge into your GitHub README to show it's tracked on ClaudeWave. Each badge links back to this page and reflects the live Trust Score.
[](https://claudewave.com/repo/ddalcu-mlx-serve)<a href="https://claudewave.com/repo/ddalcu-mlx-serve"><img src="https://claudewave.com/api/badge/ddalcu-mlx-serve" alt="Featured on ClaudeWave: ddalcu/mlx-serve" width="320" height="64" /></a>More Tools
A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.
An AI skill that provides design intelligence for building professional UI/UX across multiple platforms.
🪨 why use many token when few token do trick — Claude Code skill that cuts 65% of tokens by talking like caveman
CLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies
The fastest, litest AI Gateway. Rust core with Python SDK. Call 100+ LLM APIs in OpenAI (or native) format with cost tracking, guardrails, load balancing, and logging [Bedrock, Azure, OpenAI, Anthropic, OpenAI, VertexAI, vLLM, Nvidia NIM]
Use Claude Code, Codex, Pi, and OpenCode and more for free (1.3B+ free tokens) from your terminal, app, IDE, or phone like OpenClaw (voice supported + ToS friendly)