# mlx-serve

OpenAI- and Anthropic-compatible local inference for Apple Silicon — MLX and GGUF — faster than LM Studio on identical MLX weights. No Python. No cloud. No Electron.

- Tier: Indexed (plain plugin)
- Category: Deployment
- Page: https://www.flowy.sh/listings/ddalcu-mlx-serve
- Source: https://github.com/ddalcu/mlx-serve
- Price: free and open source

## Summary
mlx-serve is a Claude Code plugin with 2 hand-picked skills for deployment work, indexed on Flowy. Install it with the command on its page. It includes bench, release. Its skills do not fire on their own yet. Request auto-invocation to have Flowy route them as you prompt. Free and open source.

## Install (Claude Code)
```
npx -y skills add ddalcu/mlx-serve --agent claude-code
```

## Skills
- bench
- release

## FAQ

### What is mlx-serve?
OpenAI- and Anthropic-compatible local inference for Apple Silicon — MLX and GGUF — faster than LM Studio on identical MLX weights. No Python. No cloud. No Electron.

### How do I install mlx-serve?
Run these in Claude Code: npx -y skills add ddalcu/mlx-serve --agent claude-code. Then prompt normally.

### Does mlx-serve auto-invoke its skills?
Not yet. It is indexed on Flowy as a plain plugin. Request auto-invocation on its page and Flowy will route its skills for you as you prompt.

### Is mlx-serve free?
Yes. Flowy is free and open source, with nothing gated. You can read every skill in full before you install.
