Installs into .claude/skills of the current project.
Are you the author of Run Local Tool Calling Agent Inference With Rapid Mlx?
Add the live security badge to your README. It updates with every re-scan.
[](https://www.skillsdirectory.com/skills/agentskillexchange-run-local-tool-calling-agent-inference-with-rapid)
---
name: "Run local tool-calling agent inference with Rapid-MLX"
slug: "run-local-tool-calling-agent-inference-with-rapid-mlx"
description: "Serve OpenAI- and Anthropic-compatible local LLM endpoints on Apple Silicon so coding agents can run tool-calling workflows against on-device models."
github_stars: 3860
verification: "security_reviewed"
source: "https://github.com/raullenchai/Rapid-MLX"
author: "raullenchai"
publisher_type: "individual"
category: "Developer Tools"
framework: "Multi-Framework"
tool_ecosystem:
github_repo: "raullenchai/Rapid-MLX"
github_stars: 3860
---
# Run local tool-calling agent inference with Rapid-MLX
Serve OpenAI- and Anthropic-compatible local LLM endpoints on Apple Silicon so coding agents can run tool-calling workflows against on-device models.
## Prerequisites
Apple Silicon Mac, Rapid-MLX, MLX-compatible model, OpenAI- or Anthropic-compatible agent client such as Codex CLI, Claude Code, Aider, Cursor, or Continue.dev
## Installation
Install or set up from the source-backed instructions:
Install Rapid-MLX using the documented installer or package path, start the server with rapid-mlx serve, confirm the OpenAI-compatible endpoint at localhost:8000/v1, then configure the target agent client to use that local base URL and model alias.
- Source: https://github.com/raullenchai/Rapid-MLX
## Documentation
- https://rapidmlx.com/docs/
## Source
- [Agent Skill Exchange](https://agentskillexchange.com/skills/run-local-tool-calling-agent-inference-with-rapid-mlx/)