waybarrios/
vllm-mlx
waybarrios/vllm-mlxTools
High-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon. Native MLX, continuous batching, multimodal models, MCP tool calling, and Claude Code support.
1.6k
High-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon. Native MLX, continuous batching, multimodal models, MCP tool calling, and Claude Code support.

vMLX - Use MLX models easily - JANGQ (GGUF for MLX) - Not dependant on mlx_vlm