Open Issues Need Help
View All on GitHub enhancement help wanted optimization
waybarrios/vllm-mlx
1.6K
High-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon. Native MLX, continuous batching, multimodal models, MCP tool calling, and Claude Code support.
Python
#anthropic#anthropic-api#apple-silicon#claude-code#continuous-batching#inference-server#llm#local-llm#macos#mcp#mlx#multimodal-ai#openai#openai-api#openai-compatible#speech-to-text#text-to-speech#tool-calling#vision-language-model#vllm