MCP
Vllm Mlx
OpenAI and Anthropic compatible server for Apple Silicon. Run LLMs and vision-language models (Llama, Qwen-VL, LLaVA) with continuous batching, MCP tool calling, and multimodal support. Native MLX backend, 400+ tok/s. Works with Claude Code.
Diesen Eintrag beanspruchen
Verbinde dein GitHub-Konto, um zu belegen, dass du diesen Eintrag besitzt oder betreust. Wir prüfen den Repo-Zugriff automatisch — die meisten Publisher sind in Sekunden bestätigt.
1GitHub verbinden
2Anspruch einreichen
3Automatisch verifiziert oder innerhalb von 48 Std. geprüft