Skip to main content
buildradar
Sign in

waybarrios/vllm-mlx

@waybarrios

High-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon. Native MLX, continuous batching, multimodal models, MCP tool calling, and Claude Code support.

Stars
1,557
Forks
220
Language
Python
License
Apache-2.0
Last push
1 week ago
Pythonmacosmcpopenaillmtext-to-speechanthropicclaude-codevllmopenai-apilocal-llmapple-siliconmlxspeech-to-textvision-language-modelanthropic-apiopenai-compatibletool-callingmultimodal-aiinference-servercontinuous-batching

No related intel yet

This repo has not appeared in any of the sources the radar tracks. The collector runs on a schedule — check back once it covers this repo.