Zum Hauptinhalt springen
buildradar
Sign in

vllm-project/vllm

@vllm-project

Ein hocheffizientes und speichereffizientes Inferenz- und Serving-Engine für LLMs

Sterne
90.817
Forks
21.619
Sprache
Python
Lizenz
Apache-2.0
Letzter Push
vor 5 Tagen
Pythonopenaillmcudagpttransformerpytorchdeepseekqwenllamainferencekimiqwen3tpullm-servingmodel-servingmoedeepseek-v3amdgpt-ossblackwell