Aller au contenu principal
buildradar
Se connecter

amitshekhariitbhu/llm-inference-engineering

@amitshekhariitbhu

Learn LLM Inference Engineering step by step - from KV cache, PagedAttention, and continuous batching to vLLM, SGLang, and GPUs.

Étoiles
221
Bifurcations
26
Langage
Markdown
Licence
Apache-2.0
Dernier push
il y a 3 semaines
Markdownllmlarge-language-modelsllm-inferencellmsinference-optimizationllm-engineeringinference-engineering

Aucune intel associée pour le moment

Ce dépôt n'est encore apparu dans aucune des sources que le radar suit. Le collecteur suit un calendrier — revenez une fois qu'il couvre ce dépôt.