Saltar al contenido principal
buildradar
Iniciar sesión

amitshekhariitbhu/llm-inference-engineering

@amitshekhariitbhu

Learn LLM Inference Engineering step by step - from KV cache, PagedAttention, and continuous batching to vLLM, SGLang, and GPUs.

Estrellas
221
Bifurcaciones
26
Lenguaje
Markdown
Licencia
Apache-2.0
Último push
hace 4 semanas
Markdownllmlarge-language-modelsllm-inferencellmsinference-optimizationllm-engineeringinference-engineering

Aún no hay intel relacionado

Este repo no ha aparecido en ninguna de las fuentes que rastrea el radar. El recopilador funciona con un horario — vuelve cuando lo haya cubierto.