Skip to main content
buildradar
Sign in
Topic · inference-acceleration

inference-acceleration

Tracked open-source repos tagged inference-acceleration, sorted by stars.

Repos
4
Total stars
9,747
Avg. stars
2,437
Share
0.00%

Topics that frequently appear alongside inference-acceleration on the same repo.

Recent risers

Repos created in the last 90 days, tagged inference-acceleration.

No new repos tagged with this topic in the last 90 days.

  • SageAttention@thu-ml

    [ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to FlashAttention, without losing end-to-end metrics across language, image, and video models.

    3,699+14Star change over the last 7 days
  • TurboDiffusion@thu-ml

    TurboDiffusion: 100–200× Acceleration for Video Diffusion Models

    3,634+6Star change over the last 7 days
  • TeaCache@ali-vilab

    Timestep Embedding Tells: It's Time to Cache for Video Diffusion Model

    1,371+2Star change over the last 7 days
  • SpargeAttn@thu-ml

    [ICML2025] SpargeAttention: A training-free sparse attention that accelerates any model inference.

    1,045+1Star change over the last 7 days
← Back to topics