Topic · inference-acceleration
inference-acceleration
Tracked open-source repos tagged inference-acceleration, sorted by stars.
Repos
4
Total stars
9,747
Avg. stars
2,437
Share
0.00%
Related topics
Topics that frequently appear alongside inference-acceleration on the same repo.
Recent risers
Repos created in the last 90 days, tagged inference-acceleration.
No new repos tagged with this topic in the last 90 days.
- #1
[ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to FlashAttention, without losing end-to-end metrics across language, image, and video models.
★ 3,699+14Star change over the last 7 days - #2
TurboDiffusion: 100–200× Acceleration for Video Diffusion Models
★ 3,634+6Star change over the last 7 days - #3★ 1,371+2Star change over the last 7 days
- #4
[ICML2025] SpargeAttention: A training-free sparse attention that accelerates any model inference.
★ 1,045+1Star change over the last 7 days