NVlabs/Fast-dLLM
@NVlabsOfficial implementation of "Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding"
Stars
1,084
Forks
144
Language
Python
License
Apache-2.0
Last push
3 months ago
Related intel (0)
No related intel yet
This repo has not appeared in any of the sources the radar tracks. The collector runs on a schedule — check back once it covers this repo.