diffusion-models
Tracked open-source repos tagged diffusion-models, sorted by stars.
- #91
[ICLR 2026] A Training-free Iterative Framework for Long Story Visualization
★ 957+0Star change over the last 7 days - #92
[ECCV 2026] Skyfall-GS: Synthesizing Immersive 3D Urban Scenes from Satellite Imagery
★ 955+4Star change over the last 7 days - #93
one summary of diffusion-based image processing, including restoration, enhancement, coding, quality assessment
★ 954-1Star change over the last 7 days - #94
[ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation" & Causal Forcing++
★ 950+10Star change over the last 7 days - #95
Official PyTorch implementation of ODISE: Open-Vocabulary Panoptic Segmentation with Text-to-Image Diffusion Models [CVPR 2023 Highlight]
★ 943+0Star change over the last 7 days - #96
MindSpore + 🤗Huggingface: Run any Transformers/Diffusers model on MindSpore with seamless compatibility and acceleration.
★ 920+0Star change over the last 7 days - #97
Summary of key papers and blogs about diffusion models to learn about the topic. Detailed list of all published diffusion robotics papers.
★ 920+1Star change over the last 7 days - #98
App showcasing multiple real-time diffusion models pipelines with Diffusers
★ 916-1Star change over the last 7 days - #99★ 874+1Star change over the last 7 days
- #100
Official implementation of paper "MiniGPT-5: Interleaved Vision-and-Language Generation via Generative Vokens"
★ 869+0Star change over the last 7 days - #101
Unified Codebase for Advanced World Models.
★ 866+2Star change over the last 7 days - #102
Baking Gaussian Splatting into Diffusion Denoiser for Fast and Scalable Single-stage Image-to-3D Generation and Reconstruction (ICCV 2025)
★ 861+1Star change over the last 7 days - #103
[CVPR 2025 Highlight🔥] Identity-Preserving Text-to-Video Generation by Frequency Decomposition
★ 857+1Star change over the last 7 days - #104
Instruct-NeRF2NeRF: Editing 3D Scenes with Instructions (ICCV 2023)
★ 852-1Star change over the last 7 days - #105
List of papers related to neural network quantization in recent AI conferences and journals.
★ 851+6Star change over the last 7 days - #106★ 838+0Star change over the last 7 days
- #107
Official implementation of "MeshDiffusion: Score-based Generative 3D Mesh Modeling" (ICLR 2023 Spotlight)
★ 832+0Star change over the last 7 days - #108
[CVPR 2024] GaussianDreamer: Fast Generation from Text to 3D Gaussians by Bridging 2D and 3D Diffusion Models
★ 829-1Star change over the last 7 days - #109
Official implementation for "RIFLEx: A Free Lunch for Length Extrapolation in Video Diffusion Transformers" (ICML 2025) , UltraViCo (ICLR 2026) and UltraImage
★ 827+2Star change over the last 7 days - #110
A curated list of 3D Vision papers relating to Robotics domain in the era of large models i.e. LLMs/VLMs, inspired by awesome-computer-vision, including papers, codes, and related websites
★ 821+1Star change over the last 7 days - #111
[ICLR 2024] Controlling Vision-Language Models for Universal Image Restoration. 5th place in the NTIRE 2024 Restore Any Image Model in the Wild Challenge.
★ 817+1Star change over the last 7 days - #112
[CVPR 2024] Paint3D: Paint Anything 3D with Lighting-Less Texture Diffusion Models, a no lighting baked texture generative model
★ 808-1Star change over the last 7 days - #113
DeepInverse: a PyTorch library for solving imaging inverse problems using deep learning
★ 803+5Star change over the last 7 days - #114
Apply diffusion models using the new Hugging Face diffusers package to synthesize music instead of images.
★ 795+1Star change over the last 7 days - #115
Simple and readable code for training and sampling from diffusion models
★ 785+2Star change over the last 7 days - #116
A collection of awesome video generation studies.
★ 782+0Star change over the last 7 days - #117
Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities. ACM Computing Surveys, 2026.
★ 779+1Star change over the last 7 days - #118
[ICML 2024] MagicPose(also known as MagicDance): Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion
★ 776+0Star change over the last 7 days - #119★ 775+0Star change over the last 7 days
- #120
Official Implementation for "Attend-and-Excite: Attention-Based Semantic Guidance for Text-to-Image Diffusion Models" (SIGGRAPH 2023)
★ 770+0Star change over the last 7 days