diffusion-models
Tracked open-source repos tagged diffusion-models, sorted by stars.
- #61
Scalable and memory-optimized training of diffusion models
★ 1,359+0Star change over the last 7 days - #62
[ICASSP 2024] 🍵 Matcha-TTS: A fast TTS architecture with conditional flow matching
★ 1,353+3Star change over the last 7 days - #63★ 1,351+2Star change over the last 7 days
- #64
FireRed-Image-Edit is a powerful image editing foundation model achieving open-source state-of-the-art performance with precise instruction following, high-fidelity generation, superior identity consistency, and seamless multi-element fusion.
★ 1,350+3Star change over the last 7 days - #65
[AAAI 2025]👔IMAGDressing👔: Interactive Modular Apparel Generation for Virtual Dressing. It enables customizable human image generation with flexible garment, pose, and scene control, ensuring high fidelity and garment consistency for virtual dressing.
★ 1,343+0Star change over the last 7 days - #66
[TPAMI 2025🔥] MagicTime: Time-lapse Video Generation Models as Metamorphic Simulators
★ 1,337-1Star change over the last 7 days - #67
📦Portable package for running Hunyuan3D 2.0/2.1 on Windows. | 混元 3D 2.0/2.1 整合包
★ 1,326+4Star change over the last 7 days - #68
Pruna is a model optimization framework built for developers, enabling you to deliver faster, more efficient models with minimal overhead.
★ 1,274+1Star change over the last 7 days - #69
[CVPR 2025 Oral & Best Paper Finalist] Difix3D+: Improving 3D Reconstructions with Single-Step Diffusion Models
★ 1,267+1Star change over the last 7 days - #70
[CVPR2024, Highlight] Official code for DragDiffusion
★ 1,259+0Star change over the last 7 days - #71★ 1,238-1Star change over the last 7 days
- #72
A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control
★ 1,228+4Star change over the last 7 days - #73
HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
★ 1,228-1Star change over the last 7 days - #74★ 1,196-1Star change over the last 7 days
- #75
Wunjo CE: Face Swap, Lip Sync, Control Remove Objects & Text & Background, Restyling, Audio Separator, Clone Voice, Video Generation. Open Source, Local & Free.
★ 1,170+1Star change over the last 7 days - #76
[ICLR24] Official implementation of the paper “MagicDrive: Street View Generation with Diverse 3D Geometry Control”
★ 1,167+1Star change over the last 7 days - #77
A collection of resources on controllable generation with text-to-image diffusion models.
★ 1,110+0Star change over the last 7 days - #78★ 1,108+3Star change over the last 7 days
- #79
Stable Diffusion implemented from scratch in PyTorch
★ 1,078+0Star change over the last 7 days - #80
[ICLR'24 spotlight] Chinese and English Multimodal Large Model Series (Chat and Paint) | 基于CPM基础模型的中英双语多模态大模型系列
★ 1,061-1Star change over the last 7 days - #81
[ECCV 2024 Oral] MotionDirector: Motion Customization of Text-to-Video Diffusion Models.
★ 1,053+2Star change over the last 7 days - #82
[ICLR 2025 Oral] The official implementation of "Diffusion-Based Planning for Autonomous Driving with Flexible Guidance"
★ 1,050+4Star change over the last 7 days - #83
[CVPR 2025 Highlight] 3DTopia-XL: High-Quality 3D PBR Asset Generation via Primitive Diffusion
★ 1,046+0Star change over the last 7 days - #84
[ICLR 2024 Spotlight] SyncDreamer: Generating Multiview-consistent Images from a Single-view Image
★ 1,044-1Star change over the last 7 days - #85
Official implementation of the paper: "FlowEdit: Inversion-Free Text-Based Editing Using Pre-Trained Flow Models"
★ 1,013-1Star change over the last 7 days - #86★ 986+22Star change over the last 7 days
- #87
[CVPR 2024] PIA, your Personalized Image Animator. Animate your images by text prompt, combing with Dreambooth, achieving stunning videos. PIA,你的个性化图像动画生成器,利用文本提示将图像变为奇妙的动画
★ 975+0Star change over the last 7 days - #88★ 970-1Star change over the last 7 days
- #89★ 969+0Star change over the last 7 days
- #90
[ICLR 2026] A Training-free Iterative Framework for Long Story Visualization
★ 959+0Star change over the last 7 days