video-generation
Tracked open-source repos tagged video-generation, sorted by stars.
- #61
[TPAMI 2025🔥] MagicTime: Time-lapse Video Generation Models as Metamorphic Simulators
★ 1,337-1Star change over the last 7 days - #62
A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.
★ 1,334+3Star change over the last 7 days - #63
Creates short videos for TikTok, Instagram Reels, and YouTube Shorts using the Model Context Protocol (MCP) and a REST API.
★ 1,327+9Star change over the last 7 days - #64
Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer.
★ 1,300+10Star change over the last 7 days - #65
Code for Motion Representations for Articulated Animation paper
★ 1,277+0Star change over the last 7 days - #66★ 1,260+11Star change over the last 7 days
- #67
We present StableAvatar, the first end-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven avatar videos without any post-processing, conditioned on a reference image and audio.
★ 1,258-1Star change over the last 7 days - #68
[SIGGRAPH Asia 2026] 4DAnyone: Create Anyone in 4D from a Casual Monocular Video
★ 1,251+327Star change over the last 7 days - #69
半调纸拼贴 B-roll 生成 skill:三闸门审批,Gemini Omni Flash 首尾帧组装动画 | Editorial halftone paper-collage B-roll agent skill
★ 1,238+19Star change over the last 7 days - #70
HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
★ 1,228-1Star change over the last 7 days - #71
Free open-source project designed for turning youtube-viedos into viral short videos. Highlight detection, subtitles, translation, voiceover, all in one for your content.
★ 1,219—Star change over the last 7 days - #72
Code for SCIS-2025 Paper "UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation".
★ 1,189-1Star change over the last 7 days - #73
Wunjo CE: Face Swap, Lip Sync, Control Remove Objects & Text & Background, Restyling, Audio Separator, Clone Voice, Video Generation. Open Source, Local & Free.
★ 1,170+1Star change over the last 7 days - #74
[ICLR24] Official implementation of the paper “MagicDrive: Street View Generation with Diverse 3D Geometry Control”
★ 1,167+1Star change over the last 7 days - #75
Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Includes MLX Core macOS app with chat, agent mode, and tool calling.
★ 1,115+245Star change over the last 7 days - #76★ 1,110+5Star change over the last 7 days
- #77
[ECCV 2024 Oral] MotionDirector: Motion Customization of Text-to-Video Diffusion Models.
★ 1,054+3Star change over the last 7 days - #78
SCAIL: Towards Studio-Grade Character Animation via In-Context Learning of 3D-Consistent Pose Representations (CVPR 2026 Findings)
★ 1,049+4Star change over the last 7 days - #79
[ICML2025] SpargeAttention: A training-free sparse attention that accelerates any model inference.
★ 1,046+4Star change over the last 7 days - #80
[AAAI 2026] EchoMimicV3: 1.3B Parameters are All You Need for Unified Multi-Modal and Multi-Task Human Animation
★ 1,040+10Star change over the last 7 days - #81
660+ muapi-hosted generative-media models plus community-submitted third-party API tools (SEO, enrichment, social, scraping) — one YAML file per entry, browsable by capability.
★ 1,005+8Star change over the last 7 days - #82
Provider-neutral Codex Skill for producing verified AI presenter videos from a script and an authorized presenter image.
★ 997+25Star change over the last 7 days - #83
[TPAMI 2026] 3D and 4D World Modeling: A Survey
★ 982+9Star change over the last 7 days - #84
:fire: :fire: :fire: A paper list of some recent Computer Vision(CV) works
★ 973+1Star change over the last 7 days - #85
Fine-Grained Open Domain Image Animation with Motion Guidance
★ 971-1Star change over the last 7 days - #86
[ICLR 2024] SEINE: Short-to-Long Video Diffusion Model for Generative Transition and Prediction
★ 967+0Star change over the last 7 days - #87
A system of bots that collects clips automatically via custom made filters, lets you easily browse these clips, and puts them together into a compilation video ready to be uploaded straight to any social media platform. Full VPS support is provided, along with an accounts system so multiple users can use the bot at once. This bot is split up into three separate programs. The server. The client. The video generator. These programs perform different functions that when combined creates a very powerful system for auto generating compilation videos.
★ 959+3Star change over the last 7 days - #88
[ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation" & Causal Forcing++
★ 953+15Star change over the last 7 days - #89
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
★ 947+9Star change over the last 7 days - #90
Clip chaining for MiniMax H3 in ComfyUI - motion and audio genuinely continue across joins
★ 919+148Star change over the last 7 days