Skip to main content
buildradar
Sign in
Topic · video-generation

video-generation

Tracked open-source repos tagged video-generation, sorted by stars.

154 repos
  • FancyVideo@360CVGroup

    Video generation from text&image, 1st-gen

    919-2Star change over the last 7 days
  • FollowYourClick@mayuelala

    [AAAI 2025] Follow-Your-Click: This repo is the official implementation of "Follow-Your-Click: Open-domain Regional Image Animation via Short Prompts"

    908-1Star change over the last 7 days
  • Vista@OpenDriveLab

    [NeurIPS 2024] A Generalizable World Model for Autonomous Driving

    896+1Star change over the last 7 days
  • ditto-talkinghead@antgroup

    [ACM MM 2025] Ditto: Motion-Space Diffusion for Controllable Realtime Talking Head Synthesis

    881+10Star change over the last 7 days
  • SeedVR2@IceClear

    [ICLR2026] SeedVR2: One-Step Video Restoration via Diffusion Adversarial Post-Training

    867+13Star change over the last 7 days
  • ConsisID@PKU-YuanGroup

    [CVPR 2025 Highlight🔥] Identity-Preserving Text-to-Video Generation by Frequency Decomposition

    857+2Star change over the last 7 days
  • UniAnimate-DiT@ali-vilab

    UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer

    850-2Star change over the last 7 days
  • flatkey-cli@flatkey-ai

    Flatkey media generation CLI for images, videos, audio, text, credits, and model discovery.

    845+211Star change over the last 7 days
  • DiffusionAsShader@IGL-HKUST

    [SIGGRAPH 2025] Diffusion as Shader: 3D-aware Video Diffusion for Versatile Video Generation Control

    834+3Star change over the last 7 days
  • Official implementation for "RIFLEx: A Free Lunch for Length Extrapolation in Video Diffusion Transformers" (ICML 2025) , UltraViCo (ICLR 2026) and UltraImage

    827+2Star change over the last 7 days
  • lmms-engine@EvolvingLMMs-Lab

    A simple, unified multimodal models training engine. Lean, flexible, and built for hacking at scale.

    825+1Star change over the last 7 days
  • kandinsky-5@kandinskylab

    Kandinsky 5.0: A family of diffusion models for Video & Image generation

    822+6Star change over the last 7 days
  • Text-To-Video-AI@SamurAIGPT

    Generate video from text using AI

    821+7Star change over the last 7 days
  • Matrix-3D@SkyworkAI

    Generate large-scale explorable 3D scenes with high-quality panorama videos from a single image or text prompt.

    816+12Star change over the last 7 days
  • [TMLR 2025🔥] A survey for the autoregressive models in vision.

    808+1Star change over the last 7 days
  • open-webui-tools@Haervwe

    Open‑WebUI Tools is a modular toolkit designed to extend and enrich your Open WebUI instance, turning it into a powerful AI workstation. With a suite of over 15 specialized tools, function pipelines, and filters, this project supports academic research, agentic autonomy, multimodal creativity, workflows, and more

    807+6Star change over the last 7 days
  • DriveAGI@OpenDriveLab

    [CVPR 2024 Highlight] GenAD: Generalized Predictive Model for Autonomous Driving

    805+0Star change over the last 7 days
  • rcm@NVlabs

    rCM & Causal-rCM: Leading and Unified Algorithms/Infrastructures for Bidirectional/Autoregressive Video Diffusion Distillation at Scale

    800+6Star change over the last 7 days
  • InfinityStar@FoundationVision

    [NeurIPS 2025 Oral]Infinity⭐️: Unified Spacetime AutoRegressive Modeling for Visual Generation

    784+5Star change over the last 7 days
  • A collection of awesome video generation studies.

    783+1Star change over the last 7 days
  • MagicDance@Boese0601

    [ICML 2024] MagicPose(also known as MagicDance): Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion

    777+1Star change over the last 7 days
  • ima2-gen@lidge-jun

    Local-first visual generation runtime and studio for people and coding agents, with reproducible image and video workflows across multiple providers.

    751+32Star change over the last 7 days
  • A Survey on Text-to-Video Generation/Synthesis.

    744+3Star change over the last 7 days
  • cosmos-transfer2.5@nvidia-cosmos

    Cosmos-Transfer2.5, built on top of Cosmos-Predict2.5, produces high-quality world simulations conditioned on multiple spatial control inputs.

    736+7Star change over the last 7 days
  • MagicDrive-V2@flymin

    [ICCV 2025] Official implementation of the paper “MagicDrive-V2: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control”

    729+1Star change over the last 7 days
  • nano-world-model@simchowitzlabpublic

    A Minimalist, Batteries-included Repository for Advancing World Model Science.

    729+10Star change over the last 7 days
  • comfyui-mcp@artokun

    Local-first, agent-native control plane for ComfyUI — MCP server + sidebar agent that generates images, video & audio, authors and runs workflows, and edits your live graph in natural language on ANY LLM (Claude, ChatGPT, Gemini, offline Ollama, or any hosted model). 178 tools, 36 AI skills, 55 installer packs. Local, LAN, VPS, or Comfy Cloud.

    729+35Star change over the last 7 days
  • [ICML 2025] Official PyTorch Implementation of "History-Guided Video Diffusion"

    709+1Star change over the last 7 days
  • ChronoEdit@nv-tlabs

    [ICLR 2026] ChronoEdit: Towards Temporal Reasoning for Image Editing and World Simulation

    706+1Star change over the last 7 days
  • A Curated List of Awesome Video World Models with AR Diffusion: Covering Algorithms, Applications, and Infrastructure, Aimed at Serving as a Comprehensive Resource for Researchers, Practitioners, and Enthusiasts.

    702+6Star change over the last 7 days
← Back to topics