showlab
showlab's tracked open-source repos, sorted by stars.
- #1
A curated list of recent diffusion models for video generation, editing, and various other applications.
★ 5,765+3Star change over the last 7 days - #2
Automatic Video Generation from Scientific Papers
★ 2,373+3Star change over the last 7 days - #3
[ICML 2026] Video generation via code
★ 2,013+4Star change over the last 7 days - #4
[ICLR & NeurIPS 2025] Repository for Show-o series, One Single Transformer to Unify Multimodal Understanding and Generation.
★ 1,975+1Star change over the last 7 days - #5
Out-of-the-box (OOTB) GUI Agent for Windows and macOS
★ 1,958-2Star change over the last 7 days - #6
[CVPR 2025] Open-source, End-to-end, Vision-Language-Action model for GUI Agent & Computer Use.
★ 1,896+2Star change over the last 7 days - #7
💻 A curated list of papers and resources for multi-modal Graphical User Interface (GUI) agents.
★ 1,215+1Star change over the last 7 days - #8
[IJCV] Show-1: Marrying Pixel and Latent Diffusion Models for Text-to-Video Generation
★ 1,148+0Star change over the last 7 days - #9
[ECCV 2024 Oral] MotionDirector: Motion Customization of Text-to-Video Diffusion Models.
★ 1,053+2Star change over the last 7 days - #10
📖 A curated list of resources dedicated to hallucination of multimodal large language models (MLLM).
★ 1,041+3Star change over the last 7 days - #11
📖 This is a repository for organizing papers, codes and other resources related to unified multimodal models.
★ 831+1Star change over the last 7 days - #12
[CVPR 2024] X-Adapter: Adding Universal Compatibility of Plugins for Upgraded Diffusion Model
★ 769-1Star change over the last 7 days - #13
VideoLLM-online: Online Video Large Language Model for Streaming Video (CVPR 2024)
★ 683+0Star change over the last 7 days - #14★ 589+0Star change over the last 7 days
- #15
[ECCV 2024] DragAnything: Motion Control for Anything using Entity Representation
★ 506+0Star change over the last 7 days