Skip to main content
buildradar
Sign in
Topic · multi-modal

multi-modal

Tracked open-source repos tagged multi-modal, sorted by stars.

42 repos
  • VisRAG@OpenBMB

    Parsing-free RAG supported by VLMs

    979-1Star change over the last 7 days
  • farmvibes-ai@microsoft

    FarmVibes.AI: Multi-Modal GeoSpatial ML Models for Agriculture and Sustainability

    895+0Star change over the last 7 days
  • LanguageBind@PKU-YuanGroup

    【ICLR 2024🔥】 Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment

    885+0Star change over the last 7 days
  • byaldi@AnswerDotAI

    Use late-interaction multi-modal models such as ColPali in just a few lines of code.

    852+0Star change over the last 7 days
  • OmniLottie@OpenVGLab

    [CVPR 2026🔥] 🧑‍🎨 OmniLottie, an open-sourced multi-modal instructed vector animation generator that produces Lottie JSONs.

    774+3Star change over the last 7 days
  • Versatile-OCR-Program@raphael-seo

    Multi-modal OCR pipeline optimized for ML training (text, figure, math, tables, diagrams)

    676+0Star change over the last 7 days
  • UniControl@salesforce

    Unified Controllable Visual Generation Model

    662+0Star change over the last 7 days
  • Aether@InternRobotics

    [ICCV 2025 & ICCV 2025 RIWM Outstanding Paper] Aether: Geometric-Aware Unified World Modeling

    609+0Star change over the last 7 days
  • RT-2@kyegomez

    Democratization of RT-2 "RT-2: New model translates vision and language into action"

    583-1Star change over the last 7 days
  • fashion-clip@patrickjohncyh

    FashionCLIP is a CLIP-like model fine-tuned for the fashion domain.

    537+3Star change over the last 7 days
  • forge-film@F-R-L

    Multi-model DAG-driven parallel AI film generation — parallel speedup scales with scene independence; Generate film scenes simultaneously instead of one by one; "把影视生成的执行图从拓扑序变成关键路径最优调度" ; 唯一把场景叙事依赖建模为 DAG、以 CPM 算法驱动并行调度的影视生成引擎

    536+0Star change over the last 7 days
  • Robust-R1@jqtangust

    🔥🔥🔥[AAAI 2026 Oral] Official Implementation of Robust-R1: Degradation-Aware Reasoning for Robust Visual Understanding

    511+0Star change over the last 7 days
← Back to topics