Skip to main content
buildradar
Sign in
Topic · multimodal-deep-learning

multimodal-deep-learning

Tracked open-source repos tagged multimodal-deep-learning, sorted by stars.

Repos
13
Total stars
34,819
Avg. stars
2,678
Share
0.00%

Topics that frequently appear alongside multimodal-deep-learning on the same repo.

Recent risers

Repos created in the last 90 days, tagged multimodal-deep-learning.

No new repos tagged with this topic in the last 90 days.

  • LAVIS@salesforce

    LAVIS - A One-stop Library for Language-Vision Intelligence

    11,261-1Star change over the last 7 days
  • FinRobot@AI4Finance-Foundation

    FinRobot: An Open-Source AI Agent Platform for Financial Applications using Large Language Models

    7,902+33Star change over the last 7 days
  • Time-LLM@KimMeen

    [ICLR 2024] Official implementation of " 🦙 Time-LLM: Time Series Forecasting by Reprogramming Large Language Models"

    2,688+3Star change over the last 7 days
  • (ෆ`꒳´ෆ) A Survey on Text-to-Image Generation/Synthesis.

    2,444+0Star change over the last 7 days
  • BitNet@kyegomez

    Implementation of "BitNet: Scaling 1-bit Transformers for Large Language Models" in pytorch

    1,946+1Star change over the last 7 days
  • AdvancedLiterateMachinery@AlibabaResearch

    A collection of original, innovative ideas and algorithms towards Advanced Literate Machinery. This project is maintained by the OCR Team in the Language Technology Lab, Tongyi Lab, Alibaba Group.

    1,835+1Star change over the last 7 days
  • pytorch-widedeep@jrzaurin

    A flexible package for multimodal-deep-learning to combine tabular data with text and images using Wide and Deep models in Pytorch

    1,416+0Star change over the last 7 days
  • 收集 CVPR 最新的成果,包括论文、代码和demo视频等,欢迎大家推荐!Collect the latest CVPR (Conference on Computer Vision and Pattern Recognition) results, including papers, code, and demo videos, etc., and welcome recommendations from everyone!

    1,411+0Star change over the last 7 days
  • awesome grounding: A curated list of research papers in visual grounding

    1,127+0Star change over the last 7 days
  • A collection of resources on applications of multi-modal learning in medical imaging.

    975+1Star change over the last 7 days
  • blended-latent-diffusion@omriav

    Official implementation for "Blended Latent Diffusion" [SIGGRAPH 2023]

    632+0Star change over the last 7 days
  • MMMU@MMMU-Benchmark

    This repo contains evaluation code for the paper "MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI"

    593+1Star change over the last 7 days
  • VQASynth@remyxai

    Compose multimodal datasets 🎹

    589+1Star change over the last 7 days
← Back to topics