Skip to main content
buildradar
Sign in
Topic · text-to-image

text-to-image

Tracked open-source repos tagged text-to-image, sorted by stars.

62 repos
  • WebAI2API@foxhui

    WebAI2API: 基于 Camoufox 的网页 AI 转 API 工具,支持 LMArena/Gemini等,多窗口并发与账号隔离。 | Web AI to OpenAI API via Camoufox. Supports LMArena/Gemini and more, multi-window concurrency & account isolation.

    1,315+0Star change over the last 7 days
  • airunner@Capsize-Games

    Offline inference engine for art, real-time voice conversations, LLM powered chatbots and automated workflows

    1,314+0Star change over the last 7 days
  • A collection of resources on controllable generation with text-to-image diffusion models.

    1,110+0Star change over the last 7 days
  • CogView4@zai-org

    CogView4, CogView3-Plus and CogView3(ECCV 2024)

    1,101+0Star change over the last 7 days
  • awesome-agent-apis@Anil-matcha

    660+ muapi-hosted generative-media models plus community-submitted third-party API tools (SEO, enrichment, social, scraping) — one YAML file per entry, browsable by capability.

    1,005+8Star change over the last 7 days
  • muse-maskgit-pytorch@lucidrains

    Implementation of Muse: Text-to-Image Generation via Masked Generative Transformers, in Pytorch

    918+0Star change over the last 7 days
  • unfake.js@jenissimo

    Fix AI pixel art and vector images right in your browser

    916+6Star change over the last 7 days
  • Lora-for-Diffusers@haofanwang

    The most easy-to-understand tutorial for using LoRA (Low-Rank Adaptation) within diffusers framework for AI Generation Researchers🔥

    822+0Star change over the last 7 days
  • TF-ICON@Shilin-LU

    [ICCV 2023] "TF-ICON: Diffusion-Based Training-Free Cross-Domain Image Composition" (Official Implementation)

    814+0Star change over the last 7 days
  • [TMLR 2025🔥] A survey for the autoregressive models in vision.

    808+0Star change over the last 7 days
  • aphantasia@eps696

    CLIP + FFT/DWT/RGB = text to image/video

    789+0Star change over the last 7 days
  • Attend-and-Excite@yuval-alaluf

    Official Implementation for "Attend-and-Excite: Attention-Based Semantic Guidance for Text-to-Image Diffusion Models" (SIGGRAPH 2023)

    770+0Star change over the last 7 days
  • A collection of awesome text-to-image generation studies.

    765+0Star change over the last 7 days
  • A Survey on Text-to-Video Generation/Synthesis.

    744+3Star change over the last 7 days
  • PaddleMIX@PaddlePaddle

    Paddle Multimodal Integration and eXploration, supporting mainstream multi-modal tasks, including end-to-end large-scale multi-modal pretrain models and diffusion model toolbox. Equipped with high performance and flexibility.

    723-1Star change over the last 7 days
  • NanoBananaEditor@markfulton

    The most advanced Nano Banana image generator and editor application. Your central hub for AI image generation and revisions. Intuitive UI features reference images, editing with image masks, version history, and more. Powered by Gemini Flash images API.

    710+0Star change over the last 7 days
  • HunyuanImage-2.1@Tencent-Hunyuan

    HunyuanImage-2.1: An Efficient Diffusion Model for High-Resolution (2K) Text-to-Image Generation​

    674-1Star change over the last 7 days
  • OneDiffusion@lehduong

    Official implementation of OneDiffusion paper (CVPR 2025)

    665+0Star change over the last 7 days
  • flash-diffusion@gojasper

    Flash Diffusion — accelerating conditional diffusion models (AAAI 2025 Oral)

    665+2Star change over the last 7 days
  • face-to-sticker

    641+0Star change over the last 7 days
  • Liquid@FoundationVision

    (Accepted by IJCV) Liquid: Language Models are Scalable and Unified Multi-modal Generators

    640+0Star change over the last 7 days
  • blended-latent-diffusion@omriav

    Official implementation for "Blended Latent Diffusion" [SIGGRAPH 2023]

    632+0Star change over the last 7 days
  • MIGC@limuloo

    [CVPR 2024 Highlight] MIGC and [TPAMI 2024] MIGC++ (Official Implementation)

    614+0Star change over the last 7 days
  • 118+ plug-and-play JSON style packs for Nano Banana Pro, GPT Image & Midjourney. Copy one JSON, get a style. Updated daily.

    592+5Star change over the last 7 days
  • semantic-draw@ironjr

    Official code for the CVPR 2025 paper "SemanticDraw: Towards Real-Time Interactive Content Creation from Image Diffusion Models."

    588+0Star change over the last 7 days
  • blended-diffusion@omriav

    Official implementation for "Blended Diffusion for Text-driven Editing of Natural Images" [CVPR 2022]

    588-1Star change over the last 7 days
  • 🔥🔥🔥 A curated list of papers on LLMs-based multimodal generation (image, video, 3D and audio).

    552+0Star change over the last 7 days
  • payload-ai@ashbuilds

    AI Plugin is a powerful extension for the Payload CMS, integrating advanced AI capabilities to enhance content creation and management.

    547-1Star change over the last 7 days
  • UniTok@FoundationVision

    [NeurIPS 2025 Spotlight] A Unified Tokenizer for Visual Generation and Understanding

    529-1Star change over the last 7 days
  • Google-Colab_Notebooks@Isi-dev

    A Collection of Google Colab Notebooks for various projects

    521+1Star change over the last 7 days
← Back to topics