text-to-image
Tracked open-source repos tagged text-to-image, sorted by stars.
- #31
WebAI2API: 基于 Camoufox 的网页 AI 转 API 工具,支持 LMArena/Gemini等,多窗口并发与账号隔离。 | Web AI to OpenAI API via Camoufox. Supports LMArena/Gemini and more, multi-window concurrency & account isolation.
★ 1,315+0Star change over the last 7 days - #32
Offline inference engine for art, real-time voice conversations, LLM powered chatbots and automated workflows
★ 1,314+0Star change over the last 7 days - #33
A collection of resources on controllable generation with text-to-image diffusion models.
★ 1,110+0Star change over the last 7 days - #34★ 1,101+0Star change over the last 7 days
- #35
660+ muapi-hosted generative-media models plus community-submitted third-party API tools (SEO, enrichment, social, scraping) — one YAML file per entry, browsable by capability.
★ 1,005+8Star change over the last 7 days - #36
Implementation of Muse: Text-to-Image Generation via Masked Generative Transformers, in Pytorch
★ 918+0Star change over the last 7 days - #37★ 916+6Star change over the last 7 days
- #38
The most easy-to-understand tutorial for using LoRA (Low-Rank Adaptation) within diffusers framework for AI Generation Researchers🔥
★ 822+0Star change over the last 7 days - #39
[ICCV 2023] "TF-ICON: Diffusion-Based Training-Free Cross-Domain Image Composition" (Official Implementation)
★ 814+0Star change over the last 7 days - #40
[TMLR 2025🔥] A survey for the autoregressive models in vision.
★ 808+0Star change over the last 7 days - #41
CLIP + FFT/DWT/RGB = text to image/video
★ 789+0Star change over the last 7 days - #42
Official Implementation for "Attend-and-Excite: Attention-Based Semantic Guidance for Text-to-Image Diffusion Models" (SIGGRAPH 2023)
★ 770+0Star change over the last 7 days - #43
A collection of awesome text-to-image generation studies.
★ 765+0Star change over the last 7 days - #44
A Survey on Text-to-Video Generation/Synthesis.
★ 744+3Star change over the last 7 days - #45
Paddle Multimodal Integration and eXploration, supporting mainstream multi-modal tasks, including end-to-end large-scale multi-modal pretrain models and diffusion model toolbox. Equipped with high performance and flexibility.
★ 723-1Star change over the last 7 days - #46
The most advanced Nano Banana image generator and editor application. Your central hub for AI image generation and revisions. Intuitive UI features reference images, editing with image masks, version history, and more. Powered by Gemini Flash images API.
★ 710+0Star change over the last 7 days - #47
HunyuanImage-2.1: An Efficient Diffusion Model for High-Resolution (2K) Text-to-Image Generation
★ 674-1Star change over the last 7 days - #48
Official implementation of OneDiffusion paper (CVPR 2025)
★ 665+0Star change over the last 7 days - #49
Flash Diffusion — accelerating conditional diffusion models (AAAI 2025 Oral)
★ 665+2Star change over the last 7 days - #50
face-to-sticker
★ 641+0Star change over the last 7 days - #51
(Accepted by IJCV) Liquid: Language Models are Scalable and Unified Multi-modal Generators
★ 640+0Star change over the last 7 days - #52
Official implementation for "Blended Latent Diffusion" [SIGGRAPH 2023]
★ 632+0Star change over the last 7 days - #53★ 614+0Star change over the last 7 days
- #54
118+ plug-and-play JSON style packs for Nano Banana Pro, GPT Image & Midjourney. Copy one JSON, get a style. Updated daily.
★ 592+5Star change over the last 7 days - #55
Official code for the CVPR 2025 paper "SemanticDraw: Towards Real-Time Interactive Content Creation from Image Diffusion Models."
★ 588+0Star change over the last 7 days - #56
Official implementation for "Blended Diffusion for Text-driven Editing of Natural Images" [CVPR 2022]
★ 588-1Star change over the last 7 days - #57
🔥🔥🔥 A curated list of papers on LLMs-based multimodal generation (image, video, 3D and audio).
★ 552+0Star change over the last 7 days - #58
AI Plugin is a powerful extension for the Payload CMS, integrating advanced AI capabilities to enhance content creation and management.
★ 547-1Star change over the last 7 days - #59
[NeurIPS 2025 Spotlight] A Unified Tokenizer for Visual Generation and Understanding
★ 529-1Star change over the last 7 days - #60
A Collection of Google Colab Notebooks for various projects
★ 521+1Star change over the last 7 days