text-to-image
Tracked open-source repos tagged text-to-image, sorted by stars.
Related topics
Topics that frequently appear alongside text-to-image on the same repo.
Recent risers
Repos created in the last 90 days, tagged text-to-image.
No new repos tagged with this topic in the last 90 days.
- #1
Unrestricted Open-source alternative to AI video platforms — Free AI image & video generation studio with 500+ models (Flux, Midjourney, Kling, Sora, Veo). No content filters. Self-hosted, MIT licensed.
★ 27,359+557Star change over the last 7 days - #2
GPT-Image-2 API and Prompts
★ 17,006—Star change over the last 7 days - #3
Implementation of DALL-E 2, OpenAI's updated text-to-image synthesis neural network, in Pytorch
★ 11,305-2Star change over the last 7 days - #4
Implementation of Imagen, Google's Text-to-Image Neural Network, in Pytorch
★ 8,425+2Star change over the last 7 days - #5
Awesome curated collection of images and prompts generated by GPT-4o and gpt-image-1. Explore AI generated visuals created with ChatGPT and Sora, showcasing OpenAI’s advanced image generation capabilities.
★ 8,133+8Star change over the last 7 days - #6
This repository contains a hand-curated resources for Prompt Engineering with a focus on Generative Pre-trained Transformer (GPT), ChatGPT, PaLM etc
★ 6,299+27Star change over the last 7 days - #7
Implementation / replication of DALL-E, OpenAI's Text to Image Transformer, in Pytorch
★ 5,628+0Star change over the last 7 days - #8
Red Ink - A one-stop Xiaohongshu image-and-text generator based on the 🍌Nano Banana Pro🍌, "One Sentence, One Image: Generate Xiaohongshu Text and Images."
★ 5,487+11Star change over the last 7 days - #9
GPT Image 2 prompt gallery, image prompt library, agentic skill, and CLI for OpenAI image generation/editing
★ 4,987+182Star change over the last 7 days - #10
Multi-modal Generative Media Skills for AI Agents (Claude Code, Cursor, Gemini CLI). High-quality image, video, and audio generation powered by muapi.ai.
★ 4,186+75Star change over the last 7 days - #11
LTX-Video Support for ComfyUI
★ 4,105+19Star change over the last 7 days - #12
[CVPR 2026] PromptEnhancer is a prompt-rewriting tool, refining prompts into clearer, structured versions for better image generation.
★ 3,761+7Star change over the last 7 days - #13
A curated list of Generative AI tools, works, models, and references
★ 3,529+5Star change over the last 7 days - #14★ 3,494-1Star change over the last 7 days
- #15
Diffusion model papers, survey, and taxonomy
★ 3,364+0Star change over the last 7 days - #16
50+ open-source generative AI apps you can clone, deploy, and monetize — image generators, video tools, virtual try-ons, AI SaaS templates, and platform integrations. One-click Vercel deploy on every template.
★ 3,087+50Star change over the last 7 days - #17
Kandinsky 2 — multilingual text2image latent diffusion model
★ 2,814+0Star change over the last 7 days - #18
FLUX, Stable Diffusion, SDXL, SD3, LoRA, Fine Tuning, DreamBooth, Training, Automatic1111, Forge WebUI, SwarmUI, DeepFake, TTS, Animation, Text To Video, Tutorials, Guides, Lectures, Courses, ComfyUI, Google Colab, RunPod, Kaggle, NoteBooks, ControlNet, TTS, Voice Cloning, AI, AI News, ML, ML News, News, Tech, Tech News, Kohya, Midjourney, RunPod
★ 2,763+6Star change over the last 7 days - #19
A playground to generate images from any text prompt using Stable Diffusion (past: using DALL-E Mini)
★ 2,741+0Star change over the last 7 days - #20
🔥 [ICCV 2025 Highlight] InfiniteYou: Flexible Photo Recrafting While Preserving Your Identity
★ 2,686+2Star change over the last 7 days - #21
(ෆ`꒳´ෆ) A Survey on Text-to-Image Generation/Synthesis.
★ 2,444+2Star change over the last 7 days - #22
Open source implementation and extension of Google Research’s PaperBanana for automated academic figures, diagrams, and research visuals, expanded to new domains like slide generation.
★ 2,290+27Star change over the last 7 days - #23
AI magics meet Infinite draw board.
★ 1,935-1Star change over the last 7 days - #24
[ICML 2024] Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs (RPG)
★ 1,844+2Star change over the last 7 days - #25
[ECCV 2024] The official implementation of paper "BrushNet: A Plug-and-Play Image Inpainting Model with Decomposed Dual-Branch Diffusion"
★ 1,744+0Star change over the last 7 days - #26
Official Pytorch Implementation for "TokenFlow: Consistent Diffusion Features for Consistent Video Editing" presenting "TokenFlow" (ICLR 2024)
★ 1,708+0Star change over the last 7 days - #27
[CVPR 2025 Oral]Infinity ∞ : Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis
★ 1,587+0Star change over the last 7 days - #28
Official MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech, image generation and video generation APIs.
★ 1,571+7Star change over the last 7 days - #29
[ICCV 2025] 🔥🔥 UNO: A Universal Customization Method for Both Single and Multi-Subject Conditioning
★ 1,362+0Star change over the last 7 days - #30
Turn any face into a video game character, pixel art, claymation, 3D or toy
★ 1,361-2Star change over the last 7 days