image-generation
Tracked open-source repos tagged image-generation, sorted by stars.
- #31
PHP library allowing thumbnail, snapshot or PDF generation from a url or a html page. Wrapper for wkhtmltopdf/wkhtmltoimage
★ 4,474-2Star change over the last 7 days - #32★ 4,342+2Star change over the last 7 days
- #33
One delightful Ruby framework for every major AI provider. Build AI agents, chatbots, RAG apps, and multimodal workflows in beautiful, expressive code.
★ 4,335+13Star change over the last 7 days - #34
Official implementations for paper: Anydoor: zero-shot object-level image customization
★ 4,239+1Star change over the last 7 days - #35
Multi-modal Generative Media Skills for AI Agents (Claude Code, Cursor, Gemini CLI). High-quality image, video, and audio generation powered by muapi.ai.
★ 4,205+42Star change over the last 7 days - #36
基于 OpenAI gpt-image-2 API 的图片生成与编辑工具
★ 3,599+48Star change over the last 7 days - #37
✨ Reverse-engineered Python API for Google Gemini web app
★ 3,468+31Star change over the last 7 days - #38
HunyuanImage-3.0: A Powerful Native Multimodal Model for Image Generation
★ 3,256+3Star change over the last 7 days - #39
Implementation of Dreambooth (https://arxiv.org/abs/2208.12242) by way of Textual Inversion (https://arxiv.org/abs/2208.01618) for Stable Diffusion (https://arxiv.org/abs/2112.10752). Tweaks focused on training faces, objects, and styles.
★ 3,212+1Star change over the last 7 days - #40
50+ open-source generative AI apps you can clone, deploy, and monetize — image generators, video tools, virtual try-ons, AI SaaS templates, and platform integrations. One-click Vercel deploy on every template.
★ 3,122+51Star change over the last 7 days - #41
Automates PWA asset generation and image declaration. Automatically generates icon and splash screen images, favicons and mstile images. Updates manifest.json and index.html files with the generated images according to Web App Manifest specs and Apple Human Interface guidelines.
★ 3,032+2Star change over the last 7 days - #42
Render OG images and paged PDFs from JSX, HTML, and CSS. No headless browser. Runs on Node.js, Cloudflare Workers, browsers, and Rust.
★ 2,912+25Star change over the last 7 days - #43
Kandinsky 2 — multilingual text2image latent diffusion model
★ 2,814+0Star change over the last 7 days - #44
One-ink editorial print image skill — warm paper, halftone photography, active negative space, and restrained typography.
★ 2,759—Star change over the last 7 days - #45
🔥 [ICCV 2025 Highlight] InfiniteYou: Flexible Photo Recrafting While Preserving Your Identity
★ 2,684-1Star change over the last 7 days - #46
(ෆ`꒳´ෆ) A Survey on Text-to-Image Generation/Synthesis.
★ 2,444+0Star change over the last 7 days - #47
A powerful tool that translates ComfyUI workflows into executable Python code.
★ 2,377+2Star change over the last 7 days - #48
This project is the official implementation of 'DreamOmni2: Multimodal Instruction-based Editing and Generation (CVPR2026 Highlight)''
★ 1,982+4Star change over the last 7 days - #49
Autoregressive Model Beats Diffusion: 🦙 Llama for Scalable Image Generation
★ 1,966+0Star change over the last 7 days - #50
AI skill for OpenClaw & Claude Code — recommend from 10000+ Nano Banana Pro (Gemini) image prompts. Smart search by use case, content remix, sample images.
★ 1,843+7Star change over the last 7 days - #51★ 1,749+23Star change over the last 7 days
- #52
Tracking and collecting papers/projects/others related to Segment Anything.
★ 1,699+2Star change over the last 7 days - #53
[CVPR 2025 Oral]Infinity ∞ : Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis
★ 1,587+0Star change over the last 7 days - #54
Official MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech, image generation and video generation APIs.
★ 1,573+3Star change over the last 7 days - #55
A large-scale text-to-image prompt gallery dataset based on Stable Diffusion
★ 1,392+1Star change over the last 7 days - #56
中文手绘技术 PPT 整页图像生成 Skill | 21:9 封面 + 16:9 正文配图 | PNG 输出
★ 1,381+22Star change over the last 7 days - #57
Application implementation with business use cases for safely utilizing generative AI in business operations
★ 1,381+2Star change over the last 7 days - #58
[ICCV 2025] 🔥🔥 UNO: A Universal Customization Method for Both Single and Multi-Subject Conditioning
★ 1,362+0Star change over the last 7 days - #59
Plug-and-play local AI studio: uncensored chat, image & video generation, coding agent. Runs abliterated LLMs + ComfyUI 100% offline. One installer, no Docker, no cloud.
★ 1,351+201Star change over the last 7 days - #60
FireRed-Image-Edit is a powerful image editing foundation model achieving open-source state-of-the-art performance with precise instruction following, high-fidelity generation, superior identity consistency, and seamless multi-element fusion.
★ 1,350+6Star change over the last 7 days