image-generation
Tracked open-source repos tagged image-generation, sorted by stars.
- #91
[TMLR 2025🔥] A survey for the autoregressive models in vision.
★ 808+1Star change over the last 7 days - #92
Open‑WebUI Tools is a modular toolkit designed to extend and enrich your Open WebUI instance, turning it into a powerful AI workstation. With a suite of over 15 specialized tools, function pipelines, and filters, this project supports academic research, agentic autonomy, multimodal creativity, workflows, and more
★ 807+5Star change over the last 7 days - #93
Local-first visual generation runtime and studio for people and coding agents, with reproducible image and video workflows across multiple providers.
★ 751+30Star change over the last 7 days - #94
A Survey on Text-to-Video Generation/Synthesis.
★ 744+3Star change over the last 7 days - #95
A source-available Codex skill for distilling photographs into sparse editorial abstractions.
★ 743+35Star change over the last 7 days - #96
1,400+ curated trending AI image prompts from X, ranked by engagement. Works with NanoBanana, GPT Image 2, Midjourney
★ 731+7Star change over the last 7 days - #97
Local-first, agent-native control plane for ComfyUI — MCP server + sidebar agent that generates images, video & audio, authors and runs workflows, and edits your live graph in natural language on ANY LLM (Claude, ChatGPT, Gemini, offline Ollama, or any hosted model). 178 tools, 36 AI skills, 55 installer packs. Local, LAN, VPS, or Comfy Cloud.
★ 729+35Star change over the last 7 days - #98
[ICCV 2021] Focal Frequency Loss for Image Reconstruction and Synthesis
★ 714+1Star change over the last 7 days - #99
[ICLR 2026] ChronoEdit: Towards Temporal Reasoning for Image Editing and World Simulation
★ 706+1Star change over the last 7 days - #100
Image Signal Processing (ISP) Guide. Learn all about the process of converting an image/video into digital form by performing tasks like noise reduction, filtering, auto exposure, autofocus, HDR correction, and image sharpening with a Specialized type of media processor.
★ 703+1Star change over the last 7 days - #101
[ECCV 2024] OMG: Occlusion-friendly Personalized Multi-concept Generation In Diffusion Models
★ 702+1Star change over the last 7 days - #102
面向 GPT-image-2 的 AI 图片生成 WebUI 工作台,支持 Codex Responses 与 OpenAI 兼容 API 接入,内置公用图库、多类型 Chip 快捷引用、提示词模板、多任务并发和本地队列管理。An AI image generation WebUI workbench for GPT-image-2 with Codex Responses and OpenAI-compatible API support, shared gallery references, multi-type quick chips, prompt templates, concurrent tasks, and local queue management.
★ 693+9Star change over the last 7 days - #103
A unified framework for easy fine-tuning in Flow-Matching models
★ 692+7Star change over the last 7 days - #104
A Collection of Papers and Codes for CVPR2026/CVPR2025/ICCV2025/CVPR2024/ECCV2026/ECCV2024 AIGC
★ 677+1Star change over the last 7 days - #105
HunyuanImage-2.1: An Efficient Diffusion Model for High-Resolution (2K) Text-to-Image Generation
★ 675+0Star change over the last 7 days - #106
A Collection of Papers and Codes in CVPR2023/2022 about low level vision
★ 660+0Star change over the last 7 days - #107★ 660-1Star change over the last 7 days
- #108
This repository contains a pure C++ ONNX implementation of multiple offline AI models, such as StableDiffusion (1.5 and XL), ControlNet, Midas, HED and OpenPose.
★ 633+0Star change over the last 7 days - #109
Official implementation for "Blended Latent Diffusion" [SIGGRAPH 2023]
★ 632+0Star change over the last 7 days - #110
Character Select Stand Alone App with AI prompt and ComfyUI/WebUI API support for waiIllustriousSDXL, waiANIMA and others model
★ 611+1Star change over the last 7 days - #111
Yet another PyTorch implementation of Stable Diffusion (probably easy to read)
★ 594+0Star change over the last 7 days - #112
118+ plug-and-play JSON style packs for Nano Banana Pro, GPT Image & Midjourney. Copy one JSON, get a style. Updated daily.
★ 592+18Star change over the last 7 days - #113
Official code for the CVPR 2025 paper "SemanticDraw: Towards Real-Time Interactive Content Creation from Image Diffusion Models."
★ 588-1Star change over the last 7 days - #114
📚 GPT4o Prompts Dictionary | Curated Collection of AI Image Generation Prompts
★ 585+0Star change over the last 7 days - #115
Free, open-source alternative to Weavy AI, Krea Nodes, Freepik Spaces & FloraFauna AI — node-based AI workflow builder for generative image & video pipelines
★ 575+10Star change over the last 7 days - #116
Official Code for ECCV 2024 paper — One-Shot Diffusion Mimicker for Handwritten Text Generation
★ 562+1Star change over the last 7 days - #117
[AAAI2024] FontDiffuser: One-Shot Font Generation via Denoising Diffusion with Multi-Scale Content Aggregation and Style Contrastive Learning
★ 560+7Star change over the last 7 days - #118
Genblaze is an open source Python SDK for orchestrating generative AI media pipelines across video, audio, and image providers with built in provenance for every output.
★ 559-1Star change over the last 7 days - #119
Codex skill for turning everyday photos into poetic white-paper hand-drawn illustrations.
★ 557+46Star change over the last 7 days - #120
AI Plugin is a powerful extension for the Payload CMS, integrating advanced AI capabilities to enhance content creation and management.
★ 548+2Star change over the last 7 days