image-generation
Tracked open-source repos tagged image-generation, sorted by stars.
- #61
A simple standalone viewer for reading prompts from Stable Diffusion generated image outside the webui.
★ 1,340+0Star change over the last 7 days - #62
A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.
★ 1,334+3Star change over the last 7 days - #63
WebAI2API: 基于 Camoufox 的网页 AI 转 API 工具,支持 LMArena/Gemini等,多窗口并发与账号隔离。 | Web AI to OpenAI API via Camoufox. Supports LMArena/Gemini and more, multi-window concurrency & account isolation.
★ 1,315+16Star change over the last 7 days - #64
Offline inference engine for art, real-time voice conversations, LLM powered chatbots and automated workflows
★ 1,314+0Star change over the last 7 days - #65
[NeurIPS 2020] Differentiable Augmentation for Data-Efficient GAN Training
★ 1,308+0Star change over the last 7 days - #66
Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer.
★ 1,300+10Star change over the last 7 days - #67
Inpaint Anything extension performs stable diffusion inpainting on a browser UI using masks from Segment Anything.
★ 1,298+0Star change over the last 7 days - #68
[ICCV 2025] Official impl. of "MV-Adapter: Multi-view Consistent Image Generation Made Easy"
★ 1,289+1Star change over the last 7 days - #69
Easily create PDF and images in Symfony by converting html using webkit
★ 1,247+0Star change over the last 7 days - #70
Simple shell script to use OpenAI's ChatGPT and DALL-E from the terminal. No Python or JS required.
★ 1,241+3Star change over the last 7 days - #71
Collection of awesome resources on image-to-image translation.
★ 1,239+0Star change over the last 7 days - #72
Automatically find issues in image datasets and practice data-centric computer vision.
★ 1,199+0Star change over the last 7 days - #73
PyTorch - FID calculation with proper image resizing and quantization steps [CVPR 2022]
★ 1,167+0Star change over the last 7 days - #74
[ICLR24] Official implementation of the paper “MagicDrive: Street View Generation with Diverse 3D Geometry Control”
★ 1,167+1Star change over the last 7 days - #75
Official PyTorch implementation of "VITON-HD: High-Resolution Virtual Try-On via Misalignment-Aware Normalization" (CVPR 2021)
★ 1,161+0Star change over the last 7 days - #76
Uncensored local AI studio for Windows, Linux, and macOS. Zero-setup GUI for Image Generation, GGUF LLMs, Text to Speech & Speech to Text
★ 1,118+87Star change over the last 7 days - #77
归藏的材质插画 skill:生成带字解释图、图表美化和参考辅助配图。
★ 1,116+40Star change over the last 7 days - #78
Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Includes MLX Core macOS app with chat, agent mode, and tool calling.
★ 1,115+245Star change over the last 7 days - #79★ 1,102+1Star change over the last 7 days
- #80
Seamlessly extend any image in any direction with AI. Open-source web app powered by Gemini via OpenRouter, with Poisson-blended seams and best-of-3 variant picker.
★ 1,101+5Star change over the last 7 days - #81
660+ muapi-hosted generative-media models plus community-submitted third-party API tools (SEO, enrichment, social, scraping) — one YAML file per entry, browsable by capability.
★ 1,005+8Star change over the last 7 days - #82
:fire: :fire: :fire: A paper list of some recent Computer Vision(CV) works
★ 973+1Star change over the last 7 days - #83
[ICLR 2026] A Training-free Iterative Framework for Long Story Visualization
★ 957-2Star change over the last 7 days - #84
🎬 Generate images from any camera viewpoint via 3D interactive control. Drag the camera in 3D space or use sliders to set azimuth/elevation/distance, then generate. Built with Three.js + Gradio, bilingual ZH/EN UI.
★ 955+231Star change over the last 7 days - #85
Official PyTorch implementation for the paper High-Resolution Virtual Try-On with Misalignment and Occlusion-Handled Conditions (ECCV 2022).
★ 917+0Star change over the last 7 days - #86★ 899+5Star change over the last 7 days
- #87★ 881-1Star change over the last 7 days
- #88
🎨 Open-source AI slide studio inside Codex: image-native decks, every slide a full visual canvas. ⚡ 10+ high-quality slides in ~4–5 minutes — Fast mode renders every page in parallel. 🔍 Watch the whole chain live: research → outline → style → render → edit → present → export PDF/PPTX. 🖥️ Browser-first · zero API keys · durable projects.
★ 878+26Star change over the last 7 days - #89★ 875+2Star change over the last 7 days
- #90
Flatkey media generation CLI for images, videos, audio, text, credits, and model discovery.
★ 845+211Star change over the last 7 days