Skip to main content
buildradar
Sign in
Topic · image-generation

image-generation

Tracked open-source repos tagged image-generation, sorted by stars.

134 repos
  • A simple standalone viewer for reading prompts from Stable Diffusion generated image outside the webui.

    1,340+0Star change over the last 7 days
  • Lance@bytedance

    A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.

    1,334+3Star change over the last 7 days
  • WebAI2API@foxhui

    WebAI2API: 基于 Camoufox 的网页 AI 转 API 工具,支持 LMArena/Gemini等,多窗口并发与账号隔离。 | Web AI to OpenAI API via Camoufox. Supports LMArena/Gemini and more, multi-window concurrency & account isolation.

    1,315+16Star change over the last 7 days
  • airunner@Capsize-Games

    Offline inference engine for art, real-time voice conversations, LLM powered chatbots and automated workflows

    1,314+0Star change over the last 7 days
  • data-efficient-gans@mit-han-lab

    [NeurIPS 2020] Differentiable Augmentation for Data-Efficient GAN Training

    1,308+0Star change over the last 7 days
  • Bernini@bytedance

    Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer.

    1,300+10Star change over the last 7 days
  • Inpaint Anything extension performs stable diffusion inpainting on a browser UI using masks from Segment Anything.

    1,298+0Star change over the last 7 days
  • MV-Adapter@huanngzh

    [ICCV 2025] Official impl. of "MV-Adapter: Multi-view Consistent Image Generation Made Easy"

    1,289+1Star change over the last 7 days
  • Easily create PDF and images in Symfony by converting html using webkit

    1,247+0Star change over the last 7 days
  • Simple shell script to use OpenAI's ChatGPT and DALL-E from the terminal. No Python or JS required.

    1,241+3Star change over the last 7 days
  • Collection of awesome resources on image-to-image translation.

    1,239+0Star change over the last 7 days
  • cleanvision@cleanlab

    Automatically find issues in image datasets and practice data-centric computer vision.

    1,199+0Star change over the last 7 days
  • clean-fid@GaParmar

    PyTorch - FID calculation with proper image resizing and quantization steps [CVPR 2022]

    1,167+0Star change over the last 7 days
  • MagicDrive@cure-lab

    [ICLR24] Official implementation of the paper “MagicDrive: Street View Generation with Diverse 3D Geometry Control”

    1,167+1Star change over the last 7 days
  • VITON-HD@shadow2496

    Official PyTorch implementation of "VITON-HD: High-Resolution Virtual Try-On via Misalignment-Aware Normalization" (CVPR 2021)

    1,161+0Star change over the last 7 days
  • Uncensored-Local-Studio@techjarves

    Uncensored local AI studio for Windows, Linux, and macOS. Zero-setup GUI for Image Generation, GGUF LLMs, Text to Speech & Speech to Text

    1,118+87Star change over the last 7 days
  • 归藏的材质插画 skill:生成带字解释图、图表美化和参考辅助配图。

    1,116+40Star change over the last 7 days
  • mlx-serve@ddalcu

    Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Includes MLX Core macOS app with chat, agent mode, and tool calling.

    1,115+245Star change over the last 7 days
  • CogView4@zai-org

    CogView4, CogView3-Plus and CogView3(ECCV 2024)

    1,102+1Star change over the last 7 days
  • image-extender@boona13

    Seamlessly extend any image in any direction with AI. Open-source web app powered by Gemini via OpenRouter, with Poisson-blended seams and best-of-3 variant picker.

    1,101+5Star change over the last 7 days
  • awesome-agent-apis@Anil-matcha

    660+ muapi-hosted generative-media models plus community-submitted third-party API tools (SEO, enrichment, social, scraping) — one YAML file per entry, browsable by capability.

    1,005+8Star change over the last 7 days
  • :fire: :fire: :fire: A paper list of some recent Computer Vision(CV) works

    973+1Star change over the last 7 days
  • story-iter@UCSC-VLAA

    [ICLR 2026] A Training-free Iterative Framework for Long Story Visualization

    957-2Star change over the last 7 days
  • 🎬 Generate images from any camera viewpoint via 3D interactive control. Drag the camera in 3D space or use sliders to set azimuth/elevation/distance, then generate. Built with Three.js + Gradio, bilingual ZH/EN UI.

    955+231Star change over the last 7 days
  • HR-VITON@sangyun884

    Official PyTorch implementation for the paper High-Resolution Virtual Try-On with Misalignment and Occlusion-Handled Conditions (ECCV 2022).

    917+0Star change over the last 7 days
  • AMC-WebUI@yeahhe365

    面向 Gemini 的 Local-First AI 工作流 WebUI,集成多模态聊天、Canvas、文件处理、实时搜索、代码执行与高级推理。

    899+5Star change over the last 7 days
  • AnyGPT@OpenMOSS

    A unified multimodal language model based on discrete sequence modeling

    881-1Star change over the last 7 days
  • codex-slides@nexu-io

    🎨 Open-source AI slide studio inside Codex: image-native decks, every slide a full visual canvas. ⚡ 10+ high-quality slides in ~4–5 minutes — Fast mode renders every page in parallel. 🔍 Watch the whole chain live: research → outline → style → render → edit → present → export PDF/PPTX. 🖥️ Browser-first · zero API keys · durable projects.

    878+26Star change over the last 7 days
  • UniPic@SkyworkAI

    Open-source SOTA multi-image editing model

    875+2Star change over the last 7 days
  • flatkey-cli@flatkey-ai

    Flatkey media generation CLI for images, videos, audio, text, credits, and model discovery.

    845+211Star change over the last 7 days
← Back to topics