Skip to main content
buildradar
Sign in
Topic · image-generation

image-generation

Tracked open-source repos tagged image-generation, sorted by stars.

130 repos
  • snappy@KnpLabs

    PHP library allowing thumbnail, snapshot or PDF generation from a url or a html page. Wrapper for wkhtmltopdf/wkhtmltoimage

    4,474-2Star change over the last 7 days
  • OmniGen@VectorSpaceLab

    OmniGen: Unified Image Generation. https://arxiv.org/pdf/2409.11340

    4,342+2Star change over the last 7 days
  • ruby_llm@crmne

    One delightful Ruby framework for every major AI provider. Build AI agents, chatbots, RAG apps, and multimodal workflows in beautiful, expressive code.

    4,335+13Star change over the last 7 days
  • AnyDoor@ali-vilab

    Official implementations for paper: Anydoor: zero-shot object-level image customization

    4,239+1Star change over the last 7 days
  • Multi-modal Generative Media Skills for AI Agents (Claude Code, Cursor, Gemini CLI). High-quality image, video, and audio generation powered by muapi.ai.

    4,205+42Star change over the last 7 days
  • gpt_image_playground@CookSleep

    基于 OpenAI gpt-image-2 API 的图片生成与编辑工具

    3,599+48Star change over the last 7 days
  • Gemini-API@HanaokaYuzu

    ✨ Reverse-engineered Python API for Google Gemini web app

    3,468+31Star change over the last 7 days
  • HunyuanImage-3.0@Tencent-Hunyuan

    HunyuanImage-3.0: A Powerful Native Multimodal Model for Image Generation

    3,256+3Star change over the last 7 days
  • Dreambooth-Stable-Diffusion@JoePenna

    Implementation of Dreambooth (https://arxiv.org/abs/2208.12242) by way of Textual Inversion (https://arxiv.org/abs/2208.01618) for Stable Diffusion (https://arxiv.org/abs/2112.10752). Tweaks focused on training faces, objects, and styles.

    3,212+1Star change over the last 7 days
  • awesome-generative-ai-apps@Anil-matcha

    50+ open-source generative AI apps you can clone, deploy, and monetize — image generators, video tools, virtual try-ons, AI SaaS templates, and platform integrations. One-click Vercel deploy on every template.

    3,122+51Star change over the last 7 days
  • pwa-asset-generator@elegantapp

    Automates PWA asset generation and image declaration. Automatically generates icon and splash screen images, favicons and mstile images. Updates manifest.json and index.html files with the generated images according to Web App Manifest specs and Apple Human Interface guidelines.

    3,032+2Star change over the last 7 days
  • takumi@kane50613

    Render OG images and paged PDFs from JSX, HTML, and CSS. No headless browser. Runs on Node.js, Cloudflare Workers, browsers, and Rust.

    2,912+25Star change over the last 7 days
  • Kandinsky-2@ai-forever

    Kandinsky 2 — multilingual text2image latent diffusion model

    2,814+0Star change over the last 7 days
  • mono-color-skill@yanliudesign

    One-ink editorial print image skill — warm paper, halftone photography, active negative space, and restrained typography.

    2,759Star change over the last 7 days
  • InfiniteYou@bytedance

    🔥 [ICCV 2025 Highlight] InfiniteYou: Flexible Photo Recrafting While Preserving Your Identity

    2,684-1Star change over the last 7 days
  • (ෆ`꒳´ෆ) A Survey on Text-to-Image Generation/Synthesis.

    2,444+0Star change over the last 7 days
  • A powerful tool that translates ComfyUI workflows into executable Python code.

    2,377+2Star change over the last 7 days
  • DreamOmni2@JIA-Lab-research

    This project is the official implementation of 'DreamOmni2: Multimodal Instruction-based Editing and Generation (CVPR2026 Highlight)''

    1,982+4Star change over the last 7 days
  • LlamaGen@FoundationVision

    Autoregressive Model Beats Diffusion: 🦙 Llama for Scalable Image Generation

    1,966+0Star change over the last 7 days
  • AI skill for OpenClaw & Claude Code — recommend from 10000+ Nano Banana Pro (Gemini) image prompts. Smart search by use case, content remix, sample images.

    1,843+7Star change over the last 7 days
  • VideoClaw@HITsz-TMG

    🚀 AI 全自动化视频生成员工 | Your First AIGC Coworker. Chat an Idea. Get a Film. 🦞

    1,749+23Star change over the last 7 days
  • Tracking and collecting papers/projects/others related to Segment Anything.

    1,699+2Star change over the last 7 days
  • Infinity@FoundationVision

    [CVPR 2025 Oral]Infinity ∞ : Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis

    1,587+0Star change over the last 7 days
  • MiniMax-MCP@MiniMax-AI

    Official MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech, image generation and video generation APIs.

    1,573+3Star change over the last 7 days
  • diffusiondb@poloclub

    A large-scale text-to-image prompt gallery dataset based on Stable Diffusion

    1,392+1Star change over the last 7 days
  • 中文手绘技术 PPT 整页图像生成 Skill | 21:9 封面 + 16:9 正文配图 | PNG 输出

    1,381+22Star change over the last 7 days
  • generative-ai-use-cases@aws-samples

    Application implementation with business use cases for safely utilizing generative AI in business operations

    1,381+2Star change over the last 7 days
  • UNO@bytedance

    [ICCV 2025] 🔥🔥 UNO: A Universal Customization Method for Both Single and Multi-Subject Conditioning

    1,362+0Star change over the last 7 days
  • locally-uncensored@PurpleDoubleD

    Plug-and-play local AI studio: uncensored chat, image & video generation, coding agent. Runs abliterated LLMs + ComfyUI 100% offline. One installer, no Docker, no cloud.

    1,351+201Star change over the last 7 days
  • FireRed-Image-Edit@FireRedTeam

    FireRed-Image-Edit is a powerful image editing foundation model achieving open-source state-of-the-art performance with precise instruction following, high-fidelity generation, superior identity consistency, and seamless multi-element fusion.

    1,350+6Star change over the last 7 days
← Back to topics