AgriciDaniel/banana-claude

Banana Claude is a Claude Code plugin that lets you write a natural‑language request, see a detailed plan (prompt, model, cost estimate) and, after an explicit approval, call Google’s Gemini image models to generate or edit visual assets. It emphasizes offline‑first planning, privacy‑safe API key handling, and one‑time approvals to keep costs transparent.

allenk/GeminiWatermarkTool

A cross‑platform, portable executable that removes Google Gemini image (and video) watermarks using a deterministic reverse‑alpha‑blending algorithm, with optional GPU‑accelerated FDnCNN denoising, GUI/CLI interfaces, batch processing, and AI‑agent integration via MCP.

AHEKOT/ComfyUI_VNCCS

A ComfyUI pipeline for creating consistent character sprites, allowing users to manage appearance, poses, clothing, and emotions across multiple images.

KohakuBlueleaf/LyCORIS

A library implementing various parameter-efficient fine-tuning algorithms like LoHa and LoKr for Stable Diffusion, enabling high-quality model customization with minimal storage and compute.

aldegad/sprite-gen

A pipeline for turning a single image into game-ready sprite atlases, featuring identity-locking generation, background removal, and a tool to animate still frames into breathing loops.

YouMind-OpenLab/nano-banana-pro-prompts-recommend-skill

An AI agent skill that allows AI assistants to search and recommend 10,000+ curated prompts for the Nano Banana Pro (Gemini) image generation model.

Gourieff/ComfyUI-ReActor

A set of ComfyUI nodes for fast and simple face swapping in images and videos, featuring face restoration and blended face model creation.

krea-ai/krea-2

Krea 2 is an open-source image generation model featuring a base RAW model for flexible fine-tuning and a distilled Turbo model for fast, high-quality text-to-image synthesis.

VigoZhao/AI-Visual-Prompt-Cookbook

A curated library of 124 structured JSON prompt templates that allow users to maintain consistent, professional visual styles across AI image generation workflows.

nova452/Rebalance-Pack

A collection of custom ComfyUI nodes that streamline image generation and editing, featuring specialized support for Krea 2 and Ideogram 4 models.

yolain/ComfyUI-Easy-Use

An efficiency custom nodes package for ComfyUI that integrates and simplifies popular nodes to provide a faster, more streamlined image generation experience.

buluslan/gpt-image2-ecommerce

A tool for generating professional e-commerce product images using GPT-Image-2 and Codex CLI, featuring 25 specialized scene templates and reference image support.

Steve-Mr/EmojiFace

FaceMoji (EmojiFace) is an Android and web app that locally detects faces in photos using a YOLOv8‑face model and overlays them with emojis or blur effects. It runs offline, offers manual editing, supports custom emoji fonts, and can be installed as a hidden‑icon app for share‑only activation.

pangxiaobin/image-matting

A desktop application that uses the RMBG-1.4 model to provide local AI-powered background removal, batch processing, and image editing tools.

pixaroma/ComfyUI-Pixaroma

ComfyUI‑Pixaroma is a genuine open‑source extension for the ComfyUI diffusion‑workflow editor. It adds a large collection of ready‑made nodes—image loading/resizing, visual crop & paint tools, 3‑D scene builder, audio‑reactive video, LoRA stacking, text utilities, XY‑plot comparison, and workflow helpers—plus built‑in help panels and example workflows. The goal is to make everyday AI‑image generation tasks more visual and user‑friendly, letting artists and developers build, preview, and tweak pipelines without leaving ComfyUI.

ModelTC/LightX2V-Qwen-Image-Lightning

A distilled, faster version of the Qwen‑Image text‑to‑image model (4‑ or 8‑step diffusion) with fp32/bf16/fp8 checkpoints, LoRA adapters, and ready‑to‑use ComfyUI workflows. Provides 12–25× speed‑up with modest quality loss and integrates with Diffusers, ComfyUI, Nunchaku, and Cache‑dit.

nv-tlabs/PiD

PiD is a plug-and-play diffusion decoder that replaces standard VAE/RAE decoders to turn latent representations directly into super-resolved 2K to 4K pixels in a single pass.

chaiNNer-org/chaiNNer

A node-based GUI for programmatic image processing and AI upscaling that allows users to build customizable processing pipelines via a visual interface.

NVIDIA-RTX/NRD

GPU‑accelerated, API‑agnostic library for real‑time ray‑tracing denoising (REBLUR, RELAX, SIGMA).

kadevin/ilab-conjure

A local-first AI image generation workbench that unifies GPT Image and Gemini with integrated asset libraries, prompt templates, and task management.

helblazer811/Diffusion-Explorer

An interactive visualization tool for educational purposes that explains the geometric intuitions behind diffusion and flow-based generative models.

YanWenKun/ComfyUI-Windows-Portable

A ready‑to‑run Windows package that bundles ComfyUI (a node‑based UI for Stable Diffusion) with 40+ custom nodes and 300+ pre‑compiled Python packages, letting users launch AI image/video generation instantly on an NVIDIA GPU.

chflame163/ComfyUI_LayerStyle

A set of ComfyUI nodes that bring Photoshop-like layer and mask compositing functionality to the platform, reducing the need for external editing software.

cocktailpeanut/fluxgym

Flux Gym provides a Gradio web UI that lets you train FLUX diffusion LoRAs on 12‑20 GB GPUs. It wraps Kohya‑ss training scripts, offering full script options, automatic model download, optional sample‑image generation, and one‑click publishing to Hugging Face. Install via Pinokio, Docker, or manual venv setup.

leeguooooo/chatgpt-imagegen

A zero-dependency Python CLI and AI agent skill that lets users generate images using their existing ChatGPT or Gemini subscriptions instead of a paid API key.

zjx0101/ObjectClear

ObjectClear is an object removal model that jointly eliminates target objects and their associated effects, such as shadows, while preserving background consistency.

ATH-MaaS/ComfyUI-Copilot

An intelligent assistant for ComfyUI that automates the generation, debugging, and optimization of image generation workflows using an LLM-powered agent.

easydiffusion/easydiffusion

A one-click installer and user interface for Stable Diffusion that allows users to run powerful image generation AI locally without requiring technical knowledge.

hustvl/Moebius

Moebius is a lightweight 0.22B parameter image inpainting framework that achieves 10B-level quality and 15x faster inference speeds through architectural optimization and knowledge distillation.

martin-rizzo/ComfyUI-ZImagePowerNodes

A collection of ComfyUI custom nodes optimized for the Z-Image Turbo model, offering enhanced style control, low-step consistency, and compositional variety.

RoseKhlifa/Image-Studio

A cross-platform image generation and editing client that prevents connection timeouts during long inference tasks using SSE and WebSocket modes.

Comfy-Org/comfy-cli

comfy‑cli is a Python command‑line tool that installs, launches, and manages ComfyUI (an open‑source generative‑media engine). It handles environment setup, custom‑node installation, model downloads, workflow execution, and can offload jobs to the hosted Comfy Cloud service. Features include fast `uv`‑based dependency resolution, JSON‑structured output for LLM agents, shell completion, and utilities for testing PRs and managing background servers.

Code-with-Beto/snapai

A CLI tool for mobile developers to generate app icons and Google Play feature graphics using OpenAI and Google Gemini image models.

Topkill/tianruoocr

TianRuo OCR v6 is a Windows desktop app that captures screen text, runs OCR (online or offline via PaddleOCR/RapidOCR), and translates the result using many online services. It offers normal, silent, screenshot and clipboard‑listen modes, configurable shortcuts, and supports printed, table and handwritten text.

unconv-ai/Un-0

An image-generation model based on Kuramoto dynamics that generates images by integrating coupled oscillators instead of using diffusion or iterative denoising.

runpod-workers/worker-comfyui

A serverless wrapper for ComfyUI on RunPod that allows users to execute complex image generation workflows via a scalable API endpoint.

lynote-ai/ai-image-detector

AI Image Detector is an open‑source Python tool that estimates how likely an image was generated by AI. It ships with several ready‑made back‑ends (UnivFD, Sentry, Nonescape, hybrid ensembles) and lets you run a single command (`aidetect detect …`) on a file or folder, outputting a probability and label. Optional extras give a Gradio web UI or a FastAPI server. The repo includes benchmarking scripts, threshold‑calibration, and clear guidance on which backend to pick. It’s a lightweight, install‑once‑run‑anywhere utility for AI‑image detection, with documented limitations (no universal guarantee, no region‑level localization).

zhu-guli326/image2_UI_skill

Image2 UI is an OpenAI‑Codex‑powered tool that converts screenshots, sketches, or text prompts into real, editable front‑end code with interactive logic and separate image assets. It supports recreate, redesign, and create workflows, runs on Node 20+ and Python 3.10+, and is MIT‑licensed.

TheJoeFin/Text-Grab

Text Grab is a Windows‑only desktop app that captures any visible text (screenshots, PDFs, UI elements) using local OCR (WinAI, WinRT OCR, or Tesseract) and provides built‑in cleanup, spreadsheet editing, regex‑based extraction, reusable grab templates, bulk folder processing, and a Chrome/Edge extension. All processing stays on‑device, with optional NPU‑accelerated inference on Copilot+ PCs. Install via Microsoft Store, GitHub releases, or package managers; source can be built with Visual Studio or the .NET SDK.

ssitu/ComfyUI_UltimateSDUpscale

A set of ComfyUI nodes that perform image-to-image diffusion on large images using a tiled approach to improve detail and reduce hardware requirements.

yanokusnir-ai/one-node-flux-2-klein

A ComfyUI custom node that wraps the full FLUX.2 [klein] workflow into a single, self-contained UI widget to eliminate complex node graphs.

karimz1/imgcompress

A self-hosted image processing server that provides local AI-powered background removal, bulk compression, and conversion for over 70 image formats.

nregret/Comfyui-Anima-Tools

A visual prompt and LoRA selection suite for ComfyUI designed for anime AI art, featuring databases of 40k+ artists, characters, and clothing tags.

TeamMoeAI/MoeSR

An image super-resolution and restoration tool specifically optimized for ACGN illustrations, CG, and manga using deep learning models.

mittagessen/kraken

kraken is an open‑source, Python‑based OCR toolkit for historical and non‑Latin scripts. It provides trainable layout analysis, reading‑order detection, and neural‑network character recognition, supporting right‑to‑left, bidirectional and vertical scripts. The tool outputs standard XML formats (ALTO, PageXML, hOCR), offers word‑level bounding boxes, and ships with a public model repository. Install via pip/pipx on Linux or macOS, fetch a model with `kraken get …`, and run a one‑line command to binarise, segment, and OCR an image.

mrslimslim/gpt-image-canvas

A local-first AI image workspace that combines an infinite canvas with agentic planning for multi-step prompt-to-image generation.

drawthingsai/draw-things-community

A community repository providing the core image generation, sampling, and training code for the Draw Things app, including a self-hostable gRPC server for remote inference.

giriss/comfy-image-saver

A set of ComfyUI custom nodes that save images with embedded generation metadata, ensuring compatibility with Civitai and Prompthero.