photoprism/photoprism
PhotoPrism is an open‑source, self‑hosted web app that uses on‑device AI to automatically tag, search, and organize personal photos and videos while keeping all data private.
pangxiaobin/image-matting
A desktop application that uses the RMBG-1.4 model to provide local AI-powered background removal, batch processing, and image editing tools.
AgriciDaniel/banana-claude
Banana Claude is a Claude Code plugin that lets you write a natural‑language request, see a detailed plan (prompt, model, cost estimate) and, after an explicit approval, call Google’s Gemini image models to generate or edit visual assets. It emphasizes offline‑first planning, privacy‑safe API key handling, and one‑time approvals to keep costs transparent.
yolain/ComfyUI-Easy-Use
ComfyUI‑Easy‑Use is a comprehensive set of custom nodes that streamline building Stable Diffusion (and other diffusion model) workflows in the ComfyUI visual editor. It adds one‑click loaders, prompt helpers, LoRA/ControlNet stacks, loop/conditional logic, image utilities, and UI tweaks, letting users create and run complex generation pipelines with far fewer nodes.
krea-ai/krea-2
Krea 2 is an open-source image generation model featuring a base RAW model for flexible fine-tuning and a distilled Turbo model for fast, high-quality text-to-image synthesis.
palxiao/poster-design
An open-source AI poster design tool and online image editor featuring a DOM canvas, AI background removal, and a full management backend.
nuyoah-ai-works/nuyoah-xiezhen-prompt
A prompt engineering skill for AI portrait photography that enables consistent style and character recreation across multiple shots by reconstructing professional photographic plans.
nregret/Comfyui-Anima-Tools
A visual prompt and LoRA selection suite for ComfyUI designed for anime AI art, featuring databases of 40k+ artists, characters, and clothing tags.
AHEKOT/ComfyUI_VNCCS
A ComfyUI pipeline for creating consistent character sprites, allowing users to manage appearance, poses, clothing, and emotions across multiple images.
lucidrains/denoising-diffusion-pytorch
A PyTorch implementation of Denoising Diffusion Probabilistic Models (DDPM) for generating high-quality images and 1D sequences.
buluslan/gpt-image2-ecommerce
An AI-powered visual asset generator for e-commerce that uses 39 scene templates and technical pre-checks to create professional, platform-compliant product images.
wnby/photo-relic-editorial
A Codex skill that transforms photographs into vertical editorial artworks by pairing a real photo with a restrained, paper-textured printmaking version of the same image.
storytold/artcraft
ArtCraft is a desktop IDE that lets artists build 2‑D or 3‑D scenes and then fill them with AI‑generated images, video, audio, meshes and worlds. It offers visual tools for compositing, pose transfer, background removal, and prompt‑driven generation, supporting a catalog of 62 built‑in models plus external providers like Midjourney and Sora. Available for Windows/macOS (Linux from source), it targets creators who want repeatable, controllable AI art workflows.
drawthingsai/draw-things-community
A community repository providing the core image generation, sampling, and training code for the Draw Things app, including a self-hostable gRPC server for remote inference.
jtydhr88/ComfyUI-qwenmultiangle
A ComfyUI custom node that provides an interactive 3D viewport for controlling camera angles and generating compatible prompts for the Qwen-Image-Edit-2511-Multiple-Angles-LoRA.
allenk/GeminiWatermarkTool
A cross‑platform, portable executable that removes Google Gemini image (and video) watermarks using a deterministic reverse‑alpha‑blending algorithm, with optional GPU‑accelerated FDnCNN denoising, GUI/CLI interfaces, batch processing, and AI‑agent integration via MCP.
ATH-MaaS/ComfyUI-Copilot
An intelligent assistant for ComfyUI that automates the generation, debugging, and optimization of image generation workflows using an LLM-powered agent.
wzj177/ecommerce-image-suite
An AI-powered toolkit for e-commerce sellers to automatically analyze product photos and generate a complete set of professional marketing images across various styles and platforms.
TaiT-tt/tait-crt-interface-skill
A Codex image generation skill that transforms photos or text into retro illustrations with the aesthetic of early CRT computer interfaces.
Fannovel16/comfyui_controlnet_aux
A collection of ComfyUI nodes that provide preprocessors for creating ControlNet hint images, such as depth maps, canny edges, and human poses.
VigoZhao/AI-Visual-Prompt-Cookbook
A curated library of 130+ JSON‑based visual‑style prompt templates for AI image generation. Each `style.json` defines a reusable layout (e.g., poster, collage, typographic art) with placeholders you fill before pasting the whole file into a multimodal LLM (ChatGPT‑4‑Vision, Gemini‑3‑Pro, etc.). The repo includes preview images, example values, a validation script, and a copy‑prompt shortcut, aiming to make AI‑generated graphics consistent and easy to iterate.
wkentaro/labelme
labelme is a cross‑platform Python/Qt desktop app for drawing bounding boxes, polygons, masks, etc., on images or video. Annotations are saved as JSON and can be exported to VOC or COCO formats, making it a practical tool for creating training data for computer‑vision models. Install via pip, a paid standalone executable, or Linux packages; recent versions add AI‑assisted mask generation (SAM, YOLO‑World).
zuruoke/watermark-removal
A machine learning project that uses image inpainting to remove watermarks from images, creating seamless results.
dacnay816y62-hub/photo-revival
A set of AI skill instructions that transforms ordinary photos into minimalist, poetic hand-drawn illustrations with heavy white space and tiny handwritten captions.
filliptm/ComfyUI_Fill-Nodes
Fill‑Nodes is a large collection of custom ComfyUI nodes that add image‑processing, visual‑effects, captioning, AI‑model calls (GPT, DALL‑E, Hugging Face, etc.), file‑handling (PDF, Google Drive), audio‑reactive and video utilities. Install by dropping the repo into ComfyUI’s `custom_nodes` folder. Use it to build end‑to‑end workflows that load batches, apply creative effects, generate captions via LLMs, and export results as images, videos, PDFs or cloud‑stored files.
six-nut/PocketMen-with-you
A local generative tool that transforms two or more photos into consistent, animated character sprites for Codex companions using open-weight image models.
Hunyuan-PromptEnhancer/PromptEnhancer
A prompt rewriting utility that uses Chain-of-Thought rewriting to enhance text-to-image and image-to-image editing prompts while preserving original user intent.
kadevin/ilab-conjure
iLab CONJURE is a locally‑run, FastAPI‑based web UI (and CLI) for generating and editing images with multiple AI models (GPT‑Image, Gemini, Codex). It offers a task queue, persistent gallery, prompt‑template chips, SQLite history, and supports both OpenAI‑compatible APIs and a local Codex OAuth flow. Install via source or pre‑built macOS/Windows packages, then configure your API key and start generating images.
eikek/docspell
Docspell is an open‑source personal document management system that organizes scanned papers, emails and other files. It uses Stanford CoreNLP to automatically suggest tags, dates and correspondents, runs OCR when needed, and provides full‑text search via a REST API and a mobile‑friendly Elm web UI. Installable via Docker, Debian packages, Nix or Helm, it is licensed under AGPL‑v3.
SamurAIGPT/Vibe-Workflow
Vibe Workflow is an open‑source, MIT‑licensed node‑based visual editor for building generative image and video pipelines. It consists of a Next.js front‑end, a FastAPI back‑end that calls MuAPI’s generative models, and a reusable workflow‑builder UI library. You can self‑host it (Docker or manual) or use the hosted SaaS version, and you can extend it with custom nodes to call any AI model or external API.
NimaNzrii/comfyui-photoshop
A Photoshop plugin that integrates ComfyUI's AI capabilities directly into the Photoshop interface for a seamless AI-powered editing workflow.
chaiNNer-org/chaiNNer
A node-based GUI for programmatic image processing and AI upscaling that allows users to build customizable processing pipelines via a visual interface.
PrismML-Eng/Bonsai-Image-Demo
A cross‑platform demo for the 4 B‑parameter Bonsai diffusion model, offering a web‑based studio and CLI for image generation on macOS (MLX) or NVIDIA GPUs (gemlite + HQQ).
karimz1/imgcompress
A self-hosted image processing server that provides local AI-powered background removal, bulk compression, and conversion for over 70 image formats.
trueai-org/midjourney-proxy
An open-source API proxy for Midjourney that allows developers to integrate Midjourney's image generation and face-swapping capabilities into their own applications via a programmable interface.
YanWenKun/ComfyUI-Windows-Portable
A ready‑to‑run Windows package that bundles ComfyUI (a node‑based UI for Stable Diffusion) with 40+ custom nodes and 300+ pre‑compiled Python packages, letting users launch AI image/video generation instantly on an NVIDIA GPU.
Comfy-Org/comfy-cli
comfy‑cli is a Python command‑line tool that installs, launches, and manages ComfyUI (an open‑source generative‑media engine). It handles environment setup, custom‑node installation, model downloads, workflow execution, and can offload jobs to the hosted Comfy Cloud service. Features include fast `uv`‑based dependency resolution, JSON‑structured output for LLM agents, shell completion, and utilities for testing PRs and managing background servers.
FireRedTeam/FireRed-Image-Edit
A general-purpose image editing model that provides high-fidelity, consistent edits with strong identity preservation and multi-element fusion capabilities.
NVIDIA-RTX/NRD
GPU‑accelerated, API‑agnostic library for real‑time ray‑tracing denoising (REBLUR, RELAX, SIGMA).
GiMi-Xiaomi/gimi-illustration-skill
An AI Agent Skill that converts text articles and scripts into conceptual illustrations using customizable styles and consistent character IPs.
LibrePhotos/librephotos
LibrePhotos is a self‑hosted photo‑management server that automatically adds AI‑generated metadata (faces, objects, captions) and offers semantic search, map view, and multi‑user albums. It ships as Docker containers with a Django REST backend, React web UI, and optional React‑Native Android client, using ONNX‑run models for all computer‑vision tasks.
rupeshs/fastsdcpu
A high-performance implementation of Stable Diffusion optimized for CPUs, utilizing OpenVINO and Latent Consistency Models to enable near real-time image generation.
nv-tlabs/PiD
PiD is a plug-and-play diffusion decoder that replaces standard VAE/RAE decoders to turn latent representations directly into super-resolved 2K to 4K pixels in a single pass.
Gourieff/ComfyUI-ReActor
A set of ComfyUI nodes for fast and simple face swapping in images and videos, featuring face restoration and blended face model creation.
Nerogar/OneTrainer
A comprehensive training suite for diffusion models that provides tools for dataset preparation, fine-tuning, and model conversion through a GUI or CLI.
yurijmikhalevich/rclip
rclip is a local, terminal‑native image search tool that uses OpenCLIP embeddings to let you find photos by natural‑language description, example image, or weighted combinations, with an optional interactive UI and support for many image formats.
markfulton/NanoBananaEditor
A React and TypeScript-based editor for Google's Gemini image models that enables text-to-image generation, masked editing, and search-grounded prompts.
nihui/zimage-ncnn-vulkan
A portable ncnn-based implementation of the Z-Image generator that enables high-performance image generation, inpainting, and ControlNet support on any Vulkan-capable GPU.