freestylefly/awesome-gpt-image-2
A structured prompt engine and template library that converts prose-style AI image prompts into reusable, code-like protocols for stable and controllable image generation.
abi/screenshot-to-code
screenshot‑to‑code is an open‑source app that converts UI screenshots, mock‑ups, or short screen recordings into ready‑to‑run front‑end code (HTML/CSS, React, Vue, etc.) using LLMs such as Gemini, GPT‑5.5, and Claude, with optional asset extraction and visual preview.
img2threejs/img2threejs
img2threejs is an open‑source skill that lets an LLM (Claude, Codex, etc.) turn a single reference image into a fully procedural, animation‑ready Three.js model written in TypeScript. It uses a deterministic Python pipeline to build a detailed spec, runs strict quality‑gates, and only spends AI tokens on visual judgment, making the reconstruction token‑efficient and fully editable code.
yanliudesign/mono-color-skill
A design system and AI skill for creating professional editorial prints using limited-ink palettes and mechanical reproduction techniques like risograph and halftone.
s1dashu/ip-as-logo-skill
An Agent Skill that guides AI agents to generate simple, cute, and commercially viable IP mascot characters with consistent composition and style.
GangTailorUpgrade/undress-service
Dress AI Service is an open‑source, self‑hosted FastAPI application that lets you upload photos of your clothes, auto‑tags them with CLIP, generates outfit recommendations using rule‑based logic plus optional LLM reasoning, and visualises the looks with Stable Diffusion/FLUX. It runs via Docker or a local Python environment, stores data in SQLite/PostgreSQL, and includes weather‑aware styling, analytics, and a lightweight HTML/Tailwind UI.
upscayl/upscayl
A free and open-source AI image upscaler that uses Real-ESRGAN and Vulkan to enlarge and enhance low-resolution images without losing quality.
LiamGvchi/gc-minimal-zine-poster
A Codex skill that transforms themes or photos into minimal, paper-textured editorial posters using a structured prompt-compilation system.
ShuaixinHuang/image-multiple-angles-3d-camera
A 3D-controlled image generation tool that allows users to change the camera viewpoint of an uploaded image using an interactive 3D widget and the Qwen-Image-Edit-Plus model.
wuyoscar/GPT-Image2-Skill
A prompt gallery and toolset for GPT Image 2, featuring a CLI and agent skills for generating, editing, and reverse-engineering image prompts.
carolinaaafy/travel-memory-sticker-card
A Codex skill that transforms travel photos into collectible digital memory sticker cards.
HalfAI1102/anthropic-art
An agent skill that generates hand-drawn editorial illustrations in the Anthropic visual style, transforming abstract concepts into minimalist visual metaphors.
basketikun/chatgpt2api
An OpenAI-compatible API proxy that reverse-engineers ChatGPT's web capabilities to provide image generation and editing services with built-in account pool management.
shitagaki-lab/see-through
A framework that decomposes single anime illustrations into up to 23 semantically distinct, inpainted layers exported as a PSD for 2.5D animation and rigging.
T8RIN/ImageToolbox
An Android image editing tool that provides over 500 filters, AI-powered background removal, and batch processing for efficient photo manipulation.
tracefinity/tracefinity
Tracefinity is an AI‑powered web app (Docker‑ready) that turns photos of tools on a sheet of paper into custom Gridfinity 3‑D‑printable bins. It uses local ONNX models or remote Gemini/Replicate services to generate tool silhouettes, lets you edit and arrange them, and exports STL/3MF files for printing.
danielgatis/rembg
A tool for removing image backgrounds that can be used as a CLI, Python library, or HTTP server, supporting various AI models and hardware acceleration.
Comfy-Org/ComfyUI-Manager
ComfyUI‑Manager is a feature‑rich extension for the visual AI workflow tool ComfyUI. It adds a UI button that lets users browse, install, update, enable/disable, and remove custom nodes and model files from the official registry or community channels. The manager also provides snapshot/restore of installation states, one‑click sharing to workflow‑hosting sites, a CLI (`cm‑cli`) for automation, and security‑focused storage of its data. Configuration is handled via a simple `config.ini` and optional pip‑override files. Installation works via manual git clone, portable‑Windows script, `comfy-cli`, or a Linux venv script.
invoke-ai/InvokeAI
Invoke AI is an open‑source, locally‑run visual‑media generation platform. It provides a web‑based UI with a unified canvas, node‑based workflow editor, gallery management, and support for dozens of modern text‑to‑image models (Stable Diffusion, Flux, Qwen‑Image, etc.). Users install via the provided launcher, run a local server, and create or refine images through an intuitive browser interface. The project targets artists, designers, and developers who need a flexible, extensible AI image creation tool.
threerocks/hand-drawn-styles
A tool-agnostic library of 19 verified hand-drawn style prompt recipes that AI Agents can use to generate consistent, high-quality image prompts.
ciddwd/overlay-translator
Screen Translator is an Android app that captures the screen, runs OCR (ML Kit, PaddleOCR, manga‑OCR, or cloud services), translates the extracted text using local models, LLMs, or commercial MT APIs, and overlays the translation (or reads it aloud) in a floating window. It supports real‑time translation for games, comics, and visual novels, batch processing of images, word‑level lookup, looped auto‑translation, and extensive UI customization, all under an Apache‑2.0 license.
facefusion/facefusion
FaceFusion is a high-performance face manipulation platform for swapping and editing faces in images and videos.
xororz/local-dream
An Android application that enables local Stable Diffusion image generation with hardware acceleration for Snapdragon NPUs.
CookSleep/gpt_image_playground
A React/TypeScript web UI for OpenAI‑compatible image generation (gpt‑image‑2). Supports text‑to‑image, reference‑image editing, batch/streaming, transparent backgrounds, multi‑turn Agent chat, local IndexedDB history, multiple API profiles, and can be deployed on Vercel, GitHub Pages, Cloudflare Workers, Docker, or locally.
dacnay816y62-hub/photo-revival
A set of AI skill instructions that transforms ordinary photos into minimalist, poetic hand-drawn illustrations with heavy white space and tiny handwritten captions.
nuyoah-ai-works/nuyoah-xiezhen-prompt
A prompt engineering skill for AI agents to generate professional portrait and fashion photography prompts with a focus on realistic skin and lighting.
GiMi-Xiaomi/gimi-illustration-skill
An AI Agent Skill that converts text articles and scripts into conceptual illustrations using customizable styles and consistent character IPs.
lbouaraba/comfyui-krea2edit
A ComfyUI node pack for instruction-based image editing using Krea 2, enabling high-fidelity identity preservation through dual latent and semantic conditioning.
Nutlope/logocreator
LogoCreator is an open‑source web app that generates brand‑ready logos using the FLUX‑2 image model via Together AI. Built with Next.js, TypeScript, Radix, and Tailwind, it runs without an account—just a Together AI key. Features include instant generation, FLUX‑1 edits, PNG/SVG export, optional rate‑limiting and auth, and a history dashboard. The repo provides clear setup steps and lists future tasks such as size selection, cost preview, reference‑logo uploads, and a showcase of redesigned famous logos.
kijai/ComfyUI-KJNodes
ComfyUI‑KJNodes is a plug‑in for the visual Stable Diffusion UI (ComfyUI) that adds utility nodes, cross‑sub‑graph Set/Get handling, and many keyboard/mouse shortcuts to make large generation graphs easier to build and maintain.
photoprism/photoprism
PhotoPrism is an open‑source, self‑hosted web app that uses on‑device AI to automatically tag, search, and organize personal photos and videos while keeping all data private.
aldegad/sprite-gen
A pipeline for turning a single image into game-ready sprite atlases, featuring identity-locking generation, background removal, and a tool to animate still frames into breathing loops.
liyue-aigc/xianxia-visual-director
A visual direction system for Codex/Agent skills that transforms simple ideas into structured, cinematic image prompts for monumental Eastern xianxia environments.
krea-ai/krea-2
Krea 2 is an open-source image generation model featuring a base RAW model for flexible fine-tuning and a distilled Turbo model for fast, high-quality text-to-image synthesis.
zuruoke/watermark-removal
A machine learning project that uses image inpainting to remove watermarks from images, creating seamless results.
AHEKOT/ComfyUI_VNCCS
A ComfyUI pipeline for creating consistent character sprites, allowing users to manage appearance, poses, clothing, and emotions across multiple images.
AgriciDaniel/banana-claude
Banana Claude is a Claude Code plugin that lets you write a natural‑language request, see a detailed plan (prompt, model, cost estimate) and, after an explicit approval, call Google’s Gemini image models to generate or edit visual assets. It emphasizes offline‑first planning, privacy‑safe API key handling, and one‑time approvals to keep costs transparent.
willmiao/ComfyUI-Lora-Manager
A comprehensive management tool for ComfyUI that streamlines the organization, downloading from CivitAI, and application of LoRA models and checkpoints.
wnby/photo-relic-editorial
A Codex skill that transforms photographs into vertical editorial artworks by pairing a real photo with a restrained, paper-textured printmaking version of the same image.
Acly/krita-ai-diffusion
A Krita plugin that integrates generative AI diffusion models into the painting workflow, offering tools for inpainting, live painting, and precise structural control.
Hugo-Dz/spritefusion-pixel-snapper
A tool that snaps pixels to a perfect grid and quantizes colors to fix messy, inconsistent pixel art generated by AI.
VigoZhao/AI-Visual-Prompt-Cookbook
A curated library of 124 structured JSON prompt templates that allow users to maintain consistent, professional visual styles across AI image generation workflows.
TaiT-tt/tait-crt-interface-skill
A Codex image generation skill that transforms photos or text into retro illustrations with the aesthetic of early CRT computer interfaces.
wzj177/ecommerce-image-suite
An AI-powered toolkit for e-commerce sellers to automatically analyze product photos and generate a complete set of professional marketing images across various styles and platforms.
jtydhr88/ComfyUI-See-through
A ComfyUI plugin that decomposes single anime illustrations into layered 2.5D models with depth ordering, specifically designed for Live2D workflows.
KohakuBlueleaf/LyCORIS
A library implementing various parameter-efficient fine-tuning algorithms like LoHa and LoKr for Stable Diffusion, enabling high-quality model customization with minimal storage and compute.
TheJoeFin/Text-Grab
Text Grab is a Windows‑only desktop app that captures any visible text (screenshots, PDFs, UI elements) using local OCR (WinAI, WinRT OCR, or Tesseract) and provides built‑in cleanup, spreadsheet editing, regex‑based extraction, reusable grab templates, bulk folder processing, and a Chrome/Edge extension. All processing stays on‑device, with optional NPU‑accelerated inference on Copilot+ PCs. Install via Microsoft Store, GitHub releases, or package managers; source can be built with Visual Studio or the .NET SDK.
allenk/GeminiWatermarkTool
A cross‑platform, portable executable that removes Google Gemini image (and video) watermarks using a deterministic reverse‑alpha‑blending algorithm, with optional GPU‑accelerated FDnCNN denoising, GUI/CLI interfaces, batch processing, and AI‑agent integration via MCP.