AI Image Tools51 / 329 projects
AI image tool leaderboard featuring the hottest AI image processing and generation application projects on GitHub.
Stable Diffusion web UI
Your AI second brain. Self-hostable. Get answers from the web or your docs. Build custom agents, schedule automations, do deep research. Turn any online or local LLM into your personal, autonomous AI (gpt, claude, gemini, llama, qwen, mistral). Get started - free.
🚀 AI 全自动短视频引擎 | AI Fully Automated Short Video Engine
Toonflow 是开源一站式 AI 短剧创作工具,将小说、剧本快速转化为动画短剧。集成 AI 编剧、智能分镜、角色与视频生成,跨平台桌面端轻量部署,助力创作者低成本批量产出视觉内容。Toonflow is an open-source AI tool that turns stories and scripts into animated short dramas. Features AI scriptwriting, storyboarding, character and video generation. A cross-platform desktop app for efficient content creation.
Stable Diffusion built-in to Blender
SwarmUI (formerly StableSwarmUI), A Modular Stable Diffusion Web-User-Interface, with an emphasis on making powertools easily accessible, high performance, and extensibility.
基于 OpenAI gpt-image-2 API 的图片生成与编辑工具
An extensible, easy-to-use, and portable diffusion web UI 👨🎨
A simple standalone viewer for reading prompts from Stable Diffusion generated image outside the webui.
Offline inference engine for art, real-time voice conversations, LLM powered chatbots and automated workflows
High quality training free inpaint for every stable diffusion model. Supports ComfyUI
Glisp is a Lisp-based design tool that combines generative approaches with traditional design methods, empowering artists to discover new forms of expression.
Seamlessly extend any image in any direction with AI. Open-source web app powered by Gemini via OpenRouter, with Poisson-blended seams and best-of-3 variant picker.
Plug-and-play local AI studio: uncensored chat, image & video generation, coding agent. Runs abliterated LLMs + ComfyUI 100% offline. One installer, no Docker, no cloud.
Uncensored local AI studio for Windows, Linux, and macOS. Zero-setup GUI for Image Generation, GGUF LLMs, Text to Speech & Speech to Text
🎨 Open-source AI slide studio inside Codex: image-native decks, every slide a full visual canvas. ⚡ 10+ high-quality slides in ~4–5 minutes — Fast mode renders every page in parallel. 🔍 Watch the whole chain live: research → outline → style → render → edit → present → export PDF/PPTX. 🖥️ Browser-first · zero API keys · durable projects.
面向 GPT-image-2 的 AI 图片生成 WebUI 工作台,支持 Codex Responses 与 OpenAI 兼容 API 接入,内置公用图库、多类型 Chip 快捷引用、提示词模板、多任务并发和本地队列管理。An AI image generation WebUI workbench for GPT-image-2 with Codex Responses and OpenAI-compatible API support, shared gallery references, multi-type quick chips, prompt templates, concurrent tasks, and local queue management.
Minimal CLI + web UI for OpenAI GPT Image 2 generation. Dual auth: API Key (paid) or OAuth via ChatGPT (free). Text-to-image, image-to-image, parallel gen, custom sizes.
面向 GPT-image-2 的 AI 图片生成 WebUI 工作台,支持 Codex Responses 与 OpenAI 兼容 API 接入,内置公用图库、多类型 Chip 快捷引用、提示词模板、多任务并发和本地队列管理。An AI image generation WebUI workbench for GPT-image-2 with Codex Responses and OpenAI-compatible API support, shared gallery references, multi-type quick chips, prompt templates, concurrent tasks, and local queue management.
For automating the creation of large batches of AI-generated artwork locally.
Lightweight Stable Diffusion v 2.1 web UI: txt2img, img2img, depth2img, inpaint and upscale4x.
Character Select Stand Alone App with AI prompt and ComfyUI/WebUI API support for wai-il model
Free, open-source alternative to Weavy AI, Krea Nodes, Freepik Spaces & FloraFauna AI — node-based AI workflow builder for generative image & video pipelines
Multimodal AI Story Teller, built with Stable Diffusion, GPT, and neural text-to-speech
Multi-threaded GUI manager for mass creation of AI-generated art with support for multiple GPUs.
ChatGPT-Pro is an advanced application that combines the power of ChatGPT and DALL.E.
Open-source AI video workbench. Bring any model or your local ComfyUI, and let Claude Code / Codex / Cursor direct it over MCP — storyboard, references, generation, editable first cut on a real timeline. Local-first: projects, prompts, and keys stay on your machine. No account, no telemetry.
Consumer AI app for chat, image generation, video generation, and music creation powered by Ace Data Cloud APIs.
Create and customize your AI influencer open-source
Inpaint Anything performs stable diffusion inpainting on a browser UI using masks from Segment Anything.
Open-source clone of the MidJourney web interface featuring real AI image and video generation powered by Google's Gemini SDK. Use Imagen 4 to generate images and Veo 2 and 3 for image and text to video with audio.
Code for Papeg.ai
A Python library for efficient image generation using CSS Flexbox
The creative suite for character-driven AI experiences.
Edit Videos and Design Images with Claude code or Codex
AI视觉创作套件,集图解、绘画、编辑和视频生成于一体。支持Vercel一键部署
This app uses the OpenAISwift library, ChatGPTSwift library and OpenAI library to communicate with the popular ChatGPT artificial intelligence. The app allows you to have a quick message exchange with a simple and clean interface, but with useful features.
Genius - A Modern Next.js 14 SaaS AI Platform.
多模型 AI 绘图工作台,支持 Midjourney、Gemini、Flux、DALL-E、GPT-4o、Grok、通义万相等主流图像生成模型。
Free AI Image Generator API — Nano Banana Pro (Gemini 3.0 Pro). Create ultra-realistic images, edit photos, and generate stunning AI visuals in seconds — 100% free, no watermark. The ultimate AI Image solution for creators and developers.
Quote2Image is a python library for turning text quotes into graphical images
Open-source, customizable frontend for Venice AI. Chat, image gen, audio, video, embeddings + visual workflows — all in one UI. Your API key, your browser, no backend.
AI Image Generator is a powerful and user-friendly desktop application designed to help creators produce stunning, high-quality artwork, photorealistic renders, and creative concepts quickly and easily.
An AI image generation frontend focused on ease of use, versatility and capability for professional uses, packaged as a click-and-run executable.
Open-source Nano Banana image generator — production-ready Next.js SaaS for text-to-image and multi-image reference editing. Stripe billing, credits, NextAuth, and Prisma out of the box.
Generate text to image in Golang. I created this application for generating featured images for facebook or when sharing code snippet in WhatsApp.This comes with a simple web interface that lets you generate text to image.
Create high-quality images with quotes (Perfect for Instagram and Pinterest) in less than 5 seconds per 100+ images!
A Telegram bot that is capable of running stable diffusion Models (Support all model from huggingface) and generate images based on user prompts. Loaded with 20 Models .
A full prompt-writing studio in a single ComfyUI node - AI editing, style targeting, prompt browsing, and wireless prompt injection, powered by your local Ollama models. No cloud, no keys.
AI-driven image & avatar creator for WordPress, powered by DALL·E. Generate unique, royalty-free images, variations, and avatars with seamless integration to popular page builders. Compatible with Gutenberg, Elementor, Beaver Builder, WP Bakery, and Woocommerce.
Reproducible sketch-to-render studies with ControlNet: seed scouting, variation, export, replay, and a local Studio.
Source: GitHub API · Curated · Realtime