Image Editing90 / 393 projects
Image editing model leaderboard featuring the hottest AI image editing and generation projects on GitHub, including inpainting, style transfer, and upscaling.
GFPGAN aims at developing Practical Algorithms for Real-world Face Restoration.
Real-ESRGAN aims at developing Practical Algorithms for General Image/Video Restoration.
Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest AI-driven technologies. The solution offers an industry leading WebUI, and serves as the foundation for multiple commercial products.
[NeurIPS 2022] Towards Robust Blind Face Restoration with Codebook Lookup Transformer
Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything
🦙 LaMa Image Inpainting, Resolution-robust Large Mask Inpainting with Fourier Convolutions, WACV 2022
PaddlePaddle GAN library, including lots of interesting applications like First-Order motion transfer, Wav2Lip, picture repair, image editing, photo2cartoon, image style transfer, GPEN, and so on.
OpenMMLab Multimodal Advanced, Generative, and Intelligent Creation Toolbox. Unlock the magic 🪄: Generative-AI (AIGC), easy-to-use APIs, awsome model zoo, diffusion models, for text-to-image generation, image/video restoration/enhancement, etc.
a cross-platform image super-resolution tool
SUPIR aims at developing Practical Algorithms for Photo-Realistic Image Restoration In the Wild. Our new online demo is also released at suppixel.ai.
Unofficial Implementation of DragGAN - "Drag Your GAN: Interactive Point-based Manipulation on the Generative Image Manifold" (DragGAN 全功能实现,在线Demo,本地部署试用,代码、模型已全部开源,支持Windows, macOS, Linux)
🔎 Super-scale your images and run experiments with Residual Dense and Adversarial Networks.
Official implementations for paper: Anydoor: zero-shot object-level image customization
Unofficial implementation of Image Super-Resolution via Iterative Refinement by Pytorch
Outpainting with Stable Diffusion on an infinite canvas
Photo-Realistic Single Image Super-Resolution Using a Generative Adversarial Network
Official pytorch implementation of the paper: "SinGAN: Learning a Generative Model from a Single Natural Image"
Kandinsky 2 — multilingual text2image latent diffusion model
🔥 [ICCV 2025 Highlight] InfiniteYou: Flexible Photo Recrafting While Preserving Your Identity
[IJCV2024] Exploiting Diffusion Prior for Real-World Image Super-Resolution
[NeurIPS 2025] Image editing is worth a single LoRA! 0.1% training data for fantastic image editing! Surpasses GPT-4o in ID persistence~ MoE ckpt released! Only 4GB VRAM is enough to run!
This project is the official implementation of 'DreamOmni2: Multimodal Instruction-based Editing and Generation (CVPR2026 Highlight)''
A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.
Inpaint Anything extension performs stable diffusion inpainting on a browser UI using masks from Segment Anything.
Paint by Example: Exemplar-based Image Editing with Diffusion Models
A PyTorch implementation of SRGAN based on CVPR 2017 paper "Photo-Realistic Single Image Super-Resolution Using a Generative Adversarial Network"
Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer.
Nodes for better inpainting with ComfyUI: Fooocus inpaint model for SDXL, LaMa, MAT, and various other tools for pre-filling inpaint & outpaint areas.
PyTorch implementation for SDEdit: Image Synthesis and Editing with Stochastic Differential Equations
[ICCV 2023 Oral] "FateZero: Fusing Attentions for Zero-shot Text-based Video Editing"
[ECCV 2024] PowerPaint, a versatile image inpainting model that supports text-guided object inpainting, object removal, image outpainting and shape-guided object inpainting with only a single model. 一个高质量多功能的图像修补模型,可以同时支持插入物体、移除物体、图像扩展、形状可控的物体生成,只需要一个模型
Fast and Accurate One-Stage Space-Time Video Super-Resolution (accepted in CVPR 2020)
Unofficial implementation of "Image Inpainting for Irregular Holes Using Partial Convolutions". Try at: www.fixmyphoto.ai
UniWorld: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation
Open-source SOTA multi-image editing model
Tensorflow implementation of the SRGAN algorithm for single image super-resolution
[NeurIPS 2025] 4KAgent: Agentic Any Image to 4K Super-Resolution. An intelligent computer vision agent that can magically restore any image to perfect-4K!
[CVPR 2021] Anycost GANs for Interactive Image Synthesis and Editing
[ICML 2024] MagicPose(also known as MagicDance): Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion
[ECCV 2024] InstructIR: High-Quality Image Restoration Following Human Instructions https://huggingface.co/spaces/marcosv/InstructIR
[ICLR 2026] ChronoEdit: Towards Temporal Reasoning for Image Editing and World Simulation
[ECCV 2026] SparkVSR: Interactive Video Super-Resolution via Sparse Keyframe Propagation
[CVPR 2019]: Pluralistic Image Completion
[NeurIPS 2022] Denoising Diffusion Restoration Models -- Official Code Repository
Flash Diffusion — accelerating conditional diffusion models (AAAI 2025 Oral)
Code and data for "AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks" [TMLR 2024]
Official repository for the paper "High-Resolution Daytime Translation Without Domain Labels" (CVPR2020, Oral)
[CVPR2024] SeeSR: Towards Semantics-Aware Real-World Image Super-Resolution
This project is the official implementation of 'Diffir: Efficient diffusion model for image restoration', ICCV2023
ComfyUI adaptation of IDM-VTON for virtual try-on.
Official Code for DiffMorpher: Unleashing the Capability of Diffusion Models for Image Morphing (CVPR 2024)
An unified model that seamlessly integrates multimodal understanding, text-to-image generation, and image editing within a single powerful framework.
PyTorch implementation of "Improved Techniques for Training Single-Image GANs" (WACV-21)
[CVPR 2023] Collaborative Diffusion
[CVPR 2024] Official implementation of FreeDrag: Feature Dragging for Reliable Point-based Image Editing
[ICLR 2026] Official repo of paper "Reconstruction Alignment Improves Unified Multimodal Models". Unlocking the Massive Zero-shot Potential in Unified Multimodal Models through Self-supervised Learning.
Official implementation for "Stable Flow: Vital Layers for Training-Free Image Editing" [CVPR 2025]
[ICLR2024] Official repo for paper "PnP Inversion: Boosting Diffusion-based Editing with 3 Lines of Code"
[CVPR 2025] Official code repository for "Pixel-level and Semantic-level Adjustable Super-resolution: A Dual-LoRA Approach"
[ICLR 2025] Official Implementation of Meissonic: Revitalizing Masked Generative Transformers for Efficient High-Resolution Text-to-Image Synthesis
Video-Inpaint-Anything: This is the inference code for our paper CoCoCo: Improving Text-Guided Video Inpainting for Better Consistency, Controllability and Compatibility.
MinImagen: A minimal implementation of the Imagen text-to-image model
Calligrapher: Freestyle Text Image Customization
[ICML2025] An 8-step inversion and 8-step editing process works effectively with the FLUX-dev model. (3x speedup with results that are comparable or even superior to baseline methods)
Reference code for the paper HistoGAN: Controlling Colors of GAN-Generated and Real Images via Color Histograms (CVPR 2021).
A paper collection of recent diffusion models for text-image generation tasks, e,g., visual text generation, font generation, text removal, text image super resolution, text editing, handwritten generation, scene text recognition and scene text detection.
Officail Implementation for "ReNoise: Real Image Inversion Through Iterative Noising"
[CVPR 2025] FaithDiff for Classic Film Rejuvenation, Old Photo Revival, Social Media Restoration, Image Enhancement and AIGC Enhancement.
Deep Learning Inferred Multiplex ImmunoFluorescence for IHC Image Quantification (https://deepliif.org) [Nature Machine Intelligence'22, CVPR'22, MICCAI'23, Histopathology'23, MICCAI'24]
Official repository for "CFG++: manifold-constrained classifier free guidance for diffusion models" (ICLR2025)
state-of-the-art content-preserving style transfer
Codebase for performing various experiments with Stable Diffusion, supported by the diffusers library.
CoMoGAN: continuous model-guided image-to-image translation. CVPR 2021 oral.
[CVPR 2023] Ref-NPR: Reference-Based Non-PhotoRealistic Radiance Fields
This project is the official implementation of 'Basic Binary Convolution Unit for Binarized Image Restoration Network', ICLR2023
Code for You Only Cut Once: Boosting Data Augmentation with a Single Cut, ICML 2022.
Official TensorFlow code for paper "Multi-Image Super Resolution of Remotely Sensed Images Using Residual Attention Deep Neural Networks".
The official repository of BFSR: "Boosting Flow-based Generative Super-Resolution Models via Learned Prior" [CVPR 2024]
PyTorch implementation of DiffRoll, a diffusion-based generative automatic music transcription (AMT) model
HiDream O1 Image nodes + LoRA training for ComfyUI
Compose Multiplatform app generates images using Stability AI
Official Implementation for "Guide-and-Rescale: Self-Guidance Mechanism for Effective Tuning-Free Real Image Editing"
An easy-to-use image editor extension for Stable Diffusion Web UI
Self-hosted GPT Image / OpenAI-compatible WebUI for text-to-image, image editing, reference-based generation, and multilingual creative workflows.
Self-hosted GPT Image 2 / Image-2 workbench for teams: text-to-image, image-to-image, templates, quotas, history and admin dashboard.
ComfyUI custom nodes for GPT-Image-2 image generation via muapi.ai — text-to-image and image-to-image
Python wrapper for ByteDance's Seedream 5.0 Pro API — 4K text-to-image generation, natural-language image editing, precise typography, and consistent character generation.
Local AI harness/workstation: local AI image editor, Ollama image generation routing, SDXL inpainting UI, CivitAI model imports, 8GB VRAM Stable Diffusion.
[WACV 2025] I Dream My Painting: Connecting MLLMs and Diffusion Models via Prompt Generation for Text-Guided Multi-Mask Inpainting
Source: GitHub API · Curated · Realtime