计算机视觉——图像修复综述篇

目录

[1. Deterministic Image Inpainting 判别器图像修复](#1. Deterministic Image Inpainting 判别器图像修复)

[1.1. sigle-shot framework](#1.1. sigle-shot framework)

[(1) Generators](#(1) Generators)

[(2) training objects / Loss Functions](#(2) training objects / Loss Functions)

[1.2. two-stage framework](#1.2. two-stage framework)

[2. Stochastic Image Inpainting 随机图像修复](#2. Stochastic Image Inpainting 随机图像修复)

[2.1. VAE-based methods](#2.1. VAE-based methods)

[2.2. GAN-based methods](#2.2. GAN-based methods)

[2.3. Flow-based methods](#2.3. Flow-based methods)

[2.4. MLM-based methods](#2.4. MLM-based methods)

[2.5. Diffusion model-based methods](#2.5. Diffusion model-based methods)

[3. text-guided image inpainting ⽂本引导的图像修复](#3. text-guided image inpainting ⽂本引导的图像修复)

[4. Inpainting Mask 掩码机制](#4. Inpainting Mask 掩码机制)

[(1) regular mask](#(1) regular mask)

[(2) irregular mask](#(2) irregular mask)

[5. Loss Function 损失函数](#5. Loss Function 损失函数)

[6. Dataset 图像修复领域数据集](#6. Dataset 图像修复领域数据集)

[(1) faces(CelebA & CelebA-HQ)](#(1) faces(CelebA & CelebA-HQ))

[(2) real-world encountered scenes(Places2)](#(2) real-world encountered scenes(Places2))

[(3) street scenes(Paris)](#(3) street scenes(Paris))

[(4) texture(DTD)](#(4) texture(DTD))

[(5) objects (ImageNet)](#(5) objects (ImageNet))

[7. Evaluation Protocol 评估指标](#7. Evaluation Protocol 评估指标)

[7.1. pixel-aware metrics](#7.1. pixel-aware metrics)

[7.2. (human) perception-aware metriics](#7.2. (human) perception-aware metriics)

[8. Performance Evaluation 表现评估](#8. Performance Evaluation 表现评估)

[8.1 Representative Image Inpainting Methods](#8.1 Representative Image Inpainting Methods)

[8.2 Loss Functions](#8.2 Loss Functions)

[9. Inpainting-based Application 基于图像修复的领域应⽤](#9. Inpainting-based Application 基于图像修复的领域应⽤)

[(1) Object Removal](#(1) Object Removal)

[(2) Text Editing](#(2) Text Editing)

[(3) Old Photo Restoration](#(3) Old Photo Restoration)

[(4) Image Compression](#(4) Image Compression)

[(5) Text-guided image editing](#(5) Text-guided image editing)

Reference


1. Deterministic Image Inpainting 判别器图像修复

1.1. sigle-shot framework
(1) Generators
  1. mask-aware design
  2. attention mechanism
  3. multi-scale aggregation
  4. transform domain
  5. encoder-decoder connection
  6. deep prior guidance
(2) training objects / Loss Functions
  1. Pixel-wise reconstruction loss
  2. perceptual loss
  3. style loss
  4. adversarial loss
  5. prevalent training objectives
1.2. two-stage framework

(1) coarse-to-fiine methods
(2) structure-then-texture methods

2. Stochastic Image Inpainting 随机图像修复

2.1. VAE-based methods
2.2. GAN-based methods
2.3. Flow-based methods
2.4. MLM-based methods
2.5. Diffusion model-based methods

(1) sample stratage design
(2) computational cost reduction

3. text-guided image inpainting ⽂本引导的图像修复

4. Inpainting Mask 掩码机制

(1) regular mask
(2) irregular mask

5. Loss Function 损失函数

同1-1.1-(2) training objects

6. Dataset 图像修复领域数据集

(1) faces(CelebA & CelebA-HQ)
(2) real-world encountered scenes(Places2)
(3) street scenes(Paris)
(4) texture(DTD)
(5) objects (ImageNet)

7. Evaluation Protocol 评估指标

7.1. pixel-aware metrics

focus on the precision of reconstructed pixels
(1) l1 error
(1) l2 error
(3) PSNR(peak signal-to-noise ratio)
(4) SSIM(the structure similarity index)
(5) MS-SSIM(muti-scale SSIM)

7.2. (human) perception-aware metriics

the visual perception quality
(1) FID(Frechet Inception diistance)
(2) LPIPS(learned perceptual image patch similarity)
(3) P/U-IDS(pair-unpair Inception discriminative score)

8. Performance Evaluation 表现评估

8.1 Representative Image Inpainting Methods

(1) Models: RFR, MADF, DSI, CR-Fill, CoModGAN, LGNet, RePaint
(2) Dataset: CeleBA-HQ, Places2
(3) Mask: M1, M2, M3, M4, M5, M6
(4) Metrics: l1, PSNR, SSIM, MS-SSIM, FID, LP-IPS
(5) Loss: pixes reconstruction loss, perceptual loss, resnetpl loss, style loss, stylemeanstd,
percept-style loss, lsgan

8.2 Loss Functions

同1-1.1-(2) training objects

9. Inpainting-based Application 基于图像修复的领域应⽤

(1) Object Removal
(2) Text Editing
(3) Old Photo Restoration
(4) Image Compression
(5) Text-guided image editing

Reference

  1. Deep Learning-based Image and Video Inpainting: A Survey
相关推荐
龙的爹233321 分钟前
论文翻译 | The Capacity for Moral Self-Correction in Large Language Models
人工智能·深度学习·算法·机器学习·语言模型·自然语言处理·prompt
python_知世1 小时前
2024年中国金融大模型产业发展洞察报告(附完整PDF下载)
人工智能·自然语言处理·金融·llm·计算机技术·大模型微调·大模型研究报告
Fanstay9851 小时前
人工智能技术的应用前景及其对生活和工作方式的影响
人工智能·生活
lunch( ̄︶ ̄)1 小时前
《AI 使生活更美好》
人工智能·生活
Hoper.J1 小时前
用两行命令快速搭建深度学习环境(Docker/torch2.5.1+cu118/命令行美化+插件),包含完整的 Docker 安装步骤
人工智能·深度学习·docker
Shaidou_Data1 小时前
信息技术引领未来:大数据治理的实践与挑战
大数据·人工智能·数据清洗·信息技术·数据治理技术
Elastic 中国社区官方博客1 小时前
开始使用 Elastic AI Assistant 进行可观察性和 Microsoft Azure OpenAI
大数据·人工智能·elasticsearch·microsoft·搜索引擎·全文检索·azure
qq_273900232 小时前
pytorch detach方法介绍
人工智能·pytorch·python
AI狂热爱好者2 小时前
A3超级计算机虚拟机,为大型语言模型LLM和AIGC提供强大算力支持
服务器·人工智能·ai·gpu算力
边缘计算社区2 小时前
推理计算:GPT-o1 和 AI 治理
人工智能·gpt