Diffusion for Natural Image Matting

被引:0
|
作者
Hu, Yihan [1 ,2 ,5 ]
Lin, Yiheng [1 ,2 ]
Wang, Wei [1 ,2 ]
Zhao, Yao [1 ,2 ,3 ]
Wei, Yunchao [1 ,2 ,3 ]
Shi, Humphrey [4 ,5 ]
机构
[1] Beijing Jiaotong Univ, Inst Informat Sci, Beijing, Peoples R China
[2] Minist Educ, Visual Intelligence X Int Joint Lab, Beijing, Peoples R China
[3] Pengcheng Lab, Shenzhen, Peoples R China
[4] Georgia Inst Technol, Atlanta, GA 30332 USA
[5] Picsart AI Res PAIR, Atlanta, GA USA
来源
COMPUTER VISION-ECCV 2024, PT LVII | 2025年 / 15115卷
关键词
Image matting; Diffusion process; Iterative refinement;
D O I
10.1007/978-3-031-72998-0_11
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Existing natural image matting algorithms inevitably have flaws in their predictions on difficult cases, and their one-step prediction manner cannot further correct these errors. In this paper, we investigate a multi-step iterative approach for the first time to tackle the challenging natural image matting task, and achieve excellent performance by introducing a pixel-level denoising diffusion method (DiffMatte) for the alpha matte refinement. To improve iteration efficiency, we design a lightweight diffusion decoder as the only iterative component to directly denoise the alpha matte, saving the huge computational overhead of repeatedly encoding matting features. We also propose an ameliorated self-aligned strategy to consolidate the performance gains brought about by the iterative diffusion process. This allows the model to adapt to various types of errors by aligning the noisy samples used in training and inference, mitigating performance degradation caused by sampling drift. Extensive experimental results demonstrate that DiffMatte not only reaches the state-of-the-art level on the mainstream Composition-1k test set, surpassing the previous best methods by 8% and 15% in the SAD metric and MSE metric respectively, but also show stronger generalization ability in other benchmarks. The code will be open-sourced for the following research and applications. Code is available at https://github.com/YihanHu-2022/DiffMatte.
引用
收藏
页码:181 / 199
页数:19
相关论文
共 50 条
  • [1] Natural Image Matting with Attended Global Context
    Yi-Yi Zhang
    Li Niu
    Yasushi Makihara
    Jian-Fu Zhang
    Wei-Jie Zhao
    Yasushi Yagi
    Li-Qing Zhang
    Journal of Computer Science and Technology, 2023, 38 : 659 - 673
  • [2] Natural image matting based on surrogate model
    Liang, Yihui
    Gou, Hongshan
    Feng, Fujian
    Liu, Guisong
    Huang, Han
    APPLIED SOFT COMPUTING, 2023, 143
  • [3] Natural Image Matting Using HSI Framework
    Khandelwal, Vineet
    Gupta, Abhinav
    Kashyap, Manish
    Gandhi, Hitesh
    Dhawan, Aishwar
    2009 2ND IEEE INTERNATIONAL CONFERENCE ON COMPUTER SCIENCE AND INFORMATION TECHNOLOGY, VOL 2, 2009, : 141 - 144
  • [4] Natural Image Matting with Attended Global Context
    Zhang, Yi-Yi
    Niu, Li
    Makihara, Yasushi
    Zhang, Jian-Fu
    Zhao, Wei-Jie
    Yagi, Yasushi
    Zhang, Li-Qing
    JOURNAL OF COMPUTER SCIENCE AND TECHNOLOGY, 2023, 38 (03) : 659 - 673
  • [5] Robust natural image matting approach based on strokes
    Wu, Yue
    He, Fazhi
    Zhang, Dengyi
    Wei, Lingyun
    Huang, Zhiyong
    MIPPR 2007: MEDICAL IMAGING, PARALLEL PROCESSING OF IMAGES, AND OPTIMIZATION TECHNIQUES, 2007, 6789
  • [6] Natural Image Matting through Overlapping Neighborhood Propagation
    Peng, Hongjing
    Duan, Jiang
    Yuan, Jianhua
    Shao, Dinghong
    FOURTH INTERNATIONAL CONFERENCE ON MACHINE VISION (ICMV 2011): COMPUTER VISION AND IMAGE ANALYSIS: PATTERN RECOGNITION AND BASIC TECHNOLOGIES, 2012, 8350
  • [7] Cross Depth Image Filter-based Natural Image Matting
    Li, Yujie
    Lu, Huimin
    Zhang, Lifeng
    Serikawa, Seiichi
    2013 14TH ACIS INTERNATIONAL CONFERENCE ON SOFTWARE ENGINEERING, ARTIFICIAL INTELLIGENCE, NETWORKING AND PARALLEL/DISTRIBUTED COMPUTING (SNPD 2013), 2013, : 601 - 604
  • [8] Automatic framework for high-efficient natural image matting
    He, Fazhi
    Wu, Yue
    Zhang, Dengyi
    Huang, Zhiyong
    Wei, Lingyun
    Xiao, Chunxia
    MIPPR 2007: MULTISPECTRAL IMAGE PROCESSING, 2007, 6787
  • [9] A Survey on Natural Image Matting With Closed-Form Solutions
    Li, Xiaoqiang
    Li, Jide
    Lu, Hong
    IEEE ACCESS, 2019, 7 : 136658 - 136675
  • [10] NATURAL IMAGE MATTING FOR MULTIPLE WIDE-BASELINE VIEWS
    Sarim, Muhammad
    Hilton, Adrian
    Guillemaut, Jean-Yves
    Takai, Takeshi
    Kim, Hansung
    2010 IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING, 2010, : 2233 - 2236