Computer Vision Faculty Publications

Highly Accurate Dichotomous Image Segmentation

Xuebin Qin, Mohamed bin Zayed University of Artificial Intelligence
Hang Dai, Mohamed bin Zayed University of Artificial IntelligenceFollow
Xiaobin Hu, TUM Munich, Germany
Deng-Ping Fan, ETH Zurich, Switzerland
Luc Van Gool, NCAI SDAIA, KSA
Ling Shao, ETH Zurich, SwitzerlandFollow

Document Type

Conference Proceeding

Publication Title

Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)

Abstract

We present a systematic study on a new task called dichotomous image segmentation (DIS), which aims to segment highly accurate objects from natural images. To this end, we collected the first large-scale DIS dataset, called DIS5K, which contains 5,470 high-resolution (e.g., 2K, 4K or larger) images covering camouflaged, salient, or meticulous objects in various backgrounds. DIS is annotated with extremely fine-grained labels. Besides, we introduce a simple intermediate supervision baseline (IS-Net) using both feature-level and mask-level guidance for DIS model training. IS-Net outperforms various cutting-edge baselines on the proposed DIS5K, making it a general self-learned supervision network that can facilitate future research in DIS. Further, we design a new metric called human correction efforts (HCE) which approximates the number of mouse clicking operations required to correct the false positives and false negatives. HCE is utilized to measure the gap between models and real-world applications and thus can complement existing metrics. Finally, we conduct the largest-scale benchmark, evaluating 16 representative segmentation models, providing a more insightful discussion regarding object complexities, and showing several potential applications (e.g., background removal, art design, 3D reconstruction). Hoping these efforts can open up promising directions for both academic and industries. Project page: https://xuebinqin.github.io/dis/index.html. © 2022, The Author(s), under exclusive license to Springer Nature Switzerland AG.

First Page

Last Page

DOI

10.1007/978-3-031-19797-0_3

Publication Date

11-3-2022

Keywords

Benchmarking, Large dataset, Mammals, Three dimensional computer graphics, Correction efforts, Fine grained, High resolution, Highly accurate, Images segmentations, Large images, Large-scale datasets, Natural images, Simple++, Systematic study, Image segmentation, Computer Vision and Pattern Recognition (cs.CV)

Comments

IR Deposit conditions: non-described

Recommended Citation

X. Qin, H. Dai, X. Hu, D.P. Fan, L.V. Gool and L. Shao, "Highly Accurate Dichotomous Image Segmentation", in Computer Vision (ECCV 2022),, Lecture Notes in Computer Science, vol 13678, pp. 38-56, Nov. 2022, doi:10.1007/978-3-031-19797-0_3

Additional Links

Preprint available in arXiv: https://arxiv.org/abs/2203.03041

Link to Full Text

COinS

Computer Vision Faculty Publications

Highly Accurate Dichotomous Image Segmentation

Document Type

Publication Title

Abstract

First Page

Last Page

DOI

Publication Date

Keywords

Comments

Recommended Citation

Additional Links

Browse

Contribute

Links

Computer Vision Faculty Publications

Highly Accurate Dichotomous Image Segmentation

Authors

Document Type

Publication Title

Abstract

First Page

Last Page

DOI

Publication Date

Keywords

Comments

Recommended Citation

Additional Links

Share

Browse

Contribute

Links