Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Exploring the Underwater World Segmentation without Extra Training

About

Accurate segmentation of marine organisms is vital for biodiversity monitoring and ecological assessment, yet existing datasets and models remain largely limited to terrestrial scenes. To bridge this gap, we introduce \textbf{AquaOV255}, the first large-scale and fine-grained underwater segmentation dataset containing 255 categories and over 20K images, covering diverse categories for open-vocabulary (OV) evaluation. Furthermore, we establish the first underwater OV segmentation benchmark, \textbf{UOVSBench}, by integrating AquaOV255 with five additional underwater datasets to enable comprehensive evaluation. Alongside, we present \textbf{Earth2Ocean}, a training-free OV segmentation framework that transfers terrestrial vision--language models (VLMs) to underwater domains without any additional underwater training. Earth2Ocean consists of two core components: a Geometric-guided Visual Mask Generator (\textbf{GMG}) that refines visual features via self-similarity geometric priors for local structure perception, and a Category-visual Semantic Alignment (\textbf{CSA}) module that enhances text embeddings through multimodal large language model reasoning and scene-aware template construction. Extensive experiments on the UOVSBench benchmark demonstrate that Earth2Ocean achieves significant performance improvement on average while maintaining efficient inference.

Bingyu Li, Tao Huo, Da Zhang, Zhiyuan Zhao, Junyu Gao, Xuelong Li• 2025

Related benchmarks

TaskDatasetResultRank
Open-Vocabulary SegmentationDUT-Seg
aAcc74.37
18
Open-Vocabulary SegmentationMAS3K
aAcc67.26
18
Open-Vocabulary SegmentationSUIM
aAcc78.34
18
Open-Vocabulary SegmentationUSIS16K
Average Accuracy (aAcc)60.45
18
Open-Vocabulary SegmentationAquaO V255
aAcc55.98
18
Open-Vocabulary SegmentationUOVSBench Average
aAcc68.17
18
Open-Vocabulary SegmentationUSIS10K
aAcc72.62
18
Open Vocabulary Semantic SegmentationUOVSBench
aAcc54.34
6
Showing 8 of 8 rows

Other info

Follow for update