GeoCueFormer: Geometry-Guided Wavelet Representation and Prediction-Cued Dual-Stage Decoder for Underwater Semantic Segmentation

📅 2026-09-15
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
本文提出GeoCueFormer,通过几何引导的小波表示和预测提示的双阶段解码器解决水下图像因光线吸收和散射导致的视觉退化问题,提升水下语义分割性能。
📝 Abstract
Underwater semantic segmentation is essential for marine ecosystem monitoring, yet remains challenging due to severe visual degradation. Light absorption and scattering often lead to color shifts, low contrast, and blurred boundaries, making shallow detail features unreliable. Existing underwater segmentation methods improve RGB feature aggregation or boundary prediction, but still lack an explicit mechanism to distinguish structure-related details from degradation-induced responses. To address this limitation, we propose GeoCueFormer, a lightweight framework that combines geometry-constrained frequency enhancement with prediction-cued refinement. GeoCueFormer performs stage-specific wavelet enhancement on hierarchical encoder features to complement shallow boundary details while preserving deep structural semantics. A depth-derived spatial gate constrains shallow frequency enhancement toward geometry-consistent regions, and a prediction-cued dual-stage decoder further refines ambiguous high-resolution features. GeoCueFormer obtains 82.23% and 73.04% mIoU on SUIM and DUT, respectively. Under comparable model complexity and standard benchmark settings on SUIM and DUT, it achieves SOTA performance while maintaining a favorable accuracy-complexity trade-off. These results show that distinguishing structural details from degradation-induced interference is more effective for underwater segmentation.
Problem

Research questions and friction points this paper is trying to address.

underwater semantic segmentation
visual degradation
structure-related details
Innovation

Methods, ideas, or system contributions that make the work stand out.

geometry-guided wavelet enhancement
prediction-cued refinement
dual-stage decoder
shallow boundary details
deep structural semantics
🔎 Similar Papers
No similar papers found.
💼 Related Jobs
No related jobs found.
Xian Wu
Xian Wu
Hong Kong University of Science and Technology (Guangzhou)
Quantum ArchitectureCircuit Synthesis
X
Xinjin Li
Columbia University, New York, United States
Y
Yiliu Xu
Carnegie Mellon University, Pittsburgh, United States
Yining Liu
Yining Liu
Wenzhou University of Technology
VANET AuthenticationPrivacy-preserving data aggregationTrajectory privacy
Y
Yong Jiang
Southwest University of Science and Technology, Mianyang, China