article
We propose SARMA (Saliency Analysis with a Residual Mesh-based Architecture), a graph neural network designed to predict visual saliency on 3D surface meshes. Unlike traditional methods that rely on handcrafted geometric features, SARMA learns saliency patterns in an end-to-end fashion using residual Feature-Steered Graph Convolution (FeaStConv) layers. The network takes a mesh as input, represented as a graph of vertices and edges, and outputs per-vertex saliency values. To capture perceptually relevant geometry, we enrich the input features with discrete mean curvature alongside 3D coordinates. The model consists of three FeaStConv layers, each followed by a residual connection that stabilizes training and mitigates oversmoothing. We evaluate SARMA on the Schelling dataset of 3D models annotated with human saliency points, using PLCC and AUC as evaluation metrics. Our approach outperforms prior handcrafted and deep learning-based methods in terms of correlation with ground truth saliency, demonstrating the effectiveness of residual graph-based architectures for perceptual analysis of 3D shapes.
This page summarises published work. The authoritative version sits with the publisher.
DOI: 10.1109/vcip67698.2025.11396847
Is something wrong with this record? Report it or request removal.
Discussion
Have you built on this work, tried to replicate it, or seen it applied in practice? Share what you know. Verified researchers and MARATTO™ domain experts can open a discussion, and any member can reply. Contributions are reviewed before they appear.
No discussion yet. Open the first thread.
New to MARATTO™? Create a free account.