Image and Vision Computing

Papers
(The H4-Index of Image and Vision Computing is 36. The table below lists those papers that are above that threshold based on CrossRef citation counts [max. 250 papers]. The publications cover those that have been published in the past four years, i.e., from 2022-08-01 to 2026-08-01.)
ArticleCitations
ADVC: Adversarial dense video captioning with unsupervised pretraining462
Alignment and fusion for adaptive domain nighttime semantic segmentation207
Few-shot-based video generation via multimodal fusion and Fourier Spliter198
Feature decoupling and interaction network for defending against adversarial examples129
Modeling content-attribute preference for personalized image esthetics assessment128
GLMambaNet: Mamba-based decoder with local detail enhancement for semantic segmentation of remote sensing imagery107
Efficient ultra-lightweight convolutional attention network for embedded identity document recognition system97
Accurate and efficient salient object detection via position prior attention94
Multi-information guided camouflaged object detection87
G-TRACE: Grouped temporal recalibration for video object segmentation80
BF3D: Bi-directional fusion 3D detector with semantic sampling and geometric mapping78
Learning diverse and deep clues for person reidentification67
Lightweight multi-scale global attention enhancement network for image super-resolution67
DMNet: Image dehazing via Dual-Domain Modulation64
Active domain adaptation for semantic segmentation via dynamically balancing domainness and uncertainty64
Window normalization: Enhancing point cloud understanding by unifying inconsistent point densities62
Background debiased class incremental learning for video action recognition61
AI-powered trustable and explainable fall detection system using transfer learning59
RGB-T tracking by modality difference reduction and feature re-selection59
ABC: Aligning binary centers for single-stage monocular 3D object detection53
Hourglass cascaded recurrent stereo matching network52
DeepArUco++: Improved detection of square fiducial markers in challenging lighting conditions48
GAN-BodyPose: Real-time 3D human body pose data key point detection and quality assessment assisted by generative adversarial network48
HPD-Depth: High performance decoding network for self-supervised monocular depth estimation46
UCPNet: An Ultra-Lightweight Cross-Perception Network for Real-Time Semantic Segmentation44
Privacy-preserving explainable AI enable federated learning-based denoising fingerprint recognition model43
CODNet: Context-based object detection network for multimodal image captioning and virtual question answering42
Single stage architecture for improved accuracy real-time object detection on mobile devices42
MAFUNet: Mamba with adaptive fusion UNet for medical image segmentation41
SRMA-KD: Structured relational multi-scale attention knowledge distillation for effective lightweight cardiac image segmentation39
PST-Mamba: Spatio-temporal selective state fusion for effective point cloud video understanding with state space models39
Few-shot classification with multisemantic information fusion network38
Synthetic lidar point cloud generation using deep generative models for improved driving scene object recognition38
CAGS: Open-vocabulary 3D scene understanding with context-aware Gaussian splatting38
Recent advances in deterministic human motion prediction: A review36
SAGNet: Synergistic Attention-Graph Network For video salient object detection36
Two-stream transformer tracking with messengers36
Burst image super-resolution via multi-cross attention encoding and multi-scan state-space decoding36
0.70618796348572