IEEE Transactions on Image Processing

Papers
(The H4-Index of IEEE Transactions on Image Processing is 76. The table below lists those papers that are above that threshold based on CrossRef citation counts [max. 250 papers]. The publications cover those that have been published in the past four years, i.e., from 2022-08-01 to 2026-08-01.)
ArticleCitations
Variational Structured Attention Networks for Deep Visual Representation Learning1091
TSFormer: Efficient Ultra-High-Definition Image Restoration via Trusted Min- p995
An Explanation Method Based on Interpretable Linear Model With Four Key Characteristics923
Density-Guided Incremental Dominant Instance Exploration for Two-View Geometric Model Fitting830
Color Spike Camera Reconstruction via Long Short-Term Temporal Aggregation of Spike Signals810
Equivariant Local Reference Frames With Optimization for Robust Non-Rigid Point Cloud Correspondence646
Star-Shaped Multi-Person Interaction Graph Model for Group Skeleton-Based Action Recognition382
Bi-Nuclear Tensor Schatten-p Norm Minimization for Multi-View Subspace Clustering324
AdaAugment: A Tuning-Free and Adaptive Approach to Enhance Data Augmentation321
Cross-Modality Pyramid Alignment for Visual Intention Understanding313
COME: A Collaborative Optimization Framework With Low-Rank MoE for Indoor 3D Object Detection295
High-Fidelity Seismic Super-Resolution Using Prior-Informed Deep Learning With 3D Awareness292
Zero-Pose-Prior NeRF: Recursive Radiance Field Reconstruction From Unposed and Unordered Images275
Advancing Pre-Trained Teacher: Towards Robust Feature Discrepancy for Anomaly Detection240
Consensus Sparsity: Multi-Context Sparse Image Representation via L -Induced Matrix Variate220
SemiRS-COC: Semi-Supervised Classification for Complex Remote Sensing Scenes With Cross-Object Consistency219
Leveraging Feature Alignment in Grassmannian Manifold for Multi-Output Regression Tasks212
One-Class Classification Using ℓp-Norm Multiple Kernel Fisher Null Approach211
Cross-Domain Few-Shot Medical Image Segmentation via Dynamic Semantic Matching205
Pro2Diff: Proposal Propagation for Multi-Object Tracking via the Diffusion Model201
Pose-Appearance Relational Modeling for Video Action Recognition197
Uncertainty-Guided Refinement for Fine-Grained Salient Object Detection182
Spatial Frequency Modulation Network for Efficient Image Dehazing180
Information-Maximized Soft Variable Discretization for Self-Supervised Image Representation Learning167
Toward Efficient Test Time Adaptation With Hierarchical Distribution Alignment164
MaCon: A Generic Self-Supervised Framework for Unsupervised Multimodal Change Detection155
FF-LPD: A Real-Time Frame-by-Frame License Plate Detector With Knowledge Distillation and Feature Propagation154
TTVFI: Learning Trajectory-Aware Transformer for Video Frame Interpolation153
Spectral State Fusion Tree Mamba for Hyperspectral Image Classification151
Language Supervised Multi-Camera Multi-Object Tracking148
Global Modeling Matters: A Fast, Lightweight, and Effective Baseline for Efficient Image Restoration144
An Adaptive Multi-Granularity Graph Representation of Image via Granular-ball Computing128
Toward Projected Clustering With Aggregated Mapping124
Cross-Modal Retrieval With Noisy Correspondence via Consistency Refining and Mining123
Focus on Finding Deepfakes: A Robust Proactive Detection Method Based on Orthogonal Moment Watermarking118
LearnMat: Semantic-Aware Self-Supervision Fine-Grained Visual Recognition117
Revisiting Fine-Grained Image Analysis by Semantic-Part Alignment114
H 3 Former: Hypergraph-Based Semantic-Aware Aggregation via Hyperbolic Hierarchical Contrastive Loss for Fine-Grained Visual Classification113
Vision-Based UAV Self-Positioning in Low-Altitude Urban Environments112
In-Context Model Predictive Generation: Open-Vocabulary Motion Synthesis From Language Models to Physics106
OccNeRF: Advancing 3D Occupancy Prediction in LiDAR-Free Environments105
STPNet: Scale-Aware Text Prompt Network for Medical Image Segmentation105
Graph Embedding Contrastive Multi-Modal Representation Learning for Clustering104
Automatic Quaternion-Domain Color Image Stitching103
HAda: Hyper-Adaptive Parameter-Efficient Learning for Multi-View ConvNets102
LoRA-Composer: Leveraging Low-Rank Adaptation for Multi-Concept Customization in Training-Free Diffusion Models101
Harnessing Multi-Modal Large Language Models for Measuring and Interpreting Color Differences98
Multi-Granularity Contrastive Cross-Modal Collaborative Generation for End-to-End Long-Term Video Question Answering98
Fine-Grained Recognition With Learnable Semantic Data Augmentation97
Attention-Guided Neural Networks for Full-Reference and No-Reference Audio-Visual Quality Assessment96
Fast 3D Room Layout Estimation Based on Compact High-Level Representation95
HAIMNet: A Hierarchical Adaptive Interaction Modulation Network for Low-Light Image Enhancement95
Generalization Beyond Feature Alignment: Concept Activation-Guided Contrastive Learning92
Perceptually Weighted Rate Distortion Optimization for Video-Based Point Cloud Compression91
Spatial-Temporal Scene Graph Generation for Open-Vocabulary Multiple Object Tracking89
Toward Generalizable Forgery Detection and Reasoning87
ScaleNet: Scaling up Pretrained Neural Networks With Incremental Parameters85
FD-SCU: Frequency Decomposition-Based Spectrum Collaborative Upsampling for Point Cloud Color Attribute85
Vision Enhancing LLMs: Empowering Multimodal Knowledge Storage and Sharing in LLMs83
Optimization-Inspired Learning With Architecture Augmentations and Control Mechanisms for Low-Level Vision83
SharpFormer: Learning Local Feature Preserving Global Representations for Image Deblurring82
MicroSDF: Microfacet-Driven Hybrid Neural SDFs for Mixed-Reflectance Surface Reconstruction82
MCIB: Multi-Modal Complementary Information Bottleneck for Hyperspectral and LiDAR Classification81
SRS: Siamese Reconstruction-Segmentation Network Based on Dynamic-Parameter Convolution81
Stacked Deconvolutional Network for Semantic Segmentation80
Hyperspectral Meets Optical Flow: Spectral Flow Extraction for Hyperspectral Image Classification80
TSCCD: Temporal Self-Construction Cross-Domain Learning for Unsupervised Hyperspectral Change Detection80
Rethinking Sampling Strategies for Unsupervised Person Re-Identification79
NeuralDiffuser: Neuroscience-Inspired Diffusion Guidance for fMRI Visual Reconstruction79
Weakly Supervised Semantic Segmentation via Alternate Self-Dual Teaching79
Toward Robust Alignment for Video Dehazing With Temporal Lookup Table78
Point-Based Learnable Query Generator for Human–Object Interaction Detection78
Unsupervised Modality-Transferable Video Highlight Detection With Representation Activation Sequence Learning78
Decoupled Cross-Modal Phrase-Attention Network for Image-Sentence Matching77
Multi-Condition Latent Diffusion Network for Scene-Aware Neural Human Motion Prediction76
SegHSI: Semantic Segmentation of Hyperspectral Images With Limited Labeled Pixels76
Boundary-Aware Prototype in Semi-Supervised Medical Image Segmentation76
0.18151998519897