Visual Computer

Papers
(The median citation count of Visual Computer is 2. The table below lists those papers that are above that threshold based on CrossRef citation counts [max. 250 papers]. The publications cover those that have been published in the past four years, i.e., from 2022-08-01 to 2026-08-01.)
ArticleCitations
Robust object recognition via context-driven reliability assessment134
Learning shape abstraction by cropping positive cuboid primitives with negative ones109
Edge-priority-extraction network using re-parameterization for real-time super-resolution93
Enhancing green screen matting with group normalization and perceptual loss for color overflow and complex edges64
V$$^2$$MLP: an accurate and simple multi-view MLP network for fine-grained 3D shape recognition61
Multi-modal co-attention relation networks for visual question answering60
Joint attribute soft-sharing and contextual local: a multi-level features learning network for person re-identification57
Feature decomposition and structural learning for multi-diverse and multi-view data clustering55
Enhanced Temporal Representation and Spatial Alignment for High-Fidelity Talking Video Generation55
Adaptive box-level supervision with superpixel shape guidance for ultrasound image segmentation53
Monocular human depth estimation with 3D motion flow and surface normals51
Camera calibration for the surround-view system: a benchmark and dataset49
TaiChiGPT: complex sports action generation based on large language models49
A multi-target cow face detection model in complex scenes45
Self-knowledge distillation through ensemble model averaging: a novel approach for image classification45
PoseNorm-PCN: pose-normalized human point cloud completion from a single front view44
SpatioGS: spatiotemporal-aware density control for dynamic scene rendering with Gaussian splatting42
Exploring Structural Lines for Interior Floorplan Segmentation42
SES-yolov5: small object graphics detection and visualization applications42
Hybrid annotation alignment-based multi-region crop model for high-resolution image42
EMA-U-Net: efficient multi-attention U-Net for skin lesion segmentation41
MAPD: multi-receptive field and attention mechanism for multispectral pedestrian detection40
Lightweight subpixel sampling network for image super-resolution40
Developing an augmented reality framework with embedded objects and adaptive optical models for advanced lighting simulation40
Enhanced optical flow estimation via multiscale kernel selection and super-resolution integration39
Using transfer learning to determine the type of mathematical fractals image of Islamic geometric patterns37
Nighttime driver behavior prediction using taillight signal recognition via CNN-SVM classifier36
A two-stage model for spatial downscaling of daily precipitation data36
Image encryption algorithm based on cross-scrambling and rapid-mode diffusion35
Secure management of retinal imaging based on deep learning, zero-watermarking and reversible data hiding35
Weighted and truncated $$L_1$$ image smoothing based on unsupervised learning35
Uncertainty-guided time–frequency feature enhancement for emotion-aware speech-driven 3D facial animation35
Cfseg-Net: context feature extraction network for medical image segmentation34
Gaze-contingent adaptation of VR stereo parameters for cybersickness prevention33
A new chaotic image encryption algorithm based on dynamic DNA coding and RNA computing33
AttentionDIP: attention-based deep image prior model to restore satellite and aerial images from gamma distributed speckle interference33
Generative artificial intelligence for ophthalmic images: developments, applications and challenges33
Deep learning in chronic wound segmentation: a comprehensive review and meta-analysis32
PGSR-DR: high-fidelity reflective surface reconstruction with planar-based Gaussians and deferred rendering31
Depth-guided color correction and multi-scale Retinex network for underwater image enhancement31
ER: Extract-regress network for precise 3D reconstruction of interacting hands from monocular images30
SCAKD: a knowledge distillation framework based on spatial-corner attention for infrared and visible image fusion30
Road crack detection using pixel classification and intensity-based distinctive fuzzy C-means clustering30
Robust corner detection in continuous space30
MoCoSys: human motion correction based on deep learning coupled with 3D+t Laplacian motion representation30
Adaptively weighted discrete Laplacian for inverse rendering30
Self-supervised single-image 3D face reconstruction method based on attention mechanism and attribute refinement29
Preface (Vol 39. Issue 6, June 2023)28
Bridging realities: training visuo-haptic object recognition models for robots using 3D virtual simulations28
CLFormer: a unified transformer-based framework for weakly supervised crowd counting and localization28
ARP$$\Delta $$: Accelerated ray-tracing photon differentials for real-time global illumination with combined specular and diffuse solutions28
Adversarial-based refinement dual-branch network for semi-supervised salient object detection of strip steel surface defects27
Distribution-decouple learning network: an innovative approach for single image dehazing with spatial and frequency decoupling27
PL-MCT: pseudo-labeling and multi-frame consistency training for semi-supervised visual tracking27
Enhanced visual perception for underwater images based on multistage generative adversarial network27
Icg: intensity and color gradient operator on RGB images for visual object tracking27
Two-stream inter-class variation enhancement network for facial expression recognition27
DeMaskGAN: a de-masking generative adversarial network guided by semantic segmentation26
Sound signatures for images and geometric shapes26
AdverFuse: robust fusion of multimodal images based on dynamic attention and adversarial learning26
PlayNet: real-time handball play classification with Kalman embeddings and neural networks26
Compact storage of additively weighted Voronoi diagrams26
Personalized hairstyle and hair color editing based on multi-feature fusion26
Enhancing multi-scale information exchange and feature fusion for human pose estimation25
MF-SAM: enhancing multi-modal fusion with Mamba in SAM-Med3D for GPi segmentation25
An unsupervised denoising model for poisson noise using GSURE-driven deep image prior with multi-order regularization for medical imaging25
Attention-enhanced controllable disentanglement for cloth-changing person re-identification25
Correction: DFP-YOLO: a lightweight machine tool workpiece defect detection algorithm based on computer vision25
ImpRes: implicit residual diffusion models for image super-resolution25
Liver segmentation based on complementary features U-Net25
Dynamic underwater cognition: aligned detection networks for enhanced underwater object recognition24
High-fidelity facial expression transfer using part-based local–global conditional gans24
A novel robust image watermarking algorithm based on polar decomposition and image geometric correction24
Correction: 3D reconstruction method based on N-step phase unwrapping23
SSoB: searching a scene-oriented architecture for underwater object detection23
Enhancing hyperspectral image classification through spectral-spatial synergy: SSFSNet23
Hybrid Mamba-Transformer Multi-Agent Reinforcement Learning for scalable coordination in complex environments22
A point cloud self-learning network based on contrastive learning for classification and segmentation22
Boosted verification using siamese neural network with DiffBlock22
CrowdImprint: decomposing context-aware interactions22
A novel robust digital image watermarking scheme based on attention U-Net++ structure22
Double-handed dynamic gesture recognition using contour-based hand tracking and maximum mean probability ensembling (MMPE) for Indian Sign Language22
Decoupled spatio-temporal grouping transformer for skeleton-based action recognition21
Application of VR to ikebana education21
MCGFF-Net: a multi-scale context-aware and global feature fusion network for enhanced polyp and skin lesion segmentation21
Generating natural pedestrian crowds by learning real crowd trajectories through a transformer-based GAN21
Arbitrary style transfer via content consistency and style consistency21
Capturing spatiotemporal dependencies with competitive set attention for video summarization20
Study on the methods of hyperspectral image saliency detection based on MBCNN20
Anti-counterfeiting textured pattern20
CMT-UNet: enhancing remote sensing image segmentation via a hybrid CNN-Mamba-transformer architecture20
Dsf-net: a dual-stream fusion network integrating structural and detailed features for fundus-based diabetic retinopathy classification20
WeedGan: a novel generative adversarial network for cotton weed identification20
Enhancing 3D human pose estimation via spatio-temporal dual-stream fusion20
PMGAN: pretrained model-based generative adversarial network for text-to-image generation20
Vision transformers (ViT) and deep convolutional neural network (D-CNN)-based models for MRI brain primary tumors images multi-classification supported by explainable artificial intelligence (XAI)19
VGRR-Net: a reproducible visual computing framework for visibility-gap-driven collaborative 3D perception in low-altitude flying-car traffic19
Branch aware assignment for object detection19
VR-Deform: real-time visual and haptic interaction with deformable bodies using XPBD in virtual reality19
Topology-preserved human reconstruction with details19
An automatic framework for quadrilateral surface reconstruction with partitions from 3D point clouds19
Wavelet-driven meta-learning: unifying infrared-visible fusion and semantic segmentation for robust scene perception19
ConvFormer: parameter reduction in transformer models for 3D human pose estimation by leveraging dynamic multi-headed convolutional attention19
Dual-attention U-Net and multi-convolution network for single-image rain removal19
LKSMN: Large Kernel Spatial Modulation Network for Lightweight Image Super-Resolution18
OSH-Splat: optimizable semantic hyperplanes for enhanced 3D language feature Gaussian splatting18
ViT-SIR: vision transformer-based shoe image retrieval with enhanced feature representation18
Adaptive transformer-based detection: enhancing infrared image target recognition18
Cross-modal collaborative propagation for RGB–T saliency detection18
A self-attention model for viewport prediction based on distance constraint18
Light field depth estimation using occlusion-aware consistency analysis18
Blind image quality assessment by simulating the visual cortex18
Salient-aware multiple instance learning optimized network for weakly supervised object detection18
Point-voxel dual stream transformer for 3d point cloud learning17
Hash-NURF: efficient nested transparent object reconstruction using multi-resolution hash encoding17
The devil in the details: simple and effective optical flow synthetic data generation17
3D printer vision calibration system based on embedding Sobel bilateral filter in least squares filtering algorithm17
Learning to sculpt neural cityscapes17
Skin scar segmentation based on saliency detection17
Patch excitation network for boxless action recognition in still images17
TDGar-Ani: temporal motion fusion model and deformation correction network for enhancing garment animation details17
A workflow to systematically design uncertainty-aware visual analytics applications17
Virtual simulation for the dynamic response of concrete blocks under blast loading16
EdgeLF: edge-guided registration with loftr for visible and infrared images16
A mixed reality framework for microsurgery simulation with visual-tactile perception16
Multi-camera tracking of mechanically thrown objects for automated in-plant logistics by cognitive robots in Industry 4.016
Label-guided 4D Gaussian splatting for high-fidelity dynamic scene reconstruction16
Fusiform multi-scale pixel self-attention network for hyperspectral images reconstruction from a single RGB image16
An enhanced multi-scale weight assignment strategy of two-exposure fusion16
Wall segmentation in house plans: fusion of deep learning and traditional methods16
FFANet: dual attention-based flow field-aware network for wall identification16
Virtual reality support for thoracoscopic surgery design16
LVDIF: a framework for real-time interaction with large volume data16
Locality-constrained double-layer structure scaled simplex multi-view subspace clustering16
Semantic-Orthogonal Multi-modal Attention Network for RGB-D Salient Object Detection15
DICNet: achieve low-light image enhancement with image decomposition, illumination enhancement, and color restoration15
Prior-based privacy-assured compressed sensing scheme in cloud15
Occlusion-aware segmentation via RCF-Pix2Pix generative network15
HEU-Net: hybrid attention residual block-based network with external skip connections for metal corrosion semantic segmentation15
GDPNet: a hybrid GNN-Transformer with position–density-modulated attention for 3D point cloud semantic segmentation15
Gtfpose: a unified framework with double-chain GCN–transformer fusion for 3D human pose estimation15
ResNet-OSD: an optimized hybrid deep learning framework for oil spill detection in coastal drone imagery15
STVDNet: spatio-temporal interactive video de-raining network15
Latent diffusion transformer for point cloud generation15
Deep channel-spatial attention networks for enhancing super-resolution of high-magnification SEM images15
CLAC-Net: a composite medical image segmentation framework using self-attention and cross-layer asymmetric connections15
Outfit compatibility model using fully connected self-adjusting graph neural network15
Fairing-PIA: progressive-iterative approximation for fairing curve and surface generation15
A cascaded graph convolutional network for point cloud completion15
A survey on soccer player detection and tracking with videos15
Multimodal biometrics authentication using extreme learning machine with feature reduction by adaptive particle swarm optimization15
LiteMSNet: a lightweight semantic segmentation network with multi-scale feature extraction for urban streetscape scenes14
GPT-ZSS: a unified zero-shot segmentation framework leveraging GPT-generated semantic embeddings and relationship alignment14
ROMOT: Referring-expression-comprehension open-set multi-object tracking14
Generalized unsupervised functional map learning for dense correspondence14
Enhanced bridge crack segmentation via CNN–Mamba dual encoders with edge enhancement14
Editorial June 2024 ( Vol 40, Issue 6)14
Internal and external transmission encoder–decoder network for single-image deraining14
PDFT: parameter-diminish fine-tuning for transformer-based models14
Logical reasoning-enhanced interactive clustering: an efficient algorithm for large-scale datasets14
Data privacy protection domain adaptation by roughing and finishing stage14
Msc-Net: multi-stage colorization network for real-world images with specular highlights14
Facial expression recognition based on local–global information reasoning and spatial distribution of landmark features14
Advanced background learning for hyperspectral anomaly detection via synthetic spectral sample generation14
Enhancing high-vocabulary image annotation with a novel attention-based pooling14
A neural builder for spatial subdivision hierarchies14
The infinite doodler: expanding textures within tightly constrained manifolds14
RSFace: subject agnostic face swapping with expression high fidelity14
Privacy-aware Real-Time Target Person Matting in Multi-Person Scenes Using Dual Encoder-Decoder Networks14
E-FPN: an enhanced feature pyramid network for UAV scenarios detection14
Detail-aware image denoising via structure preserved network and residual diffusion model14
Multi-view clustering based on graph learning and view diversity learning14
Toward robust visual tracking for UAV with adaptive spatial-temporal weighted regularization14
MFFN: image super-resolution via multi-level features fusion network13
A hue preserving uniform illumination image enhancement via triangle similarity criterion in HSI color space13
TSNet: Task-specific network for joint diabetic retinopathy grading and lesion segmentation of ultra-wide optical coherence tomography angiography images13
Enhanced fine-grained visual classification through lightweight Transformer integration and auxiliary information fusion13
CSNet: a ConvNeXt-based Siamese network for RGB-D salient object detection13
DXAI: explaining classification by image decomposition13
Adaptive fourier-enhanced vision transformer with self-learning smoothing masks for accurate cat face recognition13
The crowd cooperation approach for formation maintenance and collision avoidance using multi-agent deep reinforcement learning13
Touching spaces: interactive physicalization for exploring spatial information13
Retina-enhanced multimodal deep learning for assessment of cardiovascular-kidney metabolic syndrome related outcomes13
AQPnP: an accurate and quaternion-based solution for the Perspective-n-Point problem13
Edge-aware texture filtering with superpixels constraint13
TRAIL: Simulating the impact of human locomotion on natural landscapes13
Adaptive frequency time-distribution network—a multiscale deblurring technique13
Recycling/upcycling graphic design: automatic design elements extraction and vectorization13
Attention-guided self-supervised distinctive region detection in point clouds13
High-frequency channel attention and contrastive learning for image super-resolution13
Enhancing the transferability and imperceptibility of adversarial attacks via rescaled variance-reduced diffusion13
Multi-channel correlated diffusion for text-driven artistic style transfer13
Semantically Enhanced Dual Visual Fusion Transformer for accurate image captioning13
OrthopedVR: clinical assessment and pre-operative planning of paediatric patients with lower limb rotational abnormalities in virtual reality13
Correction: Fast and high-quality scale-aware filtering for 3D images13
MaFIR: high-fidelity fisheye image rectification via Manhattan self-attention and dynamic feature optimization13
Adt-net: adaptive transformation-driven text-based person search network for enhancing cross-modal retrieval robustness13
ODRP: a new approach for spatial street sign detection from EXIF using deep learning-based object detection, distance estimation, rotation and projection system13
An algorithm for cross-fiber separation in yarn hairiness image processing13
Regularity-constrained point cloud reconstruction of building models via global alignment13
Psanet: prototype-guided salient attention for few-shot segmentation12
Enhancing link prediction accuracy with VG-GIN: a fusion of variational graph auto-encoders and graph isomorphism networks12
Adaptive arc area inpainting and image enhancement method based on AI-DLC model12
A new face presentation attack detection method based on face-weighted multi-color multi-level texture features12
A vision-to-decision framework for learning-augmented UAV–USV coordination in maritime search and rescue12
Digital human and embodied intelligence for sports science: advancements, opportunities and prospects12
EnvMap-GS: two-stage outdoor Gaussian reconstruction with background-to-environment map baking12
GCAENet: global-class context with advanced edge network for single human parsing12
SATD: syntax-aware handwritten mathematical expression recognition based on tree-structured transformer decoder12
Adaptive cascaded and parallel feature fusion for visual object tracking12
DPDTRN: a dynamic pixel-level difficulty-aware texture reconstruction network for document super-resolution12
Enhancing multiple-style image colorization through context-aware codebook and multi-stage learning12
Convex hull regression strategy for people detection on top-view fisheye images12
Enhancing remote sensing image segmentation with SGDC-DeepLab: a lightweight approach using Gaussian filters12
Stroke-based semantic segmentation for scene-level free-hand sketches12
Vehicle object counting network based on feature pyramid split attention mechanism12
Ellipsoid-SLAM: enhancing dynamic scene understanding through ellipsoidal object representation and trajectory tracking12
Robust 3D watermarking with high imperceptibility based on EMD on surfaces12
MPA-Det: multi-path aggregation-based object detection framework for aerial visual computing12
Boosting remote semantic segmentation using vision-and-language foundation model12
Multi-scale defect detection for plaid fabrics using scale sequence feature fusion and triple encoding12
Expression-driven monocular 3D face reconstruction based on cross-modal guidance12
Fast image recoloring for red–green anomalous trichromacy with contrast enhancement and naturalness preservation12
Group emotion recognition based on psychological principles using a fuzzy system12
Semantic guidance incremental network for efficiency video super-resolution12
Arpotcam: augmented reality-driven honeypot for enhancing security in IoT surveillance systems12
Enhancing cross-domain few-annotation object detection via memory storage-to-adaptation mechanism12
Advanced detection and segmentation of parabolic trough collector and Fresnel mirrors for CSP maintenance using YOLOv8 and segment anything model12
TransDehaze: transformer-enhanced texture attention for end-to-end single image dehaze12
Dual-domain cross-attention fusion with edge-guided frequency decoupling for RGB-D saliency object detection12
Uncertainty-aware multi-view post-aggregation for point cloud quality assessment12
M-GAN: multiattribute learning and multimodal feature fusion-based generative adversarial network for text-to-image synthesis12
Enhanced fine-grained relearning for skeleton-based action recognition12
An adaptive loss weighting multi-task network with attention-guide proposal generation for small size defect inspection11
Robust and fast QR code images deblurring via local maximum and minimum intensity prior11
Light field salient object detection based on discrete viewpoint selection and multi-feature fusion11
Image-only place recognition based on regional aggregating ConvNet features for underground parking lots11
Harnessing deep learning for faster water quality assessment: identifying bacterial contaminants in real time11
When CNN meet with ViT: decision-level feature fusion for camouflaged object detection11
ParaLkResNet: an efficient multi-scale image classification network11
Histogram equalization using a selective filter11
GenYOLO-leaf: a data-centric and open source framework for generalizable leaf instance segmentation across diverse datasets11
Enhancing scene text script identification through multi-task self-supervised learning11
Progressive region exchange: enhancing semi-supervised medical image segmentation through incremental complexity11
Coarse-to-fine blind image deblurring based on K-means clustering11
Lightweight head pose estimation without keypoints based on multi-scale lightweight neural network11
Transforming time and space: efficient video super-resolution with hybrid attention and deformable transformers11
CSI-DMT: multi-focus image fusion via cross-task semantic interaction and dual-attention mixing transformer11
Coarse-to-fine multi-scale attention-guided network for multi-exposure image fusion11
InstantTrace: fast parallel neuron tracing on GPUs11
0.25238800048828