Multimedia Systems

Papers
(The H4-Index of Multimedia Systems is 27. The table below lists those papers that are above that threshold based on CrossRef citation counts [max. 250 papers]. The publications cover those that have been published in the past four years, i.e., from 2022-08-01 to 2026-08-01.)
ArticleCitations
Pseudo-global strategy-based visual comfort assessment considering attention mechanism133
SS-CMT: a label independent cross-modal transferable adversarial video attack with sparse strategy98
Face and voice cross-modal association with learning convex feature embedding70
DiffRA: universal restorative adversarial attack based on diffusion model70
TreeSegNet: multi-scale query-based instance segmentation with frequency-aware and gated feature enhancement57
Dual-branch spectral–spatial feature extraction network for multispectral image compression56
A research for sound event localization and detection based on local–global adaptive fusion and temporal importance network53
FedMAB: adaptive multimodal federated learning with multi-armed bandits50
A visual question answering model based on image captioning47
Unsupervised deep metric learning algorithm for crop disease images based on knowledge distillation networks45
On-line monitoring of structural performance of scraper conveyor driven by digital twin44
Multi-view Isolated sign language recognition based on cross-view and multi-level transformer44
Model-based portrait video compression with spatial constraint and adaptive pose processing43
JAMD-Net: image splicing forgery detection based on JPEG compression artifacts and multi-dilated channel refinement fusion39
Segmentation-aware image super-resolution with generative adversarial networks36
CHCoT-MSLU: a coupled hierarchical chain-of-thought prompt learning model for multi-intent spoken language understanding34
Real emotion seeker: recalibrating annotation for facial expression recognition33
Towards domain adaptation underwater image enhancement and restoration32
Fast latent-feature augmentation for cross-domain face forgery detection32
SFRA: spatial fusion regression augmentation network for facial landmark detection32
A comparative study of color quantization methods using various image quality assessment indices31
Feature fusion and optimization integrated refined deep residual network for diabetic retinopathy severity classification using fundus image31
360° video quality assessment based on saliency-guided viewport extraction31
LEA-depth: a lightweight self-supervised monocular depth estimation with attention fusion and edge-aware distillation30
GVA: guided visual attention approach for automatic image caption generation30
The segmented UEC Food-100 dataset with benchmark experiment on food detection30
Mamba-driven context-aware tracking with dual prompts27
GCGV: a dual-branch hybrid network integrating graph attention, CNNs, and vision transformers for enhanced hyperspectral image classification27
0.065191984176636