Multimedia Systems

Papers
(The H4-Index of Multimedia Systems is 27. The table below lists those papers that are above that threshold based on CrossRef citation counts [max. 250 papers]. The publications cover those that have been published in the past four years, i.e., from 2022-08-01 to 2026-08-01.)
ArticleCitations
Pseudo-global strategy-based visual comfort assessment considering attention mechanism133
SS-CMT: a label independent cross-modal transferable adversarial video attack with sparse strategy98
DiffRA: universal restorative adversarial attack based on diffusion model70
Face and voice cross-modal association with learning convex feature embedding70
TreeSegNet: multi-scale query-based instance segmentation with frequency-aware and gated feature enhancement57
Dual-branch spectral–spatial feature extraction network for multispectral image compression56
A research for sound event localization and detection based on local–global adaptive fusion and temporal importance network53
FedMAB: adaptive multimodal federated learning with multi-armed bandits50
A visual question answering model based on image captioning47
Unsupervised deep metric learning algorithm for crop disease images based on knowledge distillation networks45
Multi-view Isolated sign language recognition based on cross-view and multi-level transformer44
On-line monitoring of structural performance of scraper conveyor driven by digital twin44
Model-based portrait video compression with spatial constraint and adaptive pose processing43
JAMD-Net: image splicing forgery detection based on JPEG compression artifacts and multi-dilated channel refinement fusion39
Segmentation-aware image super-resolution with generative adversarial networks36
CHCoT-MSLU: a coupled hierarchical chain-of-thought prompt learning model for multi-intent spoken language understanding34
Real emotion seeker: recalibrating annotation for facial expression recognition33
SFRA: spatial fusion regression augmentation network for facial landmark detection32
Towards domain adaptation underwater image enhancement and restoration32
Fast latent-feature augmentation for cross-domain face forgery detection32
360° video quality assessment based on saliency-guided viewport extraction31
A comparative study of color quantization methods using various image quality assessment indices31
Feature fusion and optimization integrated refined deep residual network for diabetic retinopathy severity classification using fundus image31
The segmented UEC Food-100 dataset with benchmark experiment on food detection30
LEA-depth: a lightweight self-supervised monocular depth estimation with attention fusion and edge-aware distillation30
GVA: guided visual attention approach for automatic image caption generation30
GCGV: a dual-branch hybrid network integrating graph attention, CNNs, and vision transformers for enhanced hyperspectral image classification27
Mamba-driven context-aware tracking with dual prompts27
0.22302794456482