Computer Vision and Image Understanding

Papers
(The H4-Index of Computer Vision and Image Understanding is 32. The table below lists those papers that are above that threshold based on CrossRef citation counts [max. 250 papers]. The publications cover those that have been published in the past four years, i.e., from 2022-08-01 to 2026-08-01.)
ArticleCitations
Luminance prior guided Low-Light 4C catenary image enhancement410
Editorial Board132
Editorial Board121
Improving the planarity and sharpness of monocularly estimated depth images using the Phong reflection model118
Editorial Board107
Exploring using jigsaw puzzles for out-of-distribution detection90
Extending function mixture network for improved spectral super-resolution74
MATTE: Multi-task multi-scale attention67
3D semantic segmentation based on spatial-aware convolution and shape completion for augmented reality applications56
Lightweight feature point detection network with channel enhancement56
Editorial Board56
Efficient cross-information fusion decoder for semantic segmentation56
Editorial Board56
Emerging image generation with flexible control of perceived difficulty55
Modality adaptation via feature difference learning for depth human parsing53
QB-MOTR: A simple query bootstrapping end-to-end multi-object tracking method with transformer52
Siamese self-supervised learning for fine-grained visual classification47
Deducing health cues from biometric data45
SNRD-Net: SNR-aware dual enhancement network for low-light images45
RetSeg3D: Retention-based 3D semantic segmentation for autonomous driving45
Twin-SegNet: Dynamically coupled complementary segmentation networks for generalized medical image segmentation44
Spatial Sensitive Grad-CAM++: Towards High-Quality Visual Explanations for Object Detectors via Weighted Combination of Gradient Maps43
REST: A resolution preserving network for photorealistic style transfer via semantic distillation42
Vision-based mistake analysis in procedural activities: A review of advances and challenges40
CRML-Net: Cross-Modal Reasoning and Multi-Task Learning Network for tooth image segmentation40
JEMA: Joint Embedding of Multimodal and multi-view Alignment in human-centric embedding space for manufacturing39
NaviFormer: Multimodal scene segmentation for assistive navigation39
Exploring the differences in adversarial robustness between ViT- and CNN-based models using novel metrics38
Robust Teacher: Self-correcting pseudo-label-guided semi-supervised learning for object detection36
Feature reconstruction and metric based network for few-shot object detection36
Convolutional neural network framework for deepfake detection: A diffusion-based approach34
Feature preserving 3D mesh denoising with a Dense Local Graph Neural Network32
RelFormer: Advancing contextual relations for transformer-based dense captioning32
0.097805023193359