Computational Visual Media

Papers
(The median citation count of Computational Visual Media is 3. The table below lists those papers that are above that threshold based on CrossRef citation counts [max. 250 papers]. The publications cover those that have been published in the past four years, i.e., from 2022-08-01 to 2026-08-01.)
ArticleCitations
Geometry-aware 3D pose transfer using transformer autoencoder2263
Ultra-High Resolution Facial Texture Reconstruction from a Single Image2171
MA2Net: Multi-Scale Adaptive Mixed Attention Network for Image Demoiréing885
Front Cover217
Towards harmonized regional style transfer and manipulation for facial images106
3D face recognition: A comprehensive survey in 202286
PE Loss: Perception-Enhanced Distortion-Oriented Loss for Image Restoration83
3D Indoor Scene Geometry Estimation from a Single Omnidirectional Image: A Comprehensive Survey80
Heuristic Weakly Supervised 3D Human Pose Estimation76
A Biophysical-Based Skin Model for Heterogeneous Volume Rendering70
Front Cover64
Recent advances in glinty appearance rendering55
FRNeRF: Fusion and Regularization Fields for Dynamic View Synthesis45
Controllable multi-domain semantic artwork synthesis45
Real-Time Woven Fabric Rendering Using SGGX Fitting38
A causal convolutional neural network for multi-subject motion modeling and generation32
Towards robustness and generalization of point cloud representation: A geometry coding method and a large-scale object-level dataset28
Temporal vectorized visibility for direct illumination of animated models27
Central similarity consistency hashing for asymmetric image retrieval26
Neural Scene Baking for Permutation Invariant Transparency Rendering with Real-Time Global Illumination26
Anchor-Regularized GAN Priors26
See More, Know More: Richer Prior Knowledge for Novel Class Discovery26
Practical construction of globally injective parameterizations with positional constraints22
Neural Reconstruction and Super-Resolution for Foveated Real-Time Rendering22
Multi-granularity sequence generation for hierarchical image classification22
Image-guided color mapping for categorical data visualization22
Self-Supervised Learning for Pre-Training 3D Point Clouds: A Survey21
IIDM: Image-to-Image Diffusion Model for Semantic Image Synthesis20
Contents19
D2ANet: Difference-aware attention network for multi-level change detection from satellite imagery17
MusicFace: Music-driven expressive singing face synthesis17
Photorealistic Fire Scene Video Generation via Multimodal Large Language Model and Pre-Trained Video Diffusion Model17
ARM3D: Attention-based relation module for indoor 3D object detection17
Message from Guest Editors of the CVM 2025 Special Issue16
Real-time distance field acceleration based free-viewpoint video synthesis for large sports fields16
DepthGAN: GAN-based depth generation from semantic layouts16
Front Cover15
NeuS-PIR: Learning Relightable Neural Surface Using Pre-Integrated Rendering15
A survey of urban visual analytics: Advances and future directions14
Multi3D: 3D-aware multimodal image synthesis14
Let's all dance: Enhancing amateur dance motions13
Sphere face model: A 3D morphable model with hypersphere manifold latent space using joint 2D/3D training13
Watertight surface reconstruction method for CAD models based on optimal transport12
Prediction of Scene Plausibility12
Constructing self-supporting surfaces with planar quadrilateral elements12
MMRelief: Modeling Multi-Human Relief from a Single Photograph11
Global video object segmentation with spatial constraint module11
Deep unfolding multi-scale regularizer network for image denoising11
Benchmarking visual SLAM methods in mirror environments11
Sem-iNeRF: Camera Pose Refinement by Inverting Neural Radiance Fields with Semantic Feature Consistency11
MDFP-Net: A Model-Driven Deep Neural Network for Fourier Ptychography11
Mindstorms in Natural Language-Based Societies of Mind11
Uncertainty Aware Multiple View Stereo Network with Accurate Supervision11
Addressing Missing Modality Challenges in MRI Images: A Comprehensive Review11
A Voronoi diagram approach for detecting defects in 3D printed fiber-reinforced polymers from microscope images10
Front cover10
EG-HumanNeRF: Efficient Generalizable Human NeRF Utilizing Human Prior for Sparse View10
Neural video field editing10
EFECL: Feature encoding enhancement with contrastive learning for indoor 3D object detection10
Exploring Contextual Priors for Real-World Image Super-Resolution10
Continuous Indexed Points for Multivariate Volume Visualization9
FCDFusion: A Fast, Low Color Deviation Method for Fusing Visible and Infrared Image Pairs9
Hybrid Mesh-Neural Representation for 3D Transparent Object Reconstruction9
Front cover9
SMixNet: Style Mixture Network for Exemplar-Based Image Translation9
A two-step surface-based 3D deep learning pipeline for segmentation of intracranial aneurysms9
Deep panoramic depth prediction and completion for indoor scenes9
Contents9
Contents9
SGformer: Boosting transformers for indoor lighting estimation from a single image9
Point cloud completion via structured feature maps using a feedback network8
Z-STAR+: A zero-shot style transfer method adjusting style distribution8
Focusing on your subject: Deep subject-aware image composition recommendation networks8
PuzzleSorter: Certainty-Aware Visual Restoration of Multiple Cultural Artifacts8
Front cover8
Immersive Analytics Meets Artificial Intelligence: A Systematic Review8
A survey on facial image deblurring8
Dynamic ocean inverse modeling based on differentiable rendering8
An efficient algorithm for approximate Voronoi diagram construction on triangulated surfaces8
Joint specular highlight detection and removal in single images via Unet-Transformer8
Towards uniform point distribution in feature-preserving point cloud filtering7
A Simple and Effective Filtering Scheme for Improving Neural Fields7
An anisotropic Chebyshev descriptor and its optimization for deformable shape correspondence7
Revitalizing Image Dehazing in the Real World: A High-Quality Dataset and a Customized Method7
A visual modeling method for spatiotemporal and multidimensional features in epidemiological analysis: Applied COVID-19 aggregated datasets7
Contents7
Message from the editor-in-chief7
Open-Vocabulary Camouflaged Object Segmentation with Cascaded Vision Language Models7
Cross-modal learning using privileged information for long-tailed image classification7
PVT v2: Improved baselines with pyramid vision transformer6
ImVoxelENet: Image to Voxels Epipolar Transformer for Multi-View RGB-Based 3D Object Detection6
MagicTalk: Implicit and Explicit Correlation Learning for Diffusion-Based Emotional Talking Face Generation6
Attention mechanisms in computer vision: A survey6
Front Cover6
NPRportrait 1.0: A three-level benchmark for non-photorealistic rendering of portraits6
Super-resolution reconstruction of single image for latent features6
Message from the editor-in-chief6
Exploring a Hierarchical Cross-Attention Transformer for High-Speed Tracking6
Front cover6
FilterGNN: Image feature matching with cascaded outlier filters and linear attention6
Continual few-shot patch-based learning for anime-style colorization6
Lossless Intrinsic Image Decomposition via Learning Shading Feature Filtering5
Multi-Color Compressive Hologram Synthesis with Learned Wave Propagation5
Autocompletion of repetitive stroking with image guidance5
Polygonal finite element-based content-aware image warping5
Point Mask Transformer for Outdoor Point Cloud Semantic Segmentation5
FEDNet: A Feature-Enhanced Diffusion Network for Efficient and Universal Texture Synthesis5
PraNet-V2: Dual-Supervised Reverse Attention for Medical Image Segmentation5
BoostPoint: Boosting Point Cloud Backbones with Image Pre-Training for 3D Understanding5
JNeRF: An efficient heterogeneous NeRF model zoo based on Jittor5
Full-duplex strategy for video object segmentation4
Noise4Denoise: Leveraging noise for unsupervised point cloud denoising4
Emotion Amplification of Facial Videos Using a Fine-Tuned StyleGAN4
A Comprehensive Survey on the Research and Development of RGB-T Salient Object Detection4
Message from the best paper award committee4
PMSSC: Parallelizable multi-subset based self-expressive model for subspace clustering4
SAM-driven MAE pre-training and background-aware meta-learning for unsupervised vehicle re-identification4
Pyramid-Angular-Constraint Network for Light Field Super-Resolution4
CLIP-SP: Vision-language model with adaptive prompting for scene parsing4
Multi-modal visual tracking: Review and experimental comparison4
GRIG: Data-Efficient Generative Residual Image Inpainting4
Multi-level dynamic style transfer for NeRFs4
Learning physically based material and lighting decompositions for face editing4
AR assistance for efficient dynamic target search4
LucIE: Language-Guided Local Image Editing for Fashion Images4
BLNet: Bidirectional learning network for point clouds4
Multi-Task Gradual Inference with a Single Encoder–Decoder Network for Automatic Portrait Matting4
Swin3D: A Pretrained Transformer Backbone for 3D Indoor Scene Understanding3
A benchmark for 3D mesh denoising3
Class Incremental Learning via Feature Space Calibration3
Recent advances in 3D Gaussian splatting3
Class-conditional domain adaptation for semantic segmentation3
Spatiotemporal Fusion Transformer for Video Demoiréing3
Audio-guided implicit neural representation for local image stylization3
Taming diffusion model for exemplar-based image translation3
Foundation models meet visualizations: Challenges and opportunities3
ARNet: Attribute Artifact Reduction for G-PCC Compressed Point Clouds3
LDSwap: A Semantic-Related Latent Code Disentangling Method in StyleSpace Towards High-Resolution Face Swapping3
Angle-uniform parallel coordinates3
Contents3
Contents3
FedDTR: Leveraging intra-domain global priors via domain-invariant text representation for personalized federated learning3
Self-supervised coarse-to-fine monocular depth estimation using a lightweight attention module3
Swin3D++: Effective Multi-Source Pretraining for 3D Indoor Scene Understanding3
Neural radiance fields in 3D vision: A comprehensive review3
Rethinking medical VQA models: Towards data-efficient learning3
FastMAE: Efficient Masked Autoencoder with Offline Tokenizer3
Decoupled Two-Stage Talking Head Generation via Gaussian-Landmark-Based Neural Radiance Fields3
Remote Sensing Tuning: A Survey3
Progressive edge-sensing dynamic scene deblurring3
0.16238403320312