IEEE-ACM Transactions on Audio Speech and Language Processing

Papers
(The H4-Index of IEEE-ACM Transactions on Audio Speech and Language Processing is 39. The table below lists those papers that are above that threshold based on CrossRef citation counts [max. 250 papers]. The publications cover those that have been published in the past four years, i.e., from 2022-08-01 to 2026-08-01.)
ArticleCitations
Decorrelation in Feedback Delay Networks367
CET2: Modelling Topic Transitions for Coherent and Engaging Knowledge-Grounded Conversations282
WDEA: The Structure and Semantic Fusion With Wasserstein Distance for Low-Resource Language Entity Alignment255
Towards Generating Diverse Audio Captions via Adversarial Training199
MO-Transformer: Extract High-Level Relationship Between Words for Neural Machine Translation195
DropAttack: A Random Dropped Weight Attack Adversarial Training for Natural Language Understanding181
Reverberant Source Separation Using NTF With Delayed Subsources and Spatial Priors179
Audio-Only Phonetic Segment Classification Using Embeddings Learned From Audio and Ultrasound Tongue Imaging Data142
Generalizing Speaker Verification for Spoof Awareness in the Embedding Space140
Multi-Channel to Multi-Channel Noise Reduction and Reverberant Speech Preservation in Time-Varying Acoustic Scenes for Binaural Reproduction135
Review of Methods for Automatic Speaker Verification114
Refining Synthesized Speech Using Speaker Information and Phone Masking for Data Augmentation of Speech Recognition94
Improvement of Accent Classification Models Through Grad-Transfer From Spectrograms and Gradient-Weighted Class Activation Mapping88
$\mathcal {P}$owMix: A Versatile Regularizer for Multimodal Sentiment Analysis85
Envelope-Based Multichannel Noise Reduction for Cochlear Implant Applications82
Efficient Lightweight Speaker Verification With Broadcasting CNN-Transformer and Knowledge Distillation Training of Self-Attention Maps76
A User-Centric Approach for Deep Residual-Echo Suppression in Double-Talk74
AudioLM: A Language Modeling Approach to Audio Generation66
Learning Discriminative Representations and Decision Boundaries for Open Intent Detection65
The VoxCeleb Speaker Recognition Challenge: A Retrospective65
Enhancing Robustness of Speech Watermarking Using a Transformer-Based Framework Exploiting Acoustic Features64
Attention-Based Speech Enhancement Using Human Quality Perception Modeling61
Representation Learning With Hidden Unit Clustering for Low Resource Speech Applications55
Adaptive Multi-Domain Dialogue State Tracking on Spoken Conversations52
COVID-19 Detection via Fusion of Modulation Spectrum and Linear Prediction Speech Features51
Pronunciation Dictionary-Free Multilingual Speech Synthesis Using Learned Phonetic Representations49
IEEE Signal Processing Society Information49
Emotion Prediction Oriented Method With Multiple Supervisions for Emotion-Cause Pair Extraction48
Disentangled Text Representation Learning With Information-Theoretic Perspective for Adversarial Robustness45
Complex-Domain Pitch Estimation Algorithm for Narrowband Speech Signals45
ReZero: Region-Customizable Sound Extraction45
Spherically Steerable Vector Differential Microphone Arrays43
Implicit Self-Supervised Language Representation for Spoken Language Diarization42
Enhanced Multi-Domain Dialogue State Tracker With Second-Order Slot Interactions41
Exploiting Low-Rank Tensor-Train Deep Neural Networks Based on Riemannian Gradient Descent With Illustrations of Speech Processing41
Distinctive and Natural Speaker Anonymization via Singular Value Transformation-Assisted Matrix40
Phrase-Aware Financial Sentiment Analysis Based on Constituent Syntax40
SPEC: Summary Preference Decomposition for Low-Resource Abstractive Summarization40
Predicting Level-Dependent Changes in Concurrent Vowel Scores Using the 2D-CNN Models40
Textless Unit-to-Unit Training for Many-to-Many Multilingual Speech-to-Speech Translation39
Integrated Syntactic and Semantic Tree for Targeted Sentiment Classification Using Dual-Channel Graph Convolutional Network39
0.11087107658386