Transactions of the Association for Computational Linguistics

Papers
(The median citation count of Transactions of the Association for Computational Linguistics is 3. The table below lists those papers that are above that threshold based on CrossRef citation counts [max. 250 papers]. The publications cover those that have been published in the past four years, i.e., from 2022-08-01 to 2026-08-01.)
ArticleCitations
Persona-Aware Alignment Framework for Personalized Dialogue Generation1127
Overcoming Source Object Grounding for Semantic Image Editing355
From Robustness to Improved Generalization and Calibration in Pre-trained Language Models330
Segmentation-Free Streaming Machine Translation293
The Ethics of Automating Legal Actors191
M o N a C o : More Natural and Complex Question130
How to Select Datapoints for Efficient Human Evaluation of NLG Models?123
KEFT: Knowledge-Enhanced Fine-Tuning for Large Language Models in Domain-Specific Question Answering100
DARE: Diverse Visual Question Answering with Robustness Evaluation98
Cross-functional Analysis of Generalization in Behavioral Learning92
Anthropomimetic Uncertainty: What Verbalized Uncertainty in Language Models is Missing87
Understanding and Detecting Hallucinations in Neural Machine Translation via Model Introspection85
Data Foundations of Long-Context Language Models: A Survey83
Transformers for Tabular Data Representation: A Survey of Models and Applications80
State of What Art? A Call for Multi-Prompt LLM Evaluation78
Erasure of Unaligned Attributes from Neural Representations75
Safety-Potential Pruning for Enhancing Safety Prompts Against VLM Jailbreaking Without Retraining74
Aggretriever: A Simple Approach to Aggregate Textual Representations for Robust Dense Passage Retrieval71
KisMATH: Do LLMs Have Knowledge of Implicit Structures in Mathematical Reasoning?65
Citation Failure: Definition, Analysis and Efficient Mitigation62
Revisiting Meta-evaluation for Grammatical Error Correction58
T 2 -NER: A Two-Stage Span-Based Framework for Unified Named Entity Recognition with Templates58
Bridging the Gap between Synthetic and Natural Questions via Sentence Decomposition for Semantic Parsing54
Do Multi-Document Summarization Models Synthesize?53
Federated Learning for Exploiting Annotators’ Disagreements in Natural Language Processing52
Investigating Adversarial Trigger Transfer in Large Language Models51
Context-Aware Machine Translation with Source Coreference Explanation43
Benchmarking the Generation of Fact Checking Explanations42
Learning More from Mixed Emotions: A Label Refinement Method for Emotion Recognition in Conversations38
Frame Representation Hypothesis: Multi-Token LLM Interpretability and Concept-Guided Text Generation38
Psychometric Item Validation Using Virtual Respondents with Trait-Response Mediators35
Literally Concrete or Figuratively Abstract? Multilingual Concreteness Norms for Verb-Object Expressions35
DEAR: Disentangled Event-Agnostic Representation Learning for Early Fake News Detection33
Retrieval-Pretrained Transformer: Long-range Language Modeling with Self-retrieval32
mtRAG: A Multi-Turn Conversational Benchmark for Evaluating Retrieval-Augmented Generation Systems31
Few-Shot Multilingual Open-Domain QA from Five Examples30
Adversarial Defense without Adversarial Defense : Enhancing Language Model Robustness via Instance-level Principal Component Removal29
Cross-Lingual Dialogue Dataset Creation via Outline-Based Generation28
Accelerating Language Model Workflows with Prompt Choreography27
PsyMem: Fine-grained Psychological Alignment and Explicit Memory Control for Advanced Role-Playing LLMs26
Ev2R: Evaluating Evidence Retrieval in Automated Fact-Checking25
Questions Are All You Need to Train a Dense Passage Retriever23
To Diverge or Not to Diverge: A Morphosyntactic Perspective on Machine Translation vs Human Translation23
Are Triggers Needed for Document-Level Event Extraction?23
Culturally Aware and Adapted NLP: A Taxonomy and a Survey of the State of the Art23
Aligned Probing: Relating Toxic Behavior and Model Internals22
An Energy-based Model for Word-level AutoCompletion in Computer-aided Translation22
A Systematic Assessment of Language Models with Linguistic Minimal Pairs in Chinese21
Toward Robust RALMs: Revealing the Impact of Imperfect Retrieval on Retrieval-Augmented Language Models21
In-N-Out: A Parameter-Level API Graph Dataset for Tool Agents20
Analyzing and Adapting Large Language Models for Few-Shot Multilingual NLU: Are We There Yet?19
CUS-QA: Local-Knowledge-Oriented Open-Ended Question Answering Dataset19
From Explicit to Implicit: A Theoretical Framework and Transfer Method for Preference Internalization in Language Models19
Communication Drives the Emergence of Language Universals in Neural Agents: Evidence from the Word-order/Case-marking Trade-off18
Accurate and Efficient Fine-Tuning of Quantized Large Language Models Through Optimal Balance in Adaptation18
CorefInst: Leveraging LLMs for Multilingual Coreference Resolution17
Prompt Contrastive Transformation: An Enhanced Strategy for Efficient Prompt Transfer in Natural Language Processing17
Conformal Prediction for Natural Language Processing: A Survey17
Beyond One-Size-Fits-All : Inversion Learning for Highly Effective NLG Evaluation Prompts16
Are Character-level Translations Worth the Wait? Comparing ByT5 and mT5 for Machine Translation16
Navigating Cultural Chasms: Exploring and Unlocking the Cultural POV of Text-To-Image Models16
Retrieve What You Need: A Mutual Learning Framework for Open-domain Question Answering16
Improving Probability-based Prompt Selection Through Unified Evaluation and Analysis15
OrthoEdit: Principled and Stable Knowledge Editing via Orthogonal Subspace Projection15
CRAFT Your Dataset: Task-Specific Synthetic Dataset Generation Through Corpus Retrieval and Augmentation15
Navigating the Landscape of Hint Generation Research: From the Past to the Future14
InSCIt: Information-Seeking Conversations with Mixed-Initiative Interactions14
Localizing Factual Inconsistencies in Attributable Text Generation13
OpenFact: Factuality Enhanced Open Knowledge Extraction13
Interactive Machine Teaching by Labeling Rules and Instances13
BharatBBQ: A Multilingual Bias Benchmark for Question Answering in the Indian Context12
Sense-specific Historical Word Usage Generation12
Efficient Long-Text Understanding with Short-Text Models12
Robust Pronoun Fidelity with English LLMs: Are they Reasoning, Repeating, or Just Biased?12
Salute the Classic: Revisiting Challenges of Machine Translation in the Age of Large Language Models12
Objectifying the Subjective: Cognitive Biases in Topic Interpretations12
Pre-train, Prompt, and Recommendation: A Comprehensive Survey of Language Modeling Paradigm Adaptations in Recommender Systems12
A Confidence-based Acquisition Model for Self-supervised Active Learning and Label Correction12
Safe Pruning LoRA: Robust Distance-Guided Pruning for Safety Alignment in Adaptation of LLMs10
Modeling Emotion Dynamics in Song Lyrics with State Space Models10
Addressing the Binning Problem in Calibration Assessment through Scalar Annotations10
TaxoPro: A Plug-In LoRA-based Cross-Domain Method for Low-Resource Taxonomy Completion10
NLP Security and Ethics, in the Wild10
Adding Chocolate to Mint : Mitigating Metric Interference in Machine Translation10
Human Choice Prediction in Language-based Persuasion Games: Simulation-based Off-Policy Evaluation10
MENLI: Robust Evaluation Metrics from Natural Language Inference10
Helpful Neighbors: Leveraging Neighbors in Geographic Feature Pronunciation9
Rescue Conversations from Dead-ends: Efficient Exploration for Task-oriented Dialogue Policy Optimization9
TANQ: An Open Domain Dataset of Table Answered Questions9
Self-Rationalization in the Wild: A Large-scale Out-of-Distribution Evaluation on NLI-related tasks9
MultiBLiMP 1.0: A Massively Multilingual Benchmark of Linguistic Minimal Pairs9
PaniniQA: Enhancing Patient Education Through Interactive Question Answering9
Investigating Critical Period Effects in Language Acquisition through Neural Language Models9
Time-and-Space-Efficient Weighted Deduction9
How “Real” is Your Real-Time Simultaneous Speech-to-Text Translation System?9
Towards More Realistic Extraction Attacks: An Adversarial Perspective9
xcomet : Transparent Machine Translation Evaluation through Fine-grained Error Detection8
Visual Spatial Reasoning8
Patchwise Cooperative Game-based Interpretability Method for Large Vision-language Models8
Dissecting GraphRAG: A Modular Analysis of Knowledge Structuring for Factoid Question Answering8
Sub-Character Tokenization for Chinese Pretrained Language Models8
Large Language Models Enable Few-Shot Clustering8
Erratum: Language Models Can Resolve Reference Compositionally, But It’s Not Their Native Strength: The Case of the Personal Relation Task8
Not Eliminate but Aggregate: Post-Hoc Control over Mixture-of-Experts to Address Shortcut Shifts in Natural Language Understanding8
Data-driven Parsing Evaluation for Child-Parent Interactions8
Benchmarking Large Language Models for News Summarization7
Assessing the Capacity of Transformer to Abstract Syntactic Representations: A Contrastive Analysis Based on Long-distance Agreement7
QAmeleon: Multilingual QA with Only 5 Examples7
Evaluating Transformer Models and Human Behaviors on Chinese Character Naming7
The Causal Influence of Grammatical Gender on Distributional Semantics7
Know Your Limits: A Survey of Abstention in Large Language Models7
A Cross-Linguistic Pressure for Uniform Information Density in Word Order7
Step-by-Step Unmasking for Parameter-Efficient Fine-Tuning of Large Language Models7
Bridging the Gap: A Survey on Integrating (Human) Feedback for Natural Language Generation7
On the Effect of Instruction Tuning Loss on Generalization7
Direct Speech Translation for Automatic Subtitling7
Expectations over Unspoken Alternatives Predict Pragmatic Inferences6
Conformalizing Machine Translation Evaluation6
How Abstract Is Linguistic Generalization in Large Language Models? Experiments with Argument Structure6
CROC: Evaluating and Training T2I Metrics with Pseudo- and Human-Labeled Contrastive Robustness Checks6
Speak, Read and Prompt: High-Fidelity Text-to-Speech with Minimal Supervision6
Do Large Multimodal Models Solve Caption Generation for Scientific Figures? Lessons Learned from SciCap Challenge 20236
Visual Writing Prompts: Character-Grounded Story Generation with Curated Image Sequences6
A Comparative Approach for Auditing Multilingual Phonetic Transcript Archives6
Can Large Language Models Generalize Analogy Solving Like Children Can?6
CreoleVal: Multilingual Multitask Benchmarks for Creoles6
Hallucinations in Large Multilingual Translation Models6
Scope Ambiguities in Large Language Models6
Can Authorship Representation Learning Capture Stylistic Features?6
Chinese Idiom Paraphrasing6
Bridging Auxiliary Constraints to Resolve Instruction Following in Large Reasoning Models6
QE4PE: Word-level Quality Estimation for Human Post-Editing6
Goal Alignment in LLM-Based User Simulators for Conversational AI6
Abstractive Meeting Summarization: A Survey6
Visually Grounded Speech Models Have a Mutual Exclusivity Bias6
The Parallelism Tradeoff: Limitations of Log-Precision Transformers6
Collective Human Opinions in Semantic Textual Similarity5
mGPT: Few-Shot Learners Go Multilingual5
STPar: A Structure-Aware Triaffine Parser for Screenplay Character Coreference Resolution5
Self-Consistency Falls Short! The Adverse Effects of Positional Bias on Long-Context Problems5
Fine-tuning Large Language Models with Limited Data: A Survey and Practical Guide5
Meta-Learning a Cross-lingual Manifold for Semantic Parsing5
Improving the Domain Adaptation of Retrieval Augmented Generation (RAG) Models for Open Domain Question Answering5
Cultural Adaptation of Recipes5
Preferences for Idiomatic Language are Acquired Slowly — and Forgotten Quickly: A Case Study on Swedish5
AfriSpeech-200: Pan-African Accented Speech Dataset for Clinical and General Domain ASR5
ReCOGS: How Incidental Details of a Logical Form Overshadow an Evaluation of Semantic Interpretation4
Naturalistic Causal Probing for Morpho-Syntax4
Less is More: Mitigate Spurious Correlations for Open-Domain Dialogue Response Generation Models by Causal Discovery4
Shared Lexical Items as Triggers of Code Switching4
No Shortcuts to Culture: Indonesian Multi-hop Question Answering for Complex Cultural Understanding4
Lost in the Middle: How Language Models Use Long Contexts4
How Much Semantic Information is Available in Large Language Model Tokens?4
Cross-layer Attention Sharing for Pre-trained Large Language Models4
Can Authorship Attribution Models Distinguish Speakers in Speech Transcripts?4
A Unifying Scheme for Extractive Content Selection Tasks4
Hierarchical Indexing for Retrieval-Augmented Opinion Summarization4
Exploring Contrast Consistency of Open-Domain Question Answering Systems on Minimally Edited Questions4
MAKE: Memory-Associated Knowledge Editing4
KoBBQ: Korean Bias Benchmark for Question Answering4
FoVer: First-Order Logic Verification for Natural Language Reasoning4
Comparing Humans and Large Language Models on an Experimental Protocol Inventory for Theory of Mind Evaluation (EPITOME)4
Surveying the Landscape of Image Captioning Evaluation: A Comprehensive Taxonomy, Trends, and Metrics Analysis4
RepreGuard: Detecting LLM-Generated Text by Revealing Hidden Representation Patterns4
Eliciting the Translation Ability of Large Language Models via Multilingual Finetuning with Translation Instructions3
Automatically Correcting Large Language Models: Surveying the Landscape of Diverse Automated Correction Strategies3
Self-supervised Topic Taxonomy Discovery in the Box Embedding Space3
Decision-Oriented Dialogue for Human-AI Collaboration3
Explicitly Representing Syntax Improves Sentence-to-Layout Prediction of Unexpected Situations3
An Efficient Self-Supervised Cross-View Training For Sentence Embedding3
Reasoning over Public and Private Data in Retrieval-Based Systems3
Fodor and Pylyshyn’s Systematicity Challenge Still Stands3
Do LLMs Exhibit Human-like Response Biases? A Case Study in Survey Design3
Hate Speech Classifiers Learn Normative Social Stereotypes3
Optimal Transport Posterior Alignment for Cross-lingual Semantic Parsing3
Unleashing the True Potential of Sequence-to-Sequence Models for Sequence Tagging and Structure Parsing3
Automated Essay Scoring and Language Certification: Assessing Generalizability, Agreement and Validity for French3
Tracking Brand-Associated Polarity-Bearing Topics in User Reviews3
PiKGL: Leveraging Pruned Knowledge Graphs for Explainable Stance Detection3
PASTA: A Dataset for Modeling PArticipant STAtes in Narratives3
How Often Are Errors in Natural Language Reasoning Due to Paraphrastic Variability?3
The Impact of Word Splitting on the Semantic Content of Contextualized Word Representations3
Why Does Surprisal From Larger Transformer-Based Language Models Provide a Poorer Fit to Human Reading Times?3
Do Text Simplification Systems Preserve Meaning? A Human Evaluation via Reading Comprehension3
CRVQ: Channel-Relaxed Vector Quantization for Extreme Compression of LLMs3
Intent-calibrated Self-training for Answer Selection in Open-domain Dialogues3
A Survey on Model Compression for Large Language Models3
Learning Speech Representations with Variational Predictive Coding3
FINCH: Prompt-guided Key-Value Cache Compression for Large Language Models3
0.28948783874512