ACM Transactions on Software Engineering and Methodology

Papers
(The TQCC of ACM Transactions on Software Engineering and Methodology is 11. The table below lists those papers that are above that threshold based on CrossRef citation counts [max. 250 papers]. The publications cover those that have been published in the past four years, i.e., from 2022-08-01 to 2026-08-01.)
ArticleCitations
Finding Information Leaks with Information Flow Fuzzing—RCR Report814
KAPE: k NN-based Performance Testing for Deep Code Search197
Antidote or Placebo? Unraveling the Efficacy of Neuron Coverage Criteria on Testing Transformer-based Language Models185
Horus : Accelerating Kernel Fuzzing through Efficient Host-VM Memory Access Procedures183
Deceiving Humans and Machines Alike: Search-based Test Input Generation for DNNs Using Variational Autoencoders177
Towards Reliable Generation of Executable Workflows by Foundation Models140
Reusing d-DNNFs for Efficient Feature-Model Counting133
FoC: Figure Out the Cryptographic Functions in Stripped Binaries with LLMs132
FairGenerate: Enhancing Fairness through Synthetic Data Generation and Two-Fold Biased Labels Removal131
SPENCER: Self-Adaptive Model Distillation for Efficient Code Retrieval131
Enhancing Android Malware Detection: The Influence of ChatGPT on Decision-centric Task125
Neuron Semantic-Guided Test Generation for Deep Neural Networks Fuzzing120
Causality-Driven Test Case Minimisation for Cyber-Physical Systems117
VFDelta: A Framework for Detecting Silent Vulnerability Fixes by Enhancing Code Change Learning111
Assessing the Robustness of Test Selection Methods for Deep Neural Networks111
Bounded Verification of Atomicity Violations for Interrupt-Driven Programs via Lazy Sequentialization102
Preference-wise Testing of Android Apps via Test Amplification89
M2CVD: Enhancing Vulnerability Understanding through Multi-Model Collaboration for Code Vulnerability Detection88
Understanding the OSS Communities of Deep Learning Frameworks: A Comparative Case Study of P y T orch and T ensor87
Unraveling the Key of Machine Learning-based Android Malware Detection86
A Survey on Failure Analysis and Fault Injection in AI Systems84
Securing the Ethereum from Smart Ponzi Schemes: Identification Using Static Features83
Better Supporting Human Aspects in Mobile eHealth Apps: Development and Validation of Enhanced Guidelines80
TestLoop: A Process Model Describing Human-in-the-Loop Software Test Suite Generation73
Automatic Identification of Game Stuttering via Gameplay Videos Analysis72
An empirical study on vulnerability disclosure management of open source software systems72
Test Generation Strategies for Building Failure Models and Explaining Spurious Failures69
I Depended on You and You Broke Me: An Empirical Study of Manifesting Breaking Changes in Client Packages67
An Empirical Analysis of Machine Learning Model and Dataset Documentation, Supply Chain, and Licensing Challenges on Hugging Face67
A Systematic Literature Review on Large Language Models for Automated Program Repair65
Communicating Study Design Trade-offs in Software Engineering60
An Empirical Study of the Non-Determinism of ChatGPT in Code Generation59
History-Driven Fuzzing for Deep Learning Libraries57
Toward Interpretable Graph Tensor Convolution Neural Network for Code Semantics Embedding52
Actionable Framework for Understanding and Improving Social and Human Factors that Influence the Requirements Management in Software Ecosystems52
Do Current Language Models Support Code Intelligence for R Programming Language?51
Adopting Two Supervisors for Efficient Use of Large-Scale Remote Deep Neural Networks - RCR Report51
A Comprehensive View on TD Prevention Practices and Reasons for Not Preventing It51
Stakeholder Value Criteria for Technical Debt Acquisition Decisions: An Empirical Analysis51
Towards an Oracle for Binary Decomposition Under Compilation Variance49
JavaScript SBST Heuristics to Enable Effective Fuzzing of NodeJS Web APIs49
FormatFuzzer : Effective Fuzzing of Binary File Formats48
Help Them Understand: Testing and Improving Voice User Interfaces47
Towards Automating Domain-Specific Data Generation for Text-to-SQL: A Comprehensive Approach46
Storage State Analysis and Extraction of Ethereum Blockchain Smart Contracts45
JIT-DCK: A KAN Multi-Task Model for Just-In-Time Code Defect Prediction and Localization45
FAVDisco : Modeling and Discovering File Access Vulnerabilities45
An Empirical Study on Governance in Bitcoin’s Consensus Evolution44
Surveying the Benchmarking Landscape of Large Language Models in Code Intelligence44
Fine-Tuning Large Language Models to Improve Accuracy and Comprehensibility of Automated Code Review44
PVDetector: Pretrained Vulnerability Detection on Vulnerability-enriched Code Semantic Graph42
Characterizing Deep Learning Package Supply Chains in PyPI: Domains, Clusters, and Disengagement42
Confused Deputy Attack Against Model Context Protocol42
Estimating Uncertainty in Labeled Changes by SZZ Tools on Just-In-Time Defect Prediction41
Systematic Literature Review on Software Security Vulnerability Information Extraction40
Try with Simpler - An Evaluation of Improved Principal Component Analysis in Log-based Anomaly Detection39
Enhancing Security and Acuity of Smart Contract Vulnerability Detection Based on Federated Learning and BiLSTM-Attention39
Supporting Emotional Intelligence, Productivity and Team Goals while Handling Software Requirements Changes39
An Accurate Identifier Renaming Prediction and Suggestion Approach38
Deep API Sequence Generation via Golden Solution Samples and API Seeds38
Can LLMs Hack Enterprise Networks? — RCR Report37
Single and Multi-objective Test Cases Prioritization for Self-driving Cars in Virtual Environments37
A Survey of Learning-based Automated Program Repair36
HeMiRCA: Fine-Grained Root Cause Analysis for Microservices with Heterogeneous Data Sources35
I Know What You Are Searching for: Code Snippet Recommendation from Stack Overflow Posts35
Assessing and Analyzing the Correctness of GitHub Copilot’s Code Suggestions35
Introducing Interactions in Multi-Objective Optimization of Software Architectures35
When Fine-Tuning LLMs Meets Data Privacy: An Empirical Study of Federated Learning in LLM-Based Program Repair34
APIRO: A Framework for Automated Security Tools API Recommendation34
Towards AI-Native Software Engineering (SE 3.0): A Vision and a Challenge Roadmap34
An Empirical Study on GitHub Pull Requests’ Reactions34
Contemporary Software Modernization: Strategies, Driving Forces, and Research Opportunities33
Deep Learning Framework Testing via Heuristic Guidance Based on Multiple Model Measurements33
SimADFuzz: Simulation-Feedback Fuzz Testing for Autonomous Driving Systems33
Assessing the Early Bird Heuristic (for Predicting Project Quality)33
JIT-MTL: Just-in-Time Defect Localization and Prediction with Multi-Task Learning33
GIST : Generated Inputs Sets Transferability in Deep Learning32
Editorial: Toward the Future with Eight Issues Per Year32
AutoRIC: Automated Neural Network Repairing Based on Constrained Optimization32
Editorial: ICSE and the Incredible Contradictions of Software Engineering32
A Survey of Learning-based Method Name Prediction31
F-800: Agentic Fuzzing Termination via Function Clustering: Toward Smarter Early Stopping31
SimClone: Detecting Tabular Data Clones Using Value Similarity31
Detection of Technical Debt in Java Source Code31
Towards On-the-Fly Code Performance Profiling31
On-the-Fly Generation-Quality Enhancement of Deep Code Models via Model Collaboration30
Vulnerability Repair via Concolic Execution and Code Mutations30
PatchCensor: Patch Robustness Certification for Transformers via Exhaustive Testing30
SCOPE : Performance Testing for Serverless Computing30
Exploring Data-Efficient Adaptation of Large Language Models for Code Generation30
HumanEval-V: Systematic Evaluation of Visual Reasoning in Large Multimodal Models for Code Generation30
Leveraging Reviewer Experience in Code Review Comment Generation30
Reinforcement Learning Informed Evolutionary Search for Autonomous Systems Testing29
A Systematic Literature Review of Multi-Label Learning in Software Engineering29
ADSDx : Towards Automated Accident Diagnosis for High-level Autonomous Driving Systems29
Mapping the Trust Terrain: LLMs in Software Engineering - Insights and Perspectives29
Code-Enhanced Cross-Perspective Bug Question Retrieval28
Revisiting the Identification of the Co-evolution of Production and Test Code28
An Empirical Study of Retrieval-Augmented Code Generation: Challenges and Opportunities28
SourcererJBF: A Java Build Framework For Large-Scale Compilation27
Simulator-based Explanation and Debugging of Hazard-triggering Events in DNN-based Safety-critical Systems27
Why Do GitHub Actions Workflows Fail? An Empirical Study27
Beyond Fidelity: Explaining Vulnerability Localization of Learning-Based Detectors27
Mapping NVD Records to Their Vulnerability-fixing Commits: How Hard is It?26
Towards Learning Generalizable Code Embeddings Using Task-agnostic Graph Convolutional Networks26
Demo2Test: Transfer Testing of Agent in Competitive Environment with Failure Demonstrations26
You Don’t Have to Say Where to Edit! jLED—Joint Learning to Localize and Edit Source Code26
Characterizing Installation- and Run-time Compatibility Issues in Android Benign Apps and Malware26
Ethical Prompt Engineering for AI-driven SE: Evidence-informed Interaction-time Governance Roadmap to 203025
Demystifying Hidden Sensitive Operations in Android Apps25
Graphuzz: Data-driven Seed Scheduling for Coverage-guided Greybox Fuzzing25
Test Input Prioritization for 3D Point Clouds25
Commit Messages Generation Based on Core Changes25
A Comprehensive Empirical Study of Bias Mitigation Methods for Machine Learning Classifiers24
Automatic Rule Checking for Microservices: Supporting Security Analysis with Explainability24
Security of Language Models for Code: A Systematic Literature Review24
Exploring Fine-Grained Bug Report Categorization with Large Language Models and Prompt Engineering: An Empirical Study24
Efficient Multivariate Time Series Anomaly Detection through Transfer Learning for Large-Scale Software Systems24
Digital Twin-based Out-of-Distribution Detection in Autonomous Vessels23
MR-Scout: Automated Synthesis of Metamorphic Relations from Existing Test Cases23
Cleaning Up Confounding: Accounting for Endogeneity Using Instrumental Variables and Two-Stage Models23
Exploring the Capabilities of LLMs for Code-Change-Related Tasks23
Monitoring Data for Anomaly Detection in Cloud-Based Systems: A Systematic Mapping Study23
A Characterization Study of Merge Conflicts in Java Projects23
Measure Twice, Locate Once: Mitigating Hallucinations in LLM-based Agents for Repository-Scale Fault Localization23
Actor-Driven Decomposition of Microservices through Multi-level Scalability Assessment23
Assessing and Improving Prompting Large Language Models for Software Vulnerability Analysis23
SPOLRE: Semantic Preserving Object Layout Reconstruction for Image Captioning System Testing22
Learning from Very Little Data: On the Value of Landscape Analysis for Predicting Software Project Health22
Automating TODO-missed Methods Detection and Patching22
Enhancing Task In-Progress Time Predictions through Affective and Personality Factors22
A Mixed-Method Study of Hot Fixing in Industry: Practices, Bottlenecks, and Opportunities22
Fold2Vec: Towards a Statement-Based Representation of Code for Code Comprehension21
Adaptive Modelling Languages: Abstract Syntax and Model Migration21
Certified Cost Bounds for Abstract Programs21
Variable Renaming-Based Adversarial Test Generation for Code Model: Benchmark and Enhancement21
Adopting Two Supervisors for Efficient Use of Large-Scale Remote Deep Neural Networks21
Stress Testing Control Loops in Cyber-Physical Systems—RCR Report21
Advancing LLM-Based Issue Report Classification with Explained Few-Shot Learning, Intent Extraction, Ensemble, and Summarization21
Automatically Checking Semantic Equivalence between Versions of Large-Scale C Projects21
Evolution-Aware Constraint Derivation Approach for Software Remodularization21
Bypassing Guardrails: Lessons Learned from Red Teaming ChatGPT20
Developer Perspectives on Licensing and Copyright Issues Arising from Generative AI for Software Development20
Improving Deep Assertion Generation via Fine-Tuning Retrieval-Augmented Pre-trained Language Models20
Battling against Protocol Fuzzing: Protecting Networked Embedded Devices from Dynamic Fuzzers20
Programming Smart Playtesting20
Duplicate Bug Report Detection: How Far Are We?20
Coverage-directed Differential Testing of X.509 Certificate Validation in SSL/TLS Implementations20
Efficient Management of Containers for Software Defined Vehicles20
Interpreting Deep Neural Networks via Relative Activation-Deactivation Abstractions20
Autonomous Driving System Testing via Diversity-Oriented Driving Scenario Exploration19
Testing Causality in Scientific Modelling Software19
Revisiting Vulnerability Patch Identification on Data in the Wild19
Is It Hard to Generate Holistic Commit Message?19
Fairness Concerns in App Reviews: A Study on AI-Based Mobile Apps19
Generation-based Differential Fuzzing for Deep Learning Libraries19
Measuring and Clustering Heterogeneous Chatbot Designs19
Software Vulnerabilities as Cognitive Blindspots; Assessing the Suitability of a Dual Processing Theory of Decision Making for Secure Coding19
PonziHunter: Hunting Ethereum Ponzi Contract via Static Analysis and Contrastive Learning on the Bytecode Level19
MeDeT: Medical Device Digital Twins Creation with Few-shot Meta-learning19
A Roadmap for Integrating Sustainability into Software Engineering Education19
ActRef: Enhancing the Understanding of Python Code Refactoring with Action-Based Analysis19
An In-depth Study of Java Deserialization Remote-Code Execution Exploits and Vulnerabilities19
Refactoring in Computational Notebooks18
An Interleaving Guided Metamorphic Testing Approach for Concurrent Programs18
Studying the Impact of TensorFlow and PyTorch Bindings on Machine Learning Software Quality18
Survey of Code Search Based on Deep Learning18
DiPri : Distance-Based Seed Prioritization for Greybox Fuzzing18
On the Reruns of GitHub Actions Workflows18
Differentiable Quantum Programming with Unbounded Loops18
Reference-Based Retrieval-Augmented Unit Test Generation18
Complete, Sound, and Scalable Identification of Minimal Failure-Causing Schema18
Visualization Task Taxonomy to Understand the Fuzzing Internals18
Understanding the Fundamental Design Decisions of Retrieval-Augmented Generation Systems18
On the Impact of Lower Recall and Precision in Defect Prediction for Guiding Search-based Software Testing17
On the Significance of Category Prediction for Code-Comment Synchronization17
Confidence vs. Competence: Misalignment in Judgment and Performance for Agentic Software Repair17
Exploring Development Methods for Reactive Synthesis Specifications17
PseudoBridge: Pseudo Code as the Bridge for Better Semantic and Logic Alignment in Code Retrieval17
Preparation and Utilization of Mixed States for Testing Quantum Programs17
AdaptiveLog: An Adaptive Log Analysis Framework with the Collaboration of Large and Small Language Model17
CITYWALK : Enhancing LLM-Based C++ Unit Test Generation via Project-Dependency Awareness and Language-Specific Knowledge17
MORepair : Teaching LLMs to Repair Code via Multi-Objective Fine-Tuning17
Inferring Input Grammars from Code with Symbolic Parsing16
Input Distribution Coverage: Measuring Feature Interaction Adequacy in Neural Network Testing16
LogUpdater : Automated Detection and Repair of Specific Defects in Logging Statements16
AI for DevSecOps: A Landscape and Future Opportunities16
A Large-Scale Empirical Evaluation of LLMs for Automated Self-Admitted Technical Debt Repayment16
A Comparative Study on Method Comment and Inline Comment16
Automatic Core-Developer Identification on GitHub: A Validation Study16
C2|Q>: A Robust Framework for Bridging Classical and Quantum Software Development16
Reputation Gaming in Crowd Technical Knowledge Sharing16
Obfuscated Clone Search in JavaScript based on Reinforcement Subsequence Learning16
Challenges in Testing Large Language Model Based Software: A Faceted Taxonomy16
Assessing and Advancing Benchmarks for Evaluating Large Language Models in Software Engineering Tasks16
Can GitHub Issues Help in App Review Classifications?16
The IDEA of Us: An Identity-Aware Architecture for Autonomous Systems16
The Influence of Human Aspects on Requirements Engineering-related Activities: Software Practitioners’ Perspective16
Sustainability of Machine Learning-Enabled Systems: The Machine Learning Practitioner’s Perspective15
Rise of Distributed Deep Learning Training in the Big Model Era: From a Software Engineering Perspective15
Understanding Inconsistent State Update Vulnerabilities in Smart Contracts15
The Influence of Environmental Odors on Student Programmers During Code Comprehension and Code Writing15
A Comprehensive Multi-Vocal Empirical Study of ML Cloud Service Misuses15
Large Language Models for Cyber Security: A Systematic Literature Review15
Testing RESTful APIs: A Survey15
Identifying and Explaining Safety-critical Scenarios for Autonomous Vehicles via Key Features15
Mitigating Regression Faults Induced by Feature Evolution in Deep Learning Systems15
VulDeNoise: Outlier Detection to Reduce Label Noises for Effective Vulnerability Detection15
A Hypothesis Testing-based Framework for Software Cross-modal Retrieval in Heterogeneous Semantic Spaces15
Data Complexity: A New Perspective for Analyzing the Difficulty of Defect Prediction Tasks15
PanicFI: An Infrastructure for Fixing Panic Bugs in Real-World Rust Programs15
Towards Practical Binary Code Similarity Detection: Vulnerability Verification via Patch Semantic Analysis15
Understanding Real-Time Collaborative Programming: A Study of Visual Studio Live Share15
Type-aware LLM-based Test Generation for Python Programs15
Theory of Troubleshooting: The Developer’s Cognitive Experience of Overcoming Confusion15
Booster: Effective and Efficient Web GUI Trace Reduction Based on Multi-Level State Abstraction and Time-Guided Hierarchical Delta Debugging15
Let’s Discover More API Relations: A Large Language Model-Based AI Chain for Unsupervised API Relation Inference15
NSFuzz: Towards Efficient and State-Aware Network Service Fuzzing15
Software Engineering by and for Humans in an AI Era15
Can Coverage Criteria Guide Failure Discovery for Image Classifiers? An Empirical Study15
Fairness Testing of Machine Translation Systems14
A-ProS: Towards Reliable Autonomous Programming Through Multi-Model Feedback14
Exploring JVM Garbage Collector Testing with Event-Coverage14
Arash : Token-Efficient LLM-Assisted Crash Root Cause Analysis in Fuzz Driver Generation14
Simulating Software Evolution to Evaluate the Reliability of Early Decision-making among Design Alternatives toward Maintainability14
Decision Support Model for Selecting the Optimal Blockchain Oracle Platform: An Evaluation of Key Factors14
Toward Better Comprehension of Breaking Changes in the NPM Ecosystem14
Open Problems in Fuzzing RESTful APIs: A Comparison of Tools14
Less Is More: Unlocking Semi-Supervised Deep Learning for Vulnerability Detection14
Mobile Application Online Cross-Project Just-in-Time Software Defect Prediction Framework14
Test Oracle Generation for REST APIs14
Dyn NPC: Finding More Violations Induced by ADS in Simulation Testing via Dynamic NPC Behavior Generation14
Automated Abstract Transformer Synthesis for Reduced Product Domains14
Some Seeds Are Strong: Seeding Strategies for Search-based Test Case Selection14
Representation Learning for Stack Overflow Posts: How Far Are We?14
The Havoc Paradox in Generator-Based Fuzzing—RCR Report14
Addressing OSS Community Managers’ Challenges in Contributor Retention13
JSTestCraft : Addressing Context Deficits in JavaScript Unit Test Generation via Agentic Multi-Level Contextual Analysis13
Revisiting Sentiment Analysis for Software Engineering in the Era of Large Language Models13
Verification Witnesses13
How the Quality of Maintenance Tasks is Affected by Criteria for Selecting Engineers for Collaboration13
An Empirical Study on the Relationship between Defects and Source Code’s Unnaturalness13
Large Language Model for Vulnerability Detection and Repair: Literature Review and the Road Ahead13
Analysis of EMF Meta-Model Duplication in Open Source Repositories13
What Constitutes the Deployment and Runtime Configuration System? An Empirical Study on OpenStack Projects13
Grammar Mutation for Testing Input Parsers13
Revealing the Unseen: AI Chain on LLMs for Predicting Implicit Dataflows to Generate Dataflow Graphs in Dynamically Typed Code13
Evaluating Incompatible Third-party Library API Usage in LLM-based Code Completion13
Divide-and-Conquer: Automating Code Revisions via Localization-and-Revision13
Identifying Performance Issues in Cloud Service Systems Based on Relational-Temporal Features13
0.47679400444031