IEEE Transactions on Software Engineering

Papers
(The TQCC of IEEE Transactions on Software Engineering is 16. The table below lists those papers that are above that threshold based on CrossRef citation counts [max. 250 papers]. The publications cover those that have been published in the past four years, i.e., from 2022-08-01 to 2026-08-01.)
ArticleCitations
50 Years of Transactions on Software Engineering642
Computation Tree Logic Guided Program Repair444
Confirmation Bias and Time Pressure: A Family of Experiments in Software Testing323
Shield Broken: Black-Box Adversarial Attacks on LLM-Based Vulnerability Detectors152
Towards Scalable Model Checking of Reflective Systems via Labeled Transition Systems150
Can We Trust the Phone Vendors? Comprehensive Security Measurements on the Android Firmware Ecosystem144
Prevent: An Unsupervised Approach to Predict Software Failures in Production143
Just-in-Time Prediction of Software Architectural Changes Through Commit-Level Analyses143
The Why, When, What, and How About Predictive Continuous Integration: A Simulation-Based Investigation125
A Retrospective on Whole Test Suite Generation: On the Role of SBST in the Age of LLMs120
Deobfuscation of Control Flow Flattening Based on Abstract Interpretation117
Do as You Say: Consistency Detection of Data Practice in Program Code and Privacy Policy in Mini-App109
How Composite Metamorphic Relations Enhance Test Effectiveness of DNN Testing: An Empirical Study102
Visibility of Domain Elements in the Elicitation Process Interviews: A Family of Empirical Studies102
Answering Uncertain, Under-Specified API Queries Assisted by Knowledge-Aware Human-AI Dialogue95
Efficiently Testing Distributed Systems via Abstract State Space Prioritization92
Question Selection for Multimodal Code Search Synthesis Using Probabilistic Version Spaces90
Tackling Expressive Feature-Modeling Constructs with Pseudo-Boolean d-DNNF Compilation88
MalElves: Reinforcement Learning-Driven Adversarial Example Generation for Evading Cross-Platform ELF Malware Detection87
LHGLink: LLM—Enhanced Heterogeneous Graph Learning for Issue–Issue Link Prediction86
DPSSX: Detecting and Preventing SQLi and Stored XSS Induced by Expressions in Web Application SQL Statements via Static-Dynamic Analysis85
Influence of the 1990 IEEE TSE Paper “Automated Software Test Data Generation” on Software Engineering84
To Do or Not to Do: Semantics and Patterns for Do Activities in UML PSSM State Machines83
Are Your Dependencies Code Reviewed?: Measuring Code Review Coverage in Dependency Updates82
Spotting Setting-Related UI Display Bugs in Android Apps78
Enhancing Blockchain Robustness through Efficient Feedback-Driven Chaos Engineering78
A Declarative Metamorphic Testing Framework for Autonomous Driving76
RPHunter: Unveiling Rug Pull Schemes in Crypto Token via Code-and-Transaction Fusion Analysis75
DSSDPP: Data Selection and Sampling Based Domain Programming Predictor for Cross-Project Defect Prediction72
Enhancing Protocol Fuzzing via Diverse Seed Corpus Generation71
Automatic Fairness Testing of Neural Classifiers Through Adversarial Sampling70
Advanced Smart Contract Vulnerability Detection via LLM-Powered Multi-Agent Systems69
Enhancing Mobile App Bug Reporting via Real-Time Understanding of Reproduction Steps67
Enhancing Project-Specific Code Completion by Inferring Internal API Information66
Combining Genetic Programming and Model Checking to Generate Environment Assumptions61
Mission Specification Patterns for Mobile Robots: Providing Support for Quantitative Properties61
Multi-Granularity Detector for Vulnerability Fixes60
2023 Reviewers List59
Socio-Technical Grounded Theory for Software Engineering59
T-Evos: A Large-Scale Longitudinal Study on CI Test Execution and Failure58
Detecting Malicious Packages in PyPI and NPM by Clustering Installation Scripts58
Mole: Efficient Crash Reproduction in Android Applications With Enforcing Necessary UI Events58
An Empirical Study of Refactoring Rhythms and Tactics in the Software Development Process57
Boosting Compiler Fault Localization: Getting the Best of Both Worlds by Fusing Dynamic and Historical Data57
A Theory of Pending Schemas in Combinatorial Testing57
An Empirical Study of Software Refactorings in Real-World Open-Source Java Projects55
δ-SCALPEL: Docker Image Slimming Based on Source Code Static Analysis53
Efficient State Identification for Finite State Machine-Based Testing53
Towards Automated Discovery of Asymmetric Mempool DoS in Blockchains52
Measuring the Fidelity of a Physical and a Digital Twin Using Trace Alignments52
Esale: Enhancing Code-Summary Alignment Learning for Source Code Summarization51
GenMorph: Automatically Generating Metamorphic Relations via Genetic Programming51
MASTER: Multi-Source Transfer Weighted Ensemble Learning for Multiple Sources Cross-Project Defect Prediction50
Automated Code Editing With Search-Generate-Modify50
Systematic Evaluation and Usability Analysis of Formal Methods Tools for Railway Signaling System Design49
Multimodal Fusion for Android Malware Detection Based on Large Pre-Trained Models49
P-NPR: Practical Neural Program Repair via Learning to Ensemble49
Robust Test Selection for Deep Neural Networks48
Towards a Cognitive Model of Dynamic Debugging: Does Identifier Construction Matter?48
Mutation Testing in Practice: Insights From Open-Source Software Developers48
Mask–Mediator–Wrapper Architecture as a Data Mesh Driver47
Trace Diagnostics for Signal-Based Temporal Properties47
Neural Library Recommendation by Embedding Project-Library Knowledge Graph47
A Systematic Review of IoT Systems Testing: Objectives, Approaches, Tools, and Challenges47
Multi-Objective Software Defect Prediction via Multi-Source Uncertain Information Fusion and Multi-Task Multi-View Learning43
A Survey for LLM Agent Trajectory Analysis: From Failure Attribution to Enhancement43
MBL-CPDP: A Multi-Objective Bilevel Method for Cross-Project Defect Prediction42
Mitigating False Positive Static Analysis Warnings: Progress, Challenges, and Opportunities42
How Should Software Engineering Secondary Studies Include Grey Material?41
Decision Support for Selecting Blockchain-Based Application Design Patterns With Layered Taxonomy and Quality Attributes41
A Faceted Taxonomy of Requirements Changes in Agile Contexts41
EpiTESTER: Testing Autonomous Vehicles With Epigenetic Algorithm and Attention Mechanism41
Human-in-the-Loop Automatic Program Repair41
From Open Source to Industry: A Replication Study of AAA Test Structure in Envestnet41
Program Synthesis for Cyber-Resilience40
Evaluating and Improving GPT-Based Expansion of Abbreviations39
Weighted Community Division for Automated Software Architecture Refactoring39
An Experience Report on Producing Verifiable Builds for Large-Scale Commercial Systems39
TransformCode: A Contrastive Learning Framework for Code Embedding via Subtree Transformation39
An Empirical Study of Parameter-Efficient Fine-Tuning in Code Change Learning and Beyond38
AC2Next: A Novel Model That Can Predict the Next Animation API by Fusing the Animation API Context and the UI Animation Task38
Evolutionary generation of test suites for multi-path coverage of MPI programs with non-determinism38
Generalized Coverage Criteria for Combinatorial Sequence Testing38
Context-Aware Personalized Crowdtesting Task Recommendation38
Triple Peak Day: Work Rhythms of Software Developers in Hybrid Work37
Annotative Software Product Line Analysis Using Variability-Aware Datalog36
Discovering Reusable Functional Features in Legacy Object-Oriented Systems36
Legion: Massively Composing Rankers for Improved Bug Localization at Adobe36
LLMorpheus: Mutation Testing Using Large Language Models35
Pathidea: Improving Information Retrieval-Based Bug Localization by Re-Constructing Execution Paths Using Logs35
On the Understandability of MLOps System Architectures35
API2Vec++: Boosting API Sequence Representation for Malware Detection and Classification35
Untangling Intricate Dependencies: Characterizing and Resolving Software Package Dependencies35
Leveraging Large Language Model for Automatic Patch Correctness Assessment35
Pull Request Decisions Explained: An Empirical Overview34
Automated Refactoring of Non-Idiomatic Python Code With Pythonic Idioms34
When Voice Meets Touch: Conflict Analysis in Mobile Applications33
Evaluating and Improving Unified Debugging33
Typestate-Based Fault Localization of API Usage Violations in a Deep Learning Program33
Microservice Extraction Based on a Comprehensive Evaluation of Logical Independence and Performance33
Exploring and Analyzing Software Architecture Refactoring in Practice33
From Tea Leaves to System Maps: A Survey and Framework on Context-Aware Machine Learning Monitoring32
SigRec: Automatic Recovery of Function Signatures in Smart Contracts32
Self-Admitted GenAI Usage in Open-Source Software32
“Estimating Software Project Effort Using Analogies”: Reflections After 28 Years32
Automated Commit Message Generation With Large Language Models: An Empirical Study and Beyond31
How Do Developers Structure Unit Test Cases? An Empirical Analysis of the AAA Pattern in Open Source Projects31
Empirical Validation of Automated Vulnerability Curation and Characterization31
DAppSCAN: Building Large-Scale Datasets for Smart Contract Weaknesses in DApp Projects31
DiffGAN: A Test Generation Approach for Differential Testing of Deep Neural Networks for Image Analysis31
What Drives and Sustains Self-Assignment in Agile Teams31
Retrieval-Augmented Fine-Tuning for Improving Retrieve-and-Edit Based Assertion Generation30
Practitioners’ Expectations on Log Anomaly Detection30
Test Flakiness Across Programming Languages30
Increasing the Confidence of Deep Neural Networks by Coverage Analysis30
Evaluation of Static Vulnerability Detection Tools With Java Cryptographic API Benchmarks30
Just-In-Time Obsolete Comment Detection and Update29
Not All Synthetic Vulnerabilities Are What You Need for Training Deep Vulnerability Detectors29
Cost-Effective Adversarial Attacks Against Code LLM With Model Attention28
A Study About the Knowledge and Use of Requirements Engineering Standards in Industry28
Beyond Functional Correctness: Exploring Hallucinations in LLM-Generated Code28
Detecting Continuous Integration Skip Commits Using Multi-Objective Evolutionary Search27
Nighthawk: Fully Automated Localizing UI Display Issues via Visual Understanding27
Specializing Neural Networks for Cryptographic Code Completion Applications27
Mind the Gap! A Study on the Transferability of Virtual Versus Physical-World Testing of Autonomous Driving Systems27
Towards Exploring Developers’ Struggles in Developing Upgradeable Smart Contracts26
A Systematic Study on Real-World Android App Bundles26
From Executable Specifications to Hard-to-Specify Requirements: Challenges in Describing Reactive System Behavior26
Effect of Requirements Analyst Experience on Elicitation Effectiveness: A Family of Quasi-Experiments26
The Analysis of Safety Critical Software Systems26
Reaching Software Quality for Bioinformatics Applications: How Far Are We?26
Improving Cross-Language Code Clone Detection via Code Representation Learning and Graph Neural Networks25
AdaptGen: A Problem-Adaptive Solution Template Generation Technique for Online Programming Platforms25
Cross-Language Taint Analysis: Generating Caller-Sensitive Native Code Specification for Java25
Predictive Comment Updating With Heuristics and AST-Path-Based Neural Learning: A Two-Phase Approach25
No Resource, No Benchmarks, No Problem? Evaluating and Improving LLMs for Code Generation in No-Resource Languages25
Deconstructing the Nature of Collaboration in Organizations Open Source Software Development: The Impact of Developer and Task Characteristics25
Provably Valid and Diverse Mutations of Real-World Media Data for DNN Testing25
Causality-Aware Safety Testing for Autonomous Driving Systems25
STRE: An Automated Approach to Suggesting App Developers When to Stop Reading Reviews25
Understanding the Robustness of Transformer-Based Code Intelligence via Code Transformation: Challenges and Opportunities24
FlakyFix: Using Large Language Models for Predicting Flaky Test Fix Categories and Test Code Repair24
Assessing Evaluation Metrics for Neural Test Oracle Generation24
NumScout: Unveiling Numerical Defects in Smart Contracts Using LLM-Pruning Symbolic Execution24
Does AI Code Review Lead to Code Changes? A Case Study of GitHub Actions24
CRPWarner: Warning the Risk of Contract-Related Rug Pull in DeFi Smart Contracts24
Large-Scale Empirical Analysis of Continuous Fuzzing: Insights From 1 Million Fuzzing Sessions24
On the Validity of Pre-Trained Transformers for Natural Language Processing in the Software Engineering Domain23
Diversity-Oriented Testing for Competitive Game Agent via Constraint-Guided Adversarial Agent Training23
Subgraph-Oriented Testing for Deep Learning Libraries23
Forecasting the Principal of Code Technical Debt in JavaScript Applications23
The Impact of Prompt Programming on Function-Level Code Generation23
CodeS+: Towards Assessing the Generalization Ability of Code Models Under Distribution Shift23
iTCRL: Causal-Intervention-Based Trace Contrastive Representation Learning for Microservice Systems23
A Grounded Theory of Cross-Community SECOs: Feedback Diversity Versus Synchronization23
Line-Level Defect Prediction by Capturing Code Contexts With Graph Convolutional Networks23
Parameterized Verification of Leader/Follower Systems via Arithmetic Constraints23
Automated Use-After-Free Detection and Exploit Mitigation: How Far Have We Gone?23
Towards Robust Detection for Malicious Injection Variants23
Beyond the Sum of Parts: Leveraging Entanglement for Bug Inducing Commit Localization22
DyCITO+: Scalable Deep Reinforcement Learning for Generating Class Integration Test Orders of Java Programs22
What Characterizes Pairwise Modular Smells?22
FCGHunter: Towards Evaluating Robustness of Graph-Based Android Malware Detection22
Mithra: Anomaly Detection as an Oracle for Cyberphysical Systems22
Unearthing Gas-Wasting Code Smells in Smart Contracts With Large Language Models22
Hashing Fuzzing: Introducing Input Diversity to Improve Crash Detection22
Beyond Literal Meaning: Uncover and Explain Implicit Knowledge in Code Through Wikipedia-Based Concept Linking22
Do Pretrained Language Models Indeed Understand Software Engineering Tasks?22
Stakeholder Preference Extraction From Scenarios22
Practical Mutation Testing at Scale: A view from Google21
Syntactic Versus Semantic Similarity of Artificial and Real Faults in Mutation Testing Studies21
ArchHypo: Managing Software Architecture Uncertainty Using Hypotheses Engineering21
The Power of Small LLMs: A Multi-Agent for Code Generation via Dynamic Precaution Tuning21
Understanding and Detecting Scalability Faults in Large-Scale Distributed Systems21
Range Specification Bug Detection in Flight Control System Through Fuzzing21
How Templated Requirements Specifications Inhibit Creativity in Software Engineering21
Let’s Talk With Developers, Not About Developers: A Review of Automatic Program Repair Research21
SmartOracle: Generating Smart Contract Oracle via Fine-Grained Invariant Detection21
A Variability Fault Localization Approach for Software Product Lines21
Automated Infrastructure as Code Program Testing20
Domain-Driven Design for Microservices: An Evidence-Based Investigation20
Learning to Predict User-Defined Types20
Misactivation-Aware Stealthy Backdoor Attacks on Neural Code Understanding Models20
Causes and Canonicalization of Unreproducible Builds in Java20
Boosting Generalizable Fairness With Mahalanobis Distances Guided Boltzmann Exploratory Testing20
Retrospective on: Constraint-Based Automatic Test Data Generation20
A Comparison of Natural Language Understanding Platforms for Chatbots in Software Engineering20
PopArt: Ranked Testing Efficiency20
Does Treatment Adherence Impact Experiment Results in TDD?20
The “Question Neighbourhood” Approach for Systematic Evaluation of Code-Generating LLMs19
Translating to a Low-Resource Language with Compiler Feedback: A Case Study on Cangjie19
Bridging Bug Localization and Issue Fixing: A Hierarchical Localization Framework Leveraging Large Language Models19
Onboarding Software Professionals in a Hybrid World19
Towards More Precise Coincidental Correctness Detection With Deep Semantic Learning19
Studying the Influence and Distribution of the Human Effort in a Hybrid Fitness Function for Search-Based Model-Driven Engineering19
Concretization of Abstract Traffic Scene Specifications Using Metaheuristic Search19
Boosting high-review-value warning line identification with unsupervised line-level defect prediction19
Verification of Fuzzy Decision Trees19
Clopper-Pearson Algorithms for Efficient Statistical Model Checking Estimation19
Engineering Within Boundaries When Software Has None19
Active Code Learning: Benchmarking Sample-Efficient Training of Code Models19
Generating Structurally Realistic Models With Deep Autoregressive Networks19
The Human Side of Software Engineering Teams: An Investigation of Contemporary Challenges19
Runtime Evolution of Bitcoin's Consensus Rules19
A Little Help Goes a Long Way: Tutoring LLMs in Solving Competitive Programming Through Hints19
Do Chase Your Tail! Missing Key Aspects Augmentation in Textual Vulnerability Descriptions of Long-Tail Software Through Feature Inference18
Efficient Black-Box Fault Localization for System-Level Test Code Using Large Language Models18
Emerging App Issue Identification via Online Joint Sentiment-Topic Tracing18
A Retrospective of Proving the Correctness of Multiprocess Programs18
Isolating Compiler Faults Through Differentiated Compilation Configurations18
SCAnoGenerator: Automatic Anomaly Injection for Ethereum Smart Contracts18
Stealthy Backdoor Attack for Code Models18
Accelerating Finite State Machine-Based Testing Using Reinforcement Learning18
An Assessment of Rules of Thumb for Software Phase Management, and the Relationship Between Phase Effort and Schedule Success18
Software Testing With Large Language Models: Survey, Landscape, and Vision18
PATEN: Identifying Unpatched Third-Party APIs via Fine-Grained Patch-Enhanced AST-Level Signature18
Evaluating SZZ Implementations: An Empirical Study on the Linux Kernel18
Multitask-Based Evaluation of Open-Source LLM on Software Vulnerability18
Dealing With Data Challenges When Delivering Data-Intensive Software Solutions18
Examiner-Pro: Testing Arm Emulators Across Different Privileges18
Factors Affecting On-Time Delivery in Large-Scale Agile Software Development17
A Framework for Evaluating GenAI Adoption and Use in Software Engineering17
Investigating the Feasibility of Conducting Webcam-Based Eye-Tracking Studies in Code Comprehension17
A Search-Based Testing Approach for Deep Reinforcement Learning Agents17
AddressWatcher: Sanitizer-Based Localization of Memory Leak Fixes17
RNN-Test: Towards Adversarial Testing for Recurrent Neural Network Systems17
Fast and Precise Static Null Exception Analysis With Synergistic Preprocessing17
A Framework for Emotion-Oriented Requirements Change Handling in Agile Software Engineering17
Finding Trends in Software Research17
Static Profiling of Alloy Models17
On the Workflows and Smells of Leaderboard Operations (LBOps): An Exploratory Study of Foundation Model Leaderboards17
DaNuoYi: Evolutionary Multitask Injection Testing on Web Application Firewalls17
Enforcing Correctness of Collaborative Business Processes Using Plans16
A Procedure to Continuously Evaluate Predictive Performance of Just-In-Time Software Defect Prediction Models During Software Development16
Using Symbolic States to Infer Numerical Invariants16
Let's Go to the Whiteboard (Again): Perceptions From Software Architects on Whiteboard Architecture Meetings16
HF-DGF: Hybrid Feedback Guided Directed Grey-box Fuzzing16
OpCodeBERT: A Method for Python Code Representation Learning by BERT With Opcode16
Malo in the Code Jungle: Explainable Fault Localization for Decentralized Applications16
FlexFL: Flexible and Effective Fault Localization With Open-Source Large Language Models16
MultiPL-E: A Scalable and Polyglot Approach to Benchmarking Neural Code Generation16
Active Learning of Discriminative Subgraph Patterns for API Misuse Detection16
What Makes Agile Software Development Agile?16
State of the Journal16
Darcy: Automatic Architectural Inconsistency Resolution in Java16
How Toxic Can You Get? Search-Based Toxicity Testing for Large Language Models16
Neural Transfer Learning for Repairing Security Vulnerabilities in C Code16
PackHunter: Recovering Missing Packages for C/C++ Projects16
SmarTracker: Leveraging LLMs to Automate User Behavior Tracking in Mobile Apps16
DT4LM: Differential Testing for Reliable Language Model Updates in Classification Tasks16
0.49586200714111