Software Engineering
Papers filed under cs.SE on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
241 to 300 of 1,389
A Quantitative and Qualitative Evaluation of LLM-Based Explainable Fault Localization
Sungmin Kang, Gabin An, Shin Yoo
cs.SEarXiv:2308.05487v32023How Often Do Single-Statement Bugs Occur? The ManySStuBs4J Dataset
Rafael-Michael Karampatsis, Charles Sutton
cs.SEcs.PLarXiv:1905.13334v22019Can LLMs Extract Architectural Design Decisions from Source Code Commits? - A Preliminary Exploratory Study
Amey Karan, Rudra Dhar, Mohamed Soliman +1
cs.SEcs.AIarXiv:2609.03721v12026Where Are The Gaps? A Systematic Mapping Study of Infrastructure as Code Research
Akond Rahman, Rezvan Mahdavi-Hezaveh, Laurie Williams
cs.SEarXiv:1807.04872v12018An Empirical Study of Deep Learning Models for Vulnerability Detection
Benjamin Steenhoek, Md Mahbubur Rahman, Richard Jiles +1
cs.SEcs.CRcs.LGarXiv:2212.08109v32022Stochastic Semantics and Statistical Model Checking for Networks of Priced Timed Automata
Alexandre David, Kim G. Larsen, Axel Legay +4
cs.SEarXiv:1106.3961v22011Generalization or Memorization: Data Contamination and Trustworthy Evaluation for Large Language Models
Yihong Dong, Xue Jiang, Huanyu Liu +4
cs.CLcs.AIcs.CRarXiv:2402.15938v32024The Psychological Costs of Artificial Intelligence Adoption in Software Engineering
Adam Alami, Elda Paja, Abhishek Tiwari
cs.SEcs.AIarXiv:2609.03456v12026SWIM: Synthesizing What I Mean
Mukund Raghothaman, Yi Wei, Youssef Hamadi
cs.SEarXiv:1511.08497v22015Challenges and solutions when adopting DevSecOps: A systematic review
Roshan N. Rajapakse, Mansooreh Zahedi, M. Ali Babar +1
cs.SEarXiv:2103.08266v22021Perfection Not Required? Human-AI Partnerships in Code Translation
Justin D. Weisz, Michael Muller, Stephanie Houde +5
cs.HCcs.SEarXiv:2104.03820v12021Fill in the Blank: Context-aware Automated Text Input Generation for Mobile GUI Testing
Zhe Liu, Chunyang Chen, Junjie Wang +4
cs.SEarXiv:2212.04732v12022Large Language Model for Vulnerability Detection and Repair: Literature Review and the Road Ahead
Xin Zhou, Sicong Cao, Xiaobing Sun +1
cs.SEarXiv:2404.02525v32024Microservices: How To Make Your Application Scale
Nicola Dragoni, Ivan Lanese, Stephan Thordal Larsen +3
cs.SEarXiv:1702.07149v12017A Survey on Automated Driving System Testing: Landscapes and Trends
Shuncheng Tang, Zhenya Zhang, Yi Zhang +8
cs.SEarXiv:2206.05961v22022ReCode: Robustness Evaluation of Code Generation Models
Shiqi Wang, Zheng Li, Haifeng Qian +11
cs.LGcs.CLcs.SEarXiv:2212.10264v12022You Cannot Fix What You Cannot Find! An Investigation of Fault Localization Bias in Benchmarking Automated Program Repair Systems
Kui Liu, Anil Koyuncu, Tegawendé F. Bissyandé +3
cs.SEarXiv:1812.07283v22018Automated Synthesis of Cloud Emulators
Archit Bhatnagar, Zhenning Yang, Sarah McClure +3
cs.SEcs.AIcs.DCarXiv:2608.23842v12026LongCoder: A Long-Range Pre-trained Language Model for Code Completion
Daya Guo, Canwen Xu, Nan Duan +2
cs.SEcs.AIcs.CLarXiv:2306.14893v12023Why Early-Stage Software Startups Fail: A Behavioral Framework
Carmine Giardino, Xiaofeng Wang, Pekka Abrahamsson
cs.SEarXiv:1709.04749v12017Aroma: Code Recommendation via Structural Code Search
Sifei Luan, Di Yang, Celeste Barnaby +2
cs.SEarXiv:1812.01158v42018Pynguin: Automated Unit Test Generation for Python
Stephan Lukasczyk, Gordon Fraser
cs.SEarXiv:2202.05218v12022DTM: Deterministic Approaches for Black-box Test Suite Minimization with Tree-based Similarity
Md Siam, Shartaz Sajid Nahid, Md Arif Hasan +2
cs.SEarXiv:2609.04205v12026Breaking the Alphabet: Rethinking File Ordering in Code Review
Md Shamimur Rahman, Zadia Codabux, Chanchal K. Roy
cs.SEarXiv:2609.04207v12026AI Writes Code, Humans Pay the Debt. An Empirical Study on the Sustainability and Evolution of Agent-Generated Code
Antonino Coppola, Matteo Esposito, Rick Kazman +1
cs.SEarXiv:2609.04208v12026SH-PDOPS: AI-Driven Cloud Native Enterprise Reliability Framework for Predictive Analytics and Intelligent DevOps Automation
Ayushman Bosu Roy
cs.SEarXiv:2609.04210v12026Big Questions on Software Architecture: Report of the ICSE 2026 BoF on Software Architecture
Davide Taibi, Patricia Lago, Henry Muccini
cs.SEarXiv:2609.04212v12026Engineering as Code: Bringing Software Engineering Methodology to Engineering Design
Song Difei
cs.SEarXiv:2609.04216v12026Robustness and Trade-offs for Code LLMs on Protected Code
Jin Wen, Yuejun Guo, Yujie Ma +2
cs.SEarXiv:2609.04220v12026How Developers Discuss Generative AI: A Longitudinal Study of the Visual Studio Code Community
Panida Rumriankit, Akito Monden, Hiroki Inayoshi +4
cs.SEarXiv:2609.04680v12026Data-Related Challenges and Requirements for Event Log Generation in Process Mining: A Systematic Literature Review
Ghita El Alaoui Talibi, Oleksandr Kosenkov, Anastasija Nikiforova
cs.SEarXiv:2609.04211v12026A Mixed-Method Empirical Study of LLM Assistance in Software Engineering Workflows
Pamali D. Weerasinghe, Roshan N. Rajapakse, Isuru Dharmadasa +1
cs.SEarXiv:2609.04214v12026Toward Model-Driven Digital Twin Configuration: Separating Structure Semantics and Runtime with SysML SAREF and Ditto
Andrey Sadovykh, Matthew Rusakov, Kirill Korikov
cs.SEarXiv:2609.04213v12026A Governance Methodology Layer for AI-Assisted Software Development: Defect Taxonomy, Controlled Ablation, and Process-Over-Capability Evidence
Sungjin Kwon
cs.SEarXiv:2609.04218v12026Large Language Models for Fuzz Testing in Microservices: A Systematic Literature Review
Ying Song, Ke Ping, Yuqing Wang +1
cs.SEarXiv:2609.04219v12026CPL: A Compact C-like Systems Language with Explicit Low-Level Control
Nikolay Fot, Alexander Vinarsky
cs.PLcs.SEarXiv:2609.04904v12026Ritgard: T(r)opical Islands of Socio-Technical Artifacts on GitHub
Adam Štěpánek, Marco Raglianti, Jan Byška +2
cs.SEarXiv:2609.05278v12026Software Engineering in the Agent Era From Trustworthy Change to Human Agent Software Organizations
Zhongjie Wang, Mingyi Liu
cs.SEarXiv:2609.04630v12026Automated Deployment of Real-Time Tasks for Phased Execution on Scratchpad-Based Multicore Platforms
Konstantin Dudzik, Maximilian Kirschner, Jürgen Becker
cs.SEarXiv:2609.04221v12026Adaptation Needs in Robotic Systems: Assessing Behavior Trees and Their Enhancement
Mehran Rostamnia, Gianluca Filippone, Ricardo Caldas +1
cs.ROcs.SEarXiv:2609.05331v12026The Prompt Triangle: A Registered Report on Prompts as Hybrid Artifacts
Shalini Chakraborty, Jan-Philipp Steghöfer
cs.SEarXiv:2609.04209v12026CloudGenius: Decision Support for Web Server Cloud Migration
Michael Menzel, Rajiv Ranjan
cs.DCcs.SEarXiv:1203.3997v12012STELLAR: A Search-Based Testing Framework for Large Language Model Applications
Lev Sorokin, Ivan Vasilev, Ken E. Friedl +1
cs.SEarXiv:2601.00497v22026Tokenomics: Quantifying Where Tokens Are Used in Agentic Software Engineering
Mohamad Salim, Jasmine Latendresse, SayedHassan Khatoonabadi +1
cs.SEcs.AIcs.MAarXiv:2601.14470v12026A systematic literature review on logging smell detection
Nora Madi, Manal Binkhonain
cs.SEarXiv:2609.04215v12026Security in the Age of AI Teammates: An Empirical Study of Agentic Pull Requests on GitHub
Mohammed Latif Siddiq, Xinye Zhao, Vinicius Carvalho Lopes +2
cs.CRcs.SEarXiv:2601.00477v22026Test Case Purification for Improving Fault Localization
Jifeng Xuan, Martin Monperrus
cs.SEarXiv:1409.3176v12014DevBench: A Realistic, Developer-Informed Benchmark for Code Generation Models
Adarsh Kumarappan, Pareesa Ameneh Golnari, Wen Wen +5
cs.LGcs.AIcs.SEarXiv:2601.11895v32026CircuChain: Disentangling Competence and Compliance in LLM Circuit Analysis
Mayank Ravishankara
cs.SEcs.AIarXiv:2602.15037v12026Reducing False Positives in Static Bug Detection with LLMs: An Empirical Study in Industry
Xueying Du, Jiayi Feng, Yi Zou +6
cs.SEcs.AIarXiv:2601.18844v12026SWE Context Bench: A Benchmark for Context Learning in Coding
Jiayuan Zhu, Junde Wu, Minhao Hu +9
cs.SEcs.AIarXiv:2602.08316v32026BiasScope: Towards Automated Detection of Bias in LLM-as-a-Judge Evaluation
Peng Lai, Zhihao Ou, Yong Wang +4
cs.CLcs.AIcs.SEarXiv:2602.09383v12026Engineering Trustworthy Self-Adaptive Software with Dynamic Assurance Cases
Radu Calinescu, Danny Weyns, Simos Gerasimou +3
cs.SEarXiv:1703.06350v22017WebCoderBench: Benchmarking Web Application Generation with Comprehensive and Interpretable Evaluation Metrics
Chenxu Liu, Yingjie Fu, Wei Yang +2
cs.SEcs.AIarXiv:2601.02430v32026Same Request, Different Answer: Quantization Amplifies Cache-Induced Divergence in LLM Serving
Aditi Patodiya
cs.SEcs.DCcs.LGarXiv:2609.04748v12026On the Performance of Hybrid Search Strategies for Systematic Literature Reviews in Software Engineering
Erica Mourão, João Felipe Pimentel, Leonardo Murta +3
cs.DLcs.SEarXiv:2004.09741v12020Hypothesize-Then-Verify: Speculative Root Cause Analysis for Microservices with Pathwise Parallelism
Lingzhe Zhang, Tong Jia, Yunpeng Zhai +5
cs.SEcs.AIarXiv:2601.02736v12026How AI Coding Agents Modify Code: A Large-Scale Study of GitHub Pull Requests
Daniel Ogenrwot, John Businge
cs.SEcs.AIarXiv:2601.17581v32026ScratchEval : A Multimodal Evaluation Framework for LLMs in Block-Based Programming
Yuan Si, Simeng Han, Daming Li +2
cs.SEarXiv:2602.00757v12026CodeHacker: Automated Test Case Generation for Detecting Vulnerabilities in Competitive Programming Solutions
Jingwei Shi, Xinxiang Yin, Jing Huang +2
cs.SEcs.AIcs.CRarXiv:2602.20213v22026