Software Engineering

Papers filed under cs.SE on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

241 to 300 of 1,389

  1. A Quantitative and Qualitative Evaluation of LLM-Based Explainable Fault Localization

    Sungmin Kang, Gabin An, Shin Yoo

    cs.SEarXiv:2308.05487v32023
  2. How Often Do Single-Statement Bugs Occur? The ManySStuBs4J Dataset

    Rafael-Michael Karampatsis, Charles Sutton

    cs.SEcs.PLarXiv:1905.13334v22019
  3. Can LLMs Extract Architectural Design Decisions from Source Code Commits? - A Preliminary Exploratory Study

    Amey Karan, Rudra Dhar, Mohamed Soliman +1

    cs.SEcs.AIarXiv:2609.03721v12026
  4. Where Are The Gaps? A Systematic Mapping Study of Infrastructure as Code Research

    Akond Rahman, Rezvan Mahdavi-Hezaveh, Laurie Williams

    cs.SEarXiv:1807.04872v12018
  5. An Empirical Study of Deep Learning Models for Vulnerability Detection

    Benjamin Steenhoek, Md Mahbubur Rahman, Richard Jiles +1

    cs.SEcs.CRcs.LGarXiv:2212.08109v32022
  6. Stochastic Semantics and Statistical Model Checking for Networks of Priced Timed Automata

    Alexandre David, Kim G. Larsen, Axel Legay +4

    cs.SEarXiv:1106.3961v22011
  7. Generalization or Memorization: Data Contamination and Trustworthy Evaluation for Large Language Models

    Yihong Dong, Xue Jiang, Huanyu Liu +4

    cs.CLcs.AIcs.CRarXiv:2402.15938v32024
  8. The Psychological Costs of Artificial Intelligence Adoption in Software Engineering

    Adam Alami, Elda Paja, Abhishek Tiwari

    cs.SEcs.AIarXiv:2609.03456v12026
  9. SWIM: Synthesizing What I Mean

    Mukund Raghothaman, Yi Wei, Youssef Hamadi

    cs.SEarXiv:1511.08497v22015
  10. Challenges and solutions when adopting DevSecOps: A systematic review

    Roshan N. Rajapakse, Mansooreh Zahedi, M. Ali Babar +1

    cs.SEarXiv:2103.08266v22021
  11. Perfection Not Required? Human-AI Partnerships in Code Translation

    Justin D. Weisz, Michael Muller, Stephanie Houde +5

    cs.HCcs.SEarXiv:2104.03820v12021
  12. Fill in the Blank: Context-aware Automated Text Input Generation for Mobile GUI Testing

    Zhe Liu, Chunyang Chen, Junjie Wang +4

    cs.SEarXiv:2212.04732v12022
  13. Large Language Model for Vulnerability Detection and Repair: Literature Review and the Road Ahead

    Xin Zhou, Sicong Cao, Xiaobing Sun +1

    cs.SEarXiv:2404.02525v32024
  14. Microservices: How To Make Your Application Scale

    Nicola Dragoni, Ivan Lanese, Stephan Thordal Larsen +3

    cs.SEarXiv:1702.07149v12017
  15. A Survey on Automated Driving System Testing: Landscapes and Trends

    Shuncheng Tang, Zhenya Zhang, Yi Zhang +8

    cs.SEarXiv:2206.05961v22022
  16. ReCode: Robustness Evaluation of Code Generation Models

    Shiqi Wang, Zheng Li, Haifeng Qian +11

    cs.LGcs.CLcs.SEarXiv:2212.10264v12022
  17. You Cannot Fix What You Cannot Find! An Investigation of Fault Localization Bias in Benchmarking Automated Program Repair Systems

    Kui Liu, Anil Koyuncu, Tegawendé F. Bissyandé +3

    cs.SEarXiv:1812.07283v22018
  18. Automated Synthesis of Cloud Emulators

    Archit Bhatnagar, Zhenning Yang, Sarah McClure +3

    cs.SEcs.AIcs.DCarXiv:2608.23842v12026
  19. LongCoder: A Long-Range Pre-trained Language Model for Code Completion

    Daya Guo, Canwen Xu, Nan Duan +2

    cs.SEcs.AIcs.CLarXiv:2306.14893v12023
  20. Why Early-Stage Software Startups Fail: A Behavioral Framework

    Carmine Giardino, Xiaofeng Wang, Pekka Abrahamsson

    cs.SEarXiv:1709.04749v12017
  21. Aroma: Code Recommendation via Structural Code Search

    Sifei Luan, Di Yang, Celeste Barnaby +2

    cs.SEarXiv:1812.01158v42018
  22. Pynguin: Automated Unit Test Generation for Python

    Stephan Lukasczyk, Gordon Fraser

    cs.SEarXiv:2202.05218v12022
  23. DTM: Deterministic Approaches for Black-box Test Suite Minimization with Tree-based Similarity

    Md Siam, Shartaz Sajid Nahid, Md Arif Hasan +2

    cs.SEarXiv:2609.04205v12026
  24. Breaking the Alphabet: Rethinking File Ordering in Code Review

    Md Shamimur Rahman, Zadia Codabux, Chanchal K. Roy

    cs.SEarXiv:2609.04207v12026
  25. AI Writes Code, Humans Pay the Debt. An Empirical Study on the Sustainability and Evolution of Agent-Generated Code

    Antonino Coppola, Matteo Esposito, Rick Kazman +1

    cs.SEarXiv:2609.04208v12026
  26. SH-PDOPS: AI-Driven Cloud Native Enterprise Reliability Framework for Predictive Analytics and Intelligent DevOps Automation

    Ayushman Bosu Roy

    cs.SEarXiv:2609.04210v12026
  27. Big Questions on Software Architecture: Report of the ICSE 2026 BoF on Software Architecture

    Davide Taibi, Patricia Lago, Henry Muccini

    cs.SEarXiv:2609.04212v12026
  28. Engineering as Code: Bringing Software Engineering Methodology to Engineering Design

    Song Difei

    cs.SEarXiv:2609.04216v12026
  29. Robustness and Trade-offs for Code LLMs on Protected Code

    Jin Wen, Yuejun Guo, Yujie Ma +2

    cs.SEarXiv:2609.04220v12026
  30. How Developers Discuss Generative AI: A Longitudinal Study of the Visual Studio Code Community

    Panida Rumriankit, Akito Monden, Hiroki Inayoshi +4

    cs.SEarXiv:2609.04680v12026
  31. Data-Related Challenges and Requirements for Event Log Generation in Process Mining: A Systematic Literature Review

    Ghita El Alaoui Talibi, Oleksandr Kosenkov, Anastasija Nikiforova

    cs.SEarXiv:2609.04211v12026
  32. A Mixed-Method Empirical Study of LLM Assistance in Software Engineering Workflows

    Pamali D. Weerasinghe, Roshan N. Rajapakse, Isuru Dharmadasa +1

    cs.SEarXiv:2609.04214v12026
  33. Toward Model-Driven Digital Twin Configuration: Separating Structure Semantics and Runtime with SysML SAREF and Ditto

    Andrey Sadovykh, Matthew Rusakov, Kirill Korikov

    cs.SEarXiv:2609.04213v12026
  34. A Governance Methodology Layer for AI-Assisted Software Development: Defect Taxonomy, Controlled Ablation, and Process-Over-Capability Evidence

    Sungjin Kwon

    cs.SEarXiv:2609.04218v12026
  35. Large Language Models for Fuzz Testing in Microservices: A Systematic Literature Review

    Ying Song, Ke Ping, Yuqing Wang +1

    cs.SEarXiv:2609.04219v12026
  36. CPL: A Compact C-like Systems Language with Explicit Low-Level Control

    Nikolay Fot, Alexander Vinarsky

    cs.PLcs.SEarXiv:2609.04904v12026
  37. Ritgard: T(r)opical Islands of Socio-Technical Artifacts on GitHub

    Adam Štěpánek, Marco Raglianti, Jan Byška +2

    cs.SEarXiv:2609.05278v12026
  38. Software Engineering in the Agent Era From Trustworthy Change to Human Agent Software Organizations

    Zhongjie Wang, Mingyi Liu

    cs.SEarXiv:2609.04630v12026
  39. Automated Deployment of Real-Time Tasks for Phased Execution on Scratchpad-Based Multicore Platforms

    Konstantin Dudzik, Maximilian Kirschner, Jürgen Becker

    cs.SEarXiv:2609.04221v12026
  40. Adaptation Needs in Robotic Systems: Assessing Behavior Trees and Their Enhancement

    Mehran Rostamnia, Gianluca Filippone, Ricardo Caldas +1

    cs.ROcs.SEarXiv:2609.05331v12026
  41. The Prompt Triangle: A Registered Report on Prompts as Hybrid Artifacts

    Shalini Chakraborty, Jan-Philipp Steghöfer

    cs.SEarXiv:2609.04209v12026
  42. CloudGenius: Decision Support for Web Server Cloud Migration

    Michael Menzel, Rajiv Ranjan

    cs.DCcs.SEarXiv:1203.3997v12012
  43. STELLAR: A Search-Based Testing Framework for Large Language Model Applications

    Lev Sorokin, Ivan Vasilev, Ken E. Friedl +1

    cs.SEarXiv:2601.00497v22026
  44. Tokenomics: Quantifying Where Tokens Are Used in Agentic Software Engineering

    Mohamad Salim, Jasmine Latendresse, SayedHassan Khatoonabadi +1

    cs.SEcs.AIcs.MAarXiv:2601.14470v12026
  45. A systematic literature review on logging smell detection

    Nora Madi, Manal Binkhonain

    cs.SEarXiv:2609.04215v12026
  46. Security in the Age of AI Teammates: An Empirical Study of Agentic Pull Requests on GitHub

    Mohammed Latif Siddiq, Xinye Zhao, Vinicius Carvalho Lopes +2

    cs.CRcs.SEarXiv:2601.00477v22026
  47. Test Case Purification for Improving Fault Localization

    Jifeng Xuan, Martin Monperrus

    cs.SEarXiv:1409.3176v12014
  48. DevBench: A Realistic, Developer-Informed Benchmark for Code Generation Models

    Adarsh Kumarappan, Pareesa Ameneh Golnari, Wen Wen +5

    cs.LGcs.AIcs.SEarXiv:2601.11895v32026
  49. CircuChain: Disentangling Competence and Compliance in LLM Circuit Analysis

    Mayank Ravishankara

    cs.SEcs.AIarXiv:2602.15037v12026
  50. Reducing False Positives in Static Bug Detection with LLMs: An Empirical Study in Industry

    Xueying Du, Jiayi Feng, Yi Zou +6

    cs.SEcs.AIarXiv:2601.18844v12026
  51. SWE Context Bench: A Benchmark for Context Learning in Coding

    Jiayuan Zhu, Junde Wu, Minhao Hu +9

    cs.SEcs.AIarXiv:2602.08316v32026
  52. BiasScope: Towards Automated Detection of Bias in LLM-as-a-Judge Evaluation

    Peng Lai, Zhihao Ou, Yong Wang +4

    cs.CLcs.AIcs.SEarXiv:2602.09383v12026
  53. Engineering Trustworthy Self-Adaptive Software with Dynamic Assurance Cases

    Radu Calinescu, Danny Weyns, Simos Gerasimou +3

    cs.SEarXiv:1703.06350v22017
  54. WebCoderBench: Benchmarking Web Application Generation with Comprehensive and Interpretable Evaluation Metrics

    Chenxu Liu, Yingjie Fu, Wei Yang +2

    cs.SEcs.AIarXiv:2601.02430v32026
  55. Same Request, Different Answer: Quantization Amplifies Cache-Induced Divergence in LLM Serving

    Aditi Patodiya

    cs.SEcs.DCcs.LGarXiv:2609.04748v12026
  56. On the Performance of Hybrid Search Strategies for Systematic Literature Reviews in Software Engineering

    Erica Mourão, João Felipe Pimentel, Leonardo Murta +3

    cs.DLcs.SEarXiv:2004.09741v12020
  57. Hypothesize-Then-Verify: Speculative Root Cause Analysis for Microservices with Pathwise Parallelism

    Lingzhe Zhang, Tong Jia, Yunpeng Zhai +5

    cs.SEcs.AIarXiv:2601.02736v12026
  58. How AI Coding Agents Modify Code: A Large-Scale Study of GitHub Pull Requests

    Daniel Ogenrwot, John Businge

    cs.SEcs.AIarXiv:2601.17581v32026
  59. ScratchEval : A Multimodal Evaluation Framework for LLMs in Block-Based Programming

    Yuan Si, Simeng Han, Daming Li +2

    cs.SEarXiv:2602.00757v12026
  60. CodeHacker: Automated Test Case Generation for Detecting Vulnerabilities in Competitive Programming Solutions

    Jingwei Shi, Xinxiang Yin, Jing Huang +2

    cs.SEcs.AIcs.CRarXiv:2602.20213v22026