Software Engineering
Papers filed under cs.SE on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
901 to 960 of 1,389
The Impact of Automated Parameter Optimization on Defect Prediction Models
Chakkrit Tantithamthavorn, Shane McIntosh, Ahmed E. Hassan +1
cs.SEarXiv:1801.10270v12018Spec2Vision: Contract-Guided Delivery of AI-Generated Computer Vision Pipelines
Ghfran Jabour, Sergey Ivanov
cs.SEcs.PLarXiv:2608.26400v12026Investigating Software Aging in LLM-Generated Software Systems across Generation-and-Execution Environments
Cesar Santos, Michele Vitagliano, Roberto Natella +1
cs.SEarXiv:2608.26391v12026A Catalog of User Authentication Patterns
Alex R. Mattukat, Horst Lichter
cs.CRcs.SEarXiv:2608.26955v12026Mutation Testing for Reproducibility Safeguards in Machine Learning Research Software: An Empirical Study
Ilya Shulepov
cs.SEarXiv:2608.27100v12026Software development in startup companies: A systematic mapping study
Nicolò Paternoster, Carmine Giardino, Michael Unterkalmsteiner +2
cs.SEarXiv:2307.13104v12023Successful Combination of Database Search and Snowballing for Identification of Primary Studies in Systematic Literature Studies
Claes Wohlin, Marcos Kalinowski, Katia Romero Felizardo +1
cs.SEarXiv:2307.02612v12023Evaluating Quality of Chatbots and Intelligent Conversational Agents
Nicole M. Radziwill, Morgan C. Benton
cs.CYcs.SEarXiv:1704.04579v12017sFuzz: An Efficient Adaptive Fuzzer for Solidity Smart Contracts
Tai D. Nguyen, Long H. Pham, Jun Sun +2
cs.SEarXiv:2004.08563v12020STILL: Recovering Lowered STL Semantics for LLM-assisted C++ Decompilation
Xiaohan Wang, Kevin Leach
cs.SEarXiv:2608.26408v12026ADeptS-Bench: Measuring the Trustworthiness of Computer Use Agents Across Devices
Joy Chen, Alejandro Castillejo Munoz, Pierluca D'Oro +3
cs.CRcs.AIcs.SEarXiv:2608.26204v12026"A Second Set of Eyes": The Process and Challenges of Software Documentation Review
Avinash Bhat, Ian Arawjo, Disha Shrivastava +1
cs.SEcs.HCarXiv:2608.26232v12026DeepBugs: A Learning Approach to Name-based Bug Detection
Michael Pradel, Koushik Sen
cs.SEcs.PLarXiv:1805.11683v12018Community Cloud Computing
Alexandros Marinos, Gerard Briscoe
cs.NIcs.DCcs.SEarXiv:0907.2485v32009KubeCap: A Framework for Capability Minimization in Kubernetes via Static Analysis and LLM-Assisted Rule Inference
Yuhao Liu, Yingnan Zhou, Weijie Liu +2
cs.CRcs.SEarXiv:2608.26699v12026Large Language Models are Zero-Shot Fuzzers: Fuzzing Deep-Learning Libraries via Large Language Models
Yinlin Deng, Chunqiu Steven Xia, Haoran Peng +2
cs.SEarXiv:2212.14834v42022Fairness Invariants: A Relational Approach to Explaining and Mitigating Fairness Bugs
Ranit Debnath Akash, Ashish Kumar, Gang Tan +1
cs.SEcs.AIcs.LGarXiv:2608.26209v12026The Green Software Landscape: A Systematic Mapping Study on Evolution, Applications, Software Lifecycle, and Best Practices
Max Hort, Maria Kechagia, Federica Sarro
cs.SEarXiv:2608.26229v12026Challenges and Contributions in Quality of AI-Based Software: A Systematic Mapping Study
Maryum Hamdani, Mateen Ahmed Abbasi, Marko Jäntti +1
cs.SEarXiv:2608.26215v12026NeuronFuzz: Safety Neuron Guided Fuzzing for LLM Safety Evaluation
Zhiyuan Xu, Muhammad Firhard Roslan, Joseph Gardiner +2
cs.LGcs.AIcs.CRarXiv:2608.26222v12026Knowledge Management in Software Engineering: A Systematic Review of Studied Concepts, Findings and Research Methods Used
Finn Olav Bjørnson, Torgeir Dingsøyr
cs.SEarXiv:1811.12278v12018Kale: A Transformation-Safe Spreadsheet System
Michael Coblenz, Jacob Yim, Ajinkya Bokade +11
cs.HCcs.SEarXiv:2608.26345v12026SciPy 1.0--Fundamental Algorithms for Scientific Computing in Python
Pauli Virtanen, Ralf Gommers, Travis E. Oliphant +32
cs.MScs.DScs.SEarXiv:1907.10121v12019CodeNet: A Large-Scale AI for Code Dataset for Learning a Diversity of Coding Tasks
Ruchir Puri, David S. Kung, Geert Janssen +14
cs.SEcs.AIarXiv:2105.12655v22021Bug Localization from Bug Reports: A Multi-Objective Approach
Waleed Ahmad, Mehtab Kiran Suddle, Maryam Bashir
cs.NEcs.SEarXiv:2608.27089v12026DeepMutation: Mutation Testing of Deep Learning Systems
Lei Ma, Fuyuan Zhang, Jiyuan Sun +8
cs.SEarXiv:1805.05206v22018The Thousand-Graph Hypothesis: A Testable Hypothesis of Task-Conditioned Relation Materialization in Repository-Level Code Reasoning
Fei Ding
cs.SEcs.CLarXiv:2608.26602v12026When Review Alone No Longer Scales: Layered Supervision in AI-Assisted Software Engineering
Markus Stolze, Mirco Strässle
cs.SEarXiv:2608.26316v12026When Context Gets Root: Privilege Escalation in LLM Harnesses
Xingbang He, Yuanwei Chen, Yi Qian +6
cs.CRcs.SEarXiv:2608.27299v12026Processing/p5 Defined through Practice and Learning
Kit Kuksenok, Lee Tusman
cs.SEcs.HCarXiv:2608.26614v12026VeriGen: A Large Language Model for Verilog Code Generation
Shailja Thakur, Baleegh Ahmad, Hammond Pearce +4
cs.PLcs.LGcs.SEarXiv:2308.00708v12023Evaluating prediction systems in software project estimation
Martin Shepperd, Stephen G. MacDonell
cs.SEarXiv:2101.05426v12021Manticore: A User-Friendly Symbolic Execution Framework for Binaries and Smart Contracts
Mark Mossberg, Felipe Manzano, Eric Hennenfent +5
cs.SEcs.CRarXiv:1907.03890v32019Tacet: A Language and Type System for Automatic Statistical Validity Accounting
Chiké Abuah
cs.PLcs.SEarXiv:2608.27451v12026RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems
Tianyang Liu, Canwen Xu, Julian McAuley
cs.CLcs.AIcs.SEarXiv:2306.03091v22023Characterizing the Landscape of Open-Source Satellite Software
Jinfeng Wen, Qi Liang, Yuehan Sun +4
cs.SEarXiv:2608.26211v12026AROMA+: A Study of Factors Affecting Reproducible Builds in the Maven Ecosystem
Mehdi Keshani, Amirhossein Rahmati, Mohammad Hossein Aref +1
cs.SEarXiv:2608.27125v12026Unsaid, Unsafe? Implicit Security Obligations in LLM-Based RTL Code Generation
Guang Yang, Xing Hu, Xiang Chen +1
cs.CRcs.SEarXiv:2608.26588v12026CURE: Code-Aware Neural Machine Translation for Automatic Program Repair
Nan Jiang, Thibaud Lutellier, Lin Tan
cs.SEcs.AIcs.LGarXiv:2103.00073v42021From Static to Dynamic: Benchmarking Real-World Code Review with MCR-Bench
Dewu Zheng, Yanlin Wang, Xiwen Wang +5
cs.SEcs.AIcs.CLarXiv:2608.27442v12026When Tool Outputs Become Commands: Separating Action Induction from Runtime Authorization in Tool-Augmented LLM Agents
Xiaokun Guo, Zhen Xu, Dongdong Huo +5
cs.AIcs.SEarXiv:2608.27146v12026RepairAgent: An Autonomous, LLM-Based Agent for Program Repair
Islem Bouzenia, Premkumar Devanbu, Michael Pradel
cs.SEcs.AIarXiv:2403.17134v22024An Empirical Evaluation of Using Large Language Models for Automated Model-Based Test Generation
Hafize Sanli, Onur Kilincceker, Cihat Cetinkaya
cs.SEarXiv:2608.27094v12026Scenarios for Development, Test and Validation of Automated Vehicles
Till Menzel, Gerrit Bagschik, Markus Maurer
cs.SEarXiv:1801.08598v32018AgentDV: Closed-Loop Agentic AI for Hardware Design Verification
Navya Goli, Junzhe Liu, Zhenge Jia +1
cs.SEarXiv:2608.27148v12026Improving Automatic Source Code Summarization via Deep Reinforcement Learning
Yao Wan, Zhou Zhao, Min Yang +4
cs.SEcs.CLcs.LGarXiv:1811.07234v12018Follow Me at the Edge: Mobility-Aware Dynamic Service Placement for Mobile Edge Computing
Tao Ouyang, Zhi Zhou, Xu Chen
cs.NIcs.AIcs.DCarXiv:1809.05239v12018LLMs in Digital EDA: A perspective on shifting roles from Generation to Orchestration
Matthew Youngman, Cristian Sestito, Themis Prodromakis
cs.ARcs.AIcs.SEarXiv:2608.27184v12026A storm is Coming: A Modern Probabilistic Model Checker
Christian Dehnert, Sebastian Junges, Joost-Pieter Katoen +1
cs.SEarXiv:1702.04311v12017SWE-Prime: Fewer Trajectories, Better Performance
Dewu Zheng, Ruizhe Ye, Yanlin Wang +7
cs.SEcs.AIcs.CLarXiv:2608.27449v12026VerilogEval: Evaluating Large Language Models for Verilog Code Generation
Mingjie Liu, Nathaniel Pinckney, Brucek Khailany +1
cs.LGcs.SEarXiv:2309.07544v22023Sampling in Software Engineering Research: A Critical Review and Guidelines
Sebastian Baltes, Paul Ralph
cs.SEarXiv:2002.07764v62020Fairness Testing: Testing Software for Discrimination
Sainyam Galhotra, Yuriy Brun, Alexandra Meliou
cs.SEcs.AIcs.CYarXiv:1709.03221v12017Twelve Quick Tips for Managing IT Disasters in Small Research Software Teams
Greg Wilson
cs.SEarXiv:2608.27196v120266.5% of the Neuro-Symbolic Literature Can Be Reproduced from Its Published Artifacts, a Six-Stage Audit Framework and First Instantiation
Brandon Colelough, Vladimir Martirosyan, Ishan Tamrakar +6
cs.AIcs.SEarXiv:2608.26236v12026Agent Mesh: Reliability Primitives for Non-Idempotent Agent Delegation - Identity Adequacy and Evidence Adequacy
Mazhar Shaikh, Anurag Rajkumar Bombarde, Harshal Pathak
cs.AIcs.DCcs.MAarXiv:2608.26225v12026Benchmarking AI Agents for Hardware Design Automation via MCP Tool Calling
Leonardo Liparulo, Francesco Pierri
cs.AIcs.SEarXiv:2608.26199v12026Same Model, Different Harness: Different Coding-Agent Results
Sydney Lewis
cs.AIcs.SEarXiv:2608.26218v12026Five Primitives for Governing Autonomous AI Agents at Runtime
Jiten Oswal, John Cadeddu
cs.AIcs.CRcs.SEarXiv:2608.26696v12026An Empirical Study on Learning Bug-Fixing Patches in the Wild via Neural Machine Translation
Michele Tufano, Cody Watson, Gabriele Bavota +3
cs.SEarXiv:1812.08693v22018