Software Engineering
Papers filed under cs.SE on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
961 to 1,020 of 1,406
Scenarios for Development, Test and Validation of Automated Vehicles
Till Menzel, Gerrit Bagschik, Markus Maurer
cs.SEarXiv:1801.08598v32018AgentDV: Closed-Loop Agentic AI for Hardware Design Verification
Navya Goli, Junzhe Liu, Zhenge Jia +1
cs.SEarXiv:2608.27148v12026Improving Automatic Source Code Summarization via Deep Reinforcement Learning
Yao Wan, Zhou Zhao, Min Yang +4
cs.SEcs.CLcs.LGarXiv:1811.07234v12018Follow Me at the Edge: Mobility-Aware Dynamic Service Placement for Mobile Edge Computing
Tao Ouyang, Zhi Zhou, Xu Chen
cs.NIcs.AIcs.DCarXiv:1809.05239v12018LLMs in Digital EDA: A perspective on shifting roles from Generation to Orchestration
Matthew Youngman, Cristian Sestito, Themis Prodromakis
cs.ARcs.AIcs.SEarXiv:2608.27184v12026A storm is Coming: A Modern Probabilistic Model Checker
Christian Dehnert, Sebastian Junges, Joost-Pieter Katoen +1
cs.SEarXiv:1702.04311v12017SWE-Prime: Fewer Trajectories, Better Performance
Dewu Zheng, Ruizhe Ye, Yanlin Wang +7
cs.SEcs.AIcs.CLarXiv:2608.27449v12026VerilogEval: Evaluating Large Language Models for Verilog Code Generation
Mingjie Liu, Nathaniel Pinckney, Brucek Khailany +1
cs.LGcs.SEarXiv:2309.07544v22023Sampling in Software Engineering Research: A Critical Review and Guidelines
Sebastian Baltes, Paul Ralph
cs.SEarXiv:2002.07764v62020Fairness Testing: Testing Software for Discrimination
Sainyam Galhotra, Yuriy Brun, Alexandra Meliou
cs.SEcs.AIcs.CYarXiv:1709.03221v12017Twelve Quick Tips for Managing IT Disasters in Small Research Software Teams
Greg Wilson
cs.SEarXiv:2608.27196v120266.5% of the Neuro-Symbolic Literature Can Be Reproduced from Its Published Artifacts, a Six-Stage Audit Framework and First Instantiation
Brandon Colelough, Vladimir Martirosyan, Ishan Tamrakar +6
cs.AIcs.SEarXiv:2608.26236v12026Agent Mesh: Reliability Primitives for Non-Idempotent Agent Delegation - Identity Adequacy and Evidence Adequacy
Mazhar Shaikh, Anurag Rajkumar Bombarde, Harshal Pathak
cs.AIcs.DCcs.MAarXiv:2608.26225v12026Benchmarking AI Agents for Hardware Design Automation via MCP Tool Calling
Leonardo Liparulo, Francesco Pierri
cs.AIcs.SEarXiv:2608.26199v12026Same Model, Different Harness: Different Coding-Agent Results
Sydney Lewis
cs.AIcs.SEarXiv:2608.26218v12026Five Primitives for Governing Autonomous AI Agents at Runtime
Jiten Oswal, John Cadeddu
cs.AIcs.CRcs.SEarXiv:2608.26696v12026An Empirical Study on Learning Bug-Fixing Patches in the Wild via Neural Machine Translation
Michele Tufano, Cody Watson, Gabriele Bavota +3
cs.SEarXiv:1812.08693v22018Rethinking Automated Program Repair: The Impact of Bug Complexity, Fault Localization, and LLM Cost-efficiency
Junchi Liu, Ali Bigdeli, Roya Daneshi +3
cs.SEcs.AIarXiv:2608.14065v12026A Transformer-based Approach for Source Code Summarization
Wasi Uddin Ahmad, Saikat Chakraborty, Baishakhi Ray +1
cs.SEcs.AIcs.LGarXiv:2005.00653v12020An Analysis of the Cloud Computing Security Problem
Mohamed Almorsy, John Grundy, Ingo Müller
cs.SEcs.CRarXiv:1609.01107v12016DeepRepro: State-Aware Subplanning for Paper-to-Code Reproduction in Evolving Repositories
Hongru Song, Ruqing Zhang, Jiafeng Guo +2
cs.SEarXiv:2608.26557v12026Agentic AI Containment Architecture for Security Hardening
Mohamed ElBendary
cs.SEarXiv:2608.26108v12026From State to Action: OODA-Tool for Reliable Multi-Turn Tool Use
Rongfeng Guo, Yinxuan Huang, Yusen Wu +5
cs.AIcs.SEarXiv:2608.24368v12026Revision-Aware Success Prediction from Multi-Attempt Programming Trajectories
Md Faizul Ibne Amin, Yutaka Watanobe, Daniel M. Muepu +4
cs.CYcs.SEarXiv:2608.26169v12026Agentless: Demystifying LLM-based Software Engineering Agents
Chunqiu Steven Xia, Yinlin Deng, Soren Dunn +1
cs.SEcs.AIcs.CLarXiv:2407.01489v22024Zero-Shot Self-Orchestration with Ledger-Based Control for Improved LLM Coding Performance
Victor Gao, Vida Khosrowshahi, Ali Khosrowshahi +4
cs.MAcs.AIcs.CLarXiv:2608.26480v12026Nopol: Automatic Repair of Conditional Statement Bugs in Java Programs
Jifeng Xuan, Matias Martinez, Favio Demarco +5
cs.SEarXiv:1811.04211v12018FairFuzz: Targeting Rare Branches to Rapidly Increase Greybox Fuzz Testing Coverage
Caroline Lemieux, Koushik Sen
cs.SEcs.CRarXiv:1709.07101v12017Programming Is Hard -- Or at Least It Used to Be: Educational Opportunities And Challenges of AI Code Generation
Brett A. Becker, Paul Denny, James Finnie-Ansley +3
cs.HCcs.AIcs.CYarXiv:2212.01020v12022What Does an Evaluation License? A Commit-Bound Census of Claim-Relative Inference in Inspect Evals
Xi Qin
cs.SEcs.AIarXiv:2608.19269v32026DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
DeepSeek-AI, Qihao Zhu, Daya Guo +37
cs.SEcs.AIcs.LGarXiv:2406.11931v12024Do Code Clones Matter?
Elmar Juergens, Florian Deissenboeck, Benjamin Hummel +1
cs.SEarXiv:1701.05472v12017From Subjective Judgments to Auditable Standards:Protocol-Guided AI Auditing of Website Redundancy
Ge Kong, Yongtong Cao
cs.SEcs.CVarXiv:2608.21476v12026SDR Driver for Precise Timing Applications
Fabrizio Pollastri
cs.SEarXiv:2608.23614v12026From Traceability to Justifiability: Accountability Structures in Agentic Software Engineering
Rashid Azarang
cs.SEarXiv:2608.23610v12026Adoption Telemetry: Measuring Enterprise AI Adoption from Production Signals
Damon A. Young
cs.HCcs.CYcs.SEarXiv:2608.23617v12026Rebuild Dossier: Mechanically-Enforced Specs for Agentic App Rebuilds, and What Model-Tier Failures Reveal
Parker Fawcett
cs.SEcs.AIarXiv:2608.23616v22026When May an Agent Stop? Evidence-Carrying Termination for Tool-Using LLMs
Jason Liu
cs.SEcs.AIcs.LGarXiv:2608.23623v12026Automated Test Input Generation for Android: Are We There Yet?
Shauvik Roy Choudhary, Alessandra Gorla, Alessandro Orso
cs.SEarXiv:1503.07217v22015KONTOGRAPH: Verified Point-in-Time Feature Consistency and Amortised Explanation for Real-Time Anti-Money Laundering under a 200 ms Decision Budget
Ahmed Abolfadl
cs.CRcs.AIcs.LGarXiv:2608.22389v12026Empirical Review of Automated Analysis Tools on 47,587 Ethereum Smart Contracts
Thomas Durieux, João F. Ferreira, Rui Abreu +1
cs.SEarXiv:1910.10601v22019Guiding Deep Learning System Testing using Surprise Adequacy
Jinhan Kim, Robert Feldt, Shin Yoo
cs.SEcs.NEarXiv:1808.08444v12018Automatic Generation of Programming Exercises and Code Explanations using Large Language Models
Sami Sarsa, Paul Denny, Arto Hellas +1
cs.SEcs.AIcs.CLarXiv:2206.11861v22022GitHub Copilot AI pair programmer: Asset or Liability?
Arghavan Moradi Dakhel, Vahid Majdinasab, Amin Nikanjam +4
cs.SEcs.LGarXiv:2206.15331v22022Large Language Models for Software Engineering: Survey and Open Problems
Angela Fan, Beliz Gokkaya, Mark Harman +4
cs.SEarXiv:2310.03533v42023Right-Sizing LLM-Agent Decomposition in VAT Determination: A Pilot Controlled Sweep
Pedro Santos
cs.MAcs.AIcs.SEarXiv:2608.23395v12026Beyond Executable Models: The Pufibara Agent Harness and the Modelica Agent Workflow Benchmark for Physical System Modeling
Zizhe Wang
cs.SEcs.AIarXiv:2608.23653v12026Feedback That Backfires: Why Small Language Model Agents Repeat the Call They Just Watched Fail
Esmail Gumaan
cs.SEcs.AIarXiv:2608.23651v12026Cross-Stack Validation of Language-Model Training: A Clinical Fine-Tuning Case Study
Thang Tran, Lan Dang
cs.SEarXiv:2608.24267v12026Summaries:한국어Metis: Typed Runtime Mediation for Tool-Using Software Agents
Jun Yu
cs.SEarXiv:2608.25322v12026Separating Disclosure from Authorization: Field-Tier Minimization for Agent Action Mediation
Jiten Oswal, John Cadeddu
cs.CRcs.SEarXiv:2608.25474v12026RotDroid: Cross-Orientation State Equivalence Testing for Detecting GUI Rotation Bugs in Android Apps
Mengdi Qin, Bo Jiang
cs.SEcs.AIarXiv:2608.25425v12026Point-in-Time Audit Before Alpha: Public-Archive Availability and a Negative Matched-Budget Study on BTC Perpetual Futures
Baocheng Zeng, Jinhao Yang, Peilin Han +1
cs.SEarXiv:2608.25348v12026Repair or Resample? Rethinking Failure Debugging in LLM Multi-Agent Systems
Zhongwen Luan, Xiaoyu Zhang, Ming Hu +3
cs.AIcs.SEarXiv:2608.25920v12026Summaries:简体中文XREPOTEST: Benchmarking Multilingual Repository-Level Unit Test Generation for Large Language Models
Dung Le Quang, Dong Cao Van, Nam Le Hai +3
cs.SEarXiv:2608.25939v12026Vulnerable Code Search: Transferable Attack for Code Language Models
Kaicheng Wang, Liyan Huang, Jesse Thomason +1
cs.SEcs.CRarXiv:2608.26031v12026Praxist: From Experimental Artifacts to Solution Lineages
Jin Li, Ahmed Murtadha, Zhiyu Wang +13
cs.MAcs.SEarXiv:2608.25955v12026Closing the Gap: Automated Discovery of Secure Dockerfile Reference Standards via Semantic Clustering in Enterprise Inner Source
Jessica Hösl, Benedikt Hofmann, Patrick Stöckle
cs.CRcs.SEarXiv:2608.25793v12026Answer Is Cheap, Show Me the Evidence! Augmenting Automated Vulnerability Assessment with Evidence
Shengyi Pan, Zelong Zheng, Jiayuan Zhou +3
cs.SEarXiv:2608.25905v12026Beyond the Editing Canvas: Evidence Divergence in OOXML-to-LLM Ingestion
Side Liu, Jiangpeng Liu, Jinwen Xin +2
cs.SEcs.CRarXiv:2608.25880v12026