Software Engineering

Papers filed under cs.SE on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

961 to 1,020 of 1,406

  1. Scenarios for Development, Test and Validation of Automated Vehicles

    Till Menzel, Gerrit Bagschik, Markus Maurer

    cs.SEarXiv:1801.08598v32018
  2. AgentDV: Closed-Loop Agentic AI for Hardware Design Verification

    Navya Goli, Junzhe Liu, Zhenge Jia +1

    cs.SEarXiv:2608.27148v12026
  3. Improving Automatic Source Code Summarization via Deep Reinforcement Learning

    Yao Wan, Zhou Zhao, Min Yang +4

    cs.SEcs.CLcs.LGarXiv:1811.07234v12018
  4. Follow Me at the Edge: Mobility-Aware Dynamic Service Placement for Mobile Edge Computing

    Tao Ouyang, Zhi Zhou, Xu Chen

    cs.NIcs.AIcs.DCarXiv:1809.05239v12018
  5. LLMs in Digital EDA: A perspective on shifting roles from Generation to Orchestration

    Matthew Youngman, Cristian Sestito, Themis Prodromakis

    cs.ARcs.AIcs.SEarXiv:2608.27184v12026
  6. A storm is Coming: A Modern Probabilistic Model Checker

    Christian Dehnert, Sebastian Junges, Joost-Pieter Katoen +1

    cs.SEarXiv:1702.04311v12017
  7. SWE-Prime: Fewer Trajectories, Better Performance

    Dewu Zheng, Ruizhe Ye, Yanlin Wang +7

    cs.SEcs.AIcs.CLarXiv:2608.27449v12026
  8. VerilogEval: Evaluating Large Language Models for Verilog Code Generation

    Mingjie Liu, Nathaniel Pinckney, Brucek Khailany +1

    cs.LGcs.SEarXiv:2309.07544v22023
  9. Sampling in Software Engineering Research: A Critical Review and Guidelines

    Sebastian Baltes, Paul Ralph

    cs.SEarXiv:2002.07764v62020
  10. Fairness Testing: Testing Software for Discrimination

    Sainyam Galhotra, Yuriy Brun, Alexandra Meliou

    cs.SEcs.AIcs.CYarXiv:1709.03221v12017
  11. Twelve Quick Tips for Managing IT Disasters in Small Research Software Teams

    Greg Wilson

    cs.SEarXiv:2608.27196v12026
  12. 6.5% of the Neuro-Symbolic Literature Can Be Reproduced from Its Published Artifacts, a Six-Stage Audit Framework and First Instantiation

    Brandon Colelough, Vladimir Martirosyan, Ishan Tamrakar +6

    cs.AIcs.SEarXiv:2608.26236v12026
  13. Agent Mesh: Reliability Primitives for Non-Idempotent Agent Delegation - Identity Adequacy and Evidence Adequacy

    Mazhar Shaikh, Anurag Rajkumar Bombarde, Harshal Pathak

    cs.AIcs.DCcs.MAarXiv:2608.26225v12026
  14. Benchmarking AI Agents for Hardware Design Automation via MCP Tool Calling

    Leonardo Liparulo, Francesco Pierri

    cs.AIcs.SEarXiv:2608.26199v12026
  15. Same Model, Different Harness: Different Coding-Agent Results

    Sydney Lewis

    cs.AIcs.SEarXiv:2608.26218v12026
  16. Five Primitives for Governing Autonomous AI Agents at Runtime

    Jiten Oswal, John Cadeddu

    cs.AIcs.CRcs.SEarXiv:2608.26696v12026
  17. An Empirical Study on Learning Bug-Fixing Patches in the Wild via Neural Machine Translation

    Michele Tufano, Cody Watson, Gabriele Bavota +3

    cs.SEarXiv:1812.08693v22018
  18. Rethinking Automated Program Repair: The Impact of Bug Complexity, Fault Localization, and LLM Cost-efficiency

    Junchi Liu, Ali Bigdeli, Roya Daneshi +3

    cs.SEcs.AIarXiv:2608.14065v12026
  19. A Transformer-based Approach for Source Code Summarization

    Wasi Uddin Ahmad, Saikat Chakraborty, Baishakhi Ray +1

    cs.SEcs.AIcs.LGarXiv:2005.00653v12020
  20. An Analysis of the Cloud Computing Security Problem

    Mohamed Almorsy, John Grundy, Ingo Müller

    cs.SEcs.CRarXiv:1609.01107v12016
  21. DeepRepro: State-Aware Subplanning for Paper-to-Code Reproduction in Evolving Repositories

    Hongru Song, Ruqing Zhang, Jiafeng Guo +2

    cs.SEarXiv:2608.26557v12026
  22. Agentic AI Containment Architecture for Security Hardening

    Mohamed ElBendary

    cs.SEarXiv:2608.26108v12026
  23. From State to Action: OODA-Tool for Reliable Multi-Turn Tool Use

    Rongfeng Guo, Yinxuan Huang, Yusen Wu +5

    cs.AIcs.SEarXiv:2608.24368v12026
  24. Revision-Aware Success Prediction from Multi-Attempt Programming Trajectories

    Md Faizul Ibne Amin, Yutaka Watanobe, Daniel M. Muepu +4

    cs.CYcs.SEarXiv:2608.26169v12026
  25. Agentless: Demystifying LLM-based Software Engineering Agents

    Chunqiu Steven Xia, Yinlin Deng, Soren Dunn +1

    cs.SEcs.AIcs.CLarXiv:2407.01489v22024
  26. Zero-Shot Self-Orchestration with Ledger-Based Control for Improved LLM Coding Performance

    Victor Gao, Vida Khosrowshahi, Ali Khosrowshahi +4

    cs.MAcs.AIcs.CLarXiv:2608.26480v12026
  27. Nopol: Automatic Repair of Conditional Statement Bugs in Java Programs

    Jifeng Xuan, Matias Martinez, Favio Demarco +5

    cs.SEarXiv:1811.04211v12018
  28. FairFuzz: Targeting Rare Branches to Rapidly Increase Greybox Fuzz Testing Coverage

    Caroline Lemieux, Koushik Sen

    cs.SEcs.CRarXiv:1709.07101v12017
  29. Programming Is Hard -- Or at Least It Used to Be: Educational Opportunities And Challenges of AI Code Generation

    Brett A. Becker, Paul Denny, James Finnie-Ansley +3

    cs.HCcs.AIcs.CYarXiv:2212.01020v12022
  30. What Does an Evaluation License? A Commit-Bound Census of Claim-Relative Inference in Inspect Evals

    Xi Qin

    cs.SEcs.AIarXiv:2608.19269v32026
  31. DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence

    DeepSeek-AI, Qihao Zhu, Daya Guo +37

    cs.SEcs.AIcs.LGarXiv:2406.11931v12024
  32. Do Code Clones Matter?

    Elmar Juergens, Florian Deissenboeck, Benjamin Hummel +1

    cs.SEarXiv:1701.05472v12017
  33. From Subjective Judgments to Auditable Standards:Protocol-Guided AI Auditing of Website Redundancy

    Ge Kong, Yongtong Cao

    cs.SEcs.CVarXiv:2608.21476v12026
  34. SDR Driver for Precise Timing Applications

    Fabrizio Pollastri

    cs.SEarXiv:2608.23614v12026
  35. From Traceability to Justifiability: Accountability Structures in Agentic Software Engineering

    Rashid Azarang

    cs.SEarXiv:2608.23610v12026
  36. Adoption Telemetry: Measuring Enterprise AI Adoption from Production Signals

    Damon A. Young

    cs.HCcs.CYcs.SEarXiv:2608.23617v12026
  37. Rebuild Dossier: Mechanically-Enforced Specs for Agentic App Rebuilds, and What Model-Tier Failures Reveal

    Parker Fawcett

    cs.SEcs.AIarXiv:2608.23616v22026
  38. When May an Agent Stop? Evidence-Carrying Termination for Tool-Using LLMs

    Jason Liu

    cs.SEcs.AIcs.LGarXiv:2608.23623v12026
  39. Automated Test Input Generation for Android: Are We There Yet?

    Shauvik Roy Choudhary, Alessandra Gorla, Alessandro Orso

    cs.SEarXiv:1503.07217v22015
  40. KONTOGRAPH: Verified Point-in-Time Feature Consistency and Amortised Explanation for Real-Time Anti-Money Laundering under a 200 ms Decision Budget

    Ahmed Abolfadl

    cs.CRcs.AIcs.LGarXiv:2608.22389v12026
  41. Empirical Review of Automated Analysis Tools on 47,587 Ethereum Smart Contracts

    Thomas Durieux, João F. Ferreira, Rui Abreu +1

    cs.SEarXiv:1910.10601v22019
  42. Guiding Deep Learning System Testing using Surprise Adequacy

    Jinhan Kim, Robert Feldt, Shin Yoo

    cs.SEcs.NEarXiv:1808.08444v12018
  43. Automatic Generation of Programming Exercises and Code Explanations using Large Language Models

    Sami Sarsa, Paul Denny, Arto Hellas +1

    cs.SEcs.AIcs.CLarXiv:2206.11861v22022
  44. GitHub Copilot AI pair programmer: Asset or Liability?

    Arghavan Moradi Dakhel, Vahid Majdinasab, Amin Nikanjam +4

    cs.SEcs.LGarXiv:2206.15331v22022
  45. Large Language Models for Software Engineering: Survey and Open Problems

    Angela Fan, Beliz Gokkaya, Mark Harman +4

    cs.SEarXiv:2310.03533v42023
  46. Right-Sizing LLM-Agent Decomposition in VAT Determination: A Pilot Controlled Sweep

    Pedro Santos

    cs.MAcs.AIcs.SEarXiv:2608.23395v12026
  47. Beyond Executable Models: The Pufibara Agent Harness and the Modelica Agent Workflow Benchmark for Physical System Modeling

    Zizhe Wang

    cs.SEcs.AIarXiv:2608.23653v12026
  48. Feedback That Backfires: Why Small Language Model Agents Repeat the Call They Just Watched Fail

    Esmail Gumaan

    cs.SEcs.AIarXiv:2608.23651v12026
  49. Cross-Stack Validation of Language-Model Training: A Clinical Fine-Tuning Case Study

    Thang Tran, Lan Dang

    cs.SEarXiv:2608.24267v12026
    Summaries:한국어
  50. Metis: Typed Runtime Mediation for Tool-Using Software Agents

    Jun Yu

    cs.SEarXiv:2608.25322v12026
  51. Separating Disclosure from Authorization: Field-Tier Minimization for Agent Action Mediation

    Jiten Oswal, John Cadeddu

    cs.CRcs.SEarXiv:2608.25474v12026
  52. RotDroid: Cross-Orientation State Equivalence Testing for Detecting GUI Rotation Bugs in Android Apps

    Mengdi Qin, Bo Jiang

    cs.SEcs.AIarXiv:2608.25425v12026
  53. Point-in-Time Audit Before Alpha: Public-Archive Availability and a Negative Matched-Budget Study on BTC Perpetual Futures

    Baocheng Zeng, Jinhao Yang, Peilin Han +1

    cs.SEarXiv:2608.25348v12026
  54. Repair or Resample? Rethinking Failure Debugging in LLM Multi-Agent Systems

    Zhongwen Luan, Xiaoyu Zhang, Ming Hu +3

    cs.AIcs.SEarXiv:2608.25920v12026
    Summaries:简体中文
  55. XREPOTEST: Benchmarking Multilingual Repository-Level Unit Test Generation for Large Language Models

    Dung Le Quang, Dong Cao Van, Nam Le Hai +3

    cs.SEarXiv:2608.25939v12026
  56. Vulnerable Code Search: Transferable Attack for Code Language Models

    Kaicheng Wang, Liyan Huang, Jesse Thomason +1

    cs.SEcs.CRarXiv:2608.26031v12026
  57. Praxist: From Experimental Artifacts to Solution Lineages

    Jin Li, Ahmed Murtadha, Zhiyu Wang +13

    cs.MAcs.SEarXiv:2608.25955v12026
  58. Closing the Gap: Automated Discovery of Secure Dockerfile Reference Standards via Semantic Clustering in Enterprise Inner Source

    Jessica Hösl, Benedikt Hofmann, Patrick Stöckle

    cs.CRcs.SEarXiv:2608.25793v12026
  59. Answer Is Cheap, Show Me the Evidence! Augmenting Automated Vulnerability Assessment with Evidence

    Shengyi Pan, Zelong Zheng, Jiayuan Zhou +3

    cs.SEarXiv:2608.25905v12026
  60. Beyond the Editing Canvas: Evidence Divergence in OOXML-to-LLM Ingestion

    Side Liu, Jiangpeng Liu, Jinwen Xin +2

    cs.SEcs.CRarXiv:2608.25880v12026