Software Engineering
Papers filed under cs.SE on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
661 to 720 of 1,389
ExecRetrieval: Measuring the Functional-Correctness Gap in Code-Embedding Retrieval
Aaryan Kapoor, Md Abdullah Al Hafiz Khan
cs.SEcs.CLcs.IRarXiv:2609.01865v12026HarnessDev: Can LLMs Create and Evolve Their Own Agent Harness?
Yuhao Wu, Jingyuan Zhang, Jiajun Shi +16
cs.SEcs.CLarXiv:2609.01437v12026PaperCompiler: Faithful Paper-to-Code Generation via Repository-Level Specification Compilation
Yunhao Liu, Hong Phuc Pham, Jaehong Yoon
cs.CLcs.AIcs.SEarXiv:2609.02272v12026Post-Training Language Models for Gold-Medal Performance in Coding Competitions
Aleksander Ficek, Sean Narenthiran, Mehrzad Samadi +2
cs.LGcs.AIcs.CLarXiv:2609.02849v12026A Comprehensive Study of Native Code Bugs in Python Applications
Haoran Yang, Haipeng Cai
cs.SEcs.CRarXiv:2608.29851v12026Sampling Projects in GitHub for MSR Studies
Ozren Dabic, Emad Aghajani, Gabriele Bavota
cs.SEarXiv:2103.04682v12021Why Modern Open Source Projects Fail
Jailton Coelho, Marco Tulio Valente
cs.SEcs.CYarXiv:1707.02327v12017NL2Bash: A Corpus and Semantic Parser for Natural Language Interface to the Linux Operating System
Xi Victoria Lin, Chenglong Wang, Luke Zettlemoyer +1
cs.CLcs.SEarXiv:1802.08979v22018Comparing Code Explanations Created by Students and Large Language Models
Juho Leinonen, Paul Denny, Stephen MacNeil +5
cs.CYcs.AIcs.CLarXiv:2304.03938v12023Tuning for Software Analytics: is it Really Necessary?
Wei Fu, Tim Menzies, Xipeng Shen
cs.SEarXiv:1609.01759v12016ProofPulse: Interactive Proof Coverage Analysis for Dafny
Álvaro F. Silva, Ruben Martins, Alexandra Mendes
cs.SEarXiv:2608.30818v12026Large Language Models for Test-Free Fault Localization
Aidan Z. H. Yang, Ruben Martins, Claire Le Goues +1
cs.SEcs.LGarXiv:2310.01726v12023ReACC: A Retrieval-Augmented Code Completion Framework
Shuai Lu, Nan Duan, Hojae Han +3
cs.SEcs.AIcs.CLarXiv:2203.07722v12022No More Manual Tests? Evaluating and Improving ChatGPT for Unit Test Generation
Zhiqiang Yuan, Yiling Lou, Mingwei Liu +4
cs.SEarXiv:2305.04207v32023A Phased Workflow for Operating LLM-Based Coding Agents
Ante Kapetanovic, Tomislav Duricic, Andro Mercep +1
cs.SEarXiv:2608.30701v12026MetaTool Benchmark for Large Language Models: Deciding Whether to Use Tools and Which to Use
Yue Huang, Jiawen Shi, Yuan Li +8
cs.SEcs.CLarXiv:2310.03128v62023DSEffi-Bench: Demystifying Large Language Models' Capability in Efficient Data Science Code Generation
Zhihao Gong, Junzhe Yu, Dong Huang +3
cs.SEarXiv:2608.30248v12026Open-Source Autonomous Driving System Analysis and Multi-Disciplinary Hardware-in-the-Loop Research Paradigm with Reinforcement-Learning Testing and Large Language Models
Dianjing Cheng, Yike Li, Lan Yang +23
cs.SEcs.ROarXiv:2608.30179v12026Beyond Functional Correctness: Exploring Hallucinations in LLM-Generated Code
Fang Liu, Yang Liu, Lin Shi +5
cs.SEcs.AIarXiv:2404.00971v32024Update from Hell: Can Coding Agents Survive Hidden Breakage in Dependency Upgrades?
Zijian Luo, Runzhi He, Pengfei Gao +6
cs.SEarXiv:2608.30300v12026LLM-based Hardware Development with Hierarchical IRs and End-to-End Multi-Agent Workflow
Chenyang Yin, Agasthi Haputhanthri, Aditya Anirudh Jonnalagadda +7
cs.ARcs.MAcs.SEarXiv:2608.30659v12026Reading Between the Lines: Modeling User Behavior and Costs in AI-Assisted Programming
Hussein Mozannar, Gagan Bansal, Adam Fourney +1
cs.SEcs.HCcs.LGarXiv:2210.14306v52022The Exclusion Ratchet: False-Positive Suppression Accumulates and Persists in Detection Rule Repositories
Sudaroli Dhananjeyan, Kumaran U
cs.CRcs.SEarXiv:2608.31062v12026Automated Unit Test Improvement using Large Language Models at Meta
Nadia Alshahwan, Jubin Chheda, Anastasia Finegenova +6
cs.SEarXiv:2402.09171v12024Bridge: Automatically Mining Ecosystem-Scale API Update Mappings and Client Update Instances
Kai Gao, Yu Sun, Chang-ai Sun
cs.SEarXiv:2608.30497v12026ChatGPT and Software Testing Education: Promises & Perils
Sajed Jalil, Suzzana Rafi, Thomas D. LaToza +2
cs.SEcs.HCarXiv:2302.03287v32023Practical Implementation Report on Introducing Spec-Driven Development Using AI Agents in Software Development PBL
Hidetake Tanaka, Hiroshi Igaki, Kazumasa Shimari +2
cs.SEarXiv:2608.30572v12026Federated Trust for Embodied Robot Capability Marketplaces
Xue Qin, Simin Luan, Cong Yang +1
cs.CRcs.ROcs.SEarXiv:2609.00404v12026Harness Engineering: Anatomy, Architecture, and Evolution of Coding Agents -- A Source-Code Study of Eleven Systems
Paul Barbaste, Tristan Darrigol, Germain Vu +1
cs.SEcs.MAarXiv:2609.00006v12026Harvey: A Greybox Fuzzer for Smart Contracts
Valentin Wüstholz, Maria Christakis
cs.SEcs.CRarXiv:1905.06944v12019Audit-First Rollback Semantics for Safety-Critical Deployment Pipelines
Xue Qin, Simin Luan, Cong Yang +1
cs.SEcs.DCarXiv:2609.00406v12026Bounded, Indeterminate, or a Bug: A Condition-Aware Oracle for Differential Testing of SQL Aggregates
Madhulatha Mandarapu, Sandeep Kunkunuru
cs.DBcs.SEarXiv:2609.00381v12026AI with Authority, from Application to Silicon
Jason Hickey
cs.SEcs.AIcs.ARarXiv:2608.21356v22026Summaries:简体中文Structured Neural Summarization
Patrick Fernandes, Miltiadis Allamanis, Marc Brockschmidt
cs.LGcs.CLcs.SEarXiv:1811.01824v42018Good Enough Practices in Scientific Computing
Greg Wilson, Jennifer Bryan, Karen Cranston +3
cs.SEarXiv:1609.00037v22016How Effective are Smart Contract Analysis Tools? Evaluating Smart Contract Static Analysis Tools Using Bug Injection
Asem Ghaleb, Karthik Pattabiraman
cs.SEarXiv:2005.11613v12020Database-Augmented RAG for Automated Repair of REST API Misuses
Shoei Inoue, Norihiro Yoshida, Erina Makihara +2
cs.IRcs.SEarXiv:2608.29290v12026Towards Fully Automated Medical Imaging Code Generation via Validation-based Context Engineering
Zixiao Zhao, Jing Sun, Zhe Hou +5
cs.CVcs.SEarXiv:2608.29016v12026Emergent Behavior and Uncertainty in IoT-Enhanced Business Processes: Challenges and Future Directions
Marco Pegoraro, Sara Pettinari, Ivan Compagnucci +2
cs.ETcs.SEarXiv:2608.28919v12026Elixir: Effective object-oriented program repair
Ripon K. Saha, Yingjun Lyu, Hiroaki Yoshida +1
cs.SEarXiv:2112.10915v12021DocPrompting: Generating Code by Retrieving the Docs
Shuyan Zhou, Uri Alon, Frank F. Xu +3
cs.CLcs.AIcs.SEarXiv:2207.05987v32022LLM Post-Training as Brownfield Maintenance: An Industrial Perspective on Dataware Engineering
Gopi Krishnan Rajbahadur, Amir M. Ebrahimi, Boyuan Chen +1
cs.SEcs.AIcs.LGarXiv:2608.31102v12026Natural Attack for Pre-trained Models of Code
Zhou Yang, Jieke Shi, Junda He +1
cs.SEarXiv:2201.08698v22022MontiCore: a Framework for Compositional Development of Domain Specific Languages
Holger Krahn, Bernhard Rumpe, Stefan Völkel
cs.SEarXiv:1409.2367v12014Extensible Component Based Architecture for FLASH, A Massively Parallel, Multiphysics Simulation Code
A. Dubey, L. B. Reid, K. Weide +5
cs.SEarXiv:0903.4875v22009TSExplorer: An interactive data annotation and exploration tool for time-series data
Einari Vaaras, Manu Airaksinen, Okko Räsänen
cs.HCcs.AIcs.LGarXiv:2608.30514v12026jTrans: Jump-Aware Transformer for Binary Code Similarity
Hao Wang, Wenjie Qu, Gilad Katz +5
cs.CRcs.SEarXiv:2205.12713v12022What happens when software developers are (un)happy
Daniel Graziotin, Fabian Fagerholm, Xiaofeng Wang +1
cs.SEcs.CYarXiv:1707.00432v32017Evaluating the Code Quality of AI-Assisted Code Generation Tools: An Empirical Study on GitHub Copilot, Amazon CodeWhisperer, and ChatGPT
Burak Yetiştiren, Işık Özsoy, Miray Ayerdem +1
cs.SEarXiv:2304.10778v22023Investigating Explainability of Generative AI for Code through Scenario-based Design
Jiao Sun, Q. Vera Liao, Michael Muller +4
cs.HCcs.AIcs.SEarXiv:2202.04903v12022Multi-task Learning based Pre-trained Language Model for Code Completion
Fang Liu, Ge Li, Yunfei Zhao +1
cs.SEarXiv:2012.14631v12020Large Language Models are Few-Shot Summarizers: Multi-Intent Comment Generation via In-Context Learning
Mingyang Geng, Shangwen Wang, Dezun Dong +5
cs.SEarXiv:2304.11384v32023Neural Program Repair with Execution-based Backpropagation
He Ye, Matias Martinez, Martin Monperrus
cs.SEarXiv:2105.04123v32021A Disciplined Approach to Adopting Agile Practices: The Agile Adoption Framework
Ahmed Sidky, James Arthur, Shawn Bohner
cs.SEarXiv:0704.1294v12007Big Code != Big Vocabulary: Open-Vocabulary Models for Source Code
Rafael-Michael Karampatsis, Hlib Babii, Romain Robbes +2
cs.SEarXiv:2003.07914v12020Learning Syntactic Program Transformations from Examples
Reudismam Rolim, Gustavo Soares, Loris D'Antoni +5
cs.SEcs.LGcs.PLarXiv:1608.09000v12016IFogSim2: An Extended iFogSim Simulator for Mobility, Clustering, and Microservice Management in Edge and Fog Computing Environments
Redowan Mahmud, Samodha Pallewatta, Mohammad Goudarzi +1
cs.DCcs.PFcs.SEarXiv:2109.05636v22021Mahotas: Open source software for scriptable computer vision
Luis Pedro Coelho
cs.CVcs.SEarXiv:1211.4907v22012Factors that Affect Software Systems Development Project Outcomes: A Survey of Research
Laurie McLeod, Stephen G. MacDonell
cs.SEarXiv:2101.08442v12021Preventing Repeated Real World AI Failures by Cataloging Incidents: The AI Incident Database
Sean McGregor
cs.CYcs.SEarXiv:2011.08512v12020