Software Engineering
Papers filed under cs.SE on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,081 to 1,140 of 1,389
TrustDABench: Benchmarking Reliability and Robustness of LLMs for Structured Data Analysis
Boshen Shi, Yize Liu, Chen Zhao +4
cs.CLcs.SEarXiv:2608.24145v12026The Use of Machine Learning Algorithms in Recommender Systems: A Systematic Review
Ivens Portugal, Paulo Alencar, Donald Cowan
cs.SEcs.IRcs.LGarXiv:1511.05263v42015Closing the Loop: Universal Repository Representation with RPG-Encoder
Jane Luo, Chengyu Yin, Xin Zhang +10
cs.CLcs.SEarXiv:2602.02084v22026StarHarness: Evolving Harnesses with Stratified Search for Enterprise Environments
Esakkivel Esakkiraja, Denis Akhiyarov, Vikas Yadav +4
cs.AIcs.SEarXiv:2608.24804v12026Towards LLM-Enhanced Android Taint Analysis
Nicholas Miazzo, Marco Alecci, Jordan Samhi +2
cs.SEcs.CRarXiv:2608.24269v12026Favia: Forensic Agent for Vulnerability-fix Identification and Analysis
André Storhaug, Jiamou Sun, Jingyue Li
cs.SEcs.AIcs.CRarXiv:2602.12500v12026''You Can't Open an LLM With a Screwdriver'': The De-Democratization of Software
Zixuan Feng, Italo Santos, Kostadin Damevski +1
cs.SEarXiv:2608.24720v12026SoK: ARCUS: On the Efficiency and Efficacy of Hardware Fuzzing
Alenkruth Krishnan Murali, Raghul Saravanan, Sai Manoj P D +1
cs.CRcs.SEarXiv:2608.23933v12026CL4SE: Benchmarking Context Learning on Software Engineering
Haichuan Hu, Quanjun Zhang, Ye Shang +4
cs.SEarXiv:2602.23047v32026Automatic Model Card Generation Using an LLM
Tajkia Rahman Toma, Balreet Grewal, Cor-Paul Bezemer
cs.SEcs.AIarXiv:2608.24807v12026Metamorphic Testing: A New Approach for Generating Next Test Cases
T. Y. Chen, S. C. Cheung, S. M. Yiu
cs.SEarXiv:2002.12543v12020Understanding by Reconstruction: Reversing the Software Development Process for LLM Pretraining
Zhiyuan Zeng, Yichi Zhang, Yong Shan +11
cs.SEarXiv:2603.11103v22026Automatic Generation of High-Performance RL Environments
Seth Karten, Rahul Dev Appapogu, Chi Jin
cs.LGcs.AIcs.SEarXiv:2603.12145v22026Neuro-Formal Verification: Agentic Language-Agnostic Formal Program Reasoning
Shuvendu K. Lahiri
cs.SEcs.PLarXiv:2608.21516v12026Automated Vulnerability Detection in Source Code Using Deep Representation Learning
Rebecca L. Russell, Louis Kim, Lei H. Hamilton +5
cs.LGcs.AIcs.SEarXiv:1807.04320v22018Evaluating Language Models on Cross-Language Code Functional Equivalence
Hui Sun, Anderson Uchôa, Rohit Gheyi +1
cs.SEcs.AIcs.CLarXiv:2608.23961v12026How Reliable Are NVD CWE Labels? A Large-Scale Semantic Audit with Seclometry
Yu Nong, Yao Du, Majid Behravan +1
cs.CRcs.SEarXiv:2608.21977v12026Software Testing with Large Language Models: Survey, Landscape, and Vision
Junjie Wang, Yuchao Huang, Chunyang Chen +3
cs.SEarXiv:2307.07221v32023Observability and Fault Injection for LLM-Based Multi-Agent Systems in Software Engineering
Zahra Seyedghorban, Egor Klimov, Arie van Deursen +2
cs.SEarXiv:2608.24271v12026Ontology-based Requirements Transformation
Jan Novacek, Alexander Viehl, Oliver Bringmann +1
cs.SEarXiv:2608.21945v12026Ontology-supported Design Parameter Management for Change Impact Analysis
Jan Novacek, Ali Ahari, Alessandro Cornaglia +4
cs.SEarXiv:2608.21949v12026Architecture as Capability Equalizer for Coding Agents
Arquimedes Canedo
cs.SEcs.AIcs.CLarXiv:2608.21747v12026DeepGauge: Multi-Granularity Testing Criteria for Deep Learning Systems
Lei Ma, Felix Juefei-Xu, Fuyuan Zhang +9
cs.SEcs.CRcs.LGarXiv:1803.07519v42018An AI-Assisted Migration Framework for Transforming Legacy Scientific Applications into Reusable Cloud-Based Workflows
Nafiseh Soveizi, Sven Tesselaar, Hero Robinson Brouwer +1
cs.SEarXiv:2608.23146v12026RM -RF: Reward Model for Run-Free Unit Test Evaluation
Elena Bruches, Daniil Grebenkin, Mikhail Klementev +8
cs.SEcs.LGarXiv:2601.13097v12026Benchmarking the Titans: A Multi-Dimensional Empirical Evaluation of LLM Code Generation Quality in the .NET Ecosystem
Seyed Mohammad Mahdi Ghalandarian, Majid Bazargani, Masoumeh Taromirad
cs.SEarXiv:2608.22529v12026Agile Software Development Methods: Review and Analysis
Pekka Abrahamsson, Outi Salo, Jussi Ronkainen +1
cs.SEarXiv:1709.08439v12017FullStack-Agent: Enhancing Agentic Full-Stack Web Coding via Development-Oriented Testing and Repository Back-Translation
Zimu Lu, Houxing Ren, Yunqiao Yang +4
cs.SEcs.CLcs.CVarXiv:2602.03798v12026Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study
Yi Liu, Gelei Deng, Zhengzi Xu +7
cs.SEcs.AIcs.CLarXiv:2305.13860v22023StarCoder 2 and The Stack v2: The Next Generation
Anton Lozhkov, Raymond Li, Loubna Ben Allal +63
cs.SEcs.AIarXiv:2402.19173v12024Constraint-Driven Modeling Enabling Dual Model Checking and Simulation for Discrete Event Systems
Soroosh Gholami, Hessam S. Sarjoughian
cs.SEcs.ARcs.LOarXiv:2608.22095v12026TAROT: Test-driven and Capability-adaptive Curriculum Reinforcement Fine-tuning for Code Generation with Large Language Models
Chansung Park, Juyong Jiang, Fan Wang +4
cs.CLcs.LGcs.SEarXiv:2602.15449v12026Learning from the Test: Self-Referential Differential Testing for Deep RL Agents
Junda He, Jieke Shi, Zhou Yang +2
cs.SEcs.AIcs.LGarXiv:2608.22284v12026Learning Spectral Representations of Code through Latent Graph Learning for Generalizable Cross-Language Code Clone Detection
Mohsen Hesamolhokama, Ali Sadeghi, Kousha Moeini +3
cs.SEarXiv:2608.22383v12026ARGUS: MCP-Grounded Root Cause Analysis for Kubernetes Incidents
Ergi Senja, Seyed Mohammad Reza Razavi Zadegan, Philipp Leitner
cs.SEarXiv:2608.23084v12026Towards a Neural Debugger for Python
Maximilian Beck, Jonas Gehring, Jannik Kossen +1
cs.LGcs.AIcs.SEarXiv:2603.09951v12026Think Anywhere in Code Generation
Xue Jiang, Tianyu Zhang, Ge Li +8
cs.SEcs.LGarXiv:2603.29957v32026Towards Actionable Visualization: Ten Years Later, What Generative AI Changes and What It Cannot
Leonel Merino, Mohammad Ghafari, Oscar Nierstrasz
cs.SEcs.ETarXiv:2608.22151v12026The Impact of AI on Developer Productivity: Evidence from GitHub Copilot
Sida Peng, Eirini Kalliamvakou, Peter Cihon +1
cs.SEarXiv:2302.06590v12023From Natural Language Policies to Executable Obligations: A Verification Harness for Dependable In-Car LLM Agents
Radouane Bouchekir, Damir Safin, Tomas Bueno Momcilovic
cs.SEarXiv:2608.23282v12026NESSiE: The Necessary Safety Benchmark -- Identifying Errors that should not Exist
Johannes Bertram, Jonas Geiping
cs.CRcs.SEarXiv:2602.16756v12026LLMCrater: Lifecycle-Aware FAIR Metadata Generation using Large Language Models
Dani Termaat, Nafiseh Soveizi, Zhiming Zhao +1
cs.SEarXiv:2608.23158v12026Guidelines for including grey literature and conducting multivocal literature reviews in software engineering
Vahid Garousi, Michael Felderer, Mika V. Mäntylä
cs.SEcs.DLarXiv:1707.02553v42017CodeMechanic: Bug-Property-Guided Program Mitigation
Han Zheng, Rafaila Galanopoulou, Ilia Shumailov +4
cs.SEcs.CRarXiv:2608.22275v12026QuanBench+: A Unified Multi-Framework Benchmark for LLM-Based Quantum Code Generation
Ali Slim, Haydar Hamieh, Jawad Kotaich +5
cs.LGcs.AIcs.PLarXiv:2604.08570v22026Concepts for Securing Agentic AI Coding and the Terok Environment
Jiří Vyskočil, Franz Pöschel, Andreas Knüpfer
cs.AIcs.SEarXiv:2608.22930v12026ABC-Bench: Benchmarking Agentic Backend Coding in Real-World Development
Jie Yang, Honglin Guo, Li Ji +11
cs.SEcs.AIcs.CLarXiv:2601.11077v12026SWE Refactor Bench: Can Coding Agents Complete a Long-Horizon, Whole-Repository Stack Migration?
Deyao Hong, Yizhe Chi, Wenyi Li +7
cs.CLcs.AIcs.SEarXiv:2608.23564v12026Bridging Online and Offline RL: Contextual Bandit Learning for Multi-Turn Code Generation
Ziru Chen, Dongdong Chen, Ruinan Jin +3
cs.LGcs.AIcs.CLarXiv:2602.03806v12026Deep Learning based Vulnerability Detection: Are We There Yet?
Saikat Chakraborty, Rahul Krishna, Yangruibo Ding +1
cs.SEarXiv:2009.07235v12020A Survey of Symbolic Execution Techniques
Roberto Baldoni, Emilio Coppa, Daniele Cono D'Elia +2
cs.SEcs.PLarXiv:1610.00502v32016Evaluating Inference-Time Defenses Against Package Hallucination in LLM-Generated Code
Alberick Euraste Djire, Iyiola E. Olatunji, Melissa Tessa +3
cs.SEcs.AIarXiv:2608.22652v12026Do Not Copy/Paste: Soft Barriers for Copying in AI-Assisted Programming
Iyiola E. Olatunji, Alberick Euraste Djire, Jacques Klein +1
cs.SEcs.AIarXiv:2608.22638v12026"TODO: Fix the Mess Gemini Created": Towards Understanding GenAI-Induced Self-Admitted Technical Debt
Abdullah Al Mujahid, Mia Mohammad Imran
cs.SEarXiv:2601.07786v12026DPIAgent: Divide, Protocol, Isolate for Agentic Reproduction Test Generation
Hao Liu, Steven Liu, Xin Zhang +8
cs.SEarXiv:2608.23341v12026An Interactive Agent for Requirement-Driven Candidate Sourcing
Yuanpeng He, Fangjing Li, Xiangyu Ru +10
cs.SEarXiv:2608.23501v12026SafeGround: Know When to Trust GUI Grounding Models via Uncertainty Calibration
Qingni Wang, Yue Fan, Xin Eric Wang
cs.AIcs.SEarXiv:2602.02419v22026SEVerA: Verified Synthesis of Self-Evolving Agents
Debangshu Banerjee, Changming Xu, Eugene Ie +4
cs.LGcs.PLcs.SEarXiv:2603.25111v22026Learning to Commit: Generating Organic Pull Requests via Online Repository Memory
Mo Li, L. H. Xu, Qitai Tan +2
cs.SEcs.CLarXiv:2603.26664v12026CodeCircuit: Toward Inferring LLM-Generated Code Correctness via Attribution Graphs
Yicheng He, Zheng Zhao, Zhou Kaiyu +3
cs.SEcs.AIarXiv:2602.07080v12026