Software Engineering

Papers filed under cs.SE on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

841 to 900 of 1,389

  1. Magicoder: Empowering Code Generation with OSS-Instruct

    Yuxiang Wei, Zhe Wang, Jiawei Liu +2

    cs.CLcs.AIcs.SEarXiv:2312.02120v22023
  2. A Syntax-Guided Edit Decoder for Neural Program Repair

    Qihao Zhu, Zeyu Sun, Yuan-an Xiao +4

    cs.SEcs.AIarXiv:2106.08253v62021
  3. Taxonomy of Attacks on Open-Source Software Supply Chains

    Piergiorgio Ladisa, Henrik Plate, Matias Martinez +1

    cs.CRcs.SEarXiv:2204.04008v22022
  4. Loghub: A Large Collection of System Log Datasets for AI-driven Log Analytics

    Jieming Zhu, Shilin He, Pinjia He +2

    cs.SEarXiv:2008.06448v32020
  5. Superion: Grammar-Aware Greybox Fuzzing

    Junjie Wang, Bihuan Chen, Lei Wei +1

    cs.CRcs.SEarXiv:1812.01197v32018
  6. Productivity Assessment of Neural Code Completion

    Albert Ziegler, Eirini Kalliamvakou, Shawn Simister +5

    cs.SEcs.CLcs.HCarXiv:2205.06537v12022
  7. Large Language Models for Code: Security Hardening and Adversarial Testing

    Jingxuan He, Martin Vechev

    cs.CRcs.LGcs.PLarXiv:2302.05319v52023
  8. Predictive Monitoring of Business Processes

    Fabrizio Maria Maggi, Chiara Di Francescomarino, Marlon Dumas +1

    cs.SEarXiv:1312.4874v22013
  9. A Simulation Model for the Waterfall Software Development Life Cycle

    Youssef Bassil

    cs.SEarXiv:1205.6904v12012
  10. Learnable Programming: Blocks and Beyond

    David Bau, Jeff Gray, Caitlin Kelleher +2

    cs.PLcs.CYcs.HCarXiv:1705.09413v12017
  11. DLFuzz: Differential Fuzzing Testing of Deep Learning Systems

    Jianmin Guo, Yu Jiang, Yue Zhao +2

    cs.SEarXiv:1808.09413v12018
  12. How do Data Science Workers Collaborate? Roles, Workflows, and Tools

    Amy X. Zhang, Michael Muller, Dakuo Wang

    cs.HCcs.AIcs.LGarXiv:2001.06684v32020
  13. LLM-Based Multi-Agent Systems for Software Engineering: Literature Review, Vision and the Road Ahead

    Junda He, Christoph Treude, David Lo

    cs.SEarXiv:2404.04834v42024
  14. ToolMinimize: Auditing and Rewriting LLM Agent Tool Calls to Minimize Privacy Exposure

    Wenbiao Li, Yuqiao Xu

    cs.CRcs.SEarXiv:2608.24957v12026
  15. SPIDER4TianoCore: Enhancing Patch-Propagation for the TianoCore UEFI Firmware Development Ecosystem

    Laura Baird, Devin Haggitt, Terrance E. Boult +2

    cs.SEarXiv:2608.23755v12026
  16. Vibe Coding: Practice, Performance, Productivity, and Risk -A State-of-the-Art Review

    Dominik L. Michels, Mutaz Abu Ghazaleh, Francois Lazzari +2

    cs.SEarXiv:2608.20446v12026
  17. CodeSearchNet Challenge: Evaluating the State of Semantic Code Search

    Hamel Husain, Ho-Hsiang Wu, Tiferet Gazit +2

    cs.LGcs.IRcs.SEarXiv:1909.09436v32019
  18. Large Language Models are Few-shot Testers: Exploring LLM-based General Bug Reproduction

    Sungmin Kang, Juyeon Yoon, Shin Yoo

    cs.SEarXiv:2209.11515v32022
  19. Software Engineering for AI-Based Systems: A Survey

    Silverio Martínez-Fernández, Justus Bogner, Xavier Franch +5

    cs.SEcs.AIcs.LGarXiv:2105.01984v22021
  20. An Empirical Study of the Non-determinism of ChatGPT in Code Generation

    Shuyin Ouyang, Jie M. Zhang, Mark Harman +1

    cs.SEarXiv:2308.02828v22023
  21. Precise Condition Synthesis for Program Repair

    Yingfei Xiong, Jie Wang, Runfa Yan +4

    cs.SEarXiv:1608.07754v52016
  22. Towards Accountability for Machine Learning Datasets: Practices from Software Engineering and Infrastructure

    Ben Hutchinson, Andrew Smart, Alex Hanna +5

    cs.LGcs.CYcs.DBarXiv:2010.13561v22020
  23. Improved Code Summarization via a Graph Neural Network

    Alexander LeClair, Sakib Haque, Lingfei Wu +1

    cs.SEcs.CLarXiv:2004.02843v22020
  24. AFlow: Automating Agentic Workflow Generation

    Jiayi Zhang, Jinyu Xiang, Zhaoyang Yu +11

    cs.AIcs.CLcs.LGarXiv:2410.10762v42024
  25. InferFix: End-to-End Program Repair with LLMs

    Matthew Jin, Syed Shahriar, Michele Tufano +4

    cs.SEarXiv:2303.07263v12023
  26. Autoformalization with Large Language Models

    Yuhuai Wu, Albert Q. Jiang, Wenda Li +4

    cs.LGcs.AIcs.LOarXiv:2205.12615v12022
  27. Detecting Code Clones with Graph Neural Networkand Flow-Augmented Abstract Syntax Tree

    Wenhan Wang, Ge Li, Bo Ma +2

    cs.SEcs.AIarXiv:2002.08653v12020
  28. CRUXEval: A Benchmark for Code Reasoning, Understanding and Execution

    Alex Gu, Baptiste Rozière, Hugh Leather +3

    cs.SEcs.AIcs.LGarXiv:2401.03065v12024
  29. CVEfixes: Automated Collection of Vulnerabilities and Their Fixes from Open-Source Software

    Guru Prasad Bhandari, Amara Naseer, Leon Moonen

    cs.SEcs.AIcs.CRarXiv:2107.08760v12021
  30. Fuzz4All: Universal Fuzzing with Large Language Models

    Chunqiu Steven Xia, Matteo Paltenghi, Jia Le Tian +2

    cs.SEcs.LGarXiv:2308.04748v32023
  31. Concolic Testing for Deep Neural Networks

    Youcheng Sun, Min Wu, Wenjie Ruan +3

    cs.LGcs.SEstat.MLarXiv:1805.00089v22018
  32. Understanding the Factors that Impact the Popularity of GitHub Repositories

    Hudson Borges, Andre Hora, Marco Tulio Valente

    cs.SEcs.SIarXiv:1606.04984v32016
  33. Autonomous Vehicles on the Edge: A Survey on Autonomous Vehicle Racing

    Johannes Betz, Hongrui Zheng, Alexander Liniger +5

    cs.ROcs.SEarXiv:2202.07008v12022
  34. The Adverse Effects of Code Duplication in Machine Learning Models of Code

    Miltiadis Allamanis

    cs.SEcs.LGarXiv:1812.06469v62018
  35. DiverseVul: A New Vulnerable Source Code Dataset for Deep Learning Based Vulnerability Detection

    Yizheng Chen, Zhoujie Ding, Lamya Alowain +2

    cs.CRcs.AIcs.LGarXiv:2304.00409v22023
  36. Vulnerability Detection with Fine-grained Interpretations

    Yi Li, Shaohua Wang, Tien N. Nguyen

    cs.CRcs.SEarXiv:2106.10478v12021
  37. Few-shot training LLMs for project-specific code-summarization

    Toufique Ahmed, Premkumar Devanbu

    cs.SEcs.LGarXiv:2207.04237v22022
  38. Taxonomy of Real Faults in Deep Learning Systems

    Nargiz Humbatova, Gunel Jahangirova, Gabriele Bavota +3

    cs.SEcs.AIcs.LGarXiv:1910.11015v32019
  39. A Comprehensive Study on Deep Learning Bug Characteristics

    Md Johirul Islam, Giang Nguyen, Rangeet Pan +1

    cs.SEcs.LGarXiv:1906.01388v12019
  40. Learning to Mine Aligned Code and Natural Language Pairs from Stack Overflow

    Pengcheng Yin, Bowen Deng, Edgar Chen +2

    cs.CLcs.SEarXiv:1805.08949v12018
  41. LEVER: Learning to Verify Language-to-Code Generation with Execution

    Ansong Ni, Srini Iyer, Dragomir Radev +4

    cs.LGcs.CLcs.PLarXiv:2302.08468v32023
  42. CodeAgent: Enhancing Code Generation with Tool-Integrated Agent Systems for Real-World Repo-level Coding Challenges

    Kechi Zhang, Jia Li, Ge Li +2

    cs.SEarXiv:2401.07339v22024
  43. Backstabber's Knife Collection: A Review of Open Source Software Supply Chain Attacks

    Marc Ohm, Henrik Plate, Arnold Sykosch +1

    cs.CRcs.SEarXiv:2005.09535v12020
  44. Structured Chain-of-Thought Prompting for Code Generation

    Jia Li, Ge Li, Yongmin Li +1

    cs.SEcs.CLarXiv:2305.06599v32023
  45. Co-evolution of platform architecture, platform services, and platform governance: Expanding the platform value of industrial digital platforms

    Marin Jovanovic, David Sjodin, Vinit Parida

    cs.SEarXiv:2102.04862v12021
  46. Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security

    Yuanchun Li, Hao Wen, Weijun Wang +22

    cs.HCcs.AIcs.SEarXiv:2401.05459v22024
  47. Keep the Conversation Going: Fixing 162 out of 337 bugs for $0.42 each using ChatGPT

    Chunqiu Steven Xia, Lingming Zhang

    cs.SEcs.LGarXiv:2304.00385v22023
  48. AutoCodeRover: Autonomous Program Improvement

    Yuntong Zhang, Haifeng Ruan, Zhiyu Fan +1

    cs.SEcs.AIarXiv:2404.05427v32024
  49. Cost-Utility Alignment in LLM Agent Trajectories:Profiling,Attribution,Diagnosis,Adaptation,and Evaluation

    Dan Liu, Jian Li

    cs.SEarXiv:2608.26195v12026
  50. Do Developers Update Their Library Dependencies? An Empirical Study on the Impact of Security Advisories on Library Migration

    Raula Gaikovina Kula, Daniel M. German, Ali Ouni +2

    cs.SEarXiv:1709.04621v12017
  51. Four Ways to Forge a Bundle My Own Verifier Calls Clean: Refusal-Site Mutation Testing of an Evidence-Bundle Verifier

    Erik Hill

    cs.SEcs.CRarXiv:2608.26183v12026
  52. Under-Optimized Smart Contracts Devour Your Money

    Ting Chen, Xiaoqi Li, Xiapu Luo +1

    cs.SEarXiv:1703.03994v22017
  53. SPA: Securing Persistent LLM Agents Across Queries with Plan-First Information-Flow Control

    Dylan Girrens, Guangjing Wang

    cs.CRcs.SEarXiv:2608.27234v12026
  54. A Trans-Domain Digital Twin for Bio-Aware Control of Climate and Energy in Cattle Fattening Barns Using Single-Episode Optimizer Learning

    Mansoorali Amiri

    cs.SEarXiv:2608.27185v12026
  55. BTS-AgentBench: A Deterministic, Replayable Pipeline from Read-Only Telemetry Logs to Agent Benchmarks

    Jeong-Yoon Kim

    cs.CLcs.SEarXiv:2608.27334v12026
  56. Automated Discovery of Process Models from Event Logs: Review and Benchmark

    Adriano Augusto, Raffaele Conforti, Marlon Dumas +5

    cs.SEarXiv:1705.02288v32017
  57. Beyond Execution: Auditing Experimental Fidelity in LLM-Driven Scientific Research

    Lezhi Yu, Xiaogang Xu, Yuhua Zhou +2

    cs.SEcs.AIarXiv:2608.26753v12026
  58. FaultLens: Learning Compact Behavioral Test Suites for Generated Operational Programs

    Zeming Liu, Hang Lyu, Jingtao Zhang

    cs.SEcs.AIarXiv:2608.26746v12026
  59. A Contract-Centered Architecture for Scalable and Manageable Agentic Runtimes

    Yaxiao Liu, Pengbo Liu, Yiwen Liu +3

    cs.AIcs.MAcs.SEarXiv:2608.27086v12026
  60. Harness Engineering for Predictable Agentic Systems: An Empirical Study of Deterministic Execution Constraints

    Saransh Dhage

    cs.SEarXiv:2608.26197v12026