Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

10,921 to 10,980 of 15,416

  1. An Empirical Evaluation of Using Large Language Models for Automated Unit Test Generation

    Max Schäfer, Sarah Nadi, Aryaz Eghbali +1

    cs.SEcs.AIarXiv:2302.06527v42023
  2. Exploit More, Explore Smarter for Budget-Constrained Agentic Search

    Haoyang Fang, Bernie Wang

    cs.AIcs.LGarXiv:2608.23848v12026
  3. Benchmarking Retrieval-Augmented Generation for Medicine

    Guangzhi Xiong, Qiao Jin, Zhiyong Lu +1

    cs.CLcs.AIarXiv:2402.13178v22024
  4. Serving Masked Diffusion LLMs: Characterization and Design Principles from Real Hardware

    Farhana Amin, Sabiha Afroz, Mona Moghadampanah +1

    cs.AIarXiv:2608.23807v12026
  5. T-LESS: An RGB-D Dataset for 6D Pose Estimation of Texture-less Objects

    Tomas Hodan, Pavel Haluza, Stepan Obdrzalek +3

    cs.CVcs.AIcs.ROarXiv:1701.05498v12017
  6. A Behavior-Guided Online Probabilistic Forecasting Method for Electric vehicle Charging Loads

    Chenghan Li, Qingxiang Liu, Yinliang Xu +1

    cs.AIarXiv:2608.24441v12026
  7. Do Spoken Language Models Hear Speech as They Read Text? Bridging Structural Gaps Between Speech and Text

    Hyeonyu Kim, Hwayeon Kim, Youngwon Choi +2

    cs.CLcs.AIarXiv:2608.22908v12026
  8. Function-Level Execution Feedback for Code Preference Optimization

    Idris Nechnech, Sehwan Kim, Jimin Seo +4

    cs.AIcs.SEarXiv:2608.23632v12026
  9. Machine Comprehension Using Match-LSTM and Answer Pointer

    Shuohang Wang, Jing Jiang

    cs.CLcs.AIarXiv:1608.07905v22016
  10. Agentic-MME: What Agentic Capability Really Brings to Multimodal Intelligence?

    Qianshan Wei, Yishan Yang, Siyi Wang +12

    cs.AIarXiv:2604.03016v12026
  11. Mathematical Capabilities of ChatGPT

    Simon Frieder, Luca Pinchetti, Alexis Chevalier +5

    cs.LGcs.AIcs.CLarXiv:2301.13867v22023
  12. LLMs are Few-Shot Decision-Makers: Generalized Context-Aware Microgrid Frequency Control through Prompt Decision Transformer

    Xu Yang, Chenhui Lin, Haotian Liu +3

    eess.SYcs.AIarXiv:2608.21858v12026
  13. Score-based diffusion models for accelerated MRI

    Hyungjin Chung, Jong Chul Ye

    eess.IVcs.AIcs.CVarXiv:2110.05243v32021
  14. LLM Agents Perform Controlled Experiments Using Simulation Models

    Yuchen Xia, Michael Weyrich, Nasser Jazdi +6

    cs.AIcs.CLcs.MAarXiv:2608.23622v12026
  15. Where World Models Break: Natural-Input Failure Discovery

    Zhanpeng Shi, Zi Liang, Rong Feng +3

    cs.AIarXiv:2608.22421v12026
  16. TAT-QA: A Question Answering Benchmark on a Hybrid of Tabular and Textual Content in Finance

    Fengbin Zhu, Wenqiang Lei, Youcheng Huang +5

    cs.CLcs.AIarXiv:2105.07624v22021
  17. How Much Regularization Survives Averaging? Update Masking in Federated Learning

    Wenhao Yan, Fu Kuroda, Yucheng Jin +1

    cs.LGcs.AIarXiv:2608.23286v22026
  18. Data Mixing as Mixture Experiment: Response Surface Methodology and Optimal Design for Large Language Model Pretraining

    Yicheng Mao, Hongru Du

    cs.AIstat.MLarXiv:2608.23922v12026
  19. Restoring Without Forgetting: Continual Learning Across Image Degradations

    Alif Ashrafee, Bartosz Krawczyk

    cs.CVcs.AIcs.LGarXiv:2608.23799v12026
  20. AI Agents Push Humans Out of the Loop

    Margaret Mitchell, Avijit Ghosh, Samir Passi

    cs.AIcs.HCarXiv:2608.23642v12026
  21. Beyond Static Interpretability: Anticipating Post-SFT Mechanisms from Pre-SFT Parameters for Better Tuning

    Hang Chen, Jiaying Zhu, Wenya Wang

    cs.LGcs.AIcs.CLarXiv:2608.24482v12026
  22. Simthesizer: An Agent-Driven Simulation Framework for LLM Serving Systems

    Wonung Kim, Hyunmin Choi, Minsu Kim +3

    cs.ARcs.AIarXiv:2608.24650v12026
  23. Variational Inference of Disentangled Latent Concepts from Unlabeled Observations

    Abhishek Kumar, Prasanna Sattigeri, Avinash Balakrishnan

    cs.LGcs.AIcs.CVarXiv:1711.00848v32017
  24. Predicting Radiologist Expertise from 3D Gaze Patterns During CT Interpretation

    Leila Khaertdinova, Anna Anikina, Claudia Mello-Thoms +1

    cs.CVcs.AIcs.LGarXiv:2608.23836v12026
  25. FLARE: A Systematic, Uncertainty-Aware Framework for Evidence-Based Adoption of Artificial Intelligence in Healthcare

    Jacob Idoko, Siddhartha Paudel, Mariana Bento +2

    cs.AIarXiv:2608.23643v12026
  26. Learning Invariant Representations for Reinforcement Learning without Reconstruction

    Amy Zhang, Rowan McAllister, Roberto Calandra +2

    cs.LGcs.AIstat.MLarXiv:2006.10742v22020
  27. ET-BERT: A Contextualized Datagram Representation with Pre-training Transformers for Encrypted Traffic Classification

    Xinjie Lin, Gang Xiong, Gaopeng Gou +3

    cs.CRcs.AIcs.NIarXiv:2202.06335v22022
  28. PsychJail: Exploring Psychological Jailbreaks via Multi-Turn Persuasion of LLM Policies

    Zeyu Feng, Qingyu Wu, Yuzhe Luo +1

    cs.AIarXiv:2608.23028v12026
  29. API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs

    Minghao Li, Yingxiu Zhao, Bowen Yu +6

    cs.CLcs.AIarXiv:2304.08244v22023
  30. Counterfactual Visual Explanations

    Yash Goyal, Ziyan Wu, Jan Ernst +3

    cs.LGcs.AIcs.CVarXiv:1904.07451v22019
  31. SA-RSQ: A Versatile Sparse Representation Framework for Multi-modal Recommender Systems

    Xiang Wang, Shigang Quan, Tingzhen Chang +5

    cs.AIarXiv:2608.22979v22026
  32. Improving the matrix multiplication exponent with modern optimization and AlphaEvolve

    Emilien Dupont, Marvin Eisenberger, Borislav Kozlovskii +7

    cs.DScs.AIcs.CCarXiv:2608.16884v12026
  33. LongTraceRL: Learning Long-Context Reasoning from Search Agent Trajectories with Rubric Rewards

    Nianyi Lin, Jiajie Zhang, Lei Hou +1

    cs.CLcs.AIcs.LGarXiv:2605.31584v12026
  34. Segment Anything

    Alexander Kirillov, Eric Mintun, Nikhila Ravi +9

    cs.CVcs.AIcs.LGarXiv:2304.02643v12023
  35. Averaging Weights Leads to Wider Optima and Better Generalization

    Pavel Izmailov, Dmitrii Podoprikhin, Timur Garipov +2

    cs.LGcs.AIcs.CVarXiv:1803.05407v32018
  36. Curiosity-driven Exploration by Self-supervised Prediction

    Deepak Pathak, Pulkit Agrawal, Alexei A. Efros +1

    cs.LGcs.AIcs.CVarXiv:1705.05363v12017
  37. Context Encoders: Feature Learning by Inpainting

    Deepak Pathak, Philipp Krahenbuhl, Jeff Donahue +2

    cs.CVcs.AIcs.GRarXiv:1604.07379v22016
  38. A Dual-Dimensional LLM Framework for Automated Item Incidental Content Similarity Analysis in Large-Scale Assessments

    Jing Huang, Jihong Zhang, Hua-Hua Chang

    cs.AIarXiv:2608.24825v12026
  39. Principles and Practice of Explainable Machine Learning

    Vaishak Belle, Ioannis Papantonis

    cs.LGcs.AIstat.MLarXiv:2009.11698v12020
  40. Expanding Explainability: Towards Social Transparency in AI systems

    Upol Ehsan, Q. Vera Liao, Michael Muller +2

    cs.HCcs.AIarXiv:2101.04719v12021
  41. WinCLIP: Zero-/Few-Shot Anomaly Classification and Segmentation

    Jongheon Jeong, Yang Zou, Taewan Kim +3

    cs.CVcs.AIcs.CLarXiv:2303.14814v12023
  42. PlaceSeek: Human-Centered Geospatial Retrieval of Urban Outdoor Places via Semantic Grounding and Affective Alignment

    Ziqi Cui, Shangyu Lou

    cs.CVcs.AIcs.IRarXiv:2608.24133v12026
  43. Visual Analytics in Deep Learning: An Interrogative Survey for the Next Frontiers

    Fred Hohman, Minsuk Kahng, Robert Pienta +1

    cs.HCcs.AIcs.LGarXiv:1801.06889v32018
  44. Evolutionary Recurrent Decision Model in Developing Adaptive and Maladaptive Behaviors

    Andrew Hu

    cs.AIarXiv:2608.23932v12026
  45. Hybrid Semantic Tool Discovery for Enterprise MCP Gateway: Architecture and Implementation

    Olympia Saha, Amy Wang, Srinivasan Manoharan

    cs.IRcs.AIarXiv:2608.23992v12026
  46. Safe Latent Diffusion: Mitigating Inappropriate Degeneration in Diffusion Models

    Patrick Schramowski, Manuel Brack, Björn Deiseroth +1

    cs.CVcs.AIcs.LGarXiv:2211.05105v42022
  47. Neurosymbolic AI: The 3rd Wave

    Artur d'Avila Garcez, Luis C. Lamb

    cs.AIcs.LGarXiv:2012.05876v22020
  48. MineDojo: Building Open-Ended Embodied Agents with Internet-Scale Knowledge

    Linxi Fan, Guanzhi Wang, Yunfan Jiang +7

    cs.LGcs.AIcs.CLarXiv:2206.08853v22022
  49. Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models

    Andy Zhou, Kai Yan, Michal Shlapentokh-Rothman +2

    cs.AIcs.CLcs.CVarXiv:2310.04406v32023
  50. BOLD: Dataset and Metrics for Measuring Biases in Open-Ended Language Generation

    Jwala Dhamala, Tony Sun, Varun Kumar +4

    cs.CLcs.AIcs.LGarXiv:2101.11718v12021
  51. Connecting the Dots in Trustworthy Artificial Intelligence: From AI Principles, Ethics, and Key Requirements to Responsible AI Systems and Regulation

    Natalia Díaz-Rodríguez, Javier Del Ser, Mark Coeckelbergh +3

    cs.CYcs.AIcs.LGarXiv:2305.02231v22023
  52. A Comprehensive Survey on Test-Time Adaptation under Distribution Shifts

    Jian Liang, Ran He, Tieniu Tan

    cs.LGcs.AIcs.CVarXiv:2303.15361v22023
  53. Mechanistic Circuit Identification for Controllable Data Generation

    Nakyung Lee, Sangwoo Hong, Jungwoo Lee

    cs.LGcs.AIcs.CLarXiv:2608.24065v12026
  54. Do Transformers Really Perform Bad for Graph Representation?

    Chengxuan Ying, Tianle Cai, Shengjie Luo +5

    cs.LGcs.AIarXiv:2106.05234v52021
  55. Matched Excess-Outranker Regularization for Candidate-Set Interference in Continual Knowledge Graph Embedding

    Hao Ren, Junbin Gao, Jiaojiao Jiang

    cs.AIcs.DBcs.IRarXiv:2608.24273v12026
  56. Online Planning Algorithms for POMDPs

    Stéphane Ross, Joelle Pineau, Sébastien Paquet +1

    cs.AIarXiv:1401.3436v12014
  57. The DL-Lite Family and Relations

    Alessandro Artale, Diego Calvanese, Roman Kontchakov +1

    cs.LOcs.AIarXiv:1401.3487v12014
  58. A Practical Guide to Multi-Objective Reinforcement Learning and Planning

    Conor F. Hayes, Roxana Rădulescu, Eugenio Bargiacchi +15

    cs.AIcs.LGarXiv:2103.09568v12021
  59. GPT-4V(ision) is a Generalist Web Agent, if Grounded

    Boyuan Zheng, Boyu Gou, Jihyung Kil +2

    cs.IRcs.AIcs.CLarXiv:2401.01614v22024
  60. When Youth Enter The Chat: An Epistemic Shift in the Validation of LLM-Based Measures of Student Talk

    Liliana Santos-Deonizio, James Malamut, Ramón Martínez +1

    cs.CLcs.AIcs.HCarXiv:2608.23780v12026