Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,541 to 6,600 of 15,345

  1. Learning to Invert: Signal Recovery via Deep Convolutional Networks

    Ali Mousavi, Richard G. Baraniuk

    stat.MLcs.AIcs.ITarXiv:1701.03891v12017
  2. Theory-guided Data Science: A New Paradigm for Scientific Discovery from Data

    Anuj Karpatne, Gowtham Atluri, James Faghmous +6

    cs.LGcs.AIstat.MLarXiv:1612.08544v22016
  3. Understanding Deep Neural Networks with Rectified Linear Units

    Raman Arora, Amitabh Basu, Poorya Mianjy +1

    cs.LGcond-mat.dis-nncs.AIarXiv:1611.01491v62016
  4. Deep Visual Foresight for Planning Robot Motion

    Chelsea Finn, Sergey Levine

    cs.LGcs.AIcs.CVarXiv:1610.00696v22016
  5. "Why Should I Trust You?": Explaining the Predictions of Any Classifier

    Marco Tulio Ribeiro, Sameer Singh, Carlos Guestrin

    cs.LGcs.AIstat.MLarXiv:1602.04938v32016
  6. Learning Human Activities and Object Affordances from RGB-D Videos

    Hema Swetha Koppula, Rudhir Gupta, Ashutosh Saxena

    cs.ROcs.AIcs.CVarXiv:1210.1207v22012
  7. A Linear-Programming Approximation of AC Power Flows

    Carleton Coffrin, Pascal Van Hentenryck

    cs.AImath.OCarXiv:1206.3614v32012
  8. Experimental Comparison of Representation Methods and Distance Measures for Time Series Data

    Xiaoyue Wang, Hui Ding, Goce Trajcevski +2

    cs.AIarXiv:1012.2789v12010
  9. R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization

    Jingyi Zhang, Jiaxing Huang, Huanjin Yao +4

    cs.AIcs.CLcs.CVarXiv:2503.12937v22025
  10. StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing

    StarVLA Community

    cs.ROcs.AIcs.CVarXiv:2604.05014v12026
  11. RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning

    Zihan Wang, Kangrui Wang, Qineng Wang +15

    cs.LGcs.AIcs.CLarXiv:2504.20073v22025
  12. Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning

    LASA Team, Weiwen Xu, Hou Pong Chan +16

    cs.CLcs.AIcs.CVarXiv:2506.07044v42025
  13. TradingAgents: Multi-Agents LLM Financial Trading Framework

    Yijia Xiao, Edward Sun, Di Luo +1

    q-fin.TRcs.AIcs.CEarXiv:2412.20138v72024
  14. TAVA: Template-free Animatable Volumetric Actors

    Ruilong Li, Julian Tanke, Minh Vo +4

    cs.CVcs.AIarXiv:2206.08929v22022
  15. Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence

    Diankun Wu, Fangfu Liu, Yi-Hsin Hung +1

    cs.CVcs.AIcs.LGarXiv:2505.23747v22025
    Summaries:한국어
  16. KernelBench: Can LLMs Write Efficient GPU Kernels?

    Anne Ouyang, Simon Guo, Simran Arora +4

    cs.LGcs.AIcs.PFarXiv:2502.10517v12025
  17. SWE-smith: Scaling Data for Software Engineering Agents

    John Yang, Kilian Lieret, Carlos E. Jimenez +7

    cs.SEcs.AIcs.CLarXiv:2504.21798v22025
  18. Reinforcement Learning with Verifiable Rewards Implicitly Incentivizes Correct Reasoning in Base LLMs

    Xumeng Wen, Zihan Liu, Shun Zheng +9

    cs.AIcs.CLarXiv:2506.14245v22025
  19. Tuning Hyperparameters without Grad Students: Scalable and Robust Bayesian Optimisation with Dragonfly

    Kirthevasan Kandasamy, Karun Raju Vysyaraju, Willie Neiswanger +5

    stat.MLcs.AIcs.LGarXiv:1903.06694v22019
  20. SimpleMem: Efficient Lifelong Memory for LLM Agents

    Jiaqi Liu, Yaofeng Su, Peng Xia +5

    cs.AIarXiv:2601.02553v32026
  21. TokenSkip: Controllable Chain-of-Thought Compression in LLMs

    Heming Xia, Chak Tou Leong, Wenjie Wang +2

    cs.CLcs.AIarXiv:2502.12067v32025
  22. Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions

    Yuanzhe Hu, Yu Wang, Julian McAuley

    cs.CLcs.AIarXiv:2507.05257v42025
  23. MAGI-1: Autoregressive Video Generation at Scale

    Sand. ai, Hansi Teng, Hongyu Jia +36

    cs.CVcs.AIarXiv:2505.13211v12025
  24. Chronos-2: From Univariate to Universal Forecasting

    Abdul Fatir Ansari, Oleksandr Shchur, Jaris Küken +20

    cs.LGcs.AIstat.MLarXiv:2510.15821v12025
  25. Layer by Layer: Uncovering Hidden Representations in Language Models

    Oscar Skean, Md Rifat Arefin, Dan Zhao +4

    cs.LGcs.AIcs.CLarXiv:2502.02013v22025
  26. The Offset Tree for Learning with Partial Labels

    Alina Beygelzimer, John Langford

    cs.LGcs.AIarXiv:0812.4044v32008
  27. Does Reasoning Mitigate Backdoor Attacks? A Neuro-Symbolic Perspective

    Marco Antonio Corallo, Andrea Agiollo, Mauro Conti +1

    cs.CRcs.AIarXiv:2609.00464v12026
  28. Improving Video Generation with Human Feedback

    Jie Liu, Gongye Liu, Jiajun Liang +14

    cs.CVcs.AIcs.GRarXiv:2501.13918v22025
  29. Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models

    Qiguang Chen, Libo Qin, Jinhao Liu +7

    cs.AIcs.CLarXiv:2503.09567v52025
  30. GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization

    Shih-Yang Liu, Xin Dong, Ximing Lu +10

    cs.CLcs.AIcs.LGarXiv:2601.05242v12026
  31. CHARM: Character Hallucination for Multicultural Role Play Benchmark

    Sunkyung Han, Nahyeon Park, Gaeun Seo +2

    cs.CLcs.AIarXiv:2609.01352v12026
  32. Neuro-Symbolic Geometric Abstraction (NeuSOGA): From Observations to Symbolic Mathematical Representations

    Qingde Li, Qingqi Hong, Zihan Li +1

    cs.AIcs.CVcs.GRarXiv:2609.01408v22026
  33. FlashInfer: Efficient and Customizable Attention Engine for LLM Inference Serving

    Zihao Ye, Lequn Chen, Ruihang Lai +8

    cs.DCcs.AIcs.LGarXiv:2501.01005v22025
  34. WorldBench: Culturally Grounded Benchmark for Multilingual Agents

    Leonardo Ranaldi, Sherrie Shen, Jushi Kai +1

    cs.AIcs.CLarXiv:2609.01056v12026
  35. LLaDA2.0: Scaling Up Diffusion Language Models to 100B

    Tiwei Bie, Maosong Cao, Kun Chen +28

    cs.LGcs.AIcs.CLarXiv:2512.15745v22025
  36. Attribution Patching Outperforms Automated Circuit Discovery

    Aaquib Syed, Can Rager, Arthur Conmy

    cs.LGcs.AIcs.CLarXiv:2310.10348v22023
  37. Meta-Learning without Memorization

    Mingzhang Yin, George Tucker, Mingyuan Zhou +2

    cs.LGcs.AIstat.MLarXiv:1912.03820v32019
  38. Native and Compact Structured Latents for 3D Generation

    Jianfeng Xiang, Xiaoxue Chen, Sicheng Xu +8

    cs.CVcs.AIarXiv:2512.14692v12025
  39. PaperBench: Evaluating AI's Ability to Replicate AI Research

    Giulio Starace, Oliver Jaffe, Dane Sherburn +10

    cs.AIcs.CLarXiv:2504.01848v32025
  40. SmolVLM: Redefining small and efficient multimodal models

    Andrés Marafioti, Orr Zohar, Miquel Farré +14

    cs.AIcs.CVarXiv:2504.05299v12025
  41. Trial and Error: Exploration-Based Trajectory Optimization for LLM Agents

    Yifan Song, Da Yin, Xiang Yue +3

    cs.CLcs.AIcs.LGarXiv:2403.02502v22024
  42. ReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory

    Siru Ouyang, Jun Yan, I-Hung Hsu +14

    cs.AIcs.CLarXiv:2509.25140v22025
  43. Cleaner Speech, Weaker Generalization: Revisiting Pitt-Derived Benchmarks for Alzheimer's Disease Detection

    Luqi Sun, Shreeram Suresh Chandra, Lin Zhang +5

    cs.SDcs.AIarXiv:2609.00276v12026
  44. Kimi-Audio Technical Report

    KimiTeam, Ding Ding, Zeqian Ju +37

    eess.AScs.AIcs.CLarXiv:2504.18425v12025
  45. DiffusionNFT: Online Diffusion Reinforcement with Forward Process

    Kaiwen Zheng, Huayu Chen, Haotian Ye +7

    cs.LGcs.AIcs.CVarXiv:2509.16117v22025
  46. Agentic Retrieval-Augmented Generation: A Survey on Agentic RAG

    Aditi Singh, Abul Ehtesham, Saket Kumar +2

    cs.AIcs.CLcs.IRarXiv:2501.09136v42025
  47. AReaL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning

    Wei Fu, Jiaxuan Gao, Xujie Shen +10

    cs.LGcs.AIarXiv:2505.24298v52025
  48. TabICL: A Tabular Foundation Model for In-Context Learning on Large Data

    Jingang Qu, David Holzmüller, Gaël Varoquaux +1

    cs.LGcs.AIarXiv:2502.05564v22025
  49. When Guardrails Look Effective: Construct Validity Failures in LLM Agent Commerce Evaluation

    Peiying Zhu, Sidi Chang

    cs.AIarXiv:2609.01519v12026
  50. Automated Directed Fairness Testing

    Sakshi Udeshi, Pryanshu Arora, Sudipta Chattopadhyay

    cs.LGcs.AIcs.SEarXiv:1807.00468v22018
  51. OmniGen2: Towards Instruction-Aligned Multimodal Generation

    Chenyuan Wu, Pengfei Zheng, Ruiran Yan +19

    cs.CVcs.AIcs.CLarXiv:2506.18871v42025
  52. Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation

    Bowen Baker, Joost Huizinga, Leo Gao +6

    cs.AIarXiv:2503.11926v12025
  53. Dynamic Network Embedding Survey

    Guotong Xue, Ming Zhong, Jianxin Li +3

    cs.SIcs.AIarXiv:2103.15447v12021
  54. Cascaded V-Net using ROI masks for brain tumor segmentation

    Adrià Casamitjana, Marcel Catà, Irina Sánchez +2

    cs.CVcs.AIcs.CYarXiv:1812.11588v12018
  55. A Survey of Safety and Trustworthiness of Large Language Models through the Lens of Verification and Validation

    Xiaowei Huang, Wenjie Ruan, Wei Huang +14

    cs.AIcs.LGarXiv:2305.11391v22023
  56. Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models

    Arushi Goel, Sreyan Ghosh, Jaehyeon Kim +8

    cs.SDcs.AIcs.CLarXiv:2507.08128v22025
  57. Fast3R: Towards 3D Reconstruction of 1000+ Images in One Forward Pass

    Jianing Yang, Alexander Sax, Kevin J. Liang +6

    cs.CVcs.AIcs.GRarXiv:2501.13928v22025
  58. Zep: A Temporal Knowledge Graph Architecture for Agent Memory

    Preston Rasmussen, Pavlo Paliychuk, Travis Beauvais +2

    cs.CLcs.AIcs.IRarXiv:2501.13956v12025
  59. Survey of state-of-the-art mixed data clustering algorithms

    Amir Ahmad, Shehroz S. Khan

    cs.LGcs.AIstat.MLarXiv:1811.04364v62018
  60. Smart Predict-and-Optimize for Hard Combinatorial Optimization Problems

    Jaynta Mandi, Emir Demirović, Peter. J Stuckey +1

    cs.LGcs.AImath.OCarXiv:1911.10092v12019