Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

9,841 to 9,900 of 15,404

  1. Stochastic Gradient Push for Distributed Deep Learning

    Mahmoud Assran, Nicolas Loizou, Nicolas Ballas +1

    cs.LGcs.AIcs.DCarXiv:1811.10792v32018
  2. Batch Policy Learning under Constraints

    Hoang M. Le, Cameron Voloshin, Yisong Yue

    cs.LGcs.AImath.OCarXiv:1903.08738v12019
  3. SpeechGym: An Audio-Native Gym for Training Voice Agents via Reinforcement Learning

    Jiajun Fan, Jingyuan Li, Prashanth Gurunath Shivakumar +6

    cs.SDcs.AIcs.CLarXiv:2608.26432v12026
  4. Neural Optimizer Search with Reinforcement Learning

    Irwan Bello, Barret Zoph, Vijay Vasudevan +1

    cs.AIcs.LGstat.MLarXiv:1709.07417v22017
  5. Funnel Libraries for Real-Time Robust Feedback Motion Planning

    Anirudha Majumdar, Russ Tedrake

    cs.ROcs.AIeess.SYarXiv:1601.04037v32016
  6. Dataset Security for Machine Learning: Data Poisoning, Backdoor Attacks, and Defenses

    Micah Goldblum, Dimitris Tsipras, Chulin Xie +6

    cs.LGcs.AIcs.CRarXiv:2012.10544v42020
  7. Mastering Complex Control in MOBA Games with Deep Reinforcement Learning

    Deheng Ye, Zhao Liu, Mingfei Sun +15

    cs.AIcs.LGarXiv:1912.09729v32019
  8. Great, Now Write an Article About That: The Crescendo Multi-Turn LLM Jailbreak Attack

    Mark Russinovich, Ahmed Salem, Ronen Eldan

    cs.CRcs.AIarXiv:2404.01833v32024
  9. Bayesian Optimization is Superior to Random Search for Machine Learning Hyperparameter Tuning: Analysis of the Black-Box Optimization Challenge 2020

    Ryan Turner, David Eriksson, Michael McCourt +4

    cs.LGcs.AIstat.MLarXiv:2104.10201v22021
  10. Continual Learning of Context-dependent Processing in Neural Networks

    Guanxiong Zeng, Yang Chen, Bo Cui +1

    cs.LGcs.AIcs.CVarXiv:1810.01256v32018
  11. Compositionality decomposed: how do neural networks generalise?

    Dieuwke Hupkes, Verna Dankers, Mathijs Mul +1

    cs.CLcs.AIcs.LGarXiv:1908.08351v22019
  12. Visualizing and Understanding Atari Agents

    Sam Greydanus, Anurag Koul, Jonathan Dodge +1

    cs.AIarXiv:1711.00138v52017
  13. A Survey on Anomaly Detection for Technical Systems using LSTM Networks

    Benjamin Lindemann, Benjamin Maschler, Nada Sahlab +1

    cs.LGcs.AIstat.MLarXiv:2105.13810v12021
  14. Melding the Data-Decisions Pipeline: Decision-Focused Learning for Combinatorial Optimization

    Bryan Wilder, Bistra Dilkina, Milind Tambe

    cs.LGcs.AIstat.MLarXiv:1809.05504v22018
  15. Text2Motion: From Natural Language Instructions to Feasible Plans

    Kevin Lin, Christopher Agia, Toki Migimatsu +2

    cs.ROcs.AIcs.LGarXiv:2303.12153v52023
  16. Vid2Seq: Large-Scale Pretraining of a Visual Language Model for Dense Video Captioning

    Antoine Yang, Arsha Nagrani, Paul Hongsuck Seo +5

    cs.CVcs.AIcs.CLarXiv:2302.14115v22023
  17. Safety Does Not Compose: Non-Decaying Loop State for Autonomous LLM Agents

    Chenhao Wu, Haoxuan Jia, Yang Liu +11

    cs.CRcs.AIarXiv:2608.27141v12026
  18. Emotional Preferences as Goal-Priority Regulation

    Shiqi Liu, Yihua Tan, Hu Fu +1

    cs.LGcs.AIarXiv:2608.27072v12026
  19. Robots That Ask For Help: Uncertainty Alignment for Large Language Model Planners

    Allen Z. Ren, Anushri Dixit, Alexandra Bodrova +11

    cs.ROcs.AIstat.AParXiv:2307.01928v22023
  20. Multi-Person Human Motion Forecasting in Complex Scenes

    Serdar Ozsoy, Lars Doorenbos, Juergen Gall

    cs.CVcs.AIarXiv:2608.27039v12026
  21. AudioGPT: Understanding and Generating Speech, Music, Sound, and Talking Head

    Rongjie Huang, Mingze Li, Dongchao Yang +10

    cs.CLcs.AIcs.SDarXiv:2304.12995v12023
  22. The Platonic Representation Hypothesis

    Minyoung Huh, Brian Cheung, Tongzhou Wang +1

    cs.LGcs.AIcs.CVarXiv:2405.07987v52024
  23. Improving Semantic Segmentation via Video Propagation and Label Relaxation

    Yi Zhu, Karan Sapra, Fitsum A. Reda +4

    cs.CVcs.AIcs.MMarXiv:1812.01593v32018
  24. Mapping the Landscape of Artificial Intelligence Applications against COVID-19

    Joseph Bullock, Alexandra Luccioni, Katherine Hoffmann Pham +2

    cs.CYcs.AIcs.LGarXiv:2003.11336v32020
  25. MHFormer: Multi-Hypothesis Transformer for 3D Human Pose Estimation

    Wenhao Li, Hong Liu, Hao Tang +2

    cs.CVcs.AIcs.LGarXiv:2111.12707v42021
  26. 3D Hand Shape and Pose from Images in the Wild

    Adnane Boukhayma, Rodrigo de Bem, Philip H. S. Torr

    cs.CVcs.AIcs.LGarXiv:1902.03451v12019
  27. Aequitas: A Bias and Fairness Audit Toolkit

    Pedro Saleiro, Benedict Kuester, Loren Hinkson +5

    cs.LGcs.AIcs.CYarXiv:1811.05577v22018
  28. PRECOG: PREdiction Conditioned On Goals in Visual Multi-Agent Settings

    Nicholas Rhinehart, Rowan McAllister, Kris Kitani +1

    cs.CVcs.AIcs.LGarXiv:1905.01296v32019
  29. PandaLM: An Automatic Evaluation Benchmark for LLM Instruction Tuning Optimization

    Yidong Wang, Zhuohao Yu, Zhengran Zeng +10

    cs.CLcs.AIarXiv:2306.05087v22023
  30. Vision-and-Dialog Navigation

    Jesse Thomason, Michael Murray, Maya Cakmak +1

    cs.CLcs.AIcs.CVarXiv:1907.04957v32019
  31. FaulT-Bench: Towards Benchmarking Network Troubleshooting LLM Agents under Unreliable User Tickets

    Kuan-Hao Tseng, Niruth Bogahawatta, Yasod Ginige +3

    cs.NIcs.AIarXiv:2608.27021v12026
  32. Geometric Deep Learning on Molecular Representations

    Kenneth Atz, Francesca Grisoni, Gisbert Schneider

    physics.chem-phcs.AIcs.LGarXiv:2107.12375v42021
  33. RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems

    Tianyang Liu, Canwen Xu, Julian McAuley

    cs.CLcs.AIcs.SEarXiv:2306.03091v22023
  34. How Do LLM Agents Actually Get the Flag? Trace-Level Provenance for Agentic Offensive Security Evaluation

    Kimberly Milner, Minghao Shao, Nanda Rani +8

    cs.CRcs.AIarXiv:2608.26237v12026
  35. TSMixer: Lightweight MLP-Mixer Model for Multivariate Time Series Forecasting

    Vijay Ekambaram, Arindam Jati, Nam Nguyen +2

    cs.LGcs.AIarXiv:2306.09364v42023
  36. Modality Maturity Index: A benchmark for assessing multimodal capabilities of omni models

    Rohit Patel, Dieuwke Hupkes, Sloan Strader

    cs.CVcs.AIcs.MMarXiv:2608.26317v12026
  37. NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models

    Zeqian Ju, Yuancheng Wang, Kai Shen +16

    eess.AScs.AIcs.CLarXiv:2403.03100v32024
  38. Decay-Region Group Delay as a Forensic Cue for AI-Generated Impulsive Sounds

    JaeHyeong Chang, Chengzhe Sun, Siwei Lyu

    cs.SDcs.AIarXiv:2608.26346v12026
  39. Risk-Sensitive and Robust Decision-Making: a CVaR Optimization Approach

    Yinlam Chow, Aviv Tamar, Shie Mannor +1

    cs.AImath.OCarXiv:1506.02188v12015
  40. Co-Evolving Structured Knowledge and Reasoning in Language Models

    Ryan Thomas Noonan, Linxi Zhao, Menghan Xu +6

    cs.CLcs.AIcs.LGarXiv:2608.26386v12026
  41. Efficient Online Reinforcement Learning with Offline Data

    Philip J. Ball, Laura Smith, Ilya Kostrikov +1

    cs.LGcs.AIarXiv:2302.02948v42023
  42. Training behavior of deep neural network in frequency domain

    Zhi-Qin John Xu, Yaoyu Zhang, Yanyang Xiao

    cs.LGcs.AIcs.ITarXiv:1807.01251v62018
  43. Understanding and Robustifying Differentiable Architecture Search

    Arber Zela, Thomas Elsken, Tonmoy Saikia +3

    cs.LGcs.AIcs.CVarXiv:1909.09656v22019
  44. Imitation from Observation: Learning to Imitate Behaviors from Raw Video via Context Translation

    YuXuan Liu, Abhishek Gupta, Pieter Abbeel +1

    cs.LGcs.AIcs.CVarXiv:1707.03374v22017
  45. BrailleBench: Investigating Multi-Criteria Braille Comprehension in Large Language Models

    Jinghan Zhang, Fengran Mo, Zhiyu Chen +3

    cs.AIcs.CLcs.HCarXiv:2608.27268v12026
  46. LiveVVT: High-Fidelity Video Virtual Try-On in Real Time

    Yushe Cao, Shikun Feng, Ruxiang Duan +4

    cs.CVcs.AIarXiv:2608.26714v12026
  47. PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-training

    Kimin Lee, Laura Smith, Pieter Abbeel

    cs.LGcs.AIarXiv:2106.05091v12021
  48. Multimodal Large Language Models: A Survey

    Jiayang Wu, Wensheng Gan, Zefeng Chen +2

    cs.AIarXiv:2311.13165v12023
  49. KnockGS:interaction-Grounded Calibrationof Physical Gaussian Representations

    Chenchen Ge, Hanwen Shen, Bowen Jing +6

    cs.CVcs.AIarXiv:2608.27365v12026
  50. Opportunities and Challenges for ChatGPT and Large Language Models in Biomedicine and Health

    Shubo Tian, Qiao Jin, Lana Yeganova +11

    cs.CYcs.AIcs.CLarXiv:2306.10070v22023
  51. Experience-driven Networking: A Deep Reinforcement Learning based Approach

    Zhiyuan Xu, Jian Tang, Jingsong Meng +4

    cs.NIcs.AIcs.LGarXiv:1801.05757v12018
  52. ConceptFusion: Open-set Multimodal 3D Mapping

    Krishna Murthy Jatavallabhula, Alihusein Kuwajerwala, Qiao Gu +14

    cs.CVcs.AIcs.ROarXiv:2302.07241v32023
  53. A Survey of Deep Meta-Learning

    Mike Huisman, Jan N. van Rijn, Aske Plaat

    cs.LGcs.AIstat.MLarXiv:2010.03522v22020
  54. Uncertainty-Based Offline Reinforcement Learning with Diversified Q-Ensemble

    Gaon An, Seungyong Moon, Jang-Hyun Kim +1

    cs.LGcs.AIarXiv:2110.01548v22021
  55. DeepCache: Accelerating Diffusion Models for Free

    Xinyin Ma, Gongfan Fang, Xinchao Wang

    cs.CVcs.AIarXiv:2312.00858v22023
  56. MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations

    Ajay Mandlekar, Soroush Nasiriany, Bowen Wen +5

    cs.ROcs.AIcs.CVarXiv:2310.17596v12023
  57. Beyond Parity: Fairness Objectives for Collaborative Filtering

    Sirui Yao, Bert Huang

    cs.IRcs.AIcs.LGarXiv:1705.08804v22017
  58. Trajectory-guided Control Prediction for End-to-end Autonomous Driving: A Simple yet Strong Baseline

    Penghao Wu, Xiaosong Jia, Li Chen +3

    cs.CVcs.AIcs.ROarXiv:2206.08129v22022
  59. DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models

    Yung-Sung Chuang, Yujia Xie, Hongyin Luo +3

    cs.CLcs.AIcs.LGarXiv:2309.03883v22023
  60. What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning

    Wei Liu, Weihao Zeng, Keqing He +2

    cs.CLcs.AIcs.LGarXiv:2312.15685v22023