Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
9,841 to 9,900 of 15,404
Stochastic Gradient Push for Distributed Deep Learning
Mahmoud Assran, Nicolas Loizou, Nicolas Ballas +1
cs.LGcs.AIcs.DCarXiv:1811.10792v32018Batch Policy Learning under Constraints
Hoang M. Le, Cameron Voloshin, Yisong Yue
cs.LGcs.AImath.OCarXiv:1903.08738v12019SpeechGym: An Audio-Native Gym for Training Voice Agents via Reinforcement Learning
Jiajun Fan, Jingyuan Li, Prashanth Gurunath Shivakumar +6
cs.SDcs.AIcs.CLarXiv:2608.26432v12026Neural Optimizer Search with Reinforcement Learning
Irwan Bello, Barret Zoph, Vijay Vasudevan +1
cs.AIcs.LGstat.MLarXiv:1709.07417v22017Funnel Libraries for Real-Time Robust Feedback Motion Planning
Anirudha Majumdar, Russ Tedrake
cs.ROcs.AIeess.SYarXiv:1601.04037v32016Dataset Security for Machine Learning: Data Poisoning, Backdoor Attacks, and Defenses
Micah Goldblum, Dimitris Tsipras, Chulin Xie +6
cs.LGcs.AIcs.CRarXiv:2012.10544v42020Mastering Complex Control in MOBA Games with Deep Reinforcement Learning
Deheng Ye, Zhao Liu, Mingfei Sun +15
cs.AIcs.LGarXiv:1912.09729v32019Great, Now Write an Article About That: The Crescendo Multi-Turn LLM Jailbreak Attack
Mark Russinovich, Ahmed Salem, Ronen Eldan
cs.CRcs.AIarXiv:2404.01833v32024Bayesian Optimization is Superior to Random Search for Machine Learning Hyperparameter Tuning: Analysis of the Black-Box Optimization Challenge 2020
Ryan Turner, David Eriksson, Michael McCourt +4
cs.LGcs.AIstat.MLarXiv:2104.10201v22021Continual Learning of Context-dependent Processing in Neural Networks
Guanxiong Zeng, Yang Chen, Bo Cui +1
cs.LGcs.AIcs.CVarXiv:1810.01256v32018Compositionality decomposed: how do neural networks generalise?
Dieuwke Hupkes, Verna Dankers, Mathijs Mul +1
cs.CLcs.AIcs.LGarXiv:1908.08351v22019Visualizing and Understanding Atari Agents
Sam Greydanus, Anurag Koul, Jonathan Dodge +1
cs.AIarXiv:1711.00138v52017A Survey on Anomaly Detection for Technical Systems using LSTM Networks
Benjamin Lindemann, Benjamin Maschler, Nada Sahlab +1
cs.LGcs.AIstat.MLarXiv:2105.13810v12021Melding the Data-Decisions Pipeline: Decision-Focused Learning for Combinatorial Optimization
Bryan Wilder, Bistra Dilkina, Milind Tambe
cs.LGcs.AIstat.MLarXiv:1809.05504v22018Text2Motion: From Natural Language Instructions to Feasible Plans
Kevin Lin, Christopher Agia, Toki Migimatsu +2
cs.ROcs.AIcs.LGarXiv:2303.12153v52023Vid2Seq: Large-Scale Pretraining of a Visual Language Model for Dense Video Captioning
Antoine Yang, Arsha Nagrani, Paul Hongsuck Seo +5
cs.CVcs.AIcs.CLarXiv:2302.14115v22023Safety Does Not Compose: Non-Decaying Loop State for Autonomous LLM Agents
Chenhao Wu, Haoxuan Jia, Yang Liu +11
cs.CRcs.AIarXiv:2608.27141v12026Emotional Preferences as Goal-Priority Regulation
Shiqi Liu, Yihua Tan, Hu Fu +1
cs.LGcs.AIarXiv:2608.27072v12026Robots That Ask For Help: Uncertainty Alignment for Large Language Model Planners
Allen Z. Ren, Anushri Dixit, Alexandra Bodrova +11
cs.ROcs.AIstat.AParXiv:2307.01928v22023Multi-Person Human Motion Forecasting in Complex Scenes
Serdar Ozsoy, Lars Doorenbos, Juergen Gall
cs.CVcs.AIarXiv:2608.27039v12026AudioGPT: Understanding and Generating Speech, Music, Sound, and Talking Head
Rongjie Huang, Mingze Li, Dongchao Yang +10
cs.CLcs.AIcs.SDarXiv:2304.12995v12023The Platonic Representation Hypothesis
Minyoung Huh, Brian Cheung, Tongzhou Wang +1
cs.LGcs.AIcs.CVarXiv:2405.07987v52024Improving Semantic Segmentation via Video Propagation and Label Relaxation
Yi Zhu, Karan Sapra, Fitsum A. Reda +4
cs.CVcs.AIcs.MMarXiv:1812.01593v32018Mapping the Landscape of Artificial Intelligence Applications against COVID-19
Joseph Bullock, Alexandra Luccioni, Katherine Hoffmann Pham +2
cs.CYcs.AIcs.LGarXiv:2003.11336v32020MHFormer: Multi-Hypothesis Transformer for 3D Human Pose Estimation
Wenhao Li, Hong Liu, Hao Tang +2
cs.CVcs.AIcs.LGarXiv:2111.12707v420213D Hand Shape and Pose from Images in the Wild
Adnane Boukhayma, Rodrigo de Bem, Philip H. S. Torr
cs.CVcs.AIcs.LGarXiv:1902.03451v12019Aequitas: A Bias and Fairness Audit Toolkit
Pedro Saleiro, Benedict Kuester, Loren Hinkson +5
cs.LGcs.AIcs.CYarXiv:1811.05577v22018PRECOG: PREdiction Conditioned On Goals in Visual Multi-Agent Settings
Nicholas Rhinehart, Rowan McAllister, Kris Kitani +1
cs.CVcs.AIcs.LGarXiv:1905.01296v32019PandaLM: An Automatic Evaluation Benchmark for LLM Instruction Tuning Optimization
Yidong Wang, Zhuohao Yu, Zhengran Zeng +10
cs.CLcs.AIarXiv:2306.05087v22023Vision-and-Dialog Navigation
Jesse Thomason, Michael Murray, Maya Cakmak +1
cs.CLcs.AIcs.CVarXiv:1907.04957v32019FaulT-Bench: Towards Benchmarking Network Troubleshooting LLM Agents under Unreliable User Tickets
Kuan-Hao Tseng, Niruth Bogahawatta, Yasod Ginige +3
cs.NIcs.AIarXiv:2608.27021v12026Geometric Deep Learning on Molecular Representations
Kenneth Atz, Francesca Grisoni, Gisbert Schneider
physics.chem-phcs.AIcs.LGarXiv:2107.12375v42021RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems
Tianyang Liu, Canwen Xu, Julian McAuley
cs.CLcs.AIcs.SEarXiv:2306.03091v22023How Do LLM Agents Actually Get the Flag? Trace-Level Provenance for Agentic Offensive Security Evaluation
Kimberly Milner, Minghao Shao, Nanda Rani +8
cs.CRcs.AIarXiv:2608.26237v12026TSMixer: Lightweight MLP-Mixer Model for Multivariate Time Series Forecasting
Vijay Ekambaram, Arindam Jati, Nam Nguyen +2
cs.LGcs.AIarXiv:2306.09364v42023Modality Maturity Index: A benchmark for assessing multimodal capabilities of omni models
Rohit Patel, Dieuwke Hupkes, Sloan Strader
cs.CVcs.AIcs.MMarXiv:2608.26317v12026NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models
Zeqian Ju, Yuancheng Wang, Kai Shen +16
eess.AScs.AIcs.CLarXiv:2403.03100v32024Decay-Region Group Delay as a Forensic Cue for AI-Generated Impulsive Sounds
JaeHyeong Chang, Chengzhe Sun, Siwei Lyu
cs.SDcs.AIarXiv:2608.26346v12026Risk-Sensitive and Robust Decision-Making: a CVaR Optimization Approach
Yinlam Chow, Aviv Tamar, Shie Mannor +1
cs.AImath.OCarXiv:1506.02188v12015Co-Evolving Structured Knowledge and Reasoning in Language Models
Ryan Thomas Noonan, Linxi Zhao, Menghan Xu +6
cs.CLcs.AIcs.LGarXiv:2608.26386v12026Efficient Online Reinforcement Learning with Offline Data
Philip J. Ball, Laura Smith, Ilya Kostrikov +1
cs.LGcs.AIarXiv:2302.02948v42023Training behavior of deep neural network in frequency domain
Zhi-Qin John Xu, Yaoyu Zhang, Yanyang Xiao
cs.LGcs.AIcs.ITarXiv:1807.01251v62018Understanding and Robustifying Differentiable Architecture Search
Arber Zela, Thomas Elsken, Tonmoy Saikia +3
cs.LGcs.AIcs.CVarXiv:1909.09656v22019Imitation from Observation: Learning to Imitate Behaviors from Raw Video via Context Translation
YuXuan Liu, Abhishek Gupta, Pieter Abbeel +1
cs.LGcs.AIcs.CVarXiv:1707.03374v22017BrailleBench: Investigating Multi-Criteria Braille Comprehension in Large Language Models
Jinghan Zhang, Fengran Mo, Zhiyu Chen +3
cs.AIcs.CLcs.HCarXiv:2608.27268v12026LiveVVT: High-Fidelity Video Virtual Try-On in Real Time
Yushe Cao, Shikun Feng, Ruxiang Duan +4
cs.CVcs.AIarXiv:2608.26714v12026PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-training
Kimin Lee, Laura Smith, Pieter Abbeel
cs.LGcs.AIarXiv:2106.05091v12021Multimodal Large Language Models: A Survey
Jiayang Wu, Wensheng Gan, Zefeng Chen +2
cs.AIarXiv:2311.13165v12023KnockGS:interaction-Grounded Calibrationof Physical Gaussian Representations
Chenchen Ge, Hanwen Shen, Bowen Jing +6
cs.CVcs.AIarXiv:2608.27365v12026Opportunities and Challenges for ChatGPT and Large Language Models in Biomedicine and Health
Shubo Tian, Qiao Jin, Lana Yeganova +11
cs.CYcs.AIcs.CLarXiv:2306.10070v22023Experience-driven Networking: A Deep Reinforcement Learning based Approach
Zhiyuan Xu, Jian Tang, Jingsong Meng +4
cs.NIcs.AIcs.LGarXiv:1801.05757v12018ConceptFusion: Open-set Multimodal 3D Mapping
Krishna Murthy Jatavallabhula, Alihusein Kuwajerwala, Qiao Gu +14
cs.CVcs.AIcs.ROarXiv:2302.07241v32023A Survey of Deep Meta-Learning
Mike Huisman, Jan N. van Rijn, Aske Plaat
cs.LGcs.AIstat.MLarXiv:2010.03522v22020Uncertainty-Based Offline Reinforcement Learning with Diversified Q-Ensemble
Gaon An, Seungyong Moon, Jang-Hyun Kim +1
cs.LGcs.AIarXiv:2110.01548v22021DeepCache: Accelerating Diffusion Models for Free
Xinyin Ma, Gongfan Fang, Xinchao Wang
cs.CVcs.AIarXiv:2312.00858v22023MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations
Ajay Mandlekar, Soroush Nasiriany, Bowen Wen +5
cs.ROcs.AIcs.CVarXiv:2310.17596v12023Beyond Parity: Fairness Objectives for Collaborative Filtering
Sirui Yao, Bert Huang
cs.IRcs.AIcs.LGarXiv:1705.08804v22017Trajectory-guided Control Prediction for End-to-end Autonomous Driving: A Simple yet Strong Baseline
Penghao Wu, Xiaosong Jia, Li Chen +3
cs.CVcs.AIcs.ROarXiv:2206.08129v22022DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models
Yung-Sung Chuang, Yujia Xie, Hongyin Luo +3
cs.CLcs.AIcs.LGarXiv:2309.03883v22023What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning
Wei Liu, Weihao Zeng, Keqing He +2
cs.CLcs.AIcs.LGarXiv:2312.15685v22023