Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,501 to 1,560 of 15,455
Nested conformal prediction and quantile out-of-bag ensemble methods
Chirag Gupta, Arun K. Kuchibhotla, Aaditya K. Ramdas
stat.MEcs.AImath.STarXiv:1910.10562v42019The Causal-Neural Connection: Expressiveness, Learnability, and Inference
Kevin Xia, Kai-Zhan Lee, Yoshua Bengio +1
cs.LGcs.AIarXiv:2107.00793v32021X-OPD: Cross-Modal On-Policy Distillation for Capability Alignment in Speech LLMs
Di Cao, Dongjie Fu, Hai Yu +3
eess.AScs.AIcs.CLarXiv:2603.24596v32026Few-shot In-context Learning for Knowledge Base Question Answering
Tianle Li, Xueguang Ma, Alex Zhuang +3
cs.CLcs.AIarXiv:2305.01750v22023MAD: A Scalable Dataset for Language Grounding in Videos from Movie Audio Descriptions
Mattia Soldan, Alejandro Pardo, Juan León Alcázar +4
cs.CVcs.AIarXiv:2112.00431v22021DeCap: Decoding CLIP Latents for Zero-Shot Captioning via Text-Only Training
Wei Li, Linchao Zhu, Longyin Wen +1
cs.CVcs.AIcs.CLarXiv:2303.03032v12023B-Pref: Benchmarking Preference-Based Reinforcement Learning
Kimin Lee, Laura Smith, Anca Dragan +1
cs.LGcs.AIcs.HCarXiv:2111.03026v12021Large Language Models and the Reverse Turing Test
Terrence Sejnowski
cs.CLcs.AIcs.LGarXiv:2207.14382v92022Memory-Efficient Fine-Tuning of Compressed Large Language Models via sub-4-bit Integer Quantization
Jeonghoon Kim, Jung Hyun Lee, Sungdong Kim +4
cs.LGcs.AIarXiv:2305.14152v22023Formal Verification of Autonomous Vehicle Platooning
Maryam Kamali, Louise A. Dennis, Owen McAree +2
cs.AIcs.SEarXiv:1602.01718v12016A Multi-Objective Deep Reinforcement Learning Framework
Thanh Thi Nguyen, Ngoc Duy Nguyen, Peter Vamplew +3
cs.LGcs.AIstat.MLarXiv:1803.02965v32018IPIGuard: A Novel Tool Dependency Graph-Based Defense Against Indirect Prompt Injection in LLM Agents
Hengyu An, Jinghuai Zhang, Tianyu Du +4
cs.CRcs.AIcs.CLarXiv:2508.15310v12025UNR-Explainer: Counterfactual Explanations for Unsupervised Node Representation Learning Models
Hyunju Kang, Geonhee Han, Hogun Park
cs.LGcs.AIarXiv:2605.17285v12026A spelling correction model for end-to-end speech recognition
Jinxi Guo, Tara N. Sainath, Ron J. Weiss
eess.AScs.AIcs.CLarXiv:1902.07178v12019LLM-based Agents Suffer from Hallucinations: A Survey of Taxonomy, Methods, and Directions
Xixun Lin, Yucheng Ning, Jingwen Zhang +21
cs.AIarXiv:2509.18970v22025ORGANA: A Robotic Assistant for Automated Chemistry Experimentation and Characterization
Kourosh Darvish, Marta Skreta, Yuchi Zhao +9
cs.ROcs.AIarXiv:2401.06949v22024Dynamic Multimodal Fusion
Zihui Xue, Radu Marculescu
cs.CVcs.AIcs.MMarXiv:2204.00102v22022DeepAnalyze: Agentic Large Language Models for Autonomous Data Science
Shaolei Zhang, Ju Fan, Meihao Fan +2
cs.AIcs.CLcs.DBarXiv:2510.16872v12025Theory-guided hard constraint projection (HCP): a knowledge-based data-driven scientific machine learning method
Yuntian Chen, Dou Huang, Dongxiao Zhang +4
cs.LGcs.AIarXiv:2012.06148v22020SceneGen: Learning to Generate Realistic Traffic Scenes
Shuhan Tan, Kelvin Wong, Shenlong Wang +3
cs.CVcs.AIcs.LGarXiv:2101.06541v12021OmniVinci: Enhancing Architecture and Data for Omni-Modal Understanding LLM
Hanrong Ye, Chao-Han Huck Yang, Arushi Goel +29
cs.CVcs.AIcs.CLarXiv:2510.15870v22025PEORL: Integrating Symbolic Planning and Hierarchical Reinforcement Learning for Robust Decision-Making
Fangkai Yang, Daoming Lyu, Bo Liu +1
cs.LGcs.AIstat.MLarXiv:1804.07779v32018T1: Terminal Agent Reinforcement Learning for Long-Horizon Tasks
Junyao Yang, Yucheng Shi, Zhongzhi Li +4
cs.LGcs.AIarXiv:2609.11042v12026Autonomous UAV Exploration of Dynamic Environments via Incremental Sampling and Probabilistic Roadmap
Zhefan Xu, Di Deng, Kenji Shimada
cs.ROcs.AIarXiv:2010.07429v32020Customizable Contraction Hierarchies
Julian Dibbelt, Ben Strasser, Dorothea Wagner
cs.DScs.AIarXiv:1402.0402v52014Expert Level control of Ramp Metering based on Multi-task Deep Reinforcement Learning
Francois Belletti, Daniel Haziza, Gabriel Gomes +1
cs.AIarXiv:1701.08832v12017The Vadalog System: Datalog-based Reasoning for Knowledge Graphs
Luigi Bellomarini, Georg Gottlob, Emanuel Sallinger
cs.DBcs.AIarXiv:1807.08709v12018LLM Agents can Autonomously Hack Websites
Richard Fang, Rohan Bindu, Akul Gupta +2
cs.CRcs.AIarXiv:2402.06664v32024O1 Replication Journey: A Strategic Progress Report -- Part 1
Yiwei Qin, Xuefeng Li, Haoyang Zou +8
cs.AIcs.CLarXiv:2410.18982v12024LoMar: A Local Defense Against Poisoning Attack on Federated Learning
Xingyu Li, Zhe Qu, Shangqing Zhao +3
cs.LGcs.AIcs.CRarXiv:2201.02873v12022CITRIS: Causal Identifiability from Temporal Intervened Sequences
Phillip Lippe, Sara Magliacane, Sindy Löwe +3
cs.LGcs.AIstat.MEarXiv:2202.03169v32022Reliable Conflictive Multi-View Learning
Cai Xu, Jiajun Si, Ziyu Guan +3
cs.LGcs.AIarXiv:2402.16897v22024Understanding Optical Music Recognition
Jorge Calvo-Zaragoza, Jan Hajič, Alexander Pacha
cs.CVcs.AIcs.IRarXiv:1908.03608v32019Fairlearn: Assessing and Improving Fairness of AI Systems
Hilde Weerts, Miroslav Dudík, Richard Edgar +3
cs.LGcs.AIcs.CYarXiv:2303.16626v12023A Comprehensive Evaluation of Large Language Models on Benchmark Biomedical Text Processing Tasks
Israt Jahan, Md Tahmid Rahman Laskar, Chun Peng +1
cs.CLcs.AIarXiv:2310.04270v32023BetterBench: Assessing AI Benchmarks, Uncovering Issues, and Establishing Best Practices
Anka Reuel, Amelia Hardy, Chandler Smith +3
cs.AIcs.LGarXiv:2411.12990v12024Scaling Laws for Reward Model Overoptimization in Direct Alignment Algorithms
Rafael Rafailov, Yaswanth Chittepu, Ryan Park +5
cs.LGcs.AIcs.CLarXiv:2406.02900v22024Generating SOAP Notes from Doctor-Patient Conversations Using Modular Summarization Techniques
Kundan Krishna, Sopan Khosla, Jeffrey P. Bigham +1
cs.CLcs.AIcs.LGarXiv:2005.01795v32020Automatic Sleep Staging of EEG Signals: Recent Development, Challenges, and Future Directions
Huy Phan, Kaare Mikkelsen
eess.SPcs.AIcs.LGarXiv:2111.08446v32021Improving Instruction-Following in Language Models through Activation Steering
Alessandro Stolfo, Vidhisha Balachandran, Safoora Yousefi +2
cs.CLcs.AIcs.LGarXiv:2410.12877v22024Ologs: a categorical framework for knowledge representation
David I. Spivak, Robert E. Kent
cs.LOcs.AImath.CTarXiv:1102.1889v22011Q-attention: Enabling Efficient Learning for Vision-based Robotic Manipulation
Stephen James, Andrew J. Davison
cs.ROcs.AIcs.CVarXiv:2105.14829v22021Infini-gram: Scaling Unbounded n-gram Language Models to a Trillion Tokens
Jiacheng Liu, Sewon Min, Luke Zettlemoyer +2
cs.CLcs.AIcs.IRarXiv:2401.17377v42024Large-scale Interactive Recommendation with Tree-structured Policy Gradient
Haokun Chen, Xinyi Dai, Han Cai +5
cs.LGcs.AIstat.MLarXiv:1811.05869v12018Building competitive direct acoustics-to-word models for English conversational speech recognition
Kartik Audhkhasi, Brian Kingsbury, Bhuvana Ramabhadran +2
cs.CLcs.AIcs.NEarXiv:1712.03133v12017A is for Absorption: Studying Feature Splitting and Absorption in Sparse Autoencoders
David Chanin, James Wilken-Smith, Tomáš Dulka +3
cs.CLcs.AIarXiv:2409.14507v62024HomeRobot: Open-Vocabulary Mobile Manipulation
Sriram Yenamandra, Arun Ramachandran, Karmesh Yadav +15
cs.ROcs.AIcs.CVarXiv:2306.11565v22023Embedding Uncertain Knowledge Graphs
Xuelu Chen, Muhao Chen, Weijia Shi +2
cs.AIcs.CLarXiv:1811.10667v22018Suphx: Mastering Mahjong with Deep Reinforcement Learning
Junjie Li, Sotetsu Koyamada, Qiwei Ye +7
cs.AIarXiv:2003.13590v22020The Text Anonymization Benchmark (TAB): A Dedicated Corpus and Evaluation Framework for Text Anonymization
Ildikó Pilán, Pierre Lison, Lilja Øvrelid +3
cs.CLcs.AIarXiv:2202.00443v220226D Rotation Representation For Unconstrained Head Pose Estimation
Thorsten Hempel, Ahmed A. Abdelrahman, Ayoub Al-Hamadi
cs.CVcs.AIcs.LGarXiv:2202.12555v22022VectorFusion: Text-to-SVG by Abstracting Pixel-Based Diffusion Models
Ajay Jain, Amber Xie, Pieter Abbeel
cs.CVcs.AIcs.GRarXiv:2211.11319v12022Prodigy: An Expeditiously Adaptive Parameter-Free Learner
Konstantin Mishchenko, Aaron Defazio
cs.LGcs.AImath.OCarXiv:2306.06101v42023Hierarchical Foresight: Self-Supervised Learning of Long-Horizon Tasks via Visual Subgoal Generation
Suraj Nair, Chelsea Finn
cs.LGcs.AIcs.CVarXiv:1909.05829v12019Towards Accountable AI: Hybrid Human-Machine Analyses for Characterizing System Failure
Besmira Nushi, Ece Kamar, Eric Horvitz
cs.LGcs.AIcs.HCarXiv:1809.07424v12018Investigating the Catastrophic Forgetting in Multimodal Large Language Models
Yuexiang Zhai, Shengbang Tong, Xiao Li +4
cs.CLcs.AIcs.LGarXiv:2309.10313v42023Pervasive AI for IoT applications: A Survey on Resource-efficient Distributed Artificial Intelligence
Emna Baccour, Naram Mhaisen, Alaa Awad Abdellatif +4
cs.DCcs.AIarXiv:2105.01798v22021Modeling Human Driving Behavior through Generative Adversarial Imitation Learning
Raunak Bhattacharyya, Blake Wulfe, Derek Phillips +4
cs.AIarXiv:2006.06412v22020Text and Patterns: For Effective Chain of Thought, It Takes Two to Tango
Aman Madaan, Amir Yazdanbakhsh
cs.CLcs.AIcs.LGarXiv:2209.07686v22022Improve Vision Language Model Chain-of-thought Reasoning
Ruohong Zhang, Bowen Zhang, Yanghao Li +6
cs.AIcs.CVarXiv:2410.16198v12024