Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
11,041 to 11,100 of 15,345
Multi-Goal Reinforcement Learning: Challenging Robotics Environments and Request for Research
Matthias Plappert, Marcin Andrychowicz, Alex Ray +9
cs.LGcs.AIcs.ROarXiv:1802.09464v22018Granite.Trust Policy Tools: Shareable, Actionable Policies for Generative AI Applications
Nathalie Baracaldo, Nicolas Mello, Kush R. Varshney +3
cs.AIarXiv:2608.23870v12026DAIR-V2X: A Large-Scale Dataset for Vehicle-Infrastructure Cooperative 3D Object Detection
Haibao Yu, Yizhen Luo, Mao Shu +8
cs.CVcs.AIarXiv:2204.05575v12022All are Worth Words: A ViT Backbone for Diffusion Models
Fan Bao, Shen Nie, Kaiwen Xue +4
cs.CVcs.AIcs.LGarXiv:2209.12152v42022Parameter-Efficient Self-Supervised Adaptation for EEG-FM under Fixed Computational Budgets
Meghal Dani, Stefanie Liebe
cs.LGcs.AIarXiv:2608.24727v12026Across the Loss Landscape with Progressive Growth
Paul Caillon, Christophe Cerisara, Alexandre Allauzen
cs.LGcs.AIarXiv:2608.24568v12026LumiXAI: A Modular Full-Stack Framework for Feature Attribution
Alfio Ferrara, Lorenzo Gatta, Sergio Picascia +1
cs.SEcs.AIarXiv:2608.24524v12026SonarLLM: A Native Sonar--Optical Multimodal Large Language Model for Underwater Perception
Cong Su, longxuan ma, Ling Dong +4
cs.AIarXiv:2608.24325v12026VisCache: Visual KV Cache Pruning for Efficient Vision Large Language Model Inference
Lyuke Wang, Zhuo Li, Guangxu Zhu
cs.CVcs.AIarXiv:2608.24063v12026OmniJudge or OmniBias? Diagnosing Multimodal Judges through Balanced, Decoupled Lenses
Guangzheng Hu, Ziyue Jiang, Weixu Qiao +14
cs.AIarXiv:2608.24160v12026Do LLMs Understand Limit Order Book Dynamics?
Junxiao Chen, Paul Glasserman
cs.AIarXiv:2608.23706v12026Beyond the Mandate: A Systematic Security Analysis of the Agent Payments Protocol (AP2)
Avital Aviv, Parth A. Gandh, Ron Bitton +1
cs.CRcs.AIarXiv:2608.23858v12026PoinTr: Diverse Point Cloud Completion with Geometry-Aware Transformers
Xumin Yu, Yongming Rao, Ziyi Wang +3
cs.CVcs.AIcs.LGarXiv:2108.08839v12021A Review of Cooperative Multi-Agent Deep Reinforcement Learning
Afshin OroojlooyJadid, Davood Hajinezhad
cs.LGcs.AIcs.MAarXiv:1908.03963v42019Is ChatGPT a Good NLG Evaluator? A Preliminary Study
Jiaan Wang, Yunlong Liang, Fandong Meng +6
cs.CLcs.AIarXiv:2303.04048v32023MuseGAN: Multi-track Sequential Generative Adversarial Networks for Symbolic Music Generation and Accompaniment
Hao-Wen Dong, Wen-Yi Hsiao, Li-Chia Yang +1
eess.AScs.AIcs.LGarXiv:1709.06298v22017Reinforced Cross-Modal Matching and Self-Supervised Imitation Learning for Vision-Language Navigation
Xin Wang, Qiuyuan Huang, Asli Celikyilmaz +5
cs.CVcs.AIcs.CLarXiv:1811.10092v22018What Guides the Agent? Adjudicating Unauthorized Behavior via Localizing Behavior-Guiding Instructions
Yichao Gao, Yumo Zhang, Yunhao Yao +4
cs.CRcs.AIarXiv:2608.24022v12026From Gradient-Boosted Trees to Deep Recommenders: Practical Lessons from Migrating a Production Customer Support Recommender
Sonia Sharma, Jeyendran Balakrishnan, Shreya Rajpal +3
cs.LGcs.AIarXiv:2608.24132v12026SatMAE: Pre-training Transformers for Temporal and Multi-Spectral Satellite Imagery
Yezhen Cong, Samar Khanna, Chenlin Meng +6
cs.CVcs.AIarXiv:2207.08051v32022Hallucination is Inevitable: An Innate Limitation of Large Language Models
Ziwei Xu, Sanjay Jain, Mohan Kankanhalli
cs.CLcs.AIcs.LGarXiv:2401.11817v22024Textbooks Are All You Need
Suriya Gunasekar, Yi Zhang, Jyoti Aneja +16
cs.CLcs.AIcs.LGarXiv:2306.11644v22023When Seeing Is Not Enough: Benchmarking Interactive Visual Grounding in LVLMs
Zhengxiang Wang, Owen Rambow
cs.AIcs.CVarXiv:2608.23978v12026Using an LLM to Help With Code Understanding
Daye Nam, Andrew Macvean, Vincent Hellendoorn +2
cs.SEcs.AIcs.HCarXiv:2307.08177v32023Robust Code RL via Faulty-Code-Driven Test case Synthesis and Dense Reward Shaping
Yiwen Zhang, Xiaodong Yan, Zhenyu Huang +6
cs.AIcs.SEarXiv:2608.24135v12026OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models
Anas Awadalla, Irena Gao, Josh Gardner +13
cs.CVcs.AIcs.LGarXiv:2308.01390v22023When Do Supervised UQ Ensembles Improve LLM Hallucination Detection? A Robustness Study
Mohit Singh Chauhan, Vipin Gyanchandani, Dylan Bouchard
cs.LGcs.AIcs.CLarXiv:2608.24492v12026Not All Tokens Are Equal: Region-Aware Consistency Repair of Backdoors in MLLMs
Jiali Wei, Ming Fan, Mingkun Zhang +6
cs.CRcs.AIcs.CLarXiv:2608.24354v12026OPDSearch+: On-Policy Distillation with RL Refinement for Search-Augmented Reasoning
Qinglin Ye, Zhiyuan Gu, Jingjie Xia +6
cs.AIarXiv:2608.24310v12026PeakBench: Benchmarking Resource-Aware Tool Invocation in LLM Agents
Zhi-Kai Chen, Xu-Xiang Zhong, Song-Yan Li +2
cs.AIcs.SEarXiv:2608.24509v12026Generating Visual Explanations
Lisa Anne Hendricks, Zeynep Akata, Marcus Rohrbach +3
cs.CVcs.AIcs.CLarXiv:1603.08507v12016Scalable and Versatile Identification for Hierarchical Structural Causal Models: A New Look at Project STAR
Janis Aiad, Aghiles Drali, Aymen El Ouadrhiri +6
stat.MLcs.AIarXiv:2608.24500v12026GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts
Jiahao Yu, Xingwei Lin, Zheng Yu +1
cs.AIarXiv:2309.10253v42023Adaptive Influence Graphs for Failure Attribution in Multi-Agent Systems
Yarden Bakish, Amir Dudai, Roy Ganz +4
cs.AIarXiv:2608.24361v12026Learning Synergies between Pushing and Grasping with Self-supervised Deep Reinforcement Learning
Andy Zeng, Shuran Song, Stefan Welker +3
cs.ROcs.AIcs.CVarXiv:1803.09956v32018PhysMLLMs: Spatial Priors for Unified Referring Segmentation and Grounded Reasoning of Images and Videos
Siyao Yan, Bo Han, Jisheng Dang +7
cs.AIarXiv:2608.24574v12026Video Classification with Channel-Separated Convolutional Networks
Du Tran, Heng Wang, Lorenzo Torresani +1
cs.CVcs.AIarXiv:1904.02811v42019Language Models are Multilingual Chain-of-Thought Reasoners
Freda Shi, Mirac Suzgun, Markus Freitag +9
cs.CLcs.AIcs.LGarXiv:2210.03057v12022Contrastive Branch Policy Optimization
Ying Wang, Changlin Qiu, Bang Lin +4
cs.LGcs.AIarXiv:2608.24300v12026SENSESHIFT: Continuous Sentiment-Controlled Text Generation via Encoder-based Mask Infilling
Shahed Masoudian, Markus Frohmann, Emmanouil Karystinaios +2
cs.CLcs.AIarXiv:2608.24304v12026Reinforcement Learning-Guided Evolutionary Policy Optimization for Preference-Adjustable Heterogeneous Agile Earth Observation Satellite Scheduling
He Wang, Junyu Wu, Hui Li +3
cs.AIarXiv:2608.24470v12026Metadata-Aware Adaptation of a Generative Foundation Model for Conditional CMR Synthesis
Marc Rodríguez, Grzegorz Skorupko, Nay Aung +3
cs.CVcs.AIarXiv:2608.24342v12026Maia 200: A Software Defined Dataflow System for Large-scale AI Acceleration
Sherry Xu, Marco Heddes, Jackson Peng +14
cs.ARcs.AIcs.DCarXiv:2608.24664v12026Recent Advances in Adversarial Training for Adversarial Robustness
Tao Bai, Jinqi Luo, Jun Zhao +2
cs.LGcs.AIcs.CRarXiv:2102.01356v52021Demystifying the Draft EU Artificial Intelligence Act
Michael Veale, Frederik Zuiderveen Borgesius
cs.CYcs.AIarXiv:2107.03721v42021Generating Natural Adversarial Examples
Zhengli Zhao, Dheeru Dua, Sameer Singh
cs.LGcs.AIcs.CLarXiv:1710.11342v22017COCI: Conference Organisers and Content Identifier
Angelo Salatino, Francesco Osborne, Alexis Vizcaino +2
cs.DLcs.AIarXiv:2608.24559v12026DD-PPO: Learning Near-Perfect PointGoal Navigators from 2.5 Billion Frames
Erik Wijmans, Abhishek Kadian, Ari Morcos +5
cs.CVcs.AIcs.LGarXiv:1911.00357v22019Moshi: a speech-text foundation model for real-time dialogue
Alexandre Défossez, Laurent Mazaré, Manu Orsini +5
eess.AScs.AIcs.CLarXiv:2410.00037v22024SPINAL -- Scaling-law and Preference Integration in Neural Alignment Layers
Arion Das, Partha Pratim Saha, Amit Dhanda +3
cs.LGcs.AIcs.CLarXiv:2601.06238v12026Disentangled Skill Representations for Predictive Human Modeling
Mariah Schrum, Deepak Gopinath, Srijan Srivatsa +2
cs.LGcs.AIarXiv:2608.23776v12026AgentRoom: Concurrent Multi-Agent Coding in a CRDT-Backed Shared Workspace
Seonglae Cho, Donghyun Lee
cs.AIcs.SEarXiv:2608.23740v12026The Natural Language Decathlon: Multitask Learning as Question Answering
Bryan McCann, Nitish Shirish Keskar, Caiming Xiong +1
cs.CLcs.AIcs.LGarXiv:1806.08730v12018Gold-YOLO: Efficient Object Detector via Gather-and-Distribute Mechanism
Chengcheng Wang, Wei He, Ying Nie +4
cs.CVcs.AIarXiv:2309.11331v52023Web Retrieval-Aware Chunking (W-RAC) for Efficient and Cost-Effective Retrieval-Augmented Generation Systems
Uday Allu, Sonu Kedia, Tanmay Odapally +1
cs.IRcs.AIarXiv:2604.04936v12026Controllable Memory Usage: Balancing Anchoring and Innovation in Long-Term Human-Agent Interaction
Muzhao Tian, Zisu Huang, Xiaohua Wang +8
cs.AIarXiv:2601.05107v12026On the Fallacy of Global Token Perplexity in Spoken Language Model Evaluation
Chan-Jan Hsu, Liang-Hsuan Tseng, Yi-Cheng Lin +5
cs.CLcs.AIarXiv:2601.06329v22026Data-Efficient Off-Policy Policy Evaluation for Reinforcement Learning
Philip S. Thomas, Emma Brunskill
cs.LGcs.AIarXiv:1604.00923v12016LsrIF: Enhancing Logic-Structured Instruction Following of Large Language Models
Qingyu Ren, Qianyu He, Jingwen Chang +9
cs.AIarXiv:2601.06431v32026Diverse Beam Search: Decoding Diverse Solutions from Neural Sequence Models
Ashwin K Vijayakumar, Michael Cogswell, Ramprasath R. Selvaraju +4
cs.AIcs.CLcs.CVarXiv:1610.02424v22016