Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,581 to 8,640 of 15,235
Rating the Raters: Rasch Measurement Theory for LLM Evaluation
Pratik S. Sachdeva, Nathan Boudol
cs.AIarXiv:2608.27463v12026Context Localization for Generalized Level-Based Evaluation in Knowledge-Based Systems
Ondrej Hutník, Natália Puškárová
cs.AIarXiv:2608.27482v12026Benchmarking General Mobile Assistants in Challenging Real-World Scenarios
Yiqi Zhu, Feiyu Gao, Jiaxing Fan +7
cs.AIcs.CVarXiv:2608.27477v12026Guided Conditional Diffusion for Controllable Traffic Simulation
Ziyuan Zhong, Davis Rempe, Danfei Xu +5
cs.ROcs.AIcs.LGarXiv:2210.17366v12022Hypothesize, Evaluate, Refine: A Scientific Agent for PDE Discovery with Unknown Spatial Coefficient Fields
YuJie Huang, WenWu He, ZhuoEr Lin +3
cs.AIcs.LGarXiv:2608.27475v12026Multi-view Knowledge Graph Embedding for Entity Alignment
Qingheng Zhang, Zequn Sun, Wei Hu +3
cs.AIcs.CLcs.LGarXiv:1906.02390v12019PONI: Potential Functions for ObjectGoal Navigation with Interaction-free Learning
Santhosh Kumar Ramakrishnan, Devendra Singh Chaplot, Ziad Al-Halah +2
cs.CVcs.AIarXiv:2201.10029v22022All in One: Multi-task Prompting for Graph Neural Networks
Xiangguo Sun, Hong Cheng, Jia Li +2
cs.SIcs.AIcs.LGarXiv:2307.01504v22023Table-to-text Generation by Structure-aware Seq2seq Learning
Tianyu Liu, Kexiang Wang, Lei Sha +2
cs.CLcs.AIarXiv:1711.09724v12017LLM-Augmented Causal Discovery: Probabilistic Fusion of Edge Existence and Orientation
Neville K. Kitson, Anthony Constantinou
cs.AIarXiv:2608.27472v12026Time Capsule of Testable Human Knowledge: 41 Years of Jeopardy! in a Single Free Local Model
David Noever, Forrest McKee
cs.AIarXiv:2608.27459v12026A Literate Programming Environment for Human and Machine Agents
Adam T. Burke
cs.SEcs.AIcs.PLarXiv:2608.24644v12026Hate Speech Classification In Roman Urdu: A Comparative Study On Parameter Efficient Fine-Tuning And Prompt Engineering
Toneema Zubair
cs.AIarXiv:2608.21408v12026AcademiClaw: When Students Set Challenges for AI Agents
Junjie Yu, Pengrui Lu, Weiye Si +75
cs.AIcs.CYarXiv:2605.02661v12026OpenResearcher: A Fully Open Pipeline for Long-Horizon Deep Research Trajectory Synthesis
Zhuofeng Li, Dongfu Jiang, Xueguang Ma +7
cs.IRcs.AIcs.CLarXiv:2603.20278v12026daVinci-Env: Open SWE Environment Synthesis at Scale
Dayuan Fu, Shenyu Wu, Yunze Wu +11
cs.SEcs.AIcs.CLarXiv:2603.13023v22026daVinci-Agency: Unlocking Long-Horizon Agency Data-Efficiently
Mohan Jiang, Dayuan Fu, Junhao Shi +8
cs.LGcs.AIcs.SEarXiv:2602.02619v22026TRIP-Bench: A Benchmark for Long-Horizon Interactive Agents in Real-World Scenarios
Yuanzhe Shen, Zisu Huang, Zhengyuan Wang +14
cs.AIcs.LGarXiv:2602.01675v12026AgencyBench: Benchmarking the Frontiers of Autonomous Agents in 1M-Token Real-World Contexts
Keyu Li, Junhao Shi, Yang Xiao +11
cs.AIarXiv:2601.11044v42026Large language models can accurately predict searcher preferences
Paul Thomas, Seth Spielman, Nick Craswell +1
cs.IRcs.AIcs.CLarXiv:2309.10621v32023A Perspective on Explainable Artificial Intelligence Methods: SHAP and LIME
Ahmed Salih, Zahra Raisi-Estabragh, Ilaria Boscolo Galazzo +4
stat.MLcs.AIcs.LGarXiv:2305.02012v32023VoxFormer: Sparse Voxel Transformer for Camera-based 3D Semantic Scene Completion
Yiming Li, Zhiding Yu, Christopher Choy +5
cs.CVcs.AIcs.LGarXiv:2302.12251v22023A Systematic Review of Green AI
Roberto Verdecchia, June Sallou, Luís Cruz
cs.AIarXiv:2301.11047v32023Can Foundation Models Wrangle Your Data?
Avanika Narayan, Ines Chami, Laurel Orr +2
cs.LGcs.AIcs.DBarXiv:2205.09911v22022CirCNN: Accelerating and Compressing Deep Neural Networks Using Block-CirculantWeight Matrices
Caiwen Ding, Siyu Liao, Yanzhi Wang +13
cs.CVcs.AIcs.LGarXiv:1708.08917v12017Rubric-to-Code Credit Assignment for Reinforcement Learning
Rui Jin, Jikai Chen, Yihan Chen +5
cs.AIarXiv:2608.27906v12026Explainable Artificial Intelligence: a Systematic Review
Giulia Vilone, Luca Longo
cs.AIcs.LGarXiv:2006.00093v42020PlotQA: Reasoning over Scientific Plots
Nitesh Methani, Pritha Ganguly, Mitesh M. Khapra +1
cs.CVcs.AIcs.CLarXiv:1909.00997v32019Multi-hop Reading Comprehension through Question Decomposition and Rescoring
Sewon Min, Victor Zhong, Luke Zettlemoyer +1
cs.CLcs.AIarXiv:1906.02916v22019Generalised framework for multi-criteria method selection
Jarosław Wątróbski, Jarosław Jankowski, Paweł Ziemba +2
cs.AIarXiv:1810.11078v12018Video Generative Models as Geometry Learner
Haosen Yang, Jifei Song, Zhensong Zhang +2
cs.CVcs.AIarXiv:2608.28549v12026Successor Features for Transfer in Reinforcement Learning
André Barreto, Will Dabney, Rémi Munos +4
cs.AIarXiv:1606.05312v22016Playing hard exploration games by watching YouTube
Yusuf Aytar, Tobias Pfaff, David Budden +3
cs.LGcs.AIcs.CVarXiv:1805.11592v22018"It's Weird That it Knows What I Want": Usability and Interactions with Copilot for Novice Programmers
James Prather, Brent N. Reeves, Paul Denny +6
cs.HCcs.AIarXiv:2304.02491v12023Open Problems in Cooperative AI
Allan Dafoe, Edward Hughes, Yoram Bachrach +5
cs.AIcs.MAarXiv:2012.08630v12020EdNet: A Large-Scale Hierarchical Dataset in Education
Youngduck Choi, Youngnam Lee, Dongmin Shin +7
cs.CYcs.AIcs.HCarXiv:1912.03072v32019Guiding Pretraining in Reinforcement Learning with Large Language Models
Yuqing Du, Olivia Watkins, Zihan Wang +5
cs.LGcs.AIcs.CLarXiv:2302.06692v22023Language Modeling Is Compression
Grégoire Delétang, Anian Ruoss, Paul-Ambroise Duquenne +9
cs.LGcs.AIcs.CLarXiv:2309.10668v22023End-to-End Model-Free Reinforcement Learning for Urban Driving using Implicit Affordances
Marin Toromanoff, Emilie Wirbel, Fabien Moutarde
cs.LGcs.AIcs.CVarXiv:1911.10868v22019DiffusionBERT: Improving Generative Masked Language Models with Diffusion Models
Zhengfu He, Tianxiang Sun, Kuanning Wang +2
cs.CLcs.AIcs.LGarXiv:2211.15029v22022MDFEND: Multi-domain Fake News Detection
Qiong Nan, Juan Cao, Yongchun Zhu +2
cs.CLcs.AIcs.SIarXiv:2201.00987v12022Bike Flow Prediction with Multi-Graph Convolutional Networks
Di Chai, Leye Wang, Qiang Yang
cs.LGcs.AIstat.MLarXiv:1807.10934v12018Towards Understanding Grokking: An Effective Theory of Representation Learning
Ziming Liu, Ouail Kitouni, Niklas Nolte +3
cs.LGcond-mat.dis-nncond-mat.stat-mecharXiv:2205.10343v22022MagicDrive: Street View Generation with Diverse 3D Geometry Control
Ruiyuan Gao, Kai Chen, Enze Xie +4
cs.CVcs.AIarXiv:2310.02601v72023Double Graph Based Reasoning for Document-level Relation Extraction
Shuang Zeng, Runxin Xu, Baobao Chang +1
cs.CLcs.AIcs.LGarXiv:2009.13752v12020What Can Neural Networks Reason About?
Keyulu Xu, Jingling Li, Mozhi Zhang +3
cs.LGcs.AIcs.CVarXiv:1905.13211v42019A Survey on Intelligent Internet of Things: Applications, Security, Privacy, and Future Directions
Ons Aouedi, Thai-Hoc Vu, Alessio Sacco +4
cs.NIcs.AIcs.CRarXiv:2406.03820v22024Automated Rationale Generation: A Technique for Explainable AI and its Effects on Human Perceptions
Upol Ehsan, Pradyumna Tambwekar, Larry Chan +2
cs.AIcs.HCarXiv:1901.03729v12019LongLoRA: Efficient Fine-tuning of Long-Context Large Language Models
Yukang Chen, Shengju Qian, Haotian Tang +4
cs.CLcs.AIcs.LGarXiv:2309.12307v32023DoWhy: An End-to-End Library for Causal Inference
Amit Sharma, Emre Kiciman
stat.MEcs.AIcs.MSarXiv:2011.04216v12020The future of human-AI collaboration: a taxonomy of design knowledge for hybrid intelligence systems
Dominik Dellermann, Adrian Calma, Nikolaus Lipusch +3
cs.AIcs.HCarXiv:2105.03354v12021The Science of Detecting LLM-Generated Texts
Ruixiang Tang, Yu-Neng Chuang, Xia Hu
cs.CLcs.AIarXiv:2303.07205v32023Zero-Shot Task Generalization with Multi-Task Deep Reinforcement Learning
Junhyuk Oh, Satinder Singh, Honglak Lee +1
cs.AIcs.LGarXiv:1706.05064v22017Deep Bidirectional Language-Knowledge Graph Pretraining
Michihiro Yasunaga, Antoine Bosselut, Hongyu Ren +4
cs.CLcs.AIcs.LGarXiv:2210.09338v22022LVLM-eHub: A Comprehensive Evaluation Benchmark for Large Vision-Language Models
Peng Xu, Wenqi Shao, Kaipeng Zhang +7
cs.CVcs.AIarXiv:2306.09265v12023Large Language Models as Zero-Shot Conversational Recommenders
Zhankui He, Zhouhang Xie, Rahul Jha +6
cs.IRcs.AIarXiv:2308.10053v12023TVQA+: Spatio-Temporal Grounding for Video Question Answering
Jie Lei, Licheng Yu, Tamara L. Berg +1
cs.CVcs.AIcs.CLarXiv:1904.11574v22019MiniCheck: Efficient Fact-Checking of LLMs on Grounding Documents
Liyan Tang, Philippe Laban, Greg Durrett
cs.CLcs.AIarXiv:2404.10774v22024Foundations and Trends in Multimodal Machine Learning: Principles, Challenges, and Open Questions
Paul Pu Liang, Amir Zadeh, Louis-Philippe Morency
cs.LGcs.AIcs.CLarXiv:2209.03430v22022Efficiently Trainable Text-to-Speech System Based on Deep Convolutional Networks with Guided Attention
Hideyuki Tachibana, Katsuya Uenoyama, Shunsuke Aihara
cs.SDcs.AIcs.LGarXiv:1710.08969v22017