Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
7,861 to 7,920 of 15,237
AdaPlanner: Adaptive Planning from Feedback with Language Models
Haotian Sun, Yuchen Zhuang, Lingkai Kong +2
cs.CLcs.AIcs.LGarXiv:2305.16653v12023BnB-ADOPT: An Asynchronous Branch-and-Bound DCOP Algorithm
William Yeoh, Ariel Felner, Sven Koenig
cs.AIarXiv:1401.3490v12014Graph Learning based Recommender Systems: A Review
Shoujin Wang, Liang Hu, Yan Wang +6
cs.IRcs.AIcs.LGarXiv:2105.06339v12021Selective-Supervised Contrastive Learning with Noisy Labels
Shikun Li, Xiaobo Xia, Shiming Ge +1
cs.CVcs.AIcs.LGarXiv:2203.04181v12022Generating Fact Checking Explanations
Pepa Atanasova, Jakob Grue Simonsen, Christina Lioma +1
cs.CLcs.AIcs.LGarXiv:2004.05773v12020Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming
Zhifei Xie, Changqiao Wu
cs.AIcs.CLcs.HCarXiv:2408.16725v32024Beyond Memorization: Violating Privacy Via Inference with Large Language Models
Robin Staab, Mark Vero, Mislav Balunović +1
cs.AIcs.LGarXiv:2310.07298v22023R$^2$A: Learning Persona Policies Through Persona Representation Learning and Runtime Alignment
Mohan Zhang, Chengsong You, Xiaoyu Cao +4
cs.CLcs.AIarXiv:2608.29798v12026AGIQA-3K: An Open Database for AI-Generated Image Quality Assessment
Chunyi Li, Zicheng Zhang, Haoning Wu +5
cs.CVcs.AIeess.IVarXiv:2306.04717v22023Scaling Up Dataset Distillation to ImageNet-1K with Constant Memory
Justin Cui, Ruochen Wang, Si Si +1
cs.CVcs.AIarXiv:2211.10586v42022TriDet: Temporal Action Detection with Relative Boundary Modeling
Dingfeng Shi, Yujie Zhong, Qiong Cao +3
cs.CVcs.AIcs.MMarXiv:2303.07347v22023A Review of Large Language Models and Autonomous Agents in Chemistry
Mayk Caldas Ramos, Christopher J. Collison, Andrew D. White
cs.LGcs.AIcs.CLarXiv:2407.01603v32024Dialog-based Interactive Image Retrieval
Xiaoxiao Guo, Hui Wu, Yu Cheng +3
cs.CVcs.AIarXiv:1805.00145v32018Cosine Normalization: Using Cosine Similarity Instead of Dot Product in Neural Networks
Chunjie Luo, Jianfeng Zhan, Lei Wang +1
cs.LGcs.AIstat.MLarXiv:1702.05870v52017Learning to Utilize Shaping Rewards: A New Approach of Reward Shaping
Yujing Hu, Weixun Wang, Hangtian Jia +5
cs.LGcs.AIarXiv:2011.02669v12020MI-Distillation: Selecting from Model-Interpolated Instruct-Reasoning Data Spectrum for Chain-of-Thought Distillation
Yangsong Lan, Renkai Hu, HongKai Zheng +4
cs.CLcs.AIarXiv:2608.29623v12026Rearrangement: A Challenge for Embodied AI
Dhruv Batra, Angel X. Chang, Sonia Chernova +9
cs.AIcs.CVcs.LGarXiv:2011.01975v12020Memory-First Fact-Checking: A Knowledge-Graph-Grounded Multi-Agent System for Misinformation Detection
Amelia Petrenciuc, Alexandru Lecu, Adrian Groza
cs.CLcs.AIarXiv:2608.29617v12026Multi-Scale High-Resolution Vision Transformer for Semantic Segmentation
Jiaqi Gu, Hyoukjun Kwon, Dilin Wang +6
cs.CVcs.AIcs.LGarXiv:2111.01236v22021Multi-View Spatial-Temporal Graph Convolutional Networks with Domain Generalization for Sleep Stage Classification
Ziyu Jia, Youfang Lin, Jing Wang +5
eess.SPcs.AIcs.CVarXiv:2109.01824v12021Adversarial Filters of Dataset Biases
Ronan Le Bras, Swabha Swayamdipta, Chandra Bhagavatula +4
cs.LGcs.AIcs.CLarXiv:2002.04108v32020The Emergent Symbolic Structure of Artificial Neural Networks
R. Thomas McCoy, Paul Soulos, Tal Linzen +1
cs.CLcs.AIarXiv:2608.29530v12026Algorithmic Framework for Model-based Deep Reinforcement Learning with Theoretical Guarantees
Yuping Luo, Huazhe Xu, Yuanzhi Li +3
cs.LGcs.AIstat.MLarXiv:1807.03858v52018Learning to cluster in order to transfer across domains and tasks
Yen-Chang Hsu, Zhaoyang Lv, Zsolt Kira
cs.LGcs.AIcs.CVarXiv:1711.10125v32017Progressive Prompts: Continual Learning for Language Models
Anastasia Razdaibiedina, Yuning Mao, Rui Hou +3
cs.CLcs.AIcs.LGarXiv:2301.12314v12023SUP-MIMIC: A Multi-Task Clinical Diagnosis Benchmark for Evaluating LLMs' Robustness to Contradictory Evidence
Yi Yu, Bo Wang, Chong Feng +4
cs.CLcs.AIarXiv:2608.29582v12026Large Language Models Understand and Can be Enhanced by Emotional Stimuli
Cheng Li, Jindong Wang, Yixuan Zhang +6
cs.CLcs.AIcs.HCarXiv:2307.11760v72023Variable Impedance Control in End-Effector Space: An Action Space for Reinforcement Learning in Contact-Rich Tasks
Roberto Martín-Martín, Michelle A. Lee, Rachel Gardner +3
cs.ROcs.AIcs.LGarXiv:1906.08880v22019Evaluating LLMs on Conversational Text-to-SQL under Chain Ambiguity and Intent Drift
Yujia Liu, Jiayan Lin, Zijin Hong +6
cs.CLcs.AIcs.DBarXiv:2608.29543v12026QServe: W4A8KV4 Quantization and System Co-design for Efficient LLM Serving
Yujun Lin, Haotian Tang, Shang Yang +4
cs.CLcs.AIcs.LGarXiv:2405.04532v32024SafeAtlas-VL: Beyond Binary Multimodal Safety with Large-Scale Data and Guard Models
Zongrui Wang, Xiangyang Zhu, Sicheng Wang +13
cs.AIcs.CVarXiv:2608.29098v12026Evaluating the Hidden Costs of Personalization in Large Language Models
Yumeng Wang, Yuchen Wu, Cheng Qian +6
cs.AIarXiv:2608.28833v12026Super Library Agent: Joint Generation and Maintenance of Multiple Applications Beyond the Single Codebase
Daegyu Sung, Yukyeong Lee, Geon Park +2
cs.SEcs.AIcs.CLarXiv:2608.29310v12026A Survey on Recent Advances in LLM-Based Multi-turn Dialogue Systems
Zihao Yi, Jiarui Ouyang, Zhe Xu +4
cs.CLcs.AIarXiv:2402.18013v22024Can ChatGPT Write a Good Boolean Query for Systematic Review Literature Search?
Shuai Wang, Harrisen Scells, Bevan Koopman +1
cs.IRcs.AIarXiv:2302.03495v32023Chain-of-Thought Faithfulness of Reasoning Models Varies with Where and How Preference Cues Are Delivered
Aryo Pradipta Gema, Neel Rajani, Rohit Saxena +2
cs.CLcs.AIarXiv:2608.29464v12026Graph Neural Networks Meet Neural-Symbolic Computing: A Survey and Perspective
Luis C. Lamb, Artur Garcez, Marco Gori +3
cs.AIcs.CLcs.LGarXiv:2003.00330v72020Argument-Aware Semantic Alignment of Normative Texts: A Toulmin-Based Neuro-Symbolic Approach
William Schroeder
cs.CLcs.AIarXiv:2608.29529v12026Optimal Demand Response Using Device Based Reinforcement Learning
Zheng Wen, Daniel O'Neill, Hamid Reza Maei
cs.LGcs.AIeess.SYarXiv:1401.1549v22014Foundation Models for Decision Making: Problems, Methods, and Opportunities
Sherry Yang, Ofir Nachum, Yilun Du +3
cs.AIcs.LGarXiv:2303.04129v12023Flightmare: A Flexible Quadrotor Simulator
Yunlong Song, Selim Naji, Elia Kaufmann +2
cs.ROcs.AIarXiv:2009.00563v22020The Landscape of Emerging AI Agent Architectures for Reasoning, Planning, and Tool Calling: A Survey
Tula Masterman, Sandi Besen, Mason Sawtell +1
cs.AIcs.CLarXiv:2404.11584v12024MUDDLE: Measuring Understanding of Documents under Distractor and Length Effects
Jason Luo, Saibilila Abudukelimu, Judy Song +4
cs.CLcs.AIarXiv:2608.29477v12026Repairing the Cracked Foundation: A Survey of Obstacles in Evaluation Practices for Generated Text
Sebastian Gehrmann, Elizabeth Clark, Thibault Sellam
cs.CLcs.AIcs.LGarXiv:2202.06935v12022AI Can Be Easily Persuaded in Clinical Decision Making
Jiayuan Zhu, Jiazhen Pan, Fenglin Liu +2
cs.CLcs.AIarXiv:2608.29453v12026Logic Tensor Networks for Semantic Image Interpretation
Ivan Donadello, Luciano Serafini, Artur d'Avila Garcez
cs.AIarXiv:1705.08968v12017Language as an Abstraction for Hierarchical Deep Reinforcement Learning
Yiding Jiang, Shixiang Gu, Kevin Murphy +1
cs.LGcs.AIcs.CLarXiv:1906.07343v22019Learning by Abstraction: The Neural State Machine
Drew A. Hudson, Christopher D. Manning
cs.AIcs.CLcs.CVarXiv:1907.03950v42019LightNav-0: Eliciting VLM Spatial Intelligence for Generalist Embodied Navigation
Shaoan Wang, Aocheng Luo, Fei Huang +17
cs.ROcs.AIarXiv:2608.30935v12026Key-Locked Rank One Editing for Text-to-Image Personalization
Yoad Tewel, Rinon Gal, Gal Chechik +1
cs.CVcs.AIcs.GRarXiv:2305.01644v22023FireAct: Toward Language Agent Fine-tuning
Baian Chen, Chang Shu, Ehsan Shareghi +3
cs.CLcs.AIcs.LGarXiv:2310.05915v12023LoftQ: LoRA-Fine-Tuning-Aware Quantization for Large Language Models
Yixiao Li, Yifan Yu, Chen Liang +4
cs.CLcs.AIcs.LGarXiv:2310.08659v42023Arabic Safety Alignment as Selective Refusal: An Empirical Study of SFT, DPO, and Guard Calibration
Mohamad Zbib, Ammar Mohanna
cs.CLcs.AIarXiv:2608.29378v12026MAgent: A Many-Agent Reinforcement Learning Platform for Artificial Collective Intelligence
Lianmin Zheng, Jiacheng Yang, Han Cai +3
cs.LGcs.AIcs.MAarXiv:1712.00600v12017Neural Networks and the Chomsky Hierarchy
Grégoire Delétang, Anian Ruoss, Jordi Grau-Moya +8
cs.LGcs.AIcs.CLarXiv:2207.02098v32022StageWell: A Process-Aligned Chinese Corpus for Positive-Psychology Support Dialogue
Yuxiong Wang, Ziwei Lin, Bo Wang +2
cs.CLcs.AIarXiv:2608.29326v12026Towards Conversational Recommendation over Multi-Type Dialogs
Zeming Liu, Haifeng Wang, Zheng-Yu Niu +3
cs.CLcs.AIarXiv:2005.03954v32020On Feature Decorrelation in Self-Supervised Learning
Tianyu Hua, Wenxiao Wang, Zihui Xue +3
cs.LGcs.AIcs.CVarXiv:2105.00470v22021Maximum-Likelihood Augmented Discrete Generative Adversarial Networks
Tong Che, Yanran Li, Ruixiang Zhang +4
cs.AIcs.CLcs.LGarXiv:1702.07983v12017SantaCoder: don't reach for the stars!
Loubna Ben Allal, Raymond Li, Denis Kocetkov +38
cs.SEcs.AIcs.LGarXiv:2301.03988v22023