Information Retrieval
Papers filed under cs.IR on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,441 to 1,500 of 1,639
OfficeQA Pro: An Enterprise Benchmark for End-to-End Grounded Reasoning
Krista Opsahl-Ong, Arnav Singhvi, Jasmine Collins +10
cs.AIcs.CLcs.IRarXiv:2603.08655v12026A Survey on RAG Meeting LLMs: Towards Retrieval-Augmented Large Language Models
Wenqi Fan, Yujuan Ding, Liangbo Ning +5
cs.CLcs.AIcs.IRarXiv:2405.06211v32024MiroThinker-1.7 & H1: Towards Heavy-Duty Research Agents via Verification
MiroMind Team, S. Bai, L. Bing +41
cs.CLcs.AIcs.IRarXiv:2603.15726v12026In-Context Retrieval-Augmented Language Models
Ori Ram, Yoav Levine, Itay Dalmedigos +4
cs.CLcs.IRarXiv:2302.00083v32023KoViDoRe: Korean Visual Document Retrieval
Yongbin Choi, Yongwoo Song, Mujeen Sung
cs.IRcs.CVarXiv:2608.20840v12026UniMixer: A Unified Architecture for Scaling Laws in Recommendation Systems
Mingming Ha, Guanchen Wang, Linxun Chen +9
cs.IRcs.AIarXiv:2604.00590v22026SkillX: Automatically Constructing Skill Knowledge Bases for Agents
Chenxi Wang, Zhuoyun Yu, Xin Xie +8
cs.CLcs.AIcs.IRarXiv:2604.04804v22026LightThinker++: From Reasoning Compression to Memory Management
Yuqi Zhu, Jintian Zhang, Zhenjie Wan +7
cs.CLcs.AIcs.IRarXiv:2604.03679v12026SuperLocalMemory V3.3: The Living Brain -- Biologically-Inspired Forgetting, Cognitive Quantization, and Multi-Channel Retrieval for Zero-LLM Agent Memory Systems
Varun Pratap Bhardwaj
cs.AIcs.CLcs.IRarXiv:2604.04514v12026CUE-R: Beyond the Final Answer in Retrieval-Augmented Generation
Siddharth Jain, Venkat Narayan Vedam
cs.IRcs.CLcs.LGarXiv:2604.05467v12026Beyond Hard Negatives: The Importance of Score Distribution in Knowledge Distillation for Dense Retrieval
Youngjoon Jang, Seongtae Hong, Hyeonseok Moon +1
cs.IRarXiv:2604.04734v22026Improving Semantic Proximity in Information Retrieval through Cross-Lingual Alignment
Seongtae Hong, Youngjoon Jang, Jungseob Lee +2
cs.IRarXiv:2604.05684v12026ATANT: An Evaluation Framework for AI Continuity
Samuel Sameer Tanguturi
cs.AIcs.IRarXiv:2604.06710v22026BMdataset: A Musicologically Curated LilyPond Dataset
Matteo Spanio, Ilay Guler, Antonio Rodà
cs.SDcs.CLcs.IRarXiv:2604.10628v22026PersonalAI: A Systematic Comparison of Knowledge Graph Storage and Retrieval Approaches for Personalized LLM agents
Mikhail Menschikov, Dmitry Evseev, Victoria Dochkina +5
cs.CLcs.IRarXiv:2506.17001v62025Don't Retrieve, Navigate: Distilling Enterprise Knowledge into Navigable Agent Skills for QA and RAG
Yiqun Sun, Pengfei Wei, Lawrence B. Hsieh
cs.IRcs.AIcs.CLarXiv:2604.14572v32026Code-Switching Information Retrieval: Benchmarks, Analysis, and the Limits of Current Retrievers
Qingcheng Zeng, Yuheng Lu, Zeqi Zhou +6
cs.IRarXiv:2604.17632v12026MathNet: a Global Multimodal Benchmark for Mathematical Reasoning and Retrieval
Shaden Alshammari, Kevin Wen, Abrar Zainal +5
cs.AIcs.DLcs.IRarXiv:2604.18584v22026Dual-View Training for Instruction-Following Information Retrieval
Qingcheng Zeng, Puxuan Yu, Aman Mehta +2
cs.IRarXiv:2604.18845v12026Daedalus-150M: A Convolution-Attention Hybrid Designed for CPU Inference
Christos Koutsiaris
cs.IRcs.AIcs.CLarXiv:2608.20210v12026LoopCTR: Unlocking the Loop Scaling Power for Click-Through Rate Prediction
Jiakai Tang, Runfeng Zhang, Weiqiu Wang +7
cs.IRarXiv:2604.19550v12026From a Static Multi-Level Small Semantic Codebook to a Dynamic Single-Level Large Semantic Codebook for Generative Recommendation
Tianlu Xie, Xin Ku, Mingjie Sun +8
cs.IRcs.LGarXiv:2608.21012v12026DR-Venus: Towards Frontier Edge-Scale Deep Research Agents with Only 10K Open Data
Venus Team, Sunhao Dai, Yong Deng +10
cs.LGcs.AIcs.CLarXiv:2604.19859v12026Explainable Disentangled Representation Learning for Generalizable Authorship Attribution in the Era of Generative AI
Hieu Man, Van-Cuong Pham, Nghia Trung Ngo +2
cs.CLcs.IRcs.LGarXiv:2604.21300v12026AgentSearchBench: A Benchmark for AI Agent Search in the Wild
Bin Wu, Arastun Mammadli, Xiaoyu Zhang +1
cs.AIcs.IRcs.MAarXiv:2604.22436v12026EnSI-RAG: Entity-Structure-Indexed Retrieval-Augmented Generation for Long-Document Question Answering
Xuanyu Meng, Jiashuo Sun, Jash Rajesh Parekh +1
cs.CLcs.AIcs.DBarXiv:2608.21252v12026Adapting Knowledge Graphs for Behavior Denoising in Sequential Recommendation
Zichun Jin, Zihan Zhou, Yinan Liu +2
cs.IRcs.AIarXiv:2608.21243v12026Trustworthy RAG: An Evaluation Agent for Detecting Misinformation and Knowledge Poisoning in Generative AI Systems
Balkrishna Giri, Md Toufique Hasan, Jussi Rasku +2
cs.SEcs.AIcs.CLarXiv:2608.21095v12026Profiling What Matters: Context-Aware Item Profiles from Large-Scale Metadata for LLM Recommenders
Dojun Hwang, Seunghan Lee, Cheonyoung Park +2
cs.IRcs.AIcs.CLarXiv:2608.20801v12026One Hierarchy, Two Systems: Semantic Product IDs for Discovery-Surface Ranking and Search-Page Query Reformulation
Steven Xu, Sanjyot Thete, Saathvik Dirisala +7
cs.IRcs.AIarXiv:2608.20640v12026Edge-Based Agentic Retrieval-Augmented Generation for Autonomous FHWA Bridge Inspection Compliance
Viraj Nishesh Darji, Hemaliben Rakeshkumar Darji
cs.IRcs.AIcs.MAarXiv:2608.20372v12026Clarify-Then-Search: A Clarification Benchmark for Deep Search with End-to-End Nugget Restoration
Deqiang Huang, Jingbo Zhou, Xinjiang Lu +3
cs.IRcs.AIarXiv:2608.20357v12026AutoInt: Automatic Feature Interaction Learning via Self-Attentive Neural Networks
Weiping Song, Chence Shi, Zhiping Xiao +4
cs.IRcs.AIcs.LGarXiv:1810.11921v22018Enhancing LLMs in Predictive Political QA with Semi-Structured Data
Yinan Liu, Zihan Zhou, Zichun Jin +3
cs.AIcs.CLcs.IRarXiv:2608.21218v12026RAG Deserves an Index: Why Ingest-Time Compilation Beats Query-Time Interpretation
Kyle Wild, Yusuke Takahashi, Asako Uraki
cs.AIcs.DBcs.IRarXiv:2608.20845v12026Structure for Reading, Prose for Writing: Asymmetric Structural Conditioning in Multi-Agent Document Authoring
Cheng Yu, Nikhil Mathew, Zhengjie Wang
cs.AIcs.IRarXiv:2608.20786v12026Auditable by Construction: An Ontology-Driven Framework for Trustworthy LLM Analytics in Enterprise Finance
Sergiy Lunyakin
cs.AIcs.CEcs.CLarXiv:2608.20661v12026Joint Deep Modeling of Users and Items Using Reviews for Recommendation
Lei Zheng, Vahid Noroozi, Philip S. Yu
cs.LGcs.IRarXiv:1701.04783v12017Towards Faithful Simulation of Human Shopping Behavior
Jiakai Tang, Yan Mi, Jing Yu +9
cs.IRarXiv:2608.20707v12026Knowledge Graph Convolutional Networks for Recommender Systems
Hongwei Wang, Miao Zhao, Xing Xie +2
cs.IRcs.LGstat.MLarXiv:1904.12575v12019Improving Robustness of Tabular Retrieval via Representational Stability
Kushal Raj Bhandari, Adarsh Singh, Jianxi Gao +2
cs.CLcs.AIcs.IRarXiv:2604.24040v22026Graph Engineering in the Era of LLM Agents: From Individual Intelligence to System Intelligence
Yuyuan Feng, Zhishang Xiang, Chaobin Yang +32
cs.IRcs.AIcs.ETarXiv:2608.21156v12026S^3-Rec: Self-Supervised Learning for Sequential Recommendation with Mutual Information Maximization
Kun Zhou, Hui Wang, Wayne Xin Zhao +5
cs.IRcs.LGarXiv:2008.07873v12020Large Language Models are not Fair Evaluators
Peiyi Wang, Lei Li, Liang Chen +7
cs.CLcs.AIcs.IRarXiv:2305.17926v22023FASH-iCNN: Making Editorial Fashion Identity Inspectable Through Multimodal CNN Probing
Morayo Danielle Adeyemi, Ryan A. Rossi, Franck Dernoncourt
cs.CVcs.HCcs.IRarXiv:2604.26186v12026VBPR: Visual Bayesian Personalized Ranking from Implicit Feedback
Ruining He, Julian McAuley
cs.IRcs.AIarXiv:1510.01784v12015RippleNet: Propagating User Preferences on the Knowledge Graph for Recommender Systems
Hongwei Wang, Fuzheng Zhang, Jialin Wang +4
cs.IRcs.LGstat.MLarXiv:1803.03467v42018xDeepFM: Combining Explicit and Implicit Feature Interactions for Recommender Systems
Jianxun Lian, Xiaohuan Zhou, Fuzheng Zhang +3
cs.LGcs.IRarXiv:1803.05170v32018Deep Learning for Hate Speech Detection in Tweets
Pinkesh Badjatiya, Shashank Gupta, Manish Gupta +1
cs.CLcs.IRarXiv:1706.00188v12017TabEmbed: Benchmarking and Learning Generalist Embeddings for Tabular Understanding
Minjie Qiang, Mingming Zhang, Xiaoyi Bao +5
cs.CLcs.IRarXiv:2605.04962v12026Deep Learning based Recommender System: A Survey and New Perspectives
Shuai Zhang, Lina Yao, Aixin Sun +1
cs.IRarXiv:1707.07435v72017Deep Interest Evolution Network for Click-Through Rate Prediction
Guorui Zhou, Na Mou, Ying Fan +5
stat.MLcs.IRcs.LGarXiv:1809.03672v52018DiffRetriever: Parallel Representative Tokens for Retrieval with Diffusion Language Models
Shuai Wang, Yu Yin, Shengyao Zhuang +2
cs.IRcs.CLarXiv:2605.07210v22026Determinantal point processes for machine learning
Alex Kulesza, Ben Taskar
stat.MLcs.IRcs.LGarXiv:1207.6083v42012The highD Dataset: A Drone Dataset of Naturalistic Vehicle Trajectories on German Highways for Validation of Highly Automated Driving Systems
Robert Krajewski, Julian Bock, Laurent Kloeker +1
cs.CVcs.AIcs.IRarXiv:1810.05642v12018A new ANEW: Evaluation of a word list for sentiment analysis in microblogs
Finn Årup Nielsen
cs.IRcs.CLarXiv:1103.2903v12011Urban-ImageNet: A Large-Scale Multi-Modal Dataset and Evaluation Framework for Urban Space Perception
Yiwei Ou, Chung Ching Cheung, Jun Yang Ang +5
cs.CVcs.IRcs.LGarXiv:2605.09936v12026Code-Guided Reasoning for Small Language Models: Evaluating Executable MCQA Scaffolds
Prateek Biswas, Dhaval Patel, Vedant Khandelwal +2
cs.IRcs.LGcs.PLarXiv:2605.18827v12026Towards Recursive Self-Evolving Agentic Literature Retrieval
Yuwen Du, Tian Jin, Jing Kang +8
cs.IRarXiv:2605.14306v32026Sketch-based Manga Retrieval using Manga109 Dataset
Yusuke Matsui, Kota Ito, Yuji Aramaki +2
cs.CVcs.IRcs.MMarXiv:1510.04389v12015