Information Retrieval

Papers filed under cs.IR on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,441 to 1,500 of 1,639

  1. OfficeQA Pro: An Enterprise Benchmark for End-to-End Grounded Reasoning

    Krista Opsahl-Ong, Arnav Singhvi, Jasmine Collins +10

    cs.AIcs.CLcs.IRarXiv:2603.08655v12026
  2. A Survey on RAG Meeting LLMs: Towards Retrieval-Augmented Large Language Models

    Wenqi Fan, Yujuan Ding, Liangbo Ning +5

    cs.CLcs.AIcs.IRarXiv:2405.06211v32024
  3. MiroThinker-1.7 & H1: Towards Heavy-Duty Research Agents via Verification

    MiroMind Team, S. Bai, L. Bing +41

    cs.CLcs.AIcs.IRarXiv:2603.15726v12026
  4. In-Context Retrieval-Augmented Language Models

    Ori Ram, Yoav Levine, Itay Dalmedigos +4

    cs.CLcs.IRarXiv:2302.00083v32023
  5. KoViDoRe: Korean Visual Document Retrieval

    Yongbin Choi, Yongwoo Song, Mujeen Sung

    cs.IRcs.CVarXiv:2608.20840v12026
  6. UniMixer: A Unified Architecture for Scaling Laws in Recommendation Systems

    Mingming Ha, Guanchen Wang, Linxun Chen +9

    cs.IRcs.AIarXiv:2604.00590v22026
  7. SkillX: Automatically Constructing Skill Knowledge Bases for Agents

    Chenxi Wang, Zhuoyun Yu, Xin Xie +8

    cs.CLcs.AIcs.IRarXiv:2604.04804v22026
  8. LightThinker++: From Reasoning Compression to Memory Management

    Yuqi Zhu, Jintian Zhang, Zhenjie Wan +7

    cs.CLcs.AIcs.IRarXiv:2604.03679v12026
  9. SuperLocalMemory V3.3: The Living Brain -- Biologically-Inspired Forgetting, Cognitive Quantization, and Multi-Channel Retrieval for Zero-LLM Agent Memory Systems

    Varun Pratap Bhardwaj

    cs.AIcs.CLcs.IRarXiv:2604.04514v12026
  10. CUE-R: Beyond the Final Answer in Retrieval-Augmented Generation

    Siddharth Jain, Venkat Narayan Vedam

    cs.IRcs.CLcs.LGarXiv:2604.05467v12026
  11. Beyond Hard Negatives: The Importance of Score Distribution in Knowledge Distillation for Dense Retrieval

    Youngjoon Jang, Seongtae Hong, Hyeonseok Moon +1

    cs.IRarXiv:2604.04734v22026
  12. Improving Semantic Proximity in Information Retrieval through Cross-Lingual Alignment

    Seongtae Hong, Youngjoon Jang, Jungseob Lee +2

    cs.IRarXiv:2604.05684v12026
  13. ATANT: An Evaluation Framework for AI Continuity

    Samuel Sameer Tanguturi

    cs.AIcs.IRarXiv:2604.06710v22026
  14. BMdataset: A Musicologically Curated LilyPond Dataset

    Matteo Spanio, Ilay Guler, Antonio Rodà

    cs.SDcs.CLcs.IRarXiv:2604.10628v22026
  15. PersonalAI: A Systematic Comparison of Knowledge Graph Storage and Retrieval Approaches for Personalized LLM agents

    Mikhail Menschikov, Dmitry Evseev, Victoria Dochkina +5

    cs.CLcs.IRarXiv:2506.17001v62025
  16. Don't Retrieve, Navigate: Distilling Enterprise Knowledge into Navigable Agent Skills for QA and RAG

    Yiqun Sun, Pengfei Wei, Lawrence B. Hsieh

    cs.IRcs.AIcs.CLarXiv:2604.14572v32026
  17. Code-Switching Information Retrieval: Benchmarks, Analysis, and the Limits of Current Retrievers

    Qingcheng Zeng, Yuheng Lu, Zeqi Zhou +6

    cs.IRarXiv:2604.17632v12026
  18. MathNet: a Global Multimodal Benchmark for Mathematical Reasoning and Retrieval

    Shaden Alshammari, Kevin Wen, Abrar Zainal +5

    cs.AIcs.DLcs.IRarXiv:2604.18584v22026
  19. Dual-View Training for Instruction-Following Information Retrieval

    Qingcheng Zeng, Puxuan Yu, Aman Mehta +2

    cs.IRarXiv:2604.18845v12026
  20. Daedalus-150M: A Convolution-Attention Hybrid Designed for CPU Inference

    Christos Koutsiaris

    cs.IRcs.AIcs.CLarXiv:2608.20210v12026
  21. LoopCTR: Unlocking the Loop Scaling Power for Click-Through Rate Prediction

    Jiakai Tang, Runfeng Zhang, Weiqiu Wang +7

    cs.IRarXiv:2604.19550v12026
  22. From a Static Multi-Level Small Semantic Codebook to a Dynamic Single-Level Large Semantic Codebook for Generative Recommendation

    Tianlu Xie, Xin Ku, Mingjie Sun +8

    cs.IRcs.LGarXiv:2608.21012v12026
  23. DR-Venus: Towards Frontier Edge-Scale Deep Research Agents with Only 10K Open Data

    Venus Team, Sunhao Dai, Yong Deng +10

    cs.LGcs.AIcs.CLarXiv:2604.19859v12026
  24. Explainable Disentangled Representation Learning for Generalizable Authorship Attribution in the Era of Generative AI

    Hieu Man, Van-Cuong Pham, Nghia Trung Ngo +2

    cs.CLcs.IRcs.LGarXiv:2604.21300v12026
  25. AgentSearchBench: A Benchmark for AI Agent Search in the Wild

    Bin Wu, Arastun Mammadli, Xiaoyu Zhang +1

    cs.AIcs.IRcs.MAarXiv:2604.22436v12026
  26. EnSI-RAG: Entity-Structure-Indexed Retrieval-Augmented Generation for Long-Document Question Answering

    Xuanyu Meng, Jiashuo Sun, Jash Rajesh Parekh +1

    cs.CLcs.AIcs.DBarXiv:2608.21252v12026
  27. Adapting Knowledge Graphs for Behavior Denoising in Sequential Recommendation

    Zichun Jin, Zihan Zhou, Yinan Liu +2

    cs.IRcs.AIarXiv:2608.21243v12026
  28. Trustworthy RAG: An Evaluation Agent for Detecting Misinformation and Knowledge Poisoning in Generative AI Systems

    Balkrishna Giri, Md Toufique Hasan, Jussi Rasku +2

    cs.SEcs.AIcs.CLarXiv:2608.21095v12026
  29. Profiling What Matters: Context-Aware Item Profiles from Large-Scale Metadata for LLM Recommenders

    Dojun Hwang, Seunghan Lee, Cheonyoung Park +2

    cs.IRcs.AIcs.CLarXiv:2608.20801v12026
  30. One Hierarchy, Two Systems: Semantic Product IDs for Discovery-Surface Ranking and Search-Page Query Reformulation

    Steven Xu, Sanjyot Thete, Saathvik Dirisala +7

    cs.IRcs.AIarXiv:2608.20640v12026
  31. Edge-Based Agentic Retrieval-Augmented Generation for Autonomous FHWA Bridge Inspection Compliance

    Viraj Nishesh Darji, Hemaliben Rakeshkumar Darji

    cs.IRcs.AIcs.MAarXiv:2608.20372v12026
  32. Clarify-Then-Search: A Clarification Benchmark for Deep Search with End-to-End Nugget Restoration

    Deqiang Huang, Jingbo Zhou, Xinjiang Lu +3

    cs.IRcs.AIarXiv:2608.20357v12026
  33. AutoInt: Automatic Feature Interaction Learning via Self-Attentive Neural Networks

    Weiping Song, Chence Shi, Zhiping Xiao +4

    cs.IRcs.AIcs.LGarXiv:1810.11921v22018
  34. Enhancing LLMs in Predictive Political QA with Semi-Structured Data

    Yinan Liu, Zihan Zhou, Zichun Jin +3

    cs.AIcs.CLcs.IRarXiv:2608.21218v12026
  35. RAG Deserves an Index: Why Ingest-Time Compilation Beats Query-Time Interpretation

    Kyle Wild, Yusuke Takahashi, Asako Uraki

    cs.AIcs.DBcs.IRarXiv:2608.20845v12026
  36. Structure for Reading, Prose for Writing: Asymmetric Structural Conditioning in Multi-Agent Document Authoring

    Cheng Yu, Nikhil Mathew, Zhengjie Wang

    cs.AIcs.IRarXiv:2608.20786v12026
  37. Auditable by Construction: An Ontology-Driven Framework for Trustworthy LLM Analytics in Enterprise Finance

    Sergiy Lunyakin

    cs.AIcs.CEcs.CLarXiv:2608.20661v12026
  38. Joint Deep Modeling of Users and Items Using Reviews for Recommendation

    Lei Zheng, Vahid Noroozi, Philip S. Yu

    cs.LGcs.IRarXiv:1701.04783v12017
  39. Towards Faithful Simulation of Human Shopping Behavior

    Jiakai Tang, Yan Mi, Jing Yu +9

    cs.IRarXiv:2608.20707v12026
  40. Knowledge Graph Convolutional Networks for Recommender Systems

    Hongwei Wang, Miao Zhao, Xing Xie +2

    cs.IRcs.LGstat.MLarXiv:1904.12575v12019
  41. Improving Robustness of Tabular Retrieval via Representational Stability

    Kushal Raj Bhandari, Adarsh Singh, Jianxi Gao +2

    cs.CLcs.AIcs.IRarXiv:2604.24040v22026
  42. Graph Engineering in the Era of LLM Agents: From Individual Intelligence to System Intelligence

    Yuyuan Feng, Zhishang Xiang, Chaobin Yang +32

    cs.IRcs.AIcs.ETarXiv:2608.21156v12026
  43. S^3-Rec: Self-Supervised Learning for Sequential Recommendation with Mutual Information Maximization

    Kun Zhou, Hui Wang, Wayne Xin Zhao +5

    cs.IRcs.LGarXiv:2008.07873v12020
  44. Large Language Models are not Fair Evaluators

    Peiyi Wang, Lei Li, Liang Chen +7

    cs.CLcs.AIcs.IRarXiv:2305.17926v22023
  45. FASH-iCNN: Making Editorial Fashion Identity Inspectable Through Multimodal CNN Probing

    Morayo Danielle Adeyemi, Ryan A. Rossi, Franck Dernoncourt

    cs.CVcs.HCcs.IRarXiv:2604.26186v12026
  46. VBPR: Visual Bayesian Personalized Ranking from Implicit Feedback

    Ruining He, Julian McAuley

    cs.IRcs.AIarXiv:1510.01784v12015
  47. RippleNet: Propagating User Preferences on the Knowledge Graph for Recommender Systems

    Hongwei Wang, Fuzheng Zhang, Jialin Wang +4

    cs.IRcs.LGstat.MLarXiv:1803.03467v42018
  48. xDeepFM: Combining Explicit and Implicit Feature Interactions for Recommender Systems

    Jianxun Lian, Xiaohuan Zhou, Fuzheng Zhang +3

    cs.LGcs.IRarXiv:1803.05170v32018
  49. Deep Learning for Hate Speech Detection in Tweets

    Pinkesh Badjatiya, Shashank Gupta, Manish Gupta +1

    cs.CLcs.IRarXiv:1706.00188v12017
  50. TabEmbed: Benchmarking and Learning Generalist Embeddings for Tabular Understanding

    Minjie Qiang, Mingming Zhang, Xiaoyi Bao +5

    cs.CLcs.IRarXiv:2605.04962v12026
  51. Deep Learning based Recommender System: A Survey and New Perspectives

    Shuai Zhang, Lina Yao, Aixin Sun +1

    cs.IRarXiv:1707.07435v72017
  52. Deep Interest Evolution Network for Click-Through Rate Prediction

    Guorui Zhou, Na Mou, Ying Fan +5

    stat.MLcs.IRcs.LGarXiv:1809.03672v52018
  53. DiffRetriever: Parallel Representative Tokens for Retrieval with Diffusion Language Models

    Shuai Wang, Yu Yin, Shengyao Zhuang +2

    cs.IRcs.CLarXiv:2605.07210v22026
  54. Determinantal point processes for machine learning

    Alex Kulesza, Ben Taskar

    stat.MLcs.IRcs.LGarXiv:1207.6083v42012
  55. The highD Dataset: A Drone Dataset of Naturalistic Vehicle Trajectories on German Highways for Validation of Highly Automated Driving Systems

    Robert Krajewski, Julian Bock, Laurent Kloeker +1

    cs.CVcs.AIcs.IRarXiv:1810.05642v12018
  56. A new ANEW: Evaluation of a word list for sentiment analysis in microblogs

    Finn Årup Nielsen

    cs.IRcs.CLarXiv:1103.2903v12011
  57. Urban-ImageNet: A Large-Scale Multi-Modal Dataset and Evaluation Framework for Urban Space Perception

    Yiwei Ou, Chung Ching Cheung, Jun Yang Ang +5

    cs.CVcs.IRcs.LGarXiv:2605.09936v12026
  58. Code-Guided Reasoning for Small Language Models: Evaluating Executable MCQA Scaffolds

    Prateek Biswas, Dhaval Patel, Vedant Khandelwal +2

    cs.IRcs.LGcs.PLarXiv:2605.18827v12026
  59. Towards Recursive Self-Evolving Agentic Literature Retrieval

    Yuwen Du, Tian Jin, Jing Kang +8

    cs.IRarXiv:2605.14306v32026
  60. Sketch-based Manga Retrieval using Manga109 Dataset

    Yusuke Matsui, Kota Ito, Yuji Aramaki +2

    cs.CVcs.IRcs.MMarXiv:1510.04389v12015