Information Retrieval

Papers filed under cs.IR on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,381 to 1,440 of 1,639

  1. TALLRec: An Effective and Efficient Tuning Framework to Align Large Language Model with Recommendation

    Keqin Bao, Jizhi Zhang, Yang Zhang +3

    cs.IRarXiv:2305.00447v32023
  2. Enrich-Retrieve-Rank: Scaling Capability Discovery Beyond In-Context Routing

    Nazib Sorathiya, Daniel Zhang, Bardiya Akhbari

    cs.CLcs.AIcs.IRarXiv:2608.22695v12026
  3. Modeling Online Reviews with Multi-grain Topic Models

    Ivan Titov, Ryan McDonald

    cs.IRcs.DBarXiv:0801.1063v12008
  4. SAGE: Benchmarking and Improving Retrieval for Deep Research Agents

    Tiansheng Hu, Yilun Zhao, Canyu Zhang +2

    cs.IRcs.CLarXiv:2602.05975v22026
  5. CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data

    Guillaume Wenzek, Marie-Anne Lachaux, Alexis Conneau +4

    cs.CLcs.IRcs.LGarXiv:1911.00359v22019
  6. InnoEval: On Research Idea Evaluation as a Knowledge-Grounded, Multi-Perspective Reasoning Problem

    Shuofei Qiao, Yunxiang Wei, Xuehai Wang +10

    cs.CLcs.AIcs.IRarXiv:2602.14367v22026
  7. KILT: a Benchmark for Knowledge Intensive Language Tasks

    Fabio Petroni, Aleksandra Piktus, Angela Fan +10

    cs.CLcs.AIcs.IRarXiv:2009.02252v42020
  8. Document Ranking with a Pretrained Sequence-to-Sequence Model

    Rodrigo Nogueira, Zhiying Jiang, Jimmy Lin

    cs.IRcs.LGarXiv:2003.06713v12020
  9. HyTRec: A Hybrid Temporal-Aware Attention Architecture for Long Behavior Sequential Recommendation

    Lei Xin, Yuhao Zheng, Ke Cheng +3

    cs.IRcs.AIarXiv:2602.18283v12026
  10. Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics

    Ziwen Xu, Chenyan Wu, Hengyu Sun +9

    cs.CLcs.AIcs.CVarXiv:2602.02343v32026
  11. RecGOAT: Graph Optimal Adaptive Transport for LLM-Enhanced Multimodal Recommendation with Dual Semantic Alignment

    Yuecheng Li, Hengwei Ju, Zeyu Song +4

    cs.IRcs.AIarXiv:2602.00682v22026
  12. GISA: A Benchmark for General Information-Seeking Assistant

    Yutao Zhu, Xingshuo Zhang, Maosen Zhang +9

    cs.CLcs.AIcs.IRarXiv:2602.08543v22026
  13. Self-EvolveRec: Self-Evolving Recommender Systems with LLM-based Directional Feedback

    Sein Kim, Sangwu Park, Hongseok Kang +6

    cs.IRcs.AIarXiv:2602.12612v22026
  14. Counterfactual Reasoning and Learning Systems

    Léon Bottou, Jonas Peters, Joaquin Quiñonero-Candela +6

    cs.LGcs.AIcs.IRarXiv:1209.2355v52012
  15. Recommendations as Treatments: Debiasing Learning and Evaluation

    Tobias Schnabel, Adith Swaminathan, Ashudeep Singh +2

    cs.LGcs.AIcs.IRarXiv:1602.05352v22016
  16. Towards a Densing Law for User Representation Learning at Billion-Scale Capacity

    Bin Dou, Junru Zhang, Zhaoyi Yuan +6

    cs.IRcs.AIarXiv:2608.23392v12026
  17. Learning to Retrieve from Agent Trajectories

    Yuqi Zhou, Sunhao Dai, Changle Qu +3

    cs.IRcs.AIcs.CLarXiv:2604.04949v12026
  18. Beyond the Grid: Layout-Informed Multi-Vector Retrieval with Parsed Visual Document Representations

    Yibo Yan, Mingdong Ou, Yi Cao +6

    cs.CLcs.IRarXiv:2603.01666v12026
  19. MemSifter: Offloading LLM Memory Retrieval via Outcome-Driven Proxy Reasoning

    Jiejun Tan, Zhicheng Dou, Liancheng Zhang +3

    cs.IRcs.AIarXiv:2603.03379v22026
  20. DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories

    Chenlong Deng, Mengjie Deng, Junjie Wu +10

    cs.CVcs.IRarXiv:2602.10809v22026
  21. Precise Zero-Shot Dense Retrieval without Relevance Labels

    Luyu Gao, Xueguang Ma, Jimmy Lin +1

    cs.IRcs.CLarXiv:2212.10496v12022
  22. Rethinking Generative Recommender Tokenizer: Recsys-Native Encoding and Semantic Quantization Beyond LLMs

    Yu Liang, Zhongjin Zhang, Yuxuan Zhu +10

    cs.IRcs.AIarXiv:2602.02338v12026
  23. Black-box Generation of Adversarial Text Sequences to Evade Deep Learning Classifiers

    Ji Gao, Jack Lanchantin, Mary Lou Soffa +1

    cs.CLcs.CRcs.IRarXiv:1801.04354v52018
  24. Better Retrieval, Worse Robustness:How Multi-hop RAG Amplifies Upstream ASR Errors

    Zhenghua Bao

    cs.CLcs.IReess.ASarXiv:2608.22872v12026
  25. HIRA: A Human-in-the-Loop Retrieval-Augmented Cascade for Document Classification in Regulated Industries

    Shangxuan Tian, Yanhui Chen, Carlos Queiroz

    cs.AIcs.CVcs.IRarXiv:2608.21792v12026
  26. Agentic Search in the Wild: Intents and Trajectory Dynamics from 14M+ Real Search Requests

    Jingjie Ning, João Coelho, Yibo Kong +5

    cs.IRcs.CLarXiv:2601.17617v32026
  27. PaperSearchQA: Learning to Search and Reason over Scientific Papers with RLVR

    James Burgess, Jan N. Hansen, Duo Peng +5

    cs.LGcs.AIcs.CLarXiv:2601.18207v12026
  28. Topic Modeling in Embedding Spaces

    Adji B. Dieng, Francisco J. R. Ruiz, David M. Blei

    cs.IRcs.CLcs.LGarXiv:1907.04907v12019
  29. Detection and Resolution of Rumours in Social Media: A Survey

    Arkaitz Zubiaga, Ahmet Aker, Kalina Bontcheva +2

    cs.CLcs.HCcs.IRarXiv:1704.00656v32017
  30. PEARL: Personalized Streaming Video Understanding Model

    Yuanhong Zheng, Ruichuan An, Xiaopeng Lin +10

    cs.CVcs.AIcs.IRarXiv:2603.20422v12026
  31. TAPAS: Weakly Supervised Table Parsing via Pre-training

    Jonathan Herzig, Paweł Krzysztof Nowak, Thomas Müller +2

    cs.IRcs.AIcs.CLarXiv:2004.02349v22020
  32. MSA: Memory Sparse Attention for Efficient End-to-End Memory Model Scaling to 100M Tokens

    Yu Chen, Runkai Chen, Sheng Yi +9

    cs.CLcs.AIcs.IRarXiv:2603.23516v22026
  33. Texygen: A Benchmarking Platform for Text Generation Models

    Yaoming Zhu, Sidi Lu, Lei Zheng +4

    cs.CLcs.IRcs.LGarXiv:1802.01886v12018
  34. MemoBrain: Executive Memory as an Agentic Brain for Reasoning

    Hongjin Qian, Zhao Cao, Zheng Liu

    cs.AIcs.CLcs.IRarXiv:2601.08079v12026
  35. SemEval-2017 Task 4: Sentiment Analysis in Twitter

    Sara Rosenthal, Noura Farra, Preslav Nakov

    cs.CLcs.IRcs.LGarXiv:1912.00741v12019
  36. Same Agent, Different Answers: A Repeat-Aware Audit of Corpus-Induced Answer Churn in Retrieval-Augmented QA

    Jingjie Ning, Xueqi Li

    cs.IRcs.CLarXiv:2608.22856v12026
  37. Agentic-R: Learning to Retrieve for Agentic Search

    Wenhan Liu, Xinyu Ma, Yutao Zhu +4

    cs.IRcs.CLarXiv:2601.11888v12026
  38. PyOD: A Python Toolbox for Scalable Outlier Detection

    Yue Zhao, Zain Nasrullah, Zheng Li

    cs.LGcs.IRstat.MLarXiv:1901.01588v22019
  39. Long Range Arena: A Benchmark for Efficient Transformers

    Yi Tay, Mostafa Dehghani, Samira Abnar +7

    cs.LGcs.AIcs.CLarXiv:2011.04006v12020
  40. Revisiting Text Ranking in Deep Research

    Chuan Meng, Litu Ou, Sean MacAvaney +1

    cs.IRcs.AIcs.CLarXiv:2602.21456v22026
  41. How Well Does Generative Recommendation Generalize?

    Yijie Ding, Zitian Guo, Jiacheng Li +8

    cs.IRarXiv:2603.19809v12026
  42. KnowMe-Bench: Benchmarking Person Understanding for Lifelong Digital Companions

    Tingyu Wu, Zhisheng Chen, Ziyan Weng +8

    cs.AIcs.IRarXiv:2601.04745v22026
  43. DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines

    Omar Khattab, Arnav Singhvi, Paridhi Maheshwari +10

    cs.CLcs.AIcs.IRarXiv:2310.03714v12023
  44. Bias and Debias in Recommender System: A Survey and Future Directions

    Jiawei Chen, Hande Dong, Xiang Wang +3

    cs.IRarXiv:2010.03240v22020
  45. A Deep Relevance Matching Model for Ad-hoc Retrieval

    Jiafeng Guo, Yixing Fan, Qingyao Ai +1

    cs.IRarXiv:1711.08611v12017
  46. A Survey on Large Language Models for Recommendation

    Likang Wu, Zhi Zheng, Zhaopeng Qiu +9

    cs.IRcs.AIarXiv:2305.19860v52023
  47. Multi-Vector Index Compression in Any Modality

    Hanxiang Qin, Alexander Martin, Rohan Jha +3

    cs.IRcs.CLcs.CVarXiv:2602.21202v12026
  48. Recommendation as Language Processing (RLP): A Unified Pretrain, Personalized Prompt & Predict Paradigm (P5)

    Shijie Geng, Shuchang Liu, Zuohui Fu +2

    cs.IRcs.AIcs.CLarXiv:2203.13366v72022
  49. Deep Learning Recommendation Model for Personalization and Recommendation Systems

    Maxim Naumov, Dheevatsa Mudigere, Hao-Jun Michael Shi +21

    cs.IRcs.LGarXiv:1906.00091v12019
  50. Semantic Search over 9 Million Mathematical Theorems

    Luke Alexander, Eric Leonen, Sophie Szeto +5

    cs.IRcs.AImath.HOarXiv:2602.05216v22026
  51. Vectorizing the Trie: Efficient Constrained Decoding for LLM-based Generative Retrieval on Accelerators

    Zhengyang Su, Isay Katsman, Yueqi Wang +10

    cs.IRcs.CLcs.LGarXiv:2602.22647v22026
  52. Finding statistically significant communities in networks

    Andrea Lancichinetti, Filippo Radicchi, Jose' Javier Ramasco +1

    physics.soc-phcs.IRcs.SIarXiv:1012.2363v22010
  53. OpenDecoder: Open Large Language Model Decoding to Incorporate Document Quality in RAG

    Fengran Mo, Zhan Su, Yuchen Hui +6

    cs.CLcs.AIcs.IRarXiv:2601.09028v22026
  54. A Survey on Knowledge Graph-Based Recommender Systems

    Qingyu Guo, Fuzhen Zhuang, Chuan Qin +4

    cs.IRcs.LGstat.MLarXiv:2003.00911v12020
  55. Discrimination in Online Ad Delivery

    Latanya Sweeney

    cs.IRcs.CYarXiv:1301.6822v12013
  56. Are Graph Augmentations Necessary? Simple Graph Contrastive Learning for Recommendation

    Junliang Yu, Hongzhi Yin, Xin Xia +3

    cs.IRarXiv:2112.08679v42021
  57. MTEB: Massive Text Embedding Benchmark

    Niklas Muennighoff, Nouamane Tazi, Loïc Magne +1

    cs.CLcs.IRcs.LGarXiv:2210.07316v32022
  58. LaSER: Internalizing Explicit Reasoning into Latent Space for Dense Retrieval

    Jiajie Jin, Yanzhao Zhang, Mingxin Li +4

    cs.CLcs.IRarXiv:2603.01425v12026
  59. $τ$-Knowledge: Evaluating Conversational Agents over Unstructured Knowledge

    Quan Shi, Alexandra Zytek, Pedram Razavi +2

    cs.AIcs.CLcs.IRarXiv:2603.04370v12026
  60. Fast Matrix Factorization for Online Recommendation with Implicit Feedback

    Xiangnan He, Hanwang Zhang, Min-Yen Kan +1

    cs.IRarXiv:1708.05024v12017