Information Retrieval

Papers filed under cs.IR on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,321 to 1,380 of 1,626

  1. KSE-Web: An Analysis of Hybrid Retrieval and LLM-Assisted Query Expansion for Low-Resource Khmer Semantic Search

    Nimol Thuon

    cs.CLcs.AIcs.IRarXiv:2608.21365v12026
  2. Why This, Not That? Mining User Profiles for Pair-wise Counterfactuals

    Meysam Varasteh, Veronika Bogina, Noam Koenigstein +1

    cs.IRcs.AIarXiv:2608.21662v12026
  3. Personalizing Session-based Recommendations with Hierarchical Recurrent Neural Networks

    Massimo Quadrana, Alexandros Karatzoglou, Balázs Hidasi +1

    cs.LGcs.HCcs.IRarXiv:1706.04148v52017
  4. Rethinking Composed Image Retrieval Evaluation: A Fine-Grained Benchmark from Image Editing

    Tingyu Song, Yanzhao Zhang, Mingxin Li +6

    cs.CVcs.CLcs.IRarXiv:2601.16125v12026
  5. Segment Length Matters: A Study of Segment Lengths on Audio Fingerprinting Performance

    Ziling Gong, Yunyan Ouyang, Iram Kamdar +5

    cs.SDcs.AIcs.IRarXiv:2601.17690v12026
  6. Hierarchical Exponential-Gaussian Mixtures for Watch-Time Distribution Prediction

    Sofia Gulevskaia, Mikhail Trapeznikov, Aleksandr Poslavsky +1

    cs.IRcs.LGstat.MLarXiv:2608.23356v12026
  7. WARP: Wasserstein-Aligned RAG for Population Opinions

    Aman Singh Thakur, Aditya Agrawal, Alwarappan Nakkiran +1

    cs.IRcs.CLarXiv:2608.22859v12026
  8. Clustering and Community Detection in Directed Networks: A Survey

    Fragkiskos D. Malliaros, Michalis Vazirgiannis

    cs.SIcs.IRphysics.bio-pharXiv:1308.0971v12013
  9. SQuTR: A Robustness Benchmark for Spoken Query to Text Retrieval under Acoustic Noise

    Yuejie Li, Ke Yang, Yueying Hua +4

    cs.IRcs.AIarXiv:2602.12783v32026
  10. Query as Anchor: Scenario-Adaptive User Representation via Large Language Model

    Jiahao Yuan, Yike Xu, Jinyong Wen +9

    cs.CLcs.IRarXiv:2602.14492v22026
  11. RankEvolve: Automating the Discovery of Retrieval Algorithms via LLM-Driven Evolution

    Jinming Nian, Fangchen Li, Dae Hoon Park +1

    cs.IRcs.AIarXiv:2602.16932v12026
  12. NanoKnow: How to Know What Your Language Model Knows

    Lingwei Gu, Nour Jedidi, Jimmy Lin

    cs.CLcs.AIcs.IRarXiv:2602.20122v22026
  13. Unsupervised Learning of Sentence Embeddings using Compositional n-Gram Features

    Matteo Pagliardini, Prakhar Gupta, Martin Jaggi

    cs.CLcs.AIcs.IRarXiv:1703.02507v32017
  14. Fine-grained Motion Retrieval via Joint-Angle Motion Images and Token-Patch Late Interaction

    Yao Zhang, Zhuchenyang Liu, Yanlan He +2

    cs.CVcs.IRarXiv:2603.09930v22026
  15. GRAFT: Graph-Distilled Generative Retrieval for Facet-Aware Scientific Literature Exploration

    Italo Luis da Silva, Hanqi Yan, Yujing Wang +3

    cs.IRcs.CLarXiv:2608.22381v12026
  16. Test-Time Strategies for More Efficient and Accurate Agentic RAG

    Brian Zhang, Deepti Guntur, Zhiyang Zuo +7

    cs.IRcs.AIarXiv:2603.12396v12026
  17. Training-Free Pseudo-Fusion for Composed Image Retrieval with Diffusion Models and Multimodal Large Language Models

    Fan Xu, Luis A. Leiva

    cs.CVcs.IRarXiv:2608.23102v12026
  18. Can Fairness Be Prompted? Prompt-Based Debiasing Strategies in High-Stakes Recommendations

    Mihaela Rotar, Theresia Veronika Rampisela, Maria Maistro

    cs.IRarXiv:2603.12935v12026
  19. Leaders in Social Networks, the Delicious Case

    Linyuan Lu, Yi-Cheng Zhang, Chi Ho Yeung +1

    physics.soc-phcs.IRcs.SIarXiv:1103.5231v12011
  20. Disentangled Graph Collaborative Filtering

    Xiang Wang, Hongye Jin, An Zhang +3

    cs.IRcs.LGarXiv:2007.01764v12020
  21. AgriIR: A Scalable Framework for Domain-Specific Knowledge Retrieval

    Shuvam Banerji Seal, Aheli Poddar, Alok Mishra +1

    cs.IRcs.AIarXiv:2604.16353v12026
  22. Evaluating AI-based Scientific Knowledge Synthesis with Epidemiological Systematic Reviews

    Shreyansh Padarha, Ryan Othniel Kearns, Tristan Naidoo +13

    cs.IRcs.AIcs.DLarXiv:2603.22327v22026
  23. The Emergence of Relevance Through Axiomatic Attention Patterns During LoRA Fine-Tuning

    Matthew Perlman, Atharva Nijasure, James Allan

    cs.CLcs.AIcs.IRarXiv:2608.23338v12026
  24. Mining Educational Data to Analyze Students' Performance

    Brijesh Kumar Baradwaj, Saurabh Pal

    cs.IRarXiv:1201.3417v12012
  25. ExecRubrics: Executable Tool-Augmented Rubrics for Verifiable and Efficient Long-Form Evaluation

    Kaustubh D. Dhole, Charles L. A. Clarke, Eugene Y. Agichtein

    cs.AIcs.CLcs.IRarXiv:2608.22559v12026
  26. WeMM-Embedding: WeChat Multi-Modal Embedding Technical Report

    Junjie Zhou, Ke Mei, Lei Li +3

    cs.CVcs.CLcs.IRarXiv:2608.24053v12026
    Summaries:한국어
  27. PhotoBench: Beyond Visual Matching Towards Personalized Intent-Driven Photo Retrieval

    Tianyi Xu, Rong Shan, Junjie Wu +11

    cs.IRcs.AIcs.CVarXiv:2603.01493v22026
  28. Multiple Instance Learning: A Survey of Problem Characteristics and Applications

    Marc-André Carbonneau, Veronika Cheplygina, Eric Granger +1

    cs.CVcs.AIcs.IRarXiv:1612.03365v12016
  29. VERDICT: Agreement Beats Pixel-Space Verification in Real-Document OCSR

    Yani Guan, Dengpan Dong, Shuang Luo +6

    cs.CVcs.IRcs.LGarXiv:2608.22183v12026
  30. NanoVDR: Distilling a 2B Vision-Language Retriever into a 70M Text-Only Encoder for Visual Document Retrieval

    Zhuchenyang Liu, Yao Zhang, Yu Xiao

    cs.IRcs.CVcs.LGarXiv:2603.12824v22026
  31. XR: Cross-Modal Agents for Composed Image Retrieval

    Zhongyu Yang, Wei Pang, Yingfang Yuan

    cs.IRarXiv:2601.14245v22026
  32. Deep Cross-Modal Hashing

    Qing-Yuan Jiang, Wu-Jun Li

    cs.IRarXiv:1602.02255v22016
  33. RocketQA: An Optimized Training Approach to Dense Passage Retrieval for Open-Domain Question Answering

    Yingqi Qu, Yuchen Ding, Jing Liu +6

    cs.CLcs.IRarXiv:2010.08191v22020
  34. WideSeek: Advancing Wide Research via Multi-Agent Scaling

    Ziyang Huang, Haolin Ren, Xiaowei Yuan +6

    cs.CLcs.AIcs.IRarXiv:2602.02636v12026
  35. TSWAP: A Multilingual Retrieval-Augmented Thai Wellness Advisor

    Pornthep Ukosaramig, Kobkrit Viriyayudhakorn

    cs.CLcs.IRarXiv:2608.22917v12026
  36. Pretrained Transformers for Text Ranking: BERT and Beyond

    Jimmy Lin, Rodrigo Nogueira, Andrew Yates

    cs.IRcs.CLarXiv:2010.06467v32020
  37. Adaptive Item-based Collaborative Structures via Noise Rescheduling in Diffusion for Generative Recommendation

    Jiaqi Wang, Tianying Liu, Heng Chang +3

    cs.IRcs.AIarXiv:2608.23400v12026
  38. From Data to Behavior: Predicting Unintended Model Behaviors Before Training

    Mengru Wang, Zhenqian Xu, Junfeng Fang +4

    cs.LGcs.AIcs.CLarXiv:2602.04735v12026
  39. The Compaction Cliff in Long-Running AI Agent Memory

    Saber Zerhoudi, Jelena Mitrovic, Michael Granitzer

    cs.AIcs.IRarXiv:2608.22752v12026
  40. Reasoning-Augmented Representations for Multimodal Retrieval

    Jianrui Zhang, Anirudh Sundara Rajan, Brandon Han +3

    cs.IRcs.AIcs.CVarXiv:2602.07125v12026
  41. C-Pack: Packed Resources For General Chinese Embeddings

    Shitao Xiao, Zheng Liu, Peitian Zhang +3

    cs.CLcs.AIcs.IRarXiv:2309.07597v52023
  42. Robustness of IR Models to Collection Growth

    Emmanouil Georgios Lionis, Debasis Ganguly, Sean MacAvaney

    cs.IRcs.CLarXiv:2608.23419v12026
  43. Inferring Networks of Substitutable and Complementary Products

    Julian McAuley, Rahul Pandey, Jure Leskovec

    cs.SIcs.IRarXiv:1506.08839v12015
  44. ManCAR: Manifold-Constrained Latent Reasoning with Adaptive Test-Time Computation for Sequential Recommendation

    Kun Yang, Yuxuan Zhu, Yazhe Chen +7

    cs.IRarXiv:2602.20093v12026
  45. DARE: Aligning LLM Agents with the R Statistical Ecosystem via Distribution-Aware Retrieval

    Maojun Sun, Yue Wu, Yifei Xie +5

    cs.IRcs.AIcs.CLarXiv:2603.04743v12026
  46. Aligning Biomedical Texts and Knowledge Graphs: A Systematic Comparison of Lightweight Alignment Strategies

    Artem Bisliouk, Elizaveta Nosova, Heiko Paulheim +2

    cs.CLcs.IRarXiv:2608.23214v12026
  47. ColBERTv2: Effective and Efficient Retrieval via Lightweight Late Interaction

    Keshav Santhanam, Omar Khattab, Jon Saad-Falcon +2

    cs.IRcs.CLarXiv:2112.01488v32021
  48. TALLRec: An Effective and Efficient Tuning Framework to Align Large Language Model with Recommendation

    Keqin Bao, Jizhi Zhang, Yang Zhang +3

    cs.IRarXiv:2305.00447v32023
  49. Enrich-Retrieve-Rank: Scaling Capability Discovery Beyond In-Context Routing

    Nazib Sorathiya, Daniel Zhang, Bardiya Akhbari

    cs.CLcs.AIcs.IRarXiv:2608.22695v12026
  50. Modeling Online Reviews with Multi-grain Topic Models

    Ivan Titov, Ryan McDonald

    cs.IRcs.DBarXiv:0801.1063v12008
  51. SAGE: Benchmarking and Improving Retrieval for Deep Research Agents

    Tiansheng Hu, Yilun Zhao, Canyu Zhang +2

    cs.IRcs.CLarXiv:2602.05975v22026
  52. CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data

    Guillaume Wenzek, Marie-Anne Lachaux, Alexis Conneau +4

    cs.CLcs.IRcs.LGarXiv:1911.00359v22019
  53. InnoEval: On Research Idea Evaluation as a Knowledge-Grounded, Multi-Perspective Reasoning Problem

    Shuofei Qiao, Yunxiang Wei, Xuehai Wang +10

    cs.CLcs.AIcs.IRarXiv:2602.14367v22026
  54. KILT: a Benchmark for Knowledge Intensive Language Tasks

    Fabio Petroni, Aleksandra Piktus, Angela Fan +10

    cs.CLcs.AIcs.IRarXiv:2009.02252v42020
  55. Document Ranking with a Pretrained Sequence-to-Sequence Model

    Rodrigo Nogueira, Zhiying Jiang, Jimmy Lin

    cs.IRcs.LGarXiv:2003.06713v12020
  56. HyTRec: A Hybrid Temporal-Aware Attention Architecture for Long Behavior Sequential Recommendation

    Lei Xin, Yuhao Zheng, Ke Cheng +3

    cs.IRcs.AIarXiv:2602.18283v12026
  57. Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics

    Ziwen Xu, Chenyan Wu, Hengyu Sun +9

    cs.CLcs.AIcs.CVarXiv:2602.02343v32026
  58. RecGOAT: Graph Optimal Adaptive Transport for LLM-Enhanced Multimodal Recommendation with Dual Semantic Alignment

    Yuecheng Li, Hengwei Ju, Zeyu Song +4

    cs.IRcs.AIarXiv:2602.00682v22026
  59. GISA: A Benchmark for General Information-Seeking Assistant

    Yutao Zhu, Xingshuo Zhang, Maosen Zhang +9

    cs.CLcs.AIcs.IRarXiv:2602.08543v22026
  60. Self-EvolveRec: Self-Evolving Recommender Systems with LLM-based Directional Feedback

    Sein Kim, Sangwu Park, Hongseok Kang +6

    cs.IRcs.AIarXiv:2602.12612v22026