Information Retrieval
Papers filed under cs.IR on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,321 to 1,380 of 1,626
KSE-Web: An Analysis of Hybrid Retrieval and LLM-Assisted Query Expansion for Low-Resource Khmer Semantic Search
Nimol Thuon
cs.CLcs.AIcs.IRarXiv:2608.21365v12026Why This, Not That? Mining User Profiles for Pair-wise Counterfactuals
Meysam Varasteh, Veronika Bogina, Noam Koenigstein +1
cs.IRcs.AIarXiv:2608.21662v12026Personalizing Session-based Recommendations with Hierarchical Recurrent Neural Networks
Massimo Quadrana, Alexandros Karatzoglou, Balázs Hidasi +1
cs.LGcs.HCcs.IRarXiv:1706.04148v52017Rethinking Composed Image Retrieval Evaluation: A Fine-Grained Benchmark from Image Editing
Tingyu Song, Yanzhao Zhang, Mingxin Li +6
cs.CVcs.CLcs.IRarXiv:2601.16125v12026Segment Length Matters: A Study of Segment Lengths on Audio Fingerprinting Performance
Ziling Gong, Yunyan Ouyang, Iram Kamdar +5
cs.SDcs.AIcs.IRarXiv:2601.17690v12026Hierarchical Exponential-Gaussian Mixtures for Watch-Time Distribution Prediction
Sofia Gulevskaia, Mikhail Trapeznikov, Aleksandr Poslavsky +1
cs.IRcs.LGstat.MLarXiv:2608.23356v12026WARP: Wasserstein-Aligned RAG for Population Opinions
Aman Singh Thakur, Aditya Agrawal, Alwarappan Nakkiran +1
cs.IRcs.CLarXiv:2608.22859v12026Clustering and Community Detection in Directed Networks: A Survey
Fragkiskos D. Malliaros, Michalis Vazirgiannis
cs.SIcs.IRphysics.bio-pharXiv:1308.0971v12013SQuTR: A Robustness Benchmark for Spoken Query to Text Retrieval under Acoustic Noise
Yuejie Li, Ke Yang, Yueying Hua +4
cs.IRcs.AIarXiv:2602.12783v32026Query as Anchor: Scenario-Adaptive User Representation via Large Language Model
Jiahao Yuan, Yike Xu, Jinyong Wen +9
cs.CLcs.IRarXiv:2602.14492v22026RankEvolve: Automating the Discovery of Retrieval Algorithms via LLM-Driven Evolution
Jinming Nian, Fangchen Li, Dae Hoon Park +1
cs.IRcs.AIarXiv:2602.16932v12026NanoKnow: How to Know What Your Language Model Knows
Lingwei Gu, Nour Jedidi, Jimmy Lin
cs.CLcs.AIcs.IRarXiv:2602.20122v22026Unsupervised Learning of Sentence Embeddings using Compositional n-Gram Features
Matteo Pagliardini, Prakhar Gupta, Martin Jaggi
cs.CLcs.AIcs.IRarXiv:1703.02507v32017Fine-grained Motion Retrieval via Joint-Angle Motion Images and Token-Patch Late Interaction
Yao Zhang, Zhuchenyang Liu, Yanlan He +2
cs.CVcs.IRarXiv:2603.09930v22026GRAFT: Graph-Distilled Generative Retrieval for Facet-Aware Scientific Literature Exploration
Italo Luis da Silva, Hanqi Yan, Yujing Wang +3
cs.IRcs.CLarXiv:2608.22381v12026Test-Time Strategies for More Efficient and Accurate Agentic RAG
Brian Zhang, Deepti Guntur, Zhiyang Zuo +7
cs.IRcs.AIarXiv:2603.12396v12026Training-Free Pseudo-Fusion for Composed Image Retrieval with Diffusion Models and Multimodal Large Language Models
Fan Xu, Luis A. Leiva
cs.CVcs.IRarXiv:2608.23102v12026Can Fairness Be Prompted? Prompt-Based Debiasing Strategies in High-Stakes Recommendations
Mihaela Rotar, Theresia Veronika Rampisela, Maria Maistro
cs.IRarXiv:2603.12935v12026Leaders in Social Networks, the Delicious Case
Linyuan Lu, Yi-Cheng Zhang, Chi Ho Yeung +1
physics.soc-phcs.IRcs.SIarXiv:1103.5231v12011Disentangled Graph Collaborative Filtering
Xiang Wang, Hongye Jin, An Zhang +3
cs.IRcs.LGarXiv:2007.01764v12020AgriIR: A Scalable Framework for Domain-Specific Knowledge Retrieval
Shuvam Banerji Seal, Aheli Poddar, Alok Mishra +1
cs.IRcs.AIarXiv:2604.16353v12026Evaluating AI-based Scientific Knowledge Synthesis with Epidemiological Systematic Reviews
Shreyansh Padarha, Ryan Othniel Kearns, Tristan Naidoo +13
cs.IRcs.AIcs.DLarXiv:2603.22327v22026The Emergence of Relevance Through Axiomatic Attention Patterns During LoRA Fine-Tuning
Matthew Perlman, Atharva Nijasure, James Allan
cs.CLcs.AIcs.IRarXiv:2608.23338v12026Mining Educational Data to Analyze Students' Performance
Brijesh Kumar Baradwaj, Saurabh Pal
cs.IRarXiv:1201.3417v12012ExecRubrics: Executable Tool-Augmented Rubrics for Verifiable and Efficient Long-Form Evaluation
Kaustubh D. Dhole, Charles L. A. Clarke, Eugene Y. Agichtein
cs.AIcs.CLcs.IRarXiv:2608.22559v12026WeMM-Embedding: WeChat Multi-Modal Embedding Technical Report
Junjie Zhou, Ke Mei, Lei Li +3
cs.CVcs.CLcs.IRarXiv:2608.24053v12026Summaries:한국어PhotoBench: Beyond Visual Matching Towards Personalized Intent-Driven Photo Retrieval
Tianyi Xu, Rong Shan, Junjie Wu +11
cs.IRcs.AIcs.CVarXiv:2603.01493v22026Multiple Instance Learning: A Survey of Problem Characteristics and Applications
Marc-André Carbonneau, Veronika Cheplygina, Eric Granger +1
cs.CVcs.AIcs.IRarXiv:1612.03365v12016VERDICT: Agreement Beats Pixel-Space Verification in Real-Document OCSR
Yani Guan, Dengpan Dong, Shuang Luo +6
cs.CVcs.IRcs.LGarXiv:2608.22183v12026NanoVDR: Distilling a 2B Vision-Language Retriever into a 70M Text-Only Encoder for Visual Document Retrieval
Zhuchenyang Liu, Yao Zhang, Yu Xiao
cs.IRcs.CVcs.LGarXiv:2603.12824v22026XR: Cross-Modal Agents for Composed Image Retrieval
Zhongyu Yang, Wei Pang, Yingfang Yuan
cs.IRarXiv:2601.14245v22026Deep Cross-Modal Hashing
Qing-Yuan Jiang, Wu-Jun Li
cs.IRarXiv:1602.02255v22016RocketQA: An Optimized Training Approach to Dense Passage Retrieval for Open-Domain Question Answering
Yingqi Qu, Yuchen Ding, Jing Liu +6
cs.CLcs.IRarXiv:2010.08191v22020WideSeek: Advancing Wide Research via Multi-Agent Scaling
Ziyang Huang, Haolin Ren, Xiaowei Yuan +6
cs.CLcs.AIcs.IRarXiv:2602.02636v12026TSWAP: A Multilingual Retrieval-Augmented Thai Wellness Advisor
Pornthep Ukosaramig, Kobkrit Viriyayudhakorn
cs.CLcs.IRarXiv:2608.22917v12026Pretrained Transformers for Text Ranking: BERT and Beyond
Jimmy Lin, Rodrigo Nogueira, Andrew Yates
cs.IRcs.CLarXiv:2010.06467v32020Adaptive Item-based Collaborative Structures via Noise Rescheduling in Diffusion for Generative Recommendation
Jiaqi Wang, Tianying Liu, Heng Chang +3
cs.IRcs.AIarXiv:2608.23400v12026From Data to Behavior: Predicting Unintended Model Behaviors Before Training
Mengru Wang, Zhenqian Xu, Junfeng Fang +4
cs.LGcs.AIcs.CLarXiv:2602.04735v12026The Compaction Cliff in Long-Running AI Agent Memory
Saber Zerhoudi, Jelena Mitrovic, Michael Granitzer
cs.AIcs.IRarXiv:2608.22752v12026Reasoning-Augmented Representations for Multimodal Retrieval
Jianrui Zhang, Anirudh Sundara Rajan, Brandon Han +3
cs.IRcs.AIcs.CVarXiv:2602.07125v12026C-Pack: Packed Resources For General Chinese Embeddings
Shitao Xiao, Zheng Liu, Peitian Zhang +3
cs.CLcs.AIcs.IRarXiv:2309.07597v52023Robustness of IR Models to Collection Growth
Emmanouil Georgios Lionis, Debasis Ganguly, Sean MacAvaney
cs.IRcs.CLarXiv:2608.23419v12026Inferring Networks of Substitutable and Complementary Products
Julian McAuley, Rahul Pandey, Jure Leskovec
cs.SIcs.IRarXiv:1506.08839v12015ManCAR: Manifold-Constrained Latent Reasoning with Adaptive Test-Time Computation for Sequential Recommendation
Kun Yang, Yuxuan Zhu, Yazhe Chen +7
cs.IRarXiv:2602.20093v12026DARE: Aligning LLM Agents with the R Statistical Ecosystem via Distribution-Aware Retrieval
Maojun Sun, Yue Wu, Yifei Xie +5
cs.IRcs.AIcs.CLarXiv:2603.04743v12026Aligning Biomedical Texts and Knowledge Graphs: A Systematic Comparison of Lightweight Alignment Strategies
Artem Bisliouk, Elizaveta Nosova, Heiko Paulheim +2
cs.CLcs.IRarXiv:2608.23214v12026ColBERTv2: Effective and Efficient Retrieval via Lightweight Late Interaction
Keshav Santhanam, Omar Khattab, Jon Saad-Falcon +2
cs.IRcs.CLarXiv:2112.01488v32021TALLRec: An Effective and Efficient Tuning Framework to Align Large Language Model with Recommendation
Keqin Bao, Jizhi Zhang, Yang Zhang +3
cs.IRarXiv:2305.00447v32023Enrich-Retrieve-Rank: Scaling Capability Discovery Beyond In-Context Routing
Nazib Sorathiya, Daniel Zhang, Bardiya Akhbari
cs.CLcs.AIcs.IRarXiv:2608.22695v12026Modeling Online Reviews with Multi-grain Topic Models
Ivan Titov, Ryan McDonald
cs.IRcs.DBarXiv:0801.1063v12008SAGE: Benchmarking and Improving Retrieval for Deep Research Agents
Tiansheng Hu, Yilun Zhao, Canyu Zhang +2
cs.IRcs.CLarXiv:2602.05975v22026CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data
Guillaume Wenzek, Marie-Anne Lachaux, Alexis Conneau +4
cs.CLcs.IRcs.LGarXiv:1911.00359v22019InnoEval: On Research Idea Evaluation as a Knowledge-Grounded, Multi-Perspective Reasoning Problem
Shuofei Qiao, Yunxiang Wei, Xuehai Wang +10
cs.CLcs.AIcs.IRarXiv:2602.14367v22026KILT: a Benchmark for Knowledge Intensive Language Tasks
Fabio Petroni, Aleksandra Piktus, Angela Fan +10
cs.CLcs.AIcs.IRarXiv:2009.02252v42020Document Ranking with a Pretrained Sequence-to-Sequence Model
Rodrigo Nogueira, Zhiying Jiang, Jimmy Lin
cs.IRcs.LGarXiv:2003.06713v12020HyTRec: A Hybrid Temporal-Aware Attention Architecture for Long Behavior Sequential Recommendation
Lei Xin, Yuhao Zheng, Ke Cheng +3
cs.IRcs.AIarXiv:2602.18283v12026Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics
Ziwen Xu, Chenyan Wu, Hengyu Sun +9
cs.CLcs.AIcs.CVarXiv:2602.02343v32026RecGOAT: Graph Optimal Adaptive Transport for LLM-Enhanced Multimodal Recommendation with Dual Semantic Alignment
Yuecheng Li, Hengwei Ju, Zeyu Song +4
cs.IRcs.AIarXiv:2602.00682v22026GISA: A Benchmark for General Information-Seeking Assistant
Yutao Zhu, Xingshuo Zhang, Maosen Zhang +9
cs.CLcs.AIcs.IRarXiv:2602.08543v22026Self-EvolveRec: Self-Evolving Recommender Systems with LLM-based Directional Feedback
Sein Kim, Sangwu Park, Hongseok Kang +6
cs.IRcs.AIarXiv:2602.12612v22026