Information Retrieval
Papers filed under cs.IR on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,201 to 1,260 of 1,626
Data Citation for Large Language Models: A Challenge
Gianmaria Silvello
cs.IRcs.AIcs.DBarXiv:2608.25663v12026End-to-End Open-Domain Question Answering with BERTserini
Wei Yang, Yuqing Xie, Aileen Lin +5
cs.CLcs.IRarXiv:1902.01718v22019Efficiently Teaching an Effective Dense Retriever with Balanced Topic Aware Sampling
Sebastian Hofstätter, Sheng-Chieh Lin, Jheng-Hong Yang +2
cs.IRcs.CLarXiv:2104.06967v22021FUNSD: A Dataset for Form Understanding in Noisy Scanned Documents
Guillaume Jaume, Hazim Kemal Ekenel, Jean-Philippe Thiran
cs.IRcs.CVcs.LGarXiv:1905.13538v22019Sequential Recommender Systems: Challenges, Progress and Prospects
Shoujin Wang, Liang Hu, Yan Wang +3
cs.IRcs.LGarXiv:2001.04830v12019PUMA: Post-Hoc Sparsification of Universal Multimodal Embeddings for Efficient Retrieval
Matteo Attimonelli, Alessandro De Bellis, Franco Maria Nardini +4
cs.IRarXiv:2608.25780v12026ECO: Efficient Convolutional Network for Online Video Understanding
Mohammadreza Zolfaghari, Kamaljeet Singh, Thomas Brox
cs.CVcs.AIcs.IRarXiv:1804.09066v22018Rank-Deviation Quality: A Distance-Aware Metric for Multi-Answer Retrieval and Ranking Evaluation
Xiaokun Zhou, Alessandro Moschitti, Danielle Class
cs.IRarXiv:2608.25318v12026NAIS: Neural Attentive Item Similarity Model for Recommendation
Xiangnan He, Zhankui He, Jingkuan Song +3
cs.IRarXiv:1809.07053v12018Hypergraph Embedding Indexing for Efficient Dense Vector Retrieval
Kishore Konda
cs.IRcs.AIarXiv:2608.22980v12026Less can be More: Relieving RAG Bottlenecks via Evidence Frontloading and Pressure-Adaptive Budgeting
Weibin Cai, Reza Zafarani
cs.CLcs.IRarXiv:2608.25115v12026Pointing the Way, Hiding the Destination: Practical Private Dense Retrieval at Scale
Peichun Hua, Danyang Chen, Junan Zhang +5
cs.CRcs.AIcs.IRarXiv:2608.25735v12026Multi-Task Feature Learning for Knowledge Graph Enhanced Recommendation
Hongwei Wang, Fuzheng Zhang, Miao Zhao +3
cs.IRstat.MLarXiv:1901.08907v12019Deep Learning over Multi-field Categorical Data: A Case Study on User Response Prediction
Weinan Zhang, Tianming Du, Jun Wang
cs.LGcs.IRarXiv:1601.02376v12016Query-Side Attacks on GNN-Based KGQA: Tracing Failures from Entity Linking to Answer Generation
Pankaj Kumar, Subhankar Mishra
cs.CLcs.AIcs.IRarXiv:2608.25922v12026Graph Retrieval-Augmented Generation: A Survey
Boci Peng, Yun Zhu, Yongchao Liu +5
cs.AIcs.CLcs.IRarXiv:2408.08921v22024Photo Aesthetics Ranking Network with Attributes and Content Adaptation
Shu Kong, Xiaohui Shen, Zhe Lin +2
cs.CVcs.IRcs.MMarXiv:1606.01621v22016Document Expansion by Query Prediction
Rodrigo Nogueira, Wei Yang, Jimmy Lin +1
cs.IRcs.LGarXiv:1904.08375v22019Influence of Pokémon Go on Physical Activity: Study and Implications
Tim Althoff, Ryen W. White, Eric Horvitz
cs.CYcs.HCcs.IRarXiv:1610.02085v22016VisDocAgentBench: Benchmarking Agents for Visually Rich Document Retrieval
Lexiang Hu, Yanzhao Zhang, Mingxin Li +5
cs.IRcs.AIcs.CVarXiv:2608.17889v12026Causal Intervention for Leveraging Popularity Bias in Recommendation
Yang Zhang, Fuli Feng, Xiangnan He +4
cs.IRarXiv:2105.06067v12021Item2Vec: Neural Item Embedding for Collaborative Filtering
Oren Barkan, Noam Koenigstein
cs.LGcs.AIcs.IRarXiv:1603.04259v32016FA*IR: A Fair Top-k Ranking Algorithm
Meike Zehlike, Francesco Bonchi, Carlos Castillo +3
cs.CYcs.IRarXiv:1706.06368v32017Geometric Matrix Completion with Recurrent Multi-Graph Neural Networks
Federico Monti, Michael M. Bronstein, Xavier Bresson
cs.LGcs.IRmath.NAarXiv:1704.06803v12017Retrieve, Match, Escalate: Accurate and Scalable Product Linking with VLM-Distilled Cross-Encoders and Agentic VLMs
Jian Wang, Steven Xu, Sanjyot Thete +5
cs.AIcs.CLcs.DBarXiv:2608.25037v12026CRAMER: Control via Request-Aware Masking for Editing Recommenders
Zhiyuan Julian Su, Naihe Feng, Zhen Luther Qin +1
cs.IRcs.AIcs.LGarXiv:2608.25370v12026Entire Space Multi-Task Model: An Effective Approach for Estimating Post-Click Conversion Rate
Xiao Ma, Liqin Zhao, Guan Huang +4
stat.MLcs.IRcs.LGarXiv:1804.07931v22018Practical and Optimal LSH for Angular Distance
Alexandr Andoni, Piotr Indyk, Thijs Laarhoven +2
cs.DScs.CGcs.IRarXiv:1509.02897v12015When Stale Constraints Go Unchecked: Budgeted Verification Failures in Inherited Agent Memory
Kazuki Nakayashiki
cs.IRcs.AIcs.CLarXiv:2608.25553v12026A Storage-Retrieval Gap in Parametric Knowledge Graph Memory
Martino M. L. Pulici, Cuong Xuan Chu, Evgeny Kharlamov +1
cs.LGcs.CLcs.IRarXiv:2608.25489v12026ReliableRAG: Combating Misinformation in Retrieval-Augmented Generation via Reliability-Guided Reasoning Chains
Jinpu Jiang, Xuan Wu, Wenhao Song +6
cs.CLcs.IRarXiv:2608.25487v12026The "Curse of Knowledge" in LLM Query Simulation: Concept Provenance for Tracing Answer-Side Intrusion
Chenglong Ma, Xinye Wanyan, Danula Hettiachchi +2
cs.IRcs.CLarXiv:2608.25245v12026RetrievalRouter: Joint Modality and Architecture Selection for Document Retrieval
Emre Kuru, Mehmet Onur Keskin, Reza Farahbakhsh +1
cs.IRarXiv:2608.25625v12026Equity of Attention: Amortizing Individual Fairness in Rankings
Asia J. Biega, Krishna P. Gummadi, Gerhard Weikum
cs.IRcs.CYarXiv:1805.01788v12018CheXbert: Combining Automatic Labelers and Expert Annotations for Accurate Radiology Report Labeling Using BERT
Akshay Smit, Saahil Jain, Pranav Rajpurkar +3
cs.CLcs.IRcs.LGarXiv:2004.09167v32020Temporal Relational Ranking for Stock Prediction
Fuli Feng, Xiangnan He, Xiang Wang +3
cs.CEcs.IRq-fin.GNarXiv:1809.09441v22018Offline bilingual word vectors, orthogonal transformations and the inverted softmax
Samuel L. Smith, David H. P. Turban, Steven Hamblin +1
cs.CLcs.AIcs.IRarXiv:1702.03859v12017Recommender Systems in the Era of Large Language Models (LLMs)
Zihuai Zhao, Wenqi Fan, Jiatong Li +8
cs.IRcs.AIcs.CLarXiv:2307.02046v62023A Brief Survey of Text Mining: Classification, Clustering and Extraction Techniques
Mehdi Allahyari, Seyedamin Pouriyeh, Mehdi Assefi +4
cs.CLcs.AIcs.IRarXiv:1707.02919v22017Top-K Off-Policy Correction for a REINFORCE Recommender System
Minmin Chen, Alex Beutel, Paul Covington +3
cs.LGcs.IRstat.MLarXiv:1812.02353v32018NV-Embed: Improved Techniques for Training LLMs as Generalist Embedding Models
Chankyu Lee, Rajarshi Roy, Mengyao Xu +4
cs.CLcs.AIcs.IRarXiv:2405.17428v32024Reinforcement Knowledge Graph Reasoning for Explainable Recommendation
Yikun Xian, Zuohui Fu, S. Muthukrishnan +2
cs.IRcs.LGarXiv:1906.05237v12019Multilingual E5 Text Embeddings: A Technical Report
Liang Wang, Nan Yang, Xiaolong Huang +3
cs.CLcs.IRarXiv:2402.05672v12024Billion-scale Commodity Embedding for E-commerce Recommendation in Alibaba
Jizhe Wang, Pipei Huang, Huan Zhao +3
cs.IRcs.AIarXiv:1803.02349v22018A Neural Influence Diffusion Model for Social Recommendation
Le Wu, Peijie Sun, Yanjie Fu +3
cs.IRcs.SIarXiv:1904.10322v12019Sequence-Aware Recommender Systems
Massimo Quadrana, Paolo Cremonesi, Dietmar Jannach
cs.IRcs.HCarXiv:1802.08452v12018W-RAG: Source-Aware Retrieval for Enterprise Document Generation from Heterogeneous Knowledge Bases
Hridya Dhulipala, Rajesh Ombase, Michael Wang +1
cs.SEcs.CLcs.IRarXiv:2608.22081v12026From Click Modeling to Offline and Off-Policy Evaluation in Carousel Recommendation
Jingwei Kang
cs.IRcs.HCarXiv:2608.22022v12026DREAM Technical Report
Bin Zhang, Bowen Zheng, Chao Yi +74
cs.IRarXiv:2608.09408v32026VoiceMem: Streaming Dual-Brain Memory for Real-Time Interaction
Zhifei Xie, Jiaqi Lang, Ze An +7
eess.AScs.AIcs.IRarXiv:2608.26005v12026A Survey on Conversational Recommender Systems
Dietmar Jannach, Ahtsham Manzoor, Wanling Cai +1
cs.HCcs.AIcs.IRarXiv:2004.00646v22020Retrieval Needs Multivectors: An Exponential Separation
Mihir Agarwal, Viraj Agrawal, Sabyasachi Basu +2
cs.IRcs.DBcs.LGarXiv:2608.21494v12026LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods
Haitao Li, Qian Dong, Junjie Chen +5
cs.CLcs.IRarXiv:2412.05579v22024Retrieval-Augmented Classification of Environmental Mitigations in Hydropower Licensing Documents
Hong-Jun Yoon, Tom Ruggles, Joanna Lee +1
cs.IRcs.AIarXiv:2608.23241v12026Robust Image Sentiment Analysis Using Progressively Trained and Domain Transferred Deep Networks
Quanzeng You, Jiebo Luo, Hailin Jin +1
cs.CVcs.IRcs.LGarXiv:1509.06041v12015A Multi-View Embedding Space for Modeling Internet Images, Tags, and their Semantics
Yunchao Gong, Qifa Ke, Michael Isard +1
cs.CVcs.IRcs.LGarXiv:1212.4522v22012SemEval-2014 Task 9: Sentiment Analysis in Twitter
Sara Rosenthal, Preslav Nakov, Alan Ritter +1
cs.CLcs.IRcs.LGarXiv:1912.02990v12019Training a Knowledge Base: Supervised Structure Learning for Agent-Curated Document Stores
Yu Pan, Hongfeng Yu
cs.CLcs.AIcs.IRarXiv:2608.21829v22026End-to-End Neural Ad-hoc Ranking with Kernel Pooling
Chenyan Xiong, Zhuyun Dai, Jamie Callan +2
cs.IRcs.CLarXiv:1706.06613v12017Bias in Bios: A Case Study of Semantic Representation Bias in a High-Stakes Setting
Maria De-Arteaga, Alexey Romanov, Hanna Wallach +6
cs.IRcs.LGstat.MLarXiv:1901.09451v12019