Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
7,501 to 7,560 of 11,244
XCOPA: A Multilingual Dataset for Causal Commonsense Reasoning
Edoardo Maria Ponti, Goran Glavaš, Olga Majewska +3
cs.CLarXiv:2005.00333v22020GCAN: Graph-aware Co-Attention Networks for Explainable Fake News Detection on Social Media
Yi-Ju Lu, Cheng-Te Li
cs.CLcs.LGstat.MLarXiv:2004.11648v12020Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Zhang Li, Biao Yang, Qiang Liu +6
cs.CVcs.AIcs.CLarXiv:2311.06607v42023LoRA+: Efficient Low Rank Adaptation of Large Models
Soufiane Hayou, Nikhil Ghosh, Bin Yu
cs.LGcs.AIcs.CLarXiv:2402.12354v22024Transformers as Soft Reasoners over Language
Peter Clark, Oyvind Tafjord, Kyle Richardson
cs.CLcs.AIarXiv:2002.05867v22020A Simple Method for Commonsense Reasoning
Trieu H. Trinh, Quoc V. Le
cs.AIcs.CLcs.LGarXiv:1806.02847v22018The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions
Eric Wallace, Kai Xiao, Reimar Leike +3
cs.CRcs.CLcs.LGarXiv:2404.13208v12024TelecomGPT-R1: A Unified Open-Source Reasoner for the Telecom Stack
Bohao Wang, Chenwei Wu, Haoyu Li +8
cs.CLcs.ITarXiv:2608.26126v12026An Empirical Study of Training End-to-End Vision-and-Language Transformers
Zi-Yi Dou, Yichong Xu, Zhe Gan +9
cs.CVcs.CLcs.LGarXiv:2111.02387v32021The Power of Noise: Redefining Retrieval for RAG Systems
Florin Cuconasu, Giovanni Trappolini, Federico Siciliano +5
cs.IRcs.CLarXiv:2401.14887v42024Quasi-Recurrent Neural Networks
James Bradbury, Stephen Merity, Caiming Xiong +1
cs.NEcs.AIcs.CLarXiv:1611.01576v22016MuRIL: Multilingual Representations for Indian Languages
Simran Khanuja, Diksha Bansal, Sarvesh Mehtani +11
cs.CLarXiv:2103.10730v22021RATIO: A Benchmark for Retrieval Across Typed Ideation Operations in Scientific Literature
Maayan Sharon, Tom Hope
cs.CLcs.IRarXiv:2608.27394v12026Leveraging Pre-trained Checkpoints for Sequence Generation Tasks
Sascha Rothe, Shashi Narayan, Aliaksei Severyn
cs.CLarXiv:1907.12461v22019CorporateBench: Large-Scale Q&A Benchmarking with Temporal Knowledge Bases
Sil Hamilton, Albert Yu Sun, Oscar J. Romero +4
cs.AIcs.CLcs.IRarXiv:2608.27391v12026Efficient Non-parametric Estimation of Multiple Embeddings per Word in Vector Space
Arvind Neelakantan, Jeevan Shankar, Alexandre Passos +1
cs.CLstat.MLarXiv:1504.06654v12015PullNet: Open Domain Question Answering with Iterative Retrieval on Knowledge Bases and Text
Haitian Sun, Tania Bedrax-Weiss, William W. Cohen
cs.CLcs.LGarXiv:1904.09537v12019COMET-ATOMIC 2020: On Symbolic and Neural Commonsense Knowledge Graphs
Jena D. Hwang, Chandra Bhagavatula, Ronan Le Bras +4
cs.CLarXiv:2010.05953v22020Transformation Networks for Target-Oriented Sentiment Classification
Xin Li, Lidong Bing, Wai Lam +1
cs.CLarXiv:1805.01086v12018Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers
Zhicheng Huang, Zhaoyang Zeng, Bei Liu +2
cs.CVcs.CLcs.LGarXiv:2004.00849v22020Towards Measuring the Representation of Subjective Global Opinions in Language Models
Esin Durmus, Karina Nguyen, Thomas I. Liao +15
cs.CLcs.AIarXiv:2306.16388v22023Case2Flow: Bridging Patient Cases and Guideline Flowcharts through Multimodal Retrieval
Jiale Wei, Yufan Chen, Alexander Jaus +5
cs.CLcs.IRarXiv:2608.26414v12026A Reranker for Orchestrating Heterogeneous Speech and Text Retrievers
Inho Kim, Sumyeong Ahn
cs.CLcs.AIcs.IRarXiv:2608.26194v12026Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inference
Jiaming Tang, Yilong Zhao, Kan Zhu +3
cs.CLcs.LGarXiv:2406.10774v22024Agents Don't Paginate: First-Chunk Selection for LLM Tool Responses
Tatiana Petrova, Andrei Mazniak, Radu State
cs.CLcs.IRarXiv:2608.26130v12026Evaluating Gender Bias in Machine Translation
Gabriel Stanovsky, Noah A. Smith, Luke Zettlemoyer
cs.CLarXiv:1906.00591v12019Transformer Memory as a Differentiable Search Index
Yi Tay, Vinh Q. Tran, Mostafa Dehghani +10
cs.CLcs.AIcs.IRarXiv:2202.06991v32022A Study of Generative Large Language Model for Medical Research and Healthcare
Cheng Peng, Xi Yang, Aokun Chen +16
cs.CLarXiv:2305.13523v12023UNITER: UNiversal Image-TExt Representation Learning
Yen-Chun Chen, Linjie Li, Licheng Yu +5
cs.CVcs.CLcs.LGarXiv:1909.11740v32019Gender bias and stereotypes in Large Language Models
Hadas Kotek, Rikker Dockum, David Q. Sun
cs.CLcs.CYcs.LGarXiv:2308.14921v12023Reporting Score Distributions Makes a Difference: Performance Study of LSTM-networks for Sequence Tagging
Nils Reimers, Iryna Gurevych
cs.CLstat.MLarXiv:1707.09861v12017CIFQA: A Deterministic Tool-Grounded Multi-Agent LLM Framework for Financial Query Answering
Kunjesh Parekh, Anil Kumar Tiwari, Divya Saxena
cs.AIcs.CLq-fin.CParXiv:2608.26114v12026SpinQuant: LLM quantization with learned rotations
Zechun Liu, Changsheng Zhao, Igor Fedorov +6
cs.LGcs.AIcs.CLarXiv:2405.16406v42024Gender Bias in Contextualized Word Embeddings
Jieyu Zhao, Tianlu Wang, Mark Yatskar +3
cs.CLarXiv:1904.03310v12019On Human Predictions with Explanations and Predictions of Machine Learning Models: A Case Study on Deception Detection
Vivian Lai, Chenhao Tan
cs.AIcs.CLcs.CYarXiv:1811.07901v42018Neural Legal Judgment Prediction in English
Ilias Chalkidis, Ion Androutsopoulos, Nikolaos Aletras
cs.CLarXiv:1906.02059v12019OmniQuant: Omnidirectionally Calibrated Quantization for Large Language Models
Wenqi Shao, Mengzhao Chen, Zhaoyang Zhang +7
cs.LGcs.CLarXiv:2308.13137v32023LLaVA-Video: Video Instruction Tuning With Synthetic Data
Yuanhan Zhang, Jinming Wu, Wei Li +4
cs.CVcs.CLarXiv:2410.02713v32024The Best of Both Worlds: Combining Recent Advances in Neural Machine Translation
Mia Xu Chen, Orhan Firat, Ankur Bapna +9
cs.CLcs.AIarXiv:1804.09849v22018Multimodal Intelligence: Representation Learning, Information Fusion, and Applications
Chao Zhang, Zichao Yang, Xiaodong He +1
cs.AIcs.CLcs.CVarXiv:1911.03977v32019A Survey of the State of Explainable AI for Natural Language Processing
Marina Danilevsky, Kun Qian, Ranit Aharonov +3
cs.CLcs.AIcs.LGarXiv:2010.00711v12020Lexically Constrained Decoding for Sequence Generation Using Grid Beam Search
Chris Hokamp, Qun Liu
cs.CLarXiv:1704.07138v22017BLANC: Discovering Patent White Space via Changes in Normalized Pointwise Mutual Information Between Multi-View Clusters
Shuichi Miyazawa, Kensuke Fujii
cs.IRcs.CLcs.DLarXiv:2608.26685v12026Data Recombination for Neural Semantic Parsing
Robin Jia, Percy Liang
cs.CLarXiv:1606.03622v12016IndoLEM and IndoBERT: A Benchmark Dataset and Pre-trained Language Model for Indonesian NLP
Fajri Koto, Afshin Rahimi, Jey Han Lau +1
cs.CLarXiv:2011.00677v12020Assessing the Downstream Utility of Evidence-Aware Retrieval in RAG
Utshab Kumar Ghosh, Debayan Mukhopadhyay, Shubham Chatterjee
cs.IRcs.CLarXiv:2608.26379v12026Visualizing and Measuring the Geometry of BERT
Andy Coenen, Emily Reif, Ann Yuan +4
cs.LGcs.CLstat.MLarXiv:1906.02715v22019Multimodal Explanations: Justifying Decisions and Pointing to the Evidence
Dong Huk Park, Lisa Anne Hendricks, Zeynep Akata +4
cs.AIcs.CLcs.CVarXiv:1802.08129v12018User-level sentiment analysis incorporating social networks
Chenhao Tan, Lillian Lee, Jie Tang +3
cs.CLcs.IRphysics.data-anarXiv:1109.6018v12011VirTex: Learning Visual Representations from Textual Annotations
Karan Desai, Justin Johnson
cs.CVcs.CLarXiv:2006.06666v32020Span-based Joint Entity and Relation Extraction with Transformer Pre-training
Markus Eberts, Adrian Ulges
cs.CLcs.LGarXiv:1909.07755v42019Simple BERT Models for Relation Extraction and Semantic Role Labeling
Peng Shi, Jimmy Lin
cs.CLarXiv:1904.05255v12019Towards Emotional Support Dialog Systems
Siyang Liu, Chujie Zheng, Orianna Demasi +5
cs.CLarXiv:2106.01144v12021Transformer Accelerator (TFA): A Macro-Op INT8 Hardware Chip for Transformer Inference and Machine Translation
Shashank
cs.ARcs.CLcs.LGarXiv:2608.23582v12026Deep Active Learning for Named Entity Recognition
Yanyao Shen, Hyokun Yun, Zachary C. Lipton +2
cs.CLarXiv:1707.05928v32017CyrillicQA: The Influence of Phonetically Encoded Secret Language on LLM Performance
Erik Thureck
cs.CLcs.AIcs.LGarXiv:2608.21462v12026Minimum Risk Training for Neural Machine Translation
Shiqi Shen, Yong Cheng, Zhongjun He +4
cs.CLarXiv:1512.02433v32015TTPO: Test-Time Policy Optimization
Aozhe Wang, Zhengxi Lu, Jianze Wang +8
cs.CLarXiv:2608.27448v12026A Survey of Paraphrasing and Textual Entailment Methods
Ion Androutsopoulos, Prodromos Malakasiotis
cs.CLcs.AIarXiv:0912.3747v32009WikiSkill: Compiling Agent Experience into Persistent Knowledge for Skill Evolution
Liyan Tang, Cyrus Rashtchian, Chun-Sung Ferng +3
cs.AIcs.CLarXiv:2608.27454v12026