Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

7,501 to 7,560 of 11,244

  1. XCOPA: A Multilingual Dataset for Causal Commonsense Reasoning

    Edoardo Maria Ponti, Goran Glavaš, Olga Majewska +3

    cs.CLarXiv:2005.00333v22020
  2. GCAN: Graph-aware Co-Attention Networks for Explainable Fake News Detection on Social Media

    Yi-Ju Lu, Cheng-Te Li

    cs.CLcs.LGstat.MLarXiv:2004.11648v12020
  3. Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models

    Zhang Li, Biao Yang, Qiang Liu +6

    cs.CVcs.AIcs.CLarXiv:2311.06607v42023
  4. LoRA+: Efficient Low Rank Adaptation of Large Models

    Soufiane Hayou, Nikhil Ghosh, Bin Yu

    cs.LGcs.AIcs.CLarXiv:2402.12354v22024
  5. Transformers as Soft Reasoners over Language

    Peter Clark, Oyvind Tafjord, Kyle Richardson

    cs.CLcs.AIarXiv:2002.05867v22020
  6. A Simple Method for Commonsense Reasoning

    Trieu H. Trinh, Quoc V. Le

    cs.AIcs.CLcs.LGarXiv:1806.02847v22018
  7. The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions

    Eric Wallace, Kai Xiao, Reimar Leike +3

    cs.CRcs.CLcs.LGarXiv:2404.13208v12024
  8. TelecomGPT-R1: A Unified Open-Source Reasoner for the Telecom Stack

    Bohao Wang, Chenwei Wu, Haoyu Li +8

    cs.CLcs.ITarXiv:2608.26126v12026
  9. An Empirical Study of Training End-to-End Vision-and-Language Transformers

    Zi-Yi Dou, Yichong Xu, Zhe Gan +9

    cs.CVcs.CLcs.LGarXiv:2111.02387v32021
  10. The Power of Noise: Redefining Retrieval for RAG Systems

    Florin Cuconasu, Giovanni Trappolini, Federico Siciliano +5

    cs.IRcs.CLarXiv:2401.14887v42024
  11. Quasi-Recurrent Neural Networks

    James Bradbury, Stephen Merity, Caiming Xiong +1

    cs.NEcs.AIcs.CLarXiv:1611.01576v22016
  12. MuRIL: Multilingual Representations for Indian Languages

    Simran Khanuja, Diksha Bansal, Sarvesh Mehtani +11

    cs.CLarXiv:2103.10730v22021
  13. RATIO: A Benchmark for Retrieval Across Typed Ideation Operations in Scientific Literature

    Maayan Sharon, Tom Hope

    cs.CLcs.IRarXiv:2608.27394v12026
  14. Leveraging Pre-trained Checkpoints for Sequence Generation Tasks

    Sascha Rothe, Shashi Narayan, Aliaksei Severyn

    cs.CLarXiv:1907.12461v22019
  15. CorporateBench: Large-Scale Q&A Benchmarking with Temporal Knowledge Bases

    Sil Hamilton, Albert Yu Sun, Oscar J. Romero +4

    cs.AIcs.CLcs.IRarXiv:2608.27391v12026
  16. Efficient Non-parametric Estimation of Multiple Embeddings per Word in Vector Space

    Arvind Neelakantan, Jeevan Shankar, Alexandre Passos +1

    cs.CLstat.MLarXiv:1504.06654v12015
  17. PullNet: Open Domain Question Answering with Iterative Retrieval on Knowledge Bases and Text

    Haitian Sun, Tania Bedrax-Weiss, William W. Cohen

    cs.CLcs.LGarXiv:1904.09537v12019
  18. COMET-ATOMIC 2020: On Symbolic and Neural Commonsense Knowledge Graphs

    Jena D. Hwang, Chandra Bhagavatula, Ronan Le Bras +4

    cs.CLarXiv:2010.05953v22020
  19. Transformation Networks for Target-Oriented Sentiment Classification

    Xin Li, Lidong Bing, Wai Lam +1

    cs.CLarXiv:1805.01086v12018
  20. Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers

    Zhicheng Huang, Zhaoyang Zeng, Bei Liu +2

    cs.CVcs.CLcs.LGarXiv:2004.00849v22020
  21. Towards Measuring the Representation of Subjective Global Opinions in Language Models

    Esin Durmus, Karina Nguyen, Thomas I. Liao +15

    cs.CLcs.AIarXiv:2306.16388v22023
  22. Case2Flow: Bridging Patient Cases and Guideline Flowcharts through Multimodal Retrieval

    Jiale Wei, Yufan Chen, Alexander Jaus +5

    cs.CLcs.IRarXiv:2608.26414v12026
  23. A Reranker for Orchestrating Heterogeneous Speech and Text Retrievers

    Inho Kim, Sumyeong Ahn

    cs.CLcs.AIcs.IRarXiv:2608.26194v12026
  24. Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inference

    Jiaming Tang, Yilong Zhao, Kan Zhu +3

    cs.CLcs.LGarXiv:2406.10774v22024
  25. Agents Don't Paginate: First-Chunk Selection for LLM Tool Responses

    Tatiana Petrova, Andrei Mazniak, Radu State

    cs.CLcs.IRarXiv:2608.26130v12026
  26. Evaluating Gender Bias in Machine Translation

    Gabriel Stanovsky, Noah A. Smith, Luke Zettlemoyer

    cs.CLarXiv:1906.00591v12019
  27. Transformer Memory as a Differentiable Search Index

    Yi Tay, Vinh Q. Tran, Mostafa Dehghani +10

    cs.CLcs.AIcs.IRarXiv:2202.06991v32022
  28. A Study of Generative Large Language Model for Medical Research and Healthcare

    Cheng Peng, Xi Yang, Aokun Chen +16

    cs.CLarXiv:2305.13523v12023
  29. UNITER: UNiversal Image-TExt Representation Learning

    Yen-Chun Chen, Linjie Li, Licheng Yu +5

    cs.CVcs.CLcs.LGarXiv:1909.11740v32019
  30. Gender bias and stereotypes in Large Language Models

    Hadas Kotek, Rikker Dockum, David Q. Sun

    cs.CLcs.CYcs.LGarXiv:2308.14921v12023
  31. Reporting Score Distributions Makes a Difference: Performance Study of LSTM-networks for Sequence Tagging

    Nils Reimers, Iryna Gurevych

    cs.CLstat.MLarXiv:1707.09861v12017
  32. CIFQA: A Deterministic Tool-Grounded Multi-Agent LLM Framework for Financial Query Answering

    Kunjesh Parekh, Anil Kumar Tiwari, Divya Saxena

    cs.AIcs.CLq-fin.CParXiv:2608.26114v12026
  33. SpinQuant: LLM quantization with learned rotations

    Zechun Liu, Changsheng Zhao, Igor Fedorov +6

    cs.LGcs.AIcs.CLarXiv:2405.16406v42024
  34. Gender Bias in Contextualized Word Embeddings

    Jieyu Zhao, Tianlu Wang, Mark Yatskar +3

    cs.CLarXiv:1904.03310v12019
  35. On Human Predictions with Explanations and Predictions of Machine Learning Models: A Case Study on Deception Detection

    Vivian Lai, Chenhao Tan

    cs.AIcs.CLcs.CYarXiv:1811.07901v42018
  36. Neural Legal Judgment Prediction in English

    Ilias Chalkidis, Ion Androutsopoulos, Nikolaos Aletras

    cs.CLarXiv:1906.02059v12019
  37. OmniQuant: Omnidirectionally Calibrated Quantization for Large Language Models

    Wenqi Shao, Mengzhao Chen, Zhaoyang Zhang +7

    cs.LGcs.CLarXiv:2308.13137v32023
  38. LLaVA-Video: Video Instruction Tuning With Synthetic Data

    Yuanhan Zhang, Jinming Wu, Wei Li +4

    cs.CVcs.CLarXiv:2410.02713v32024
  39. The Best of Both Worlds: Combining Recent Advances in Neural Machine Translation

    Mia Xu Chen, Orhan Firat, Ankur Bapna +9

    cs.CLcs.AIarXiv:1804.09849v22018
  40. Multimodal Intelligence: Representation Learning, Information Fusion, and Applications

    Chao Zhang, Zichao Yang, Xiaodong He +1

    cs.AIcs.CLcs.CVarXiv:1911.03977v32019
  41. A Survey of the State of Explainable AI for Natural Language Processing

    Marina Danilevsky, Kun Qian, Ranit Aharonov +3

    cs.CLcs.AIcs.LGarXiv:2010.00711v12020
  42. Lexically Constrained Decoding for Sequence Generation Using Grid Beam Search

    Chris Hokamp, Qun Liu

    cs.CLarXiv:1704.07138v22017
  43. BLANC: Discovering Patent White Space via Changes in Normalized Pointwise Mutual Information Between Multi-View Clusters

    Shuichi Miyazawa, Kensuke Fujii

    cs.IRcs.CLcs.DLarXiv:2608.26685v12026
  44. Data Recombination for Neural Semantic Parsing

    Robin Jia, Percy Liang

    cs.CLarXiv:1606.03622v12016
  45. IndoLEM and IndoBERT: A Benchmark Dataset and Pre-trained Language Model for Indonesian NLP

    Fajri Koto, Afshin Rahimi, Jey Han Lau +1

    cs.CLarXiv:2011.00677v12020
  46. Assessing the Downstream Utility of Evidence-Aware Retrieval in RAG

    Utshab Kumar Ghosh, Debayan Mukhopadhyay, Shubham Chatterjee

    cs.IRcs.CLarXiv:2608.26379v12026
  47. Visualizing and Measuring the Geometry of BERT

    Andy Coenen, Emily Reif, Ann Yuan +4

    cs.LGcs.CLstat.MLarXiv:1906.02715v22019
  48. Multimodal Explanations: Justifying Decisions and Pointing to the Evidence

    Dong Huk Park, Lisa Anne Hendricks, Zeynep Akata +4

    cs.AIcs.CLcs.CVarXiv:1802.08129v12018
  49. User-level sentiment analysis incorporating social networks

    Chenhao Tan, Lillian Lee, Jie Tang +3

    cs.CLcs.IRphysics.data-anarXiv:1109.6018v12011
  50. VirTex: Learning Visual Representations from Textual Annotations

    Karan Desai, Justin Johnson

    cs.CVcs.CLarXiv:2006.06666v32020
  51. Span-based Joint Entity and Relation Extraction with Transformer Pre-training

    Markus Eberts, Adrian Ulges

    cs.CLcs.LGarXiv:1909.07755v42019
  52. Simple BERT Models for Relation Extraction and Semantic Role Labeling

    Peng Shi, Jimmy Lin

    cs.CLarXiv:1904.05255v12019
  53. Towards Emotional Support Dialog Systems

    Siyang Liu, Chujie Zheng, Orianna Demasi +5

    cs.CLarXiv:2106.01144v12021
  54. Transformer Accelerator (TFA): A Macro-Op INT8 Hardware Chip for Transformer Inference and Machine Translation

    Shashank

    cs.ARcs.CLcs.LGarXiv:2608.23582v12026
  55. Deep Active Learning for Named Entity Recognition

    Yanyao Shen, Hyokun Yun, Zachary C. Lipton +2

    cs.CLarXiv:1707.05928v32017
  56. CyrillicQA: The Influence of Phonetically Encoded Secret Language on LLM Performance

    Erik Thureck

    cs.CLcs.AIcs.LGarXiv:2608.21462v12026
  57. Minimum Risk Training for Neural Machine Translation

    Shiqi Shen, Yong Cheng, Zhongjun He +4

    cs.CLarXiv:1512.02433v32015
  58. TTPO: Test-Time Policy Optimization

    Aozhe Wang, Zhengxi Lu, Jianze Wang +8

    cs.CLarXiv:2608.27448v12026
  59. A Survey of Paraphrasing and Textual Entailment Methods

    Ion Androutsopoulos, Prodromos Malakasiotis

    cs.CLcs.AIarXiv:0912.3747v32009
  60. WikiSkill: Compiling Agent Experience into Persistent Knowledge for Skill Evolution

    Liyan Tang, Cyrus Rashtchian, Chun-Sung Ferng +3

    cs.AIcs.CLarXiv:2608.27454v12026