Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

7,621 to 7,680 of 11,332

  1. SpinQuant: LLM quantization with learned rotations

    Zechun Liu, Changsheng Zhao, Igor Fedorov +6

    cs.LGcs.AIcs.CLarXiv:2405.16406v42024
  2. Gender Bias in Contextualized Word Embeddings

    Jieyu Zhao, Tianlu Wang, Mark Yatskar +3

    cs.CLarXiv:1904.03310v12019
  3. On Human Predictions with Explanations and Predictions of Machine Learning Models: A Case Study on Deception Detection

    Vivian Lai, Chenhao Tan

    cs.AIcs.CLcs.CYarXiv:1811.07901v42018
  4. Neural Legal Judgment Prediction in English

    Ilias Chalkidis, Ion Androutsopoulos, Nikolaos Aletras

    cs.CLarXiv:1906.02059v12019
  5. OmniQuant: Omnidirectionally Calibrated Quantization for Large Language Models

    Wenqi Shao, Mengzhao Chen, Zhaoyang Zhang +7

    cs.LGcs.CLarXiv:2308.13137v32023
  6. LLaVA-Video: Video Instruction Tuning With Synthetic Data

    Yuanhan Zhang, Jinming Wu, Wei Li +4

    cs.CVcs.CLarXiv:2410.02713v32024
  7. The Best of Both Worlds: Combining Recent Advances in Neural Machine Translation

    Mia Xu Chen, Orhan Firat, Ankur Bapna +9

    cs.CLcs.AIarXiv:1804.09849v22018
  8. Multimodal Intelligence: Representation Learning, Information Fusion, and Applications

    Chao Zhang, Zichao Yang, Xiaodong He +1

    cs.AIcs.CLcs.CVarXiv:1911.03977v32019
  9. A Survey of the State of Explainable AI for Natural Language Processing

    Marina Danilevsky, Kun Qian, Ranit Aharonov +3

    cs.CLcs.AIcs.LGarXiv:2010.00711v12020
  10. Lexically Constrained Decoding for Sequence Generation Using Grid Beam Search

    Chris Hokamp, Qun Liu

    cs.CLarXiv:1704.07138v22017
  11. BLANC: Discovering Patent White Space via Changes in Normalized Pointwise Mutual Information Between Multi-View Clusters

    Shuichi Miyazawa, Kensuke Fujii

    cs.IRcs.CLcs.DLarXiv:2608.26685v12026
  12. Data Recombination for Neural Semantic Parsing

    Robin Jia, Percy Liang

    cs.CLarXiv:1606.03622v12016
  13. IndoLEM and IndoBERT: A Benchmark Dataset and Pre-trained Language Model for Indonesian NLP

    Fajri Koto, Afshin Rahimi, Jey Han Lau +1

    cs.CLarXiv:2011.00677v12020
  14. Assessing the Downstream Utility of Evidence-Aware Retrieval in RAG

    Utshab Kumar Ghosh, Debayan Mukhopadhyay, Shubham Chatterjee

    cs.IRcs.CLarXiv:2608.26379v12026
  15. Visualizing and Measuring the Geometry of BERT

    Andy Coenen, Emily Reif, Ann Yuan +4

    cs.LGcs.CLstat.MLarXiv:1906.02715v22019
  16. Multimodal Explanations: Justifying Decisions and Pointing to the Evidence

    Dong Huk Park, Lisa Anne Hendricks, Zeynep Akata +4

    cs.AIcs.CLcs.CVarXiv:1802.08129v12018
  17. User-level sentiment analysis incorporating social networks

    Chenhao Tan, Lillian Lee, Jie Tang +3

    cs.CLcs.IRphysics.data-anarXiv:1109.6018v12011
  18. VirTex: Learning Visual Representations from Textual Annotations

    Karan Desai, Justin Johnson

    cs.CVcs.CLarXiv:2006.06666v32020
  19. Span-based Joint Entity and Relation Extraction with Transformer Pre-training

    Markus Eberts, Adrian Ulges

    cs.CLcs.LGarXiv:1909.07755v42019
  20. Simple BERT Models for Relation Extraction and Semantic Role Labeling

    Peng Shi, Jimmy Lin

    cs.CLarXiv:1904.05255v12019
  21. Towards Emotional Support Dialog Systems

    Siyang Liu, Chujie Zheng, Orianna Demasi +5

    cs.CLarXiv:2106.01144v12021
  22. Transformer Accelerator (TFA): A Macro-Op INT8 Hardware Chip for Transformer Inference and Machine Translation

    Shashank

    cs.ARcs.CLcs.LGarXiv:2608.23582v12026
  23. Deep Active Learning for Named Entity Recognition

    Yanyao Shen, Hyokun Yun, Zachary C. Lipton +2

    cs.CLarXiv:1707.05928v32017
  24. CyrillicQA: The Influence of Phonetically Encoded Secret Language on LLM Performance

    Erik Thureck

    cs.CLcs.AIcs.LGarXiv:2608.21462v12026
  25. Minimum Risk Training for Neural Machine Translation

    Shiqi Shen, Yong Cheng, Zhongjun He +4

    cs.CLarXiv:1512.02433v32015
  26. TTPO: Test-Time Policy Optimization

    Aozhe Wang, Zhengxi Lu, Jianze Wang +8

    cs.CLarXiv:2608.27448v12026
  27. A Survey of Paraphrasing and Textual Entailment Methods

    Ion Androutsopoulos, Prodromos Malakasiotis

    cs.CLcs.AIarXiv:0912.3747v32009
  28. WikiSkill: Compiling Agent Experience into Persistent Knowledge for Skill Evolution

    Liyan Tang, Cyrus Rashtchian, Chun-Sung Ferng +3

    cs.AIcs.CLarXiv:2608.27454v12026
  29. Low-Dimensional Hyperbolic Knowledge Graph Embeddings

    Ines Chami, Adva Wolf, Da-Cheng Juan +3

    cs.LGcs.AIcs.CLarXiv:2005.00545v12020
  30. Evaluation of Text Generation: A Survey

    Asli Celikyilmaz, Elizabeth Clark, Jianfeng Gao

    cs.CLcs.LGarXiv:2006.14799v22020
  31. Massively Multilingual Neural Machine Translation in the Wild: Findings and Challenges

    Naveen Arivazhagan, Ankur Bapna, Orhan Firat +10

    cs.CLcs.LGarXiv:1907.05019v12019
  32. Simple and Effective Multi-Paragraph Reading Comprehension

    Christopher Clark, Matt Gardner

    cs.CLarXiv:1710.10723v22017
  33. PhoBERT: Pre-trained language models for Vietnamese

    Dat Quoc Nguyen, Anh Tuan Nguyen

    cs.CLcs.AIarXiv:2003.00744v32020
  34. Transferable Multi-Domain State Generator for Task-Oriented Dialogue Systems

    Chien-Sheng Wu, Andrea Madotto, Ehsan Hosseini-Asl +3

    cs.CLcs.AIarXiv:1905.08743v22019
  35. RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

    Tianyu Yu, Yuan Yao, Haoye Zhang +8

    cs.CLcs.CVarXiv:2312.00849v22023
  36. ProofWriter: Generating Implications, Proofs, and Abductive Statements over Natural Language

    Oyvind Tafjord, Bhavana Dalvi Mishra, Peter Clark

    cs.CLcs.AIarXiv:2012.13048v22020
  37. Targeted Syntactic Evaluation of Language Models

    Rebecca Marvin, Tal Linzen

    cs.CLarXiv:1808.09031v12018
  38. Large Language Models are Effective Text Rankers with Pairwise Ranking Prompting

    Zhen Qin, Rolf Jagerman, Kai Hui +9

    cs.IRcs.CLcs.LGarXiv:2306.17563v22023
  39. Visual-Language Prompt Tuning with Knowledge-guided Context Optimization

    Hantao Yao, Rui Zhang, Changsheng Xu

    cs.CVcs.CLarXiv:2303.13283v12023
  40. Selection-Inference: Exploiting Large Language Models for Interpretable Logical Reasoning

    Antonia Creswell, Murray Shanahan, Irina Higgins

    cs.AIcs.CLarXiv:2205.09712v12022
  41. Llemma: An Open Language Model For Mathematics

    Zhangir Azerbayev, Hailey Schoelkopf, Keiran Paster +6

    cs.CLcs.AIcs.LOarXiv:2310.10631v32023
  42. Editing Large Language Models: Problems, Methods, and Opportunities

    Yunzhi Yao, Peng Wang, Bozhong Tian +5

    cs.CLcs.AIcs.CVarXiv:2305.13172v32023
  43. LipNet: End-to-End Sentence-level Lipreading

    Yannis M. Assael, Brendan Shillingford, Shimon Whiteson +1

    cs.LGcs.CLcs.CVarXiv:1611.01599v22016
  44. Analyzing the Structure of Attention in a Transformer Language Model

    Jesse Vig, Yonatan Belinkov

    cs.CLcs.LGstat.MLarXiv:1906.04284v22019
  45. To Tune or Not to Tune? Adapting Pretrained Representations to Diverse Tasks

    Matthew E. Peters, Sebastian Ruder, Noah A. Smith

    cs.CLcs.LGarXiv:1903.05987v22019
  46. Reinforced Self-Training (ReST) for Language Modeling

    Caglar Gulcehre, Tom Le Paine, Srivatsan Srinivasan +11

    cs.CLcs.LGarXiv:2308.08998v22023
  47. Rasa: Open Source Language Understanding and Dialogue Management

    Tom Bocklisch, Joey Faulkner, Nick Pawlowski +1

    cs.CLcs.AIcs.LGarXiv:1712.05181v22017
  48. A Survey on Neural Speech Synthesis

    Xu Tan, Tao Qin, Frank Soong +1

    eess.AScs.CLcs.LGarXiv:2106.15561v32021
  49. Improving Question Answering by Commonsense-Based Pre-Training

    Wanjun Zhong, Duyu Tang, Nan Duan +3

    cs.CLarXiv:1809.03568v32018
  50. FEQA: A Question Answering Evaluation Framework for Faithfulness Assessment in Abstractive Summarization

    Esin Durmus, He He, Mona Diab

    cs.CLarXiv:2005.03754v12020
  51. Towards Deep Conversational Recommendations

    Raymond Li, Samira Kahou, Hannes Schulz +3

    cs.LGcs.CLcs.IRarXiv:1812.07617v22018
  52. A Survey on Data Augmentation for Text Classification

    Markus Bayer, Marc-André Kaufhold, Christian Reuter

    cs.CLcs.AIarXiv:2107.03158v62021
  53. Generative Spoken Language Modeling from Raw Audio

    Kushal Lakhotia, Evgeny Kharitonov, Wei-Ning Hsu +8

    cs.CLarXiv:2102.01192v22021
  54. Words Can Shift: Dynamically Adjusting Word Representations Using Nonverbal Behaviors

    Yansen Wang, Ying Shen, Zhun Liu +3

    cs.CLcs.AIarXiv:1811.09362v22018
  55. Stance and Sentiment in Tweets

    Saif M. Mohammad, Parinaz Sobhani, Svetlana Kiritchenko

    cs.CLarXiv:1605.01655v12016
  56. Do Prompt-Based Models Really Understand the Meaning of their Prompts?

    Albert Webson, Ellie Pavlick

    cs.CLarXiv:2109.01247v22021
  57. The Microsoft 2017 Conversational Speech Recognition System

    W. Xiong, L. Wu, F. Alleva +3

    cs.CLarXiv:1708.06073v22017
  58. HDLTex: Hierarchical Deep Learning for Text Classification

    Kamran Kowsari, Donald E. Brown, Mojtaba Heidarysafa +3

    cs.LGcs.AIcs.CLarXiv:1709.08267v22017
  59. Sentence Encoders on STILTs: Supplementary Training on Intermediate Labeled-data Tasks

    Jason Phang, Thibault Févry, Samuel R. Bowman

    cs.CLarXiv:1811.01088v22018
  60. Mixture-of-Agents Enhances Large Language Model Capabilities

    Junlin Wang, Jue Wang, Ben Athiwaratkun +2

    cs.CLarXiv:2406.04692v12024