Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
7,621 to 7,680 of 11,332
SpinQuant: LLM quantization with learned rotations
Zechun Liu, Changsheng Zhao, Igor Fedorov +6
cs.LGcs.AIcs.CLarXiv:2405.16406v42024Gender Bias in Contextualized Word Embeddings
Jieyu Zhao, Tianlu Wang, Mark Yatskar +3
cs.CLarXiv:1904.03310v12019On Human Predictions with Explanations and Predictions of Machine Learning Models: A Case Study on Deception Detection
Vivian Lai, Chenhao Tan
cs.AIcs.CLcs.CYarXiv:1811.07901v42018Neural Legal Judgment Prediction in English
Ilias Chalkidis, Ion Androutsopoulos, Nikolaos Aletras
cs.CLarXiv:1906.02059v12019OmniQuant: Omnidirectionally Calibrated Quantization for Large Language Models
Wenqi Shao, Mengzhao Chen, Zhaoyang Zhang +7
cs.LGcs.CLarXiv:2308.13137v32023LLaVA-Video: Video Instruction Tuning With Synthetic Data
Yuanhan Zhang, Jinming Wu, Wei Li +4
cs.CVcs.CLarXiv:2410.02713v32024The Best of Both Worlds: Combining Recent Advances in Neural Machine Translation
Mia Xu Chen, Orhan Firat, Ankur Bapna +9
cs.CLcs.AIarXiv:1804.09849v22018Multimodal Intelligence: Representation Learning, Information Fusion, and Applications
Chao Zhang, Zichao Yang, Xiaodong He +1
cs.AIcs.CLcs.CVarXiv:1911.03977v32019A Survey of the State of Explainable AI for Natural Language Processing
Marina Danilevsky, Kun Qian, Ranit Aharonov +3
cs.CLcs.AIcs.LGarXiv:2010.00711v12020Lexically Constrained Decoding for Sequence Generation Using Grid Beam Search
Chris Hokamp, Qun Liu
cs.CLarXiv:1704.07138v22017BLANC: Discovering Patent White Space via Changes in Normalized Pointwise Mutual Information Between Multi-View Clusters
Shuichi Miyazawa, Kensuke Fujii
cs.IRcs.CLcs.DLarXiv:2608.26685v12026Data Recombination for Neural Semantic Parsing
Robin Jia, Percy Liang
cs.CLarXiv:1606.03622v12016IndoLEM and IndoBERT: A Benchmark Dataset and Pre-trained Language Model for Indonesian NLP
Fajri Koto, Afshin Rahimi, Jey Han Lau +1
cs.CLarXiv:2011.00677v12020Assessing the Downstream Utility of Evidence-Aware Retrieval in RAG
Utshab Kumar Ghosh, Debayan Mukhopadhyay, Shubham Chatterjee
cs.IRcs.CLarXiv:2608.26379v12026Visualizing and Measuring the Geometry of BERT
Andy Coenen, Emily Reif, Ann Yuan +4
cs.LGcs.CLstat.MLarXiv:1906.02715v22019Multimodal Explanations: Justifying Decisions and Pointing to the Evidence
Dong Huk Park, Lisa Anne Hendricks, Zeynep Akata +4
cs.AIcs.CLcs.CVarXiv:1802.08129v12018User-level sentiment analysis incorporating social networks
Chenhao Tan, Lillian Lee, Jie Tang +3
cs.CLcs.IRphysics.data-anarXiv:1109.6018v12011VirTex: Learning Visual Representations from Textual Annotations
Karan Desai, Justin Johnson
cs.CVcs.CLarXiv:2006.06666v32020Span-based Joint Entity and Relation Extraction with Transformer Pre-training
Markus Eberts, Adrian Ulges
cs.CLcs.LGarXiv:1909.07755v42019Simple BERT Models for Relation Extraction and Semantic Role Labeling
Peng Shi, Jimmy Lin
cs.CLarXiv:1904.05255v12019Towards Emotional Support Dialog Systems
Siyang Liu, Chujie Zheng, Orianna Demasi +5
cs.CLarXiv:2106.01144v12021Transformer Accelerator (TFA): A Macro-Op INT8 Hardware Chip for Transformer Inference and Machine Translation
Shashank
cs.ARcs.CLcs.LGarXiv:2608.23582v12026Deep Active Learning for Named Entity Recognition
Yanyao Shen, Hyokun Yun, Zachary C. Lipton +2
cs.CLarXiv:1707.05928v32017CyrillicQA: The Influence of Phonetically Encoded Secret Language on LLM Performance
Erik Thureck
cs.CLcs.AIcs.LGarXiv:2608.21462v12026Minimum Risk Training for Neural Machine Translation
Shiqi Shen, Yong Cheng, Zhongjun He +4
cs.CLarXiv:1512.02433v32015TTPO: Test-Time Policy Optimization
Aozhe Wang, Zhengxi Lu, Jianze Wang +8
cs.CLarXiv:2608.27448v12026A Survey of Paraphrasing and Textual Entailment Methods
Ion Androutsopoulos, Prodromos Malakasiotis
cs.CLcs.AIarXiv:0912.3747v32009WikiSkill: Compiling Agent Experience into Persistent Knowledge for Skill Evolution
Liyan Tang, Cyrus Rashtchian, Chun-Sung Ferng +3
cs.AIcs.CLarXiv:2608.27454v12026Low-Dimensional Hyperbolic Knowledge Graph Embeddings
Ines Chami, Adva Wolf, Da-Cheng Juan +3
cs.LGcs.AIcs.CLarXiv:2005.00545v12020Evaluation of Text Generation: A Survey
Asli Celikyilmaz, Elizabeth Clark, Jianfeng Gao
cs.CLcs.LGarXiv:2006.14799v22020Massively Multilingual Neural Machine Translation in the Wild: Findings and Challenges
Naveen Arivazhagan, Ankur Bapna, Orhan Firat +10
cs.CLcs.LGarXiv:1907.05019v12019Simple and Effective Multi-Paragraph Reading Comprehension
Christopher Clark, Matt Gardner
cs.CLarXiv:1710.10723v22017PhoBERT: Pre-trained language models for Vietnamese
Dat Quoc Nguyen, Anh Tuan Nguyen
cs.CLcs.AIarXiv:2003.00744v32020Transferable Multi-Domain State Generator for Task-Oriented Dialogue Systems
Chien-Sheng Wu, Andrea Madotto, Ehsan Hosseini-Asl +3
cs.CLcs.AIarXiv:1905.08743v22019RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback
Tianyu Yu, Yuan Yao, Haoye Zhang +8
cs.CLcs.CVarXiv:2312.00849v22023ProofWriter: Generating Implications, Proofs, and Abductive Statements over Natural Language
Oyvind Tafjord, Bhavana Dalvi Mishra, Peter Clark
cs.CLcs.AIarXiv:2012.13048v22020Targeted Syntactic Evaluation of Language Models
Rebecca Marvin, Tal Linzen
cs.CLarXiv:1808.09031v12018Large Language Models are Effective Text Rankers with Pairwise Ranking Prompting
Zhen Qin, Rolf Jagerman, Kai Hui +9
cs.IRcs.CLcs.LGarXiv:2306.17563v22023Visual-Language Prompt Tuning with Knowledge-guided Context Optimization
Hantao Yao, Rui Zhang, Changsheng Xu
cs.CVcs.CLarXiv:2303.13283v12023Selection-Inference: Exploiting Large Language Models for Interpretable Logical Reasoning
Antonia Creswell, Murray Shanahan, Irina Higgins
cs.AIcs.CLarXiv:2205.09712v12022Llemma: An Open Language Model For Mathematics
Zhangir Azerbayev, Hailey Schoelkopf, Keiran Paster +6
cs.CLcs.AIcs.LOarXiv:2310.10631v32023Editing Large Language Models: Problems, Methods, and Opportunities
Yunzhi Yao, Peng Wang, Bozhong Tian +5
cs.CLcs.AIcs.CVarXiv:2305.13172v32023LipNet: End-to-End Sentence-level Lipreading
Yannis M. Assael, Brendan Shillingford, Shimon Whiteson +1
cs.LGcs.CLcs.CVarXiv:1611.01599v22016Analyzing the Structure of Attention in a Transformer Language Model
Jesse Vig, Yonatan Belinkov
cs.CLcs.LGstat.MLarXiv:1906.04284v22019To Tune or Not to Tune? Adapting Pretrained Representations to Diverse Tasks
Matthew E. Peters, Sebastian Ruder, Noah A. Smith
cs.CLcs.LGarXiv:1903.05987v22019Reinforced Self-Training (ReST) for Language Modeling
Caglar Gulcehre, Tom Le Paine, Srivatsan Srinivasan +11
cs.CLcs.LGarXiv:2308.08998v22023Rasa: Open Source Language Understanding and Dialogue Management
Tom Bocklisch, Joey Faulkner, Nick Pawlowski +1
cs.CLcs.AIcs.LGarXiv:1712.05181v22017A Survey on Neural Speech Synthesis
Xu Tan, Tao Qin, Frank Soong +1
eess.AScs.CLcs.LGarXiv:2106.15561v32021Improving Question Answering by Commonsense-Based Pre-Training
Wanjun Zhong, Duyu Tang, Nan Duan +3
cs.CLarXiv:1809.03568v32018FEQA: A Question Answering Evaluation Framework for Faithfulness Assessment in Abstractive Summarization
Esin Durmus, He He, Mona Diab
cs.CLarXiv:2005.03754v12020Towards Deep Conversational Recommendations
Raymond Li, Samira Kahou, Hannes Schulz +3
cs.LGcs.CLcs.IRarXiv:1812.07617v22018A Survey on Data Augmentation for Text Classification
Markus Bayer, Marc-André Kaufhold, Christian Reuter
cs.CLcs.AIarXiv:2107.03158v62021Generative Spoken Language Modeling from Raw Audio
Kushal Lakhotia, Evgeny Kharitonov, Wei-Ning Hsu +8
cs.CLarXiv:2102.01192v22021Words Can Shift: Dynamically Adjusting Word Representations Using Nonverbal Behaviors
Yansen Wang, Ying Shen, Zhun Liu +3
cs.CLcs.AIarXiv:1811.09362v22018Stance and Sentiment in Tweets
Saif M. Mohammad, Parinaz Sobhani, Svetlana Kiritchenko
cs.CLarXiv:1605.01655v12016Do Prompt-Based Models Really Understand the Meaning of their Prompts?
Albert Webson, Ellie Pavlick
cs.CLarXiv:2109.01247v22021The Microsoft 2017 Conversational Speech Recognition System
W. Xiong, L. Wu, F. Alleva +3
cs.CLarXiv:1708.06073v22017HDLTex: Hierarchical Deep Learning for Text Classification
Kamran Kowsari, Donald E. Brown, Mojtaba Heidarysafa +3
cs.LGcs.AIcs.CLarXiv:1709.08267v22017Sentence Encoders on STILTs: Supplementary Training on Intermediate Labeled-data Tasks
Jason Phang, Thibault Févry, Samuel R. Bowman
cs.CLarXiv:1811.01088v22018Mixture-of-Agents Enhances Large Language Model Capabilities
Junlin Wang, Jue Wang, Ben Athiwaratkun +2
cs.CLarXiv:2406.04692v12024