Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,001 to 6,060 of 11,259
Predicting Turn-Taking Outcomes in Multi-Party Conversation: Interpretable Modelling of Speech and Gaze Dynamics with Interpersonal Closeness
Mark Dourado, Karim Haddad, Henrik G. Hassager +1
cs.CLcs.SDarXiv:2608.27988v12026MUSE: Machine Unlearning Six-Way Evaluation for Language Models
Weijia Shi, Jaechan Lee, Yangsibo Huang +7
cs.CLcs.AIarXiv:2407.06460v22024QUORUM: QUality-Optimized Routing Using Multiple annotators
Antonio Purificato, Maria Sofia Bucarelli, Andrea Bacciu +2
cs.CLarXiv:2608.27974v12026Lexically conditioned realization ambiguity in Korean predicate morphology
Wonjun Oh, KyungTae Lim, Jungyeul Park
cs.CLarXiv:2608.27966v12026Entity-Memory Graph Retrieval Improves Evidence Coverage in Long-Conversation Question Answering
Shumao Sun
cs.CLarXiv:2608.27925v12026GeneGPT: Augmenting Large Language Models with Domain Tools for Improved Access to Biomedical Information
Qiao Jin, Yifan Yang, Qingyu Chen +1
cs.CLcs.AIq-bio.GNarXiv:2304.09667v32023PersonaEdit: Representative Sample Selection for Personalized Model Editing
You-Mei Huang, Chung-Chi Chen, An-Zi Yen
cs.CLarXiv:2608.27816v12026SGPT: GPT Sentence Embeddings for Semantic Search
Niklas Muennighoff
cs.CLcs.AIcs.IRarXiv:2202.08904v52022Bad Actor, Good Advisor: Exploring the Role of Large Language Models in Fake News Detection
Beizhe Hu, Qiang Sheng, Juan Cao +4
cs.CLcs.AIcs.CYarXiv:2309.12247v22023Did you hear that? Adversarial Examples Against Automatic Speech Recognition
Moustafa Alzantot, Bharathan Balaji, Mani Srivastava
cs.CLcs.CRarXiv:1801.00554v12018Massive Activations in Large Language Models
Mingjie Sun, Xinlei Chen, J. Zico Kolter +1
cs.CLcs.LGarXiv:2402.17762v22024On the Limitations of Unsupervised Bilingual Dictionary Induction
Anders Søgaard, Sebastian Ruder, Ivan Vulić
cs.CLcs.LGstat.MLarXiv:1805.03620v12018Dialogue Natural Language Inference
Sean Welleck, Jason Weston, Arthur Szlam +1
cs.CLcs.AIarXiv:1811.00671v22018SugarCrepe: Fixing Hackable Benchmarks for Vision-Language Compositionality
Cheng-Yu Hsieh, Jieyu Zhang, Zixian Ma +2
cs.CVcs.CLcs.LGarXiv:2306.14610v12023Working Memory Connections for LSTM
Federico Landi, Lorenzo Baraldi, Marcella Cornia +1
cs.LGcs.CLcs.CVarXiv:2109.00020v12021Process for Adapting Language Models to Society (PALMS) with Values-Targeted Datasets
Irene Solaiman, Christy Dennison
cs.CLcs.CYarXiv:2106.10328v22021Load-Bearing Context: The Question Damage Score for Evaluating Context Reliance in Linguistic Reasoning
Neh Majmudar, Elena Filatova
cs.CLarXiv:2608.27756v12026When Tokenizers Fail: Byte-Level Chunking for Zero-Shot Transfer to Low-Resource Languages
Sanjeev Kumar, Atsuki Yamaguchi, Nikolaos Aletras
cs.CLarXiv:2608.27658v12026Fast and accurate sentiment classification using an enhanced Naive Bayes model
Vivek Narayanan, Ishan Arora, Arjun Bhatia
cs.CLcs.IRcs.LGarXiv:1305.6143v22013Modelling Context with User Embeddings for Sarcasm Detection in Social Media
Silvio Amir, Byron C. Wallace, Hao Lyu +1
cs.CLcs.AIarXiv:1607.00976v22016Representation of syntax in LLMs through the lens of linear distance and similarity-aware entropy
Juan Pablo Vigneaux, Mary Kennedy, Khalil Iskarous +2
cs.CLarXiv:2608.27813v12026Informational Antilocality and the Locality Bias in LLMs
Andrew McInnerney, Shane Storks, Steven Abney +1
cs.CLarXiv:2608.27760v12026Below the Noise Floor: Bimodal Seed Collapse and Distinct Failure Modes in Small-Model Knowledge Distillation
Dipto Sumit, Sakib Ul Haque, Farig Sadeque
cs.CLarXiv:2608.27729v12026How Do Linear Probes Emerge? A Circuit-Tracing Framework with Concept-Targeted Attribution
Vedant Palit, Florent Draye, Terry Jingchen Zhang +2
cs.CLcs.LGarXiv:2608.27510v12026Corrective Retrieval Augmented Generation
Shi-Qi Yan, Jia-Chen Gu, Yun Zhu +1
cs.CLarXiv:2401.15884v32024Valley: Video Assistant with Large Language model Enhanced abilitY
Ruipu Luo, Ziwang Zhao, Min Yang +6
cs.CVcs.AIcs.CLarXiv:2306.07207v32023Leveraging Large Language Models for Multiple Choice Question Answering
Joshua Robinson, Christopher Michael Rytting, David Wingate
cs.CLcs.LGarXiv:2210.12353v32022Accelerating LLM Inference via Vector Index Based Output Embeddings
Martin Loretz, Sepp Hochreiter
cs.CLcs.LGarXiv:2608.27460v12026Recall and Learn: Fine-tuning Deep Pretrained Language Models with Less Forgetting
Sanyuan Chen, Yutai Hou, Yiming Cui +3
cs.CLarXiv:2004.12651v12020Large Language Models are Versatile Decomposers: Decompose Evidence and Questions for Table-based Reasoning
Yunhu Ye, Binyuan Hui, Min Yang +3
cs.CLarXiv:2301.13808v32023LingxiDiagBench: A Multi-Agent Framework for Benchmarking LLMs in Chinese Psychiatric Consultation and Diagnosis
Shihao Xu, Tiancheng Zhou, Jiatong Ma +8
cs.AIcs.CLarXiv:2602.09379v32026Navigate through Enigmatic Labyrinth A Survey of Chain of Thought Reasoning: Advances, Frontiers and Future
Zheng Chu, Jingchang Chen, Qianglong Chen +7
cs.CLcs.AIarXiv:2309.15402v32023NL2AGBench: Benchmarking LLM Auto-Formalization for AlphaGeometry
Samuel Xiao, Judy Song, Rory Hu +1
cs.CLcs.AIarXiv:2608.28481v12026Prioritized Training on Points that are Learnable, Worth Learning, and Not Yet Learnt
Sören Mindermann, Jan Brauner, Muhammed Razzak +8
cs.LGcs.AIcs.CLarXiv:2206.07137v32022Continual Lifelong Learning in Natural Language Processing: A Survey
Magdalena Biesialska, Katarzyna Biesialska, Marta R. Costa-jussà
cs.CLcs.AIcs.LGarXiv:2012.09823v12020Fidelity Is Not Enough: Dispatch-Level Instrumentation for Agentic Datasheet Extraction
Qing Ye, Meng-Hsuan Lin
cs.CLcs.AIarXiv:2608.28439v12026TI-CNN: Convolutional Neural Networks for Fake News Detection
Yang Yang, Lei Zheng, Jiawei Zhang +3
cs.CLcs.SIarXiv:1806.00749v32018How Context Affects Language Models' Factual Predictions
Fabio Petroni, Patrick Lewis, Aleksandra Piktus +4
cs.CLarXiv:2005.04611v12020Language Model Tokenizers Introduce Unfairness Between Languages
Aleksandar Petrov, Emanuele La Malfa, Philip H. S. Torr +1
cs.CLcs.LGarXiv:2305.15425v22023BanglaMed-QA: A Question Answering System for Healthcare Support in Bangla
Rowzatul Zannat, Abdullah Al Shafi, K. M. Azharul Hasan +1
cs.CLcs.AIcs.LGarXiv:2608.28329v12026A Survey on Data Selection for Language Models
Alon Albalak, Yanai Elazar, Sang Michael Xie +11
cs.CLcs.LGarXiv:2402.16827v32024When Linguistic and Internal Confidence Diverge in Large Language Models
Hefan Zhang, Bingquan Zhang, Ming Cheng +3
cs.CLcs.AIarXiv:2608.28382v12026DeepNet: Scaling Transformers to 1,000 Layers
Hongyu Wang, Shuming Ma, Li Dong +3
cs.CLcs.LGarXiv:2203.00555v12022Neural Paraphrase Generation with Stacked Residual LSTM Networks
Aaditya Prakash, Sadid A. Hasan, Kathy Lee +4
cs.CLarXiv:1610.03098v32016Simultaneously Self-Attending to All Mentions for Full-Abstract Biological Relation Extraction
Patrick Verga, Emma Strubell, Andrew McCallum
cs.CLarXiv:1802.10569v12018SimVerb-3500: A Large-Scale Evaluation Set of Verb Similarity
Daniela Gerz, Ivan Vulić, Felix Hill +2
cs.CLarXiv:1608.00869v42016BIGPATENT: A Large-Scale Dataset for Abstractive and Coherent Summarization
Eva Sharma, Chen Li, Lu Wang
cs.CLcs.LGarXiv:1906.03741v12019TaskMatrix.AI: Completing Tasks by Connecting Foundation Models with Millions of APIs
Yaobo Liang, Chenfei Wu, Ting Song +11
cs.AIcs.CLarXiv:2303.16434v12023RLHF Workflow: From Reward Modeling to Online RLHF
Hanze Dong, Wei Xiong, Bo Pang +7
cs.LGcs.AIcs.CLarXiv:2405.07863v32024General Facial Representation Learning in a Visual-Linguistic Manner
Yinglin Zheng, Hao Yang, Ting Zhang +7
cs.CVcs.CLarXiv:2112.03109v32021VISTA: Verifier-Informed Student-to-Teacher Adaptation for On-Policy Self-Distillation
Zewen Ding, Zezhong Wu, Zhou Tao +5
cs.LGcs.AIcs.CLarXiv:2608.28306v12026An Introductory Survey on Attention Mechanisms in NLP Problems
Dichao Hu
cs.CLcs.LGstat.MLarXiv:1811.05544v12018A Probabilistic Interpretation of KV Cache Eviction
Renato Geh, Alex Chen, Daniel Israel +2
cs.CLcs.AIarXiv:2608.28293v12026Automatic Sarcasm Detection: A Survey
Aditya Joshi, Pushpak Bhattacharyya, Mark James Carman
cs.CLarXiv:1602.03426v22016Embedding Models for Stance-Aware Argument Retrieval
Angelo Sparacino, Francesca Toni, Adam Dejl
cs.CLcs.AIarXiv:2608.28283v12026Siamese CBOW: Optimizing Word Embeddings for Sentence Representations
Tom Kenter, Alexey Borisov, Maarten de Rijke
cs.CLarXiv:1606.04640v12016Leveraging BERT for Extractive Text Summarization on Lectures
Derek Miller
cs.CLcs.LGcs.SDarXiv:1906.04165v12019Text Restoration of Ancient Documents with Language Models
Shibingfeng Zhang, Edoardo Caraffa, Annafelicia Zuffrano +2
cs.CLcs.AIarXiv:2608.28170v12026OneLLM: One Framework to Align All Modalities with Language
Jiaming Han, Kaixiong Gong, Yiyuan Zhang +6
cs.CVcs.AIcs.CLarXiv:2312.03700v22023ConvFinQA: Exploring the Chain of Numerical Reasoning in Conversational Finance Question Answering
Zhiyu Chen, Shiyang Li, Charese Smiley +3
cs.CLarXiv:2210.03849v12022