Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,001 to 6,060 of 11,259

  1. Predicting Turn-Taking Outcomes in Multi-Party Conversation: Interpretable Modelling of Speech and Gaze Dynamics with Interpersonal Closeness

    Mark Dourado, Karim Haddad, Henrik G. Hassager +1

    cs.CLcs.SDarXiv:2608.27988v12026
  2. MUSE: Machine Unlearning Six-Way Evaluation for Language Models

    Weijia Shi, Jaechan Lee, Yangsibo Huang +7

    cs.CLcs.AIarXiv:2407.06460v22024
  3. QUORUM: QUality-Optimized Routing Using Multiple annotators

    Antonio Purificato, Maria Sofia Bucarelli, Andrea Bacciu +2

    cs.CLarXiv:2608.27974v12026
  4. Lexically conditioned realization ambiguity in Korean predicate morphology

    Wonjun Oh, KyungTae Lim, Jungyeul Park

    cs.CLarXiv:2608.27966v12026
  5. Entity-Memory Graph Retrieval Improves Evidence Coverage in Long-Conversation Question Answering

    Shumao Sun

    cs.CLarXiv:2608.27925v12026
  6. GeneGPT: Augmenting Large Language Models with Domain Tools for Improved Access to Biomedical Information

    Qiao Jin, Yifan Yang, Qingyu Chen +1

    cs.CLcs.AIq-bio.GNarXiv:2304.09667v32023
  7. PersonaEdit: Representative Sample Selection for Personalized Model Editing

    You-Mei Huang, Chung-Chi Chen, An-Zi Yen

    cs.CLarXiv:2608.27816v12026
  8. SGPT: GPT Sentence Embeddings for Semantic Search

    Niklas Muennighoff

    cs.CLcs.AIcs.IRarXiv:2202.08904v52022
  9. Bad Actor, Good Advisor: Exploring the Role of Large Language Models in Fake News Detection

    Beizhe Hu, Qiang Sheng, Juan Cao +4

    cs.CLcs.AIcs.CYarXiv:2309.12247v22023
  10. Did you hear that? Adversarial Examples Against Automatic Speech Recognition

    Moustafa Alzantot, Bharathan Balaji, Mani Srivastava

    cs.CLcs.CRarXiv:1801.00554v12018
  11. Massive Activations in Large Language Models

    Mingjie Sun, Xinlei Chen, J. Zico Kolter +1

    cs.CLcs.LGarXiv:2402.17762v22024
  12. On the Limitations of Unsupervised Bilingual Dictionary Induction

    Anders Søgaard, Sebastian Ruder, Ivan Vulić

    cs.CLcs.LGstat.MLarXiv:1805.03620v12018
  13. Dialogue Natural Language Inference

    Sean Welleck, Jason Weston, Arthur Szlam +1

    cs.CLcs.AIarXiv:1811.00671v22018
  14. SugarCrepe: Fixing Hackable Benchmarks for Vision-Language Compositionality

    Cheng-Yu Hsieh, Jieyu Zhang, Zixian Ma +2

    cs.CVcs.CLcs.LGarXiv:2306.14610v12023
  15. Working Memory Connections for LSTM

    Federico Landi, Lorenzo Baraldi, Marcella Cornia +1

    cs.LGcs.CLcs.CVarXiv:2109.00020v12021
  16. Process for Adapting Language Models to Society (PALMS) with Values-Targeted Datasets

    Irene Solaiman, Christy Dennison

    cs.CLcs.CYarXiv:2106.10328v22021
  17. Load-Bearing Context: The Question Damage Score for Evaluating Context Reliance in Linguistic Reasoning

    Neh Majmudar, Elena Filatova

    cs.CLarXiv:2608.27756v12026
  18. When Tokenizers Fail: Byte-Level Chunking for Zero-Shot Transfer to Low-Resource Languages

    Sanjeev Kumar, Atsuki Yamaguchi, Nikolaos Aletras

    cs.CLarXiv:2608.27658v12026
  19. Fast and accurate sentiment classification using an enhanced Naive Bayes model

    Vivek Narayanan, Ishan Arora, Arjun Bhatia

    cs.CLcs.IRcs.LGarXiv:1305.6143v22013
  20. Modelling Context with User Embeddings for Sarcasm Detection in Social Media

    Silvio Amir, Byron C. Wallace, Hao Lyu +1

    cs.CLcs.AIarXiv:1607.00976v22016
  21. Representation of syntax in LLMs through the lens of linear distance and similarity-aware entropy

    Juan Pablo Vigneaux, Mary Kennedy, Khalil Iskarous +2

    cs.CLarXiv:2608.27813v12026
  22. Informational Antilocality and the Locality Bias in LLMs

    Andrew McInnerney, Shane Storks, Steven Abney +1

    cs.CLarXiv:2608.27760v12026
  23. Below the Noise Floor: Bimodal Seed Collapse and Distinct Failure Modes in Small-Model Knowledge Distillation

    Dipto Sumit, Sakib Ul Haque, Farig Sadeque

    cs.CLarXiv:2608.27729v12026
  24. How Do Linear Probes Emerge? A Circuit-Tracing Framework with Concept-Targeted Attribution

    Vedant Palit, Florent Draye, Terry Jingchen Zhang +2

    cs.CLcs.LGarXiv:2608.27510v12026
  25. Corrective Retrieval Augmented Generation

    Shi-Qi Yan, Jia-Chen Gu, Yun Zhu +1

    cs.CLarXiv:2401.15884v32024
  26. Valley: Video Assistant with Large Language model Enhanced abilitY

    Ruipu Luo, Ziwang Zhao, Min Yang +6

    cs.CVcs.AIcs.CLarXiv:2306.07207v32023
  27. Leveraging Large Language Models for Multiple Choice Question Answering

    Joshua Robinson, Christopher Michael Rytting, David Wingate

    cs.CLcs.LGarXiv:2210.12353v32022
  28. Accelerating LLM Inference via Vector Index Based Output Embeddings

    Martin Loretz, Sepp Hochreiter

    cs.CLcs.LGarXiv:2608.27460v12026
  29. Recall and Learn: Fine-tuning Deep Pretrained Language Models with Less Forgetting

    Sanyuan Chen, Yutai Hou, Yiming Cui +3

    cs.CLarXiv:2004.12651v12020
  30. Large Language Models are Versatile Decomposers: Decompose Evidence and Questions for Table-based Reasoning

    Yunhu Ye, Binyuan Hui, Min Yang +3

    cs.CLarXiv:2301.13808v32023
  31. LingxiDiagBench: A Multi-Agent Framework for Benchmarking LLMs in Chinese Psychiatric Consultation and Diagnosis

    Shihao Xu, Tiancheng Zhou, Jiatong Ma +8

    cs.AIcs.CLarXiv:2602.09379v32026
  32. Navigate through Enigmatic Labyrinth A Survey of Chain of Thought Reasoning: Advances, Frontiers and Future

    Zheng Chu, Jingchang Chen, Qianglong Chen +7

    cs.CLcs.AIarXiv:2309.15402v32023
  33. NL2AGBench: Benchmarking LLM Auto-Formalization for AlphaGeometry

    Samuel Xiao, Judy Song, Rory Hu +1

    cs.CLcs.AIarXiv:2608.28481v12026
  34. Prioritized Training on Points that are Learnable, Worth Learning, and Not Yet Learnt

    Sören Mindermann, Jan Brauner, Muhammed Razzak +8

    cs.LGcs.AIcs.CLarXiv:2206.07137v32022
  35. Continual Lifelong Learning in Natural Language Processing: A Survey

    Magdalena Biesialska, Katarzyna Biesialska, Marta R. Costa-jussà

    cs.CLcs.AIcs.LGarXiv:2012.09823v12020
  36. Fidelity Is Not Enough: Dispatch-Level Instrumentation for Agentic Datasheet Extraction

    Qing Ye, Meng-Hsuan Lin

    cs.CLcs.AIarXiv:2608.28439v12026
  37. TI-CNN: Convolutional Neural Networks for Fake News Detection

    Yang Yang, Lei Zheng, Jiawei Zhang +3

    cs.CLcs.SIarXiv:1806.00749v32018
  38. How Context Affects Language Models' Factual Predictions

    Fabio Petroni, Patrick Lewis, Aleksandra Piktus +4

    cs.CLarXiv:2005.04611v12020
  39. Language Model Tokenizers Introduce Unfairness Between Languages

    Aleksandar Petrov, Emanuele La Malfa, Philip H. S. Torr +1

    cs.CLcs.LGarXiv:2305.15425v22023
  40. BanglaMed-QA: A Question Answering System for Healthcare Support in Bangla

    Rowzatul Zannat, Abdullah Al Shafi, K. M. Azharul Hasan +1

    cs.CLcs.AIcs.LGarXiv:2608.28329v12026
  41. A Survey on Data Selection for Language Models

    Alon Albalak, Yanai Elazar, Sang Michael Xie +11

    cs.CLcs.LGarXiv:2402.16827v32024
  42. When Linguistic and Internal Confidence Diverge in Large Language Models

    Hefan Zhang, Bingquan Zhang, Ming Cheng +3

    cs.CLcs.AIarXiv:2608.28382v12026
  43. DeepNet: Scaling Transformers to 1,000 Layers

    Hongyu Wang, Shuming Ma, Li Dong +3

    cs.CLcs.LGarXiv:2203.00555v12022
  44. Neural Paraphrase Generation with Stacked Residual LSTM Networks

    Aaditya Prakash, Sadid A. Hasan, Kathy Lee +4

    cs.CLarXiv:1610.03098v32016
  45. Simultaneously Self-Attending to All Mentions for Full-Abstract Biological Relation Extraction

    Patrick Verga, Emma Strubell, Andrew McCallum

    cs.CLarXiv:1802.10569v12018
  46. SimVerb-3500: A Large-Scale Evaluation Set of Verb Similarity

    Daniela Gerz, Ivan Vulić, Felix Hill +2

    cs.CLarXiv:1608.00869v42016
  47. BIGPATENT: A Large-Scale Dataset for Abstractive and Coherent Summarization

    Eva Sharma, Chen Li, Lu Wang

    cs.CLcs.LGarXiv:1906.03741v12019
  48. TaskMatrix.AI: Completing Tasks by Connecting Foundation Models with Millions of APIs

    Yaobo Liang, Chenfei Wu, Ting Song +11

    cs.AIcs.CLarXiv:2303.16434v12023
  49. RLHF Workflow: From Reward Modeling to Online RLHF

    Hanze Dong, Wei Xiong, Bo Pang +7

    cs.LGcs.AIcs.CLarXiv:2405.07863v32024
  50. General Facial Representation Learning in a Visual-Linguistic Manner

    Yinglin Zheng, Hao Yang, Ting Zhang +7

    cs.CVcs.CLarXiv:2112.03109v32021
  51. VISTA: Verifier-Informed Student-to-Teacher Adaptation for On-Policy Self-Distillation

    Zewen Ding, Zezhong Wu, Zhou Tao +5

    cs.LGcs.AIcs.CLarXiv:2608.28306v12026
  52. An Introductory Survey on Attention Mechanisms in NLP Problems

    Dichao Hu

    cs.CLcs.LGstat.MLarXiv:1811.05544v12018
  53. A Probabilistic Interpretation of KV Cache Eviction

    Renato Geh, Alex Chen, Daniel Israel +2

    cs.CLcs.AIarXiv:2608.28293v12026
  54. Automatic Sarcasm Detection: A Survey

    Aditya Joshi, Pushpak Bhattacharyya, Mark James Carman

    cs.CLarXiv:1602.03426v22016
  55. Embedding Models for Stance-Aware Argument Retrieval

    Angelo Sparacino, Francesca Toni, Adam Dejl

    cs.CLcs.AIarXiv:2608.28283v12026
  56. Siamese CBOW: Optimizing Word Embeddings for Sentence Representations

    Tom Kenter, Alexey Borisov, Maarten de Rijke

    cs.CLarXiv:1606.04640v12016
  57. Leveraging BERT for Extractive Text Summarization on Lectures

    Derek Miller

    cs.CLcs.LGcs.SDarXiv:1906.04165v12019
  58. Text Restoration of Ancient Documents with Language Models

    Shibingfeng Zhang, Edoardo Caraffa, Annafelicia Zuffrano +2

    cs.CLcs.AIarXiv:2608.28170v12026
  59. OneLLM: One Framework to Align All Modalities with Language

    Jiaming Han, Kaixiong Gong, Yiyuan Zhang +6

    cs.CVcs.AIcs.CLarXiv:2312.03700v22023
  60. ConvFinQA: Exploring the Chain of Numerical Reasoning in Conversational Finance Question Answering

    Zhiyu Chen, Shiyang Li, Charese Smiley +3

    cs.CLarXiv:2210.03849v12022