Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,801 to 1,860 of 11,250

  1. Sequence-to-Sequence Neural Net Models for Grapheme-to-Phoneme Conversion

    Kaisheng Yao, Geoffrey Zweig

    cs.CLarXiv:1506.00196v32015
  2. On NMT Search Errors and Model Errors: Cat Got Your Tongue?

    Felix Stahlberg, Bill Byrne

    cs.CLarXiv:1908.10090v12019
  3. The Impact of Synthetic Data Augmentation on Discourse-Pragmatic Function Classification

    Sara Sorahi, Kevin Tang, Reza Kazemian

    cs.CLarXiv:2609.03652v12026
  4. EditNTS: An Neural Programmer-Interpreter Model for Sentence Simplification through Explicit Editing

    Yue Dong, Zichao Li, Mehdi Rezagholizadeh +1

    cs.CLarXiv:1906.08104v12019
  5. Rephrasing the Web: A Recipe for Compute and Data-Efficient Language Modeling

    Pratyush Maini, Skyler Seto, He Bai +3

    cs.CLarXiv:2401.16380v12024
  6. Miles v0.1: Production-Level Post-Training

    RadixArk, :, Tom Chen +11

    cs.LGcs.CLarXiv:2609.08368v12026
  7. A Circuit for Plural Reference: How LLMs Represent and Retrieve Singular and Plural Entities

    Anh Danh, Rick Nouwen, Massimo Poesio

    cs.CLarXiv:2609.03687v12026
  8. Opening mind by opening architecture: analysis strategies

    Francesco Vitucci, Giuseppe Silvi, Daniele Giuseppe Annese +2

    cs.CLarXiv:2609.03719v12026
  9. Typological Feature Prediction with Large Language Models: An In-Context Learning Approach

    Qianwen Wang, York Hay Ng, Aditya Khan +1

    cs.CLarXiv:2609.03775v12026
  10. Is MAP Decoding All You Need? The Inadequacy of the Mode in Neural Machine Translation

    Bryan Eikema, Wilker Aziz

    cs.CLarXiv:2005.10283v22020
  11. How to Train Your DRAGON: Diverse Augmentation Towards Generalizable Dense Retrieval

    Sheng-Chieh Lin, Akari Asai, Minghan Li +5

    cs.IRcs.CLarXiv:2302.07452v12023
  12. TREC CAsT 2019: The Conversational Assistance Track Overview

    Jeffrey Dalton, Chenyan Xiong, Jamie Callan

    cs.IRcs.CLcs.LGarXiv:2003.13624v12020
  13. Environments as Scaffold: Enriching Feedback to Bootstrap Self-Evolving Agents in Long-Horizon Tasks

    Hongbang Yuan, Zhuoran Jin, Yixin Cao

    cs.LGcs.AIcs.CLarXiv:2609.08404v12026
  14. Biomedical Entity Representations with Synonym Marginalization

    Mujeen Sung, Hwisang Jeon, Jinhyuk Lee +1

    cs.CLcs.LGarXiv:2005.00239v12020
  15. Lost in Reordering: Structural Sensitivity of Multilingual LLMs under Semantics-Preserving Perturbations

    Karthika Nhayakkat, Rajat Verma, Maharaj Brahma +4

    cs.CLarXiv:2609.03511v12026
  16. KhatianDoc: A Human-Verified Benchmark Diagnosing Multimodal LLM Failure on Bengali Legal Land Records

    Tasmiad Hasan, Arafat Zaman Ratul, Sarker Sadman Saalim +3

    cs.CLarXiv:2609.03597v12026
  17. Language, Language Models, and What We're Talking About

    Malvina Nissim

    cs.CLarXiv:2609.03577v12026
  18. Fine-tuning large language models for domain adaptation: Exploration of training strategies, scaling, model merging and synergistic capabilities

    Wei Lu, Rachel K. Luu, Markus J. Buehler

    cs.CLcond-mat.mtrl-scics.AIarXiv:2409.03444v12024
  19. MOLE: Detecting Insider Threats in AI Agents

    Aashiq Muhamed, Virginia Smith

    cs.LGcs.CLcs.CRarXiv:2609.06966v12026
  20. Text Classification using Capsules

    Jaeyoung Kim, Sion Jang, Sungchul Choi +1

    cs.CLarXiv:1808.03976v22018
  21. Teach Me to Explain: A Review of Datasets for Explainable Natural Language Processing

    Sarah Wiegreffe, Ana Marasović

    cs.CLcs.AIcs.LGarXiv:2102.12060v42021
  22. Trojaning Language Models for Fun and Profit

    Xinyang Zhang, Zheng Zhang, Shouling Ji +1

    cs.CRcs.CLcs.LGarXiv:2008.00312v22020
  23. BLEU might be Guilty but References are not Innocent

    Markus Freitag, David Grangier, Isaac Caswell

    cs.CLcs.AIcs.LGarXiv:2004.06063v22020
  24. Lngram v2: Latent N-Gram Memory with Interpretable Discrete Representations

    Yunao Zheng, Bin Wen, Xiaojie Wang

    cs.CLarXiv:2609.03426v12026
  25. To What Extent Do Large Language Models Understand Bangla Idioms?

    Mousumi Akter, Md. Faiyaz Abdullah Sayeedi, Nurul Labib Sayeedi +1

    cs.CLarXiv:2609.03410v12026
  26. Continual Pre-Training of Large Language Models: How to (re)warm your model?

    Kshitij Gupta, Benjamin Thérien, Adam Ibrahim +5

    cs.CLcs.LGarXiv:2308.04014v22023
  27. Chiaroscuro for Emotions: A Contrastive Emotion Benchmark Grounded in Appraisal Theory

    Divyesh Bommana, Mohammad Saim, Tianyu Jiang

    cs.CLarXiv:2609.03394v12026
  28. Accountable AI with Grounded, Faithful, Consistent, Actionable Rationales: A Case Study in Clinical Trial Matching with VERDICT

    Zikai Zhou, Yufei Jin, Yilin Xu +3

    cs.CLcs.CYcs.LOarXiv:2609.03366v12026
  29. Less Is Moral: A CHARMing Framework for Moral Foundations Detection in Endorsement Behaviour

    Huixiang Fu, Marian-Andrei Rizoiu

    cs.CLcs.CYcs.SIarXiv:2609.03330v22026
  30. How Perturbations Propagate: A Multi-Level Analysis of Robustness in Large Language Models

    Dun Li Chan, Emily Liu, Niyathi Allu +1

    cs.CLstat.MLarXiv:2609.03322v12026
  31. FrameBench:A Language Understanding Benchmark Based on Frame Semantics

    Chihiro Yano, Ryohei Sasano

    cs.CLarXiv:2609.03370v12026
  32. Decoupling Turn-Taking from Semantics: A Decoupled Data Approach for Finite-State-Machine-Based Full-Duplex Dialogue

    Yihang Li, Chenhui Chu

    cs.CLarXiv:2609.03321v12026
  33. Contextual Tamil Spelling and Grammar Correction Using Progressively Fine-Tuned Sequence-to-Sequence Transformers

    Karthikeyan A, Jaya Nirmala S, Sangeetha Sivanesan +4

    cs.CLarXiv:2609.03273v12026
  34. Sequential Beats Joint: On the Interplay between On-Policy Distillation and RLVR

    Boyan Li, Bingsen Chen, Chenghao Yang +3

    cs.CLcs.AIcs.LGarXiv:2609.04108v22026
  35. Divergent discourse between protests and counter-protests: #BlackLivesMatter and #AllLivesMatter

    Ryan J. Gallagher, Andrew J. Reagan, Christopher M. Danforth +1

    cs.CLcs.CYcs.SIarXiv:1606.06820v52016
  36. TeraPipe: Token-Level Pipeline Parallelism for Training Large-Scale Language Models

    Zhuohan Li, Siyuan Zhuang, Shiyuan Guo +4

    cs.LGcs.CLcs.DCarXiv:2102.07988v22021
  37. ESPO: Error-Structured Prompt Optimization via Diagnose, Diversify, and Stabilize

    Lihao Liu, Peng Tang, Kunwar Yashraj Singh +1

    cs.CLcs.AIarXiv:2609.04197v12026
  38. Knowledge Acquisition During Pre-training? Large Language Models Learn Better With Auxiliary Views

    Joseph Lee, Yidi Huang, Dokyoon Kim +2

    cs.CLcs.AIarXiv:2609.04180v12026
  39. A Combined CNN and LSTM Model for Arabic Sentiment Analysis

    Abdulaziz M. Alayba, Vasile Palade, Matthew England +1

    cs.CLarXiv:1807.02911v32018
  40. Who is GPT-3? An Exploration of Personality, Values and Demographics

    Marilù Miotto, Nicola Rossberg, Bennett Kleinberg

    cs.CLarXiv:2209.14338v22022
  41. Representational alignment yields generalizable safety in language models

    Lingyu Li, Yan Teng, Yingchun Wang +1

    cs.CLcs.AIarXiv:2609.04022v12026
  42. Translation as a Decision Space: A Multi-Agent Perspective on Low-Resource Dialect Generation

    Hasan Alkhder, Mohammad Abboush, Igor Tchappi +2

    cs.CLcs.AIarXiv:2609.04048v12026
  43. Investigating the Ability of Large Language Models to Analyze Recipes for Diabetes

    Revathy Venkataramanan, Aditya Luthra, Venkatesan Nadimuthu +1

    cs.CLcs.AIarXiv:2609.03967v12026
  44. Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning

    Sriyash Poddar, Yanming Wan, Hamish Ivison +2

    cs.LGcs.AIcs.CLarXiv:2408.10075v12024
  45. THE-X: Privacy-Preserving Transformer Inference with Homomorphic Encryption

    Tianyu Chen, Hangbo Bao, Shaohan Huang +6

    cs.CRcs.CLarXiv:2206.00216v22022
  46. Knowledge-Driven CoT: Exploring Faithful Reasoning in LLMs for Knowledge-intensive Question Answering

    Keheng Wang, Feiyu Duan, Sirui Wang +5

    cs.CLcs.AIarXiv:2308.13259v22023
  47. Hypothesis Search: Inductive Reasoning with Language Models

    Ruocheng Wang, Eric Zelikman, Gabriel Poesia +3

    cs.LGcs.AIcs.CLarXiv:2309.05660v22023
  48. A Survey on LLM-Generated Text Detection: Necessity, Methods, and Future Directions

    Junchao Wu, Shu Yang, Runzhe Zhan +3

    cs.CLcs.AIarXiv:2310.14724v32023
  49. Pretraining task diversity and the emergence of non-Bayesian in-context learning for regression

    Allan Raventós, Mansheej Paul, Feng Chen +1

    cs.LGcs.AIcs.CLarXiv:2306.15063v22023
  50. N-ary Relation Extraction using Graph State LSTM

    Linfeng Song, Yue Zhang, Zhiguo Wang +1

    cs.CLarXiv:1808.09101v12018
  51. MediaSum: A Large-scale Media Interview Dataset for Dialogue Summarization

    Chenguang Zhu, Yang Liu, Jie Mei +1

    cs.CLarXiv:2103.06410v22021
  52. GSM-Plus: A Comprehensive Benchmark for Evaluating the Robustness of LLMs as Mathematical Problem Solvers

    Qintong Li, Leyang Cui, Xueliang Zhao +2

    cs.CLarXiv:2402.19255v22024
  53. E-BERT: Efficient-Yet-Effective Entity Embeddings for BERT

    Nina Poerner, Ulli Waltinger, Hinrich Schütze

    cs.CLarXiv:1911.03681v22019
  54. GLiNER: Generalist Model for Named Entity Recognition using Bidirectional Transformer

    Urchade Zaratiana, Nadi Tomeh, Pierre Holat +1

    cs.CLcs.AIcs.LGarXiv:2311.08526v12023
  55. A Partition Filter Network for Joint Entity and Relation Extraction

    Zhiheng Yan, Chong Zhang, Jinlan Fu +2

    cs.CLarXiv:2108.12202v82021
  56. A Discrete Hard EM Approach for Weakly Supervised Question Answering

    Sewon Min, Danqi Chen, Hannaneh Hajishirzi +1

    cs.CLcs.AIarXiv:1909.04849v12019
  57. Text Relatedness Based on a Word Thesaurus

    George Tsatsaronis, Iraklis Varlamis, Michalis Vazirgiannis

    cs.CLarXiv:1401.5699v12014
  58. AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs

    Anselm Paulus, Arman Zharmagambetov, Chuan Guo +2

    cs.CRcs.AIcs.CLarXiv:2404.16873v22024
  59. Headroom-Drift Replay: A Primitive for Principled Replay Control in GRPO

    Hyun Bin Park, Du-Seong Chang

    cs.LGcs.AIcs.CLarXiv:2609.03941v12026
  60. Multilingual LAMA: Investigating Knowledge in Multilingual Pretrained Language Models

    Nora Kassner, Philipp Dufter, Hinrich Schütze

    cs.CLarXiv:2102.00894v12021