Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,801 to 1,860 of 11,250
Sequence-to-Sequence Neural Net Models for Grapheme-to-Phoneme Conversion
Kaisheng Yao, Geoffrey Zweig
cs.CLarXiv:1506.00196v32015On NMT Search Errors and Model Errors: Cat Got Your Tongue?
Felix Stahlberg, Bill Byrne
cs.CLarXiv:1908.10090v12019The Impact of Synthetic Data Augmentation on Discourse-Pragmatic Function Classification
Sara Sorahi, Kevin Tang, Reza Kazemian
cs.CLarXiv:2609.03652v12026EditNTS: An Neural Programmer-Interpreter Model for Sentence Simplification through Explicit Editing
Yue Dong, Zichao Li, Mehdi Rezagholizadeh +1
cs.CLarXiv:1906.08104v12019Rephrasing the Web: A Recipe for Compute and Data-Efficient Language Modeling
Pratyush Maini, Skyler Seto, He Bai +3
cs.CLarXiv:2401.16380v12024Miles v0.1: Production-Level Post-Training
RadixArk, :, Tom Chen +11
cs.LGcs.CLarXiv:2609.08368v12026A Circuit for Plural Reference: How LLMs Represent and Retrieve Singular and Plural Entities
Anh Danh, Rick Nouwen, Massimo Poesio
cs.CLarXiv:2609.03687v12026Opening mind by opening architecture: analysis strategies
Francesco Vitucci, Giuseppe Silvi, Daniele Giuseppe Annese +2
cs.CLarXiv:2609.03719v12026Typological Feature Prediction with Large Language Models: An In-Context Learning Approach
Qianwen Wang, York Hay Ng, Aditya Khan +1
cs.CLarXiv:2609.03775v12026Is MAP Decoding All You Need? The Inadequacy of the Mode in Neural Machine Translation
Bryan Eikema, Wilker Aziz
cs.CLarXiv:2005.10283v22020How to Train Your DRAGON: Diverse Augmentation Towards Generalizable Dense Retrieval
Sheng-Chieh Lin, Akari Asai, Minghan Li +5
cs.IRcs.CLarXiv:2302.07452v12023TREC CAsT 2019: The Conversational Assistance Track Overview
Jeffrey Dalton, Chenyan Xiong, Jamie Callan
cs.IRcs.CLcs.LGarXiv:2003.13624v12020Environments as Scaffold: Enriching Feedback to Bootstrap Self-Evolving Agents in Long-Horizon Tasks
Hongbang Yuan, Zhuoran Jin, Yixin Cao
cs.LGcs.AIcs.CLarXiv:2609.08404v12026Biomedical Entity Representations with Synonym Marginalization
Mujeen Sung, Hwisang Jeon, Jinhyuk Lee +1
cs.CLcs.LGarXiv:2005.00239v12020Lost in Reordering: Structural Sensitivity of Multilingual LLMs under Semantics-Preserving Perturbations
Karthika Nhayakkat, Rajat Verma, Maharaj Brahma +4
cs.CLarXiv:2609.03511v12026KhatianDoc: A Human-Verified Benchmark Diagnosing Multimodal LLM Failure on Bengali Legal Land Records
Tasmiad Hasan, Arafat Zaman Ratul, Sarker Sadman Saalim +3
cs.CLarXiv:2609.03597v12026Language, Language Models, and What We're Talking About
Malvina Nissim
cs.CLarXiv:2609.03577v12026Fine-tuning large language models for domain adaptation: Exploration of training strategies, scaling, model merging and synergistic capabilities
Wei Lu, Rachel K. Luu, Markus J. Buehler
cs.CLcond-mat.mtrl-scics.AIarXiv:2409.03444v12024MOLE: Detecting Insider Threats in AI Agents
Aashiq Muhamed, Virginia Smith
cs.LGcs.CLcs.CRarXiv:2609.06966v12026Text Classification using Capsules
Jaeyoung Kim, Sion Jang, Sungchul Choi +1
cs.CLarXiv:1808.03976v22018Teach Me to Explain: A Review of Datasets for Explainable Natural Language Processing
Sarah Wiegreffe, Ana Marasović
cs.CLcs.AIcs.LGarXiv:2102.12060v42021Trojaning Language Models for Fun and Profit
Xinyang Zhang, Zheng Zhang, Shouling Ji +1
cs.CRcs.CLcs.LGarXiv:2008.00312v22020BLEU might be Guilty but References are not Innocent
Markus Freitag, David Grangier, Isaac Caswell
cs.CLcs.AIcs.LGarXiv:2004.06063v22020Lngram v2: Latent N-Gram Memory with Interpretable Discrete Representations
Yunao Zheng, Bin Wen, Xiaojie Wang
cs.CLarXiv:2609.03426v12026To What Extent Do Large Language Models Understand Bangla Idioms?
Mousumi Akter, Md. Faiyaz Abdullah Sayeedi, Nurul Labib Sayeedi +1
cs.CLarXiv:2609.03410v12026Continual Pre-Training of Large Language Models: How to (re)warm your model?
Kshitij Gupta, Benjamin Thérien, Adam Ibrahim +5
cs.CLcs.LGarXiv:2308.04014v22023Chiaroscuro for Emotions: A Contrastive Emotion Benchmark Grounded in Appraisal Theory
Divyesh Bommana, Mohammad Saim, Tianyu Jiang
cs.CLarXiv:2609.03394v12026Accountable AI with Grounded, Faithful, Consistent, Actionable Rationales: A Case Study in Clinical Trial Matching with VERDICT
Zikai Zhou, Yufei Jin, Yilin Xu +3
cs.CLcs.CYcs.LOarXiv:2609.03366v12026Less Is Moral: A CHARMing Framework for Moral Foundations Detection in Endorsement Behaviour
Huixiang Fu, Marian-Andrei Rizoiu
cs.CLcs.CYcs.SIarXiv:2609.03330v22026How Perturbations Propagate: A Multi-Level Analysis of Robustness in Large Language Models
Dun Li Chan, Emily Liu, Niyathi Allu +1
cs.CLstat.MLarXiv:2609.03322v12026FrameBench:A Language Understanding Benchmark Based on Frame Semantics
Chihiro Yano, Ryohei Sasano
cs.CLarXiv:2609.03370v12026Decoupling Turn-Taking from Semantics: A Decoupled Data Approach for Finite-State-Machine-Based Full-Duplex Dialogue
Yihang Li, Chenhui Chu
cs.CLarXiv:2609.03321v12026Contextual Tamil Spelling and Grammar Correction Using Progressively Fine-Tuned Sequence-to-Sequence Transformers
Karthikeyan A, Jaya Nirmala S, Sangeetha Sivanesan +4
cs.CLarXiv:2609.03273v12026Sequential Beats Joint: On the Interplay between On-Policy Distillation and RLVR
Boyan Li, Bingsen Chen, Chenghao Yang +3
cs.CLcs.AIcs.LGarXiv:2609.04108v22026Divergent discourse between protests and counter-protests: #BlackLivesMatter and #AllLivesMatter
Ryan J. Gallagher, Andrew J. Reagan, Christopher M. Danforth +1
cs.CLcs.CYcs.SIarXiv:1606.06820v52016TeraPipe: Token-Level Pipeline Parallelism for Training Large-Scale Language Models
Zhuohan Li, Siyuan Zhuang, Shiyuan Guo +4
cs.LGcs.CLcs.DCarXiv:2102.07988v22021ESPO: Error-Structured Prompt Optimization via Diagnose, Diversify, and Stabilize
Lihao Liu, Peng Tang, Kunwar Yashraj Singh +1
cs.CLcs.AIarXiv:2609.04197v12026Knowledge Acquisition During Pre-training? Large Language Models Learn Better With Auxiliary Views
Joseph Lee, Yidi Huang, Dokyoon Kim +2
cs.CLcs.AIarXiv:2609.04180v12026A Combined CNN and LSTM Model for Arabic Sentiment Analysis
Abdulaziz M. Alayba, Vasile Palade, Matthew England +1
cs.CLarXiv:1807.02911v32018Who is GPT-3? An Exploration of Personality, Values and Demographics
Marilù Miotto, Nicola Rossberg, Bennett Kleinberg
cs.CLarXiv:2209.14338v22022Representational alignment yields generalizable safety in language models
Lingyu Li, Yan Teng, Yingchun Wang +1
cs.CLcs.AIarXiv:2609.04022v12026Translation as a Decision Space: A Multi-Agent Perspective on Low-Resource Dialect Generation
Hasan Alkhder, Mohammad Abboush, Igor Tchappi +2
cs.CLcs.AIarXiv:2609.04048v12026Investigating the Ability of Large Language Models to Analyze Recipes for Diabetes
Revathy Venkataramanan, Aditya Luthra, Venkatesan Nadimuthu +1
cs.CLcs.AIarXiv:2609.03967v12026Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
Sriyash Poddar, Yanming Wan, Hamish Ivison +2
cs.LGcs.AIcs.CLarXiv:2408.10075v12024THE-X: Privacy-Preserving Transformer Inference with Homomorphic Encryption
Tianyu Chen, Hangbo Bao, Shaohan Huang +6
cs.CRcs.CLarXiv:2206.00216v22022Knowledge-Driven CoT: Exploring Faithful Reasoning in LLMs for Knowledge-intensive Question Answering
Keheng Wang, Feiyu Duan, Sirui Wang +5
cs.CLcs.AIarXiv:2308.13259v22023Hypothesis Search: Inductive Reasoning with Language Models
Ruocheng Wang, Eric Zelikman, Gabriel Poesia +3
cs.LGcs.AIcs.CLarXiv:2309.05660v22023A Survey on LLM-Generated Text Detection: Necessity, Methods, and Future Directions
Junchao Wu, Shu Yang, Runzhe Zhan +3
cs.CLcs.AIarXiv:2310.14724v32023Pretraining task diversity and the emergence of non-Bayesian in-context learning for regression
Allan Raventós, Mansheej Paul, Feng Chen +1
cs.LGcs.AIcs.CLarXiv:2306.15063v22023N-ary Relation Extraction using Graph State LSTM
Linfeng Song, Yue Zhang, Zhiguo Wang +1
cs.CLarXiv:1808.09101v12018MediaSum: A Large-scale Media Interview Dataset for Dialogue Summarization
Chenguang Zhu, Yang Liu, Jie Mei +1
cs.CLarXiv:2103.06410v22021GSM-Plus: A Comprehensive Benchmark for Evaluating the Robustness of LLMs as Mathematical Problem Solvers
Qintong Li, Leyang Cui, Xueliang Zhao +2
cs.CLarXiv:2402.19255v22024E-BERT: Efficient-Yet-Effective Entity Embeddings for BERT
Nina Poerner, Ulli Waltinger, Hinrich Schütze
cs.CLarXiv:1911.03681v22019GLiNER: Generalist Model for Named Entity Recognition using Bidirectional Transformer
Urchade Zaratiana, Nadi Tomeh, Pierre Holat +1
cs.CLcs.AIcs.LGarXiv:2311.08526v12023A Partition Filter Network for Joint Entity and Relation Extraction
Zhiheng Yan, Chong Zhang, Jinlan Fu +2
cs.CLarXiv:2108.12202v82021A Discrete Hard EM Approach for Weakly Supervised Question Answering
Sewon Min, Danqi Chen, Hannaneh Hajishirzi +1
cs.CLcs.AIarXiv:1909.04849v12019Text Relatedness Based on a Word Thesaurus
George Tsatsaronis, Iraklis Varlamis, Michalis Vazirgiannis
cs.CLarXiv:1401.5699v12014AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs
Anselm Paulus, Arman Zharmagambetov, Chuan Guo +2
cs.CRcs.AIcs.CLarXiv:2404.16873v22024Headroom-Drift Replay: A Primitive for Principled Replay Control in GRPO
Hyun Bin Park, Du-Seong Chang
cs.LGcs.AIcs.CLarXiv:2609.03941v12026Multilingual LAMA: Investigating Knowledge in Multilingual Pretrained Language Models
Nora Kassner, Philipp Dufter, Hinrich Schütze
cs.CLarXiv:2102.00894v12021