Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
2,221 to 2,280 of 11,310
A Pre-training Based Personalized Dialogue Generation Model with Persona-sparse Data
Yinhe Zheng, Rongsheng Zhang, Xiaoxi Mao +1
cs.CLcs.AIcs.LGarXiv:1911.04700v12019Vectorizing Classical Tamil: Representation Learning for Verse-Commentary Pairs
Amrit Gopinath, Sangeetha Sivanesan
cs.CLarXiv:2609.04755v12026Prompt-Learning for Fine-Grained Entity Typing
Ning Ding, Yulin Chen, Xu Han +6
cs.CLcs.AIarXiv:2108.10604v12021How Do Language Models Represent and Use Phonological Information for Allomorph Selection?
Sangwoo Kim, Sangah Lee
cs.CLarXiv:2609.04708v12026Ask the Right Questions: Active Question Reformulation with Reinforcement Learning
Christian Buck, Jannis Bulian, Massimiliano Ciaramita +4
cs.CLcs.AIarXiv:1705.07830v32017Explain in Your Own Words: Improving Reasoning via Token-Selective Dual Knowledge Distillation
Minsang Kim, Seung Jun Baek
cs.CLcs.AIcs.LGarXiv:2603.13260v12026Choosing the Right Language Mode at Inference Time for Multilingual Reliability
Ekata Mitra, Ameeta Agrawal
cs.CLarXiv:2609.04653v12026Self-Play Only Evolves When Self-Synthetic Pipeline Ensures Learnable Information Gain
Wei Liu, Siya Qi, Yali Du +1
cs.LGcs.AIcs.CLarXiv:2603.02218v22026Text summarization via global structure awareness
Jiaquan Zhang, Chaoning Zhang, Shuxu Chen +9
cs.CLcs.AIarXiv:2602.09821v22026ConsensusBench: Benchmark of Consensus Nodes for LLM Reasoning via Outcome Reward Densifying
Shi-Qi Yan, Chao-Hong Tan, Qian Chen +3
cs.CLarXiv:2609.04648v12026FinCARDS: Card-Based Analyst Reranking for Financial Document Question Answering
Yixi Zhou, Fan Zhang, Yu Chen +3
cs.IRcs.AIcs.CLarXiv:2601.06992v22026CAGE: Coherence-Aware Graph Encoding for Retrieval-Augmented Generation
Tong Qi, Jingyu Wu, Youbing Yin +3
cs.CLcs.IRarXiv:2609.04647v12026All Circuits Lead to Rome: Rethinking Functional Anisotropy in Circuit and Sheaf Discovery for LLMs
Xi Chen, Mingyu Jin, Jingcheng Niu +7
cs.CLarXiv:2605.12671v12026Evaluating Long-Horizon Memory for Multi-Party Collaborative Dialogues
Chuanrui Hu, Tong Li, Xingze Gao +8
cs.CLcs.AIarXiv:2602.01313v32026Can We Predict Before Executing Machine Learning Agents?
Jingsheng Zheng, Jintian Zhang, Yujie Luo +5
cs.CLcs.AIcs.LGarXiv:2601.05930v22026From Verbatim to Gist: Distilling Pyramidal Multimodal Memory via Semantic Information Bottleneck for Long-Horizon Video Agents
Niu Lian, Yuting Wang, Hanshu Yao +5
cs.CVcs.AIcs.CLarXiv:2603.01455v32026Interactively Picking Real-World Objects with Unconstrained Spoken Language Instructions
Jun Hatori, Yuta Kikuchi, Sosuke Kobayashi +5
cs.ROcs.CLarXiv:1710.06280v22017TALON: Confidence-Aware Speculative Decoding with Adaptive Token Trees
Tianyu Liu, Qitan Lv, Yuhao Shen +2
cs.CLarXiv:2601.07353v12026BackdoorAgent: A Unified Framework for Backdoor Attacks on LLM-based Agents
Yunhao Feng, Yige Li, Yutao Wu +6
cs.AIcs.CLarXiv:2601.04566v22026Retrieval as Generation: A Unified Framework with Self-Triggered Information Planning
Bo Li, Mingda Wang, Gexiang Fang +2
cs.CLcs.AIarXiv:2604.11407v22026Adversarial Multi-Criteria Learning for Chinese Word Segmentation
Xinchi Chen, Zhan Shi, Xipeng Qiu +1
cs.CLarXiv:1704.07556v12017Strong Baselines for Neural Semi-supervised Learning under Domain Shift
Sebastian Ruder, Barbara Plank
cs.CLcs.LGstat.MLarXiv:1804.09530v12018Depth-Recurrent Attention Mixtures: Giving Latent Reasoning the Attention it Deserves
Jonas Knupp, Jan Hendrik Metzen, Jeremias Bohn +2
cs.AIcs.CLcs.LGarXiv:2601.21582v12026Thought-Retriever: Don't Just Retrieve Raw Data, Retrieve Thoughts for Memory-Augmented Agentic Systems
Tao Feng, Pengrui Han, Guanyu Lin +2
cs.CLcs.IRarXiv:2604.12231v12026Entities as Experts: Sparse Memory Access with Entity Supervision
Thibault Févry, Livio Baldini Soares, Nicholas FitzGerald +2
cs.CLcs.LGarXiv:2004.07202v22020A Verifier-Guided Explainable Reasoning Framework with Gold-Anchored QLoRA, Task-Aware Mixture-of-Experts, and Group-Relative RLVR
Thi Kim Trang Vo, Nam Tien Le, Thi Kim Nguyet Vo +2
cs.CLcs.AIcs.LGarXiv:2609.05221v12026CharacterBERT: Reconciling ELMo and BERT for Word-Level Open-Vocabulary Representations From Characters
Hicham El Boukkouri, Olivier Ferret, Thomas Lavergne +3
cs.CLarXiv:2010.10392v32020The Shape of Beliefs: Geometry, Dynamics, and Interventions along Representation Manifolds of Language Models' Posteriors
Raphaël Sarfati, Eric Bigelow, Daniel Wurgaft +6
cs.CLarXiv:2602.02315v22026The Good, The Bad, and The Greedy: Evaluation of LLMs Should Not Ignore Non-Determinism
Yifan Song, Guoyin Wang, Sujian Li +1
cs.CLcs.AIarXiv:2407.10457v12024RC-GRPO: Reward-Conditioned Group Relative Policy Optimization for Multi-Turn Tool Calling Agents
Haitian Zhong, Jixiu Zhai, Lei Song +3
cs.AIcs.CLarXiv:2602.03025v12026BioMegatron: Larger Biomedical Domain Language Model
Hoo-Chang Shin, Yang Zhang, Evelina Bakhturina +4
cs.CLarXiv:2010.06060v22020Triplets Better Than Pairs: Towards Stable and Effective Self-Play Fine-Tuning for LLMs
Yibo Wang, Hai-Long Sun, Qing-Guo Chen +4
cs.CLcs.LGarXiv:2601.08198v12026Multimodal Multi-Agent Empowered Legal Judgment Prediction
Zhaolu Kang, Junhao Gong, Qingxi Chen +7
cs.CLcs.AIcs.CYarXiv:2601.12815v52026Towards Leaving No Indic Language Behind: Building Monolingual Corpora, Benchmark and Models for Indic Languages
Sumanth Doddapaneni, Rahul Aralikatte, Gowtham Ramesh +4
cs.CLarXiv:2212.05409v32022Improving Natural Language Inference Using External Knowledge in the Science Questions Domain
Xiaoyan Wang, Pavan Kapanipathi, Ryan Musa +8
cs.AIcs.CLcs.LGarXiv:1809.05724v22018A Structured Debate-Mixture-of-Agents Framework for Complex Clinical Diagnostic Decision Support
Chang Xia, Leilei Ouyang, Huimin Wang +2
cs.CLcs.AIarXiv:2609.05069v12026An Analysis of Neural Language Modeling at Multiple Scales
Stephen Merity, Nitish Shirish Keskar, Richard Socher
cs.CLcs.AIcs.NEarXiv:1803.08240v12018Are We Done with MMLU?
Aryo Pradipta Gema, Joshua Ong Jun Leang, Giwon Hong +13
cs.CLcs.AIarXiv:2406.04127v32024A Human-in-the-Loop Framework for AI-Assisted Scoring in Large-Scale Writing Assessment
María Eugenia Curi, Germán Capdehourat, Isabel Amigo +4
cs.CLcs.AIarXiv:2609.05143v12026MAGIC: A Co-Evolving Attacker-Defender Adversarial Game for Robust LLM Safety
Xiaoyu Wen, Zhida He, Han Qi +7
cs.AIcs.CLcs.LGarXiv:2602.01539v22026COLD: A Benchmark for Chinese Offensive Language Detection
Jiawen Deng, Jingyan Zhou, Hao Sun +4
cs.CLcs.AIarXiv:2201.06025v22022OP-Bench: Benchmarking Over-Personalization for Memory-Augmented Personalized Conversational Agents
Yulin Hu, Zimo Long, Jiahe Guo +5
cs.CLcs.AIarXiv:2601.13722v12026Finding Universal Grammatical Relations in Multilingual BERT
Ethan A. Chi, John Hewitt, Christopher D. Manning
cs.CLcs.LGarXiv:2005.04511v22020Improving Reasoning Capabilities in Small Models through Mixture-of-Layers Distillation with Stepwise Attention on Key Information
Yao Chen, Jiawei Sheng, Wenyuan Zhang +1
cs.CLarXiv:2604.15701v12026ECG-R1: Protocol-Guided and Modality-Agnostic MLLM for Reliable ECG Interpretation
Jiarui Jin, Haoyu Wang, Xingliang Wu +9
cs.CLarXiv:2602.04279v32026How do LLMs Evaluate Perceived Moral Agency? Investigating Moral Decision-Making in Human-Artificial Agents Interactions
Fernanda Mansilla, Aloysius Tok, Bahia Guellaï +2
cs.CLcs.AIarXiv:2609.05037v12026Evaluating Factuality in Generation with Dependency-level Entailment
Tanya Goyal, Greg Durrett
cs.CLarXiv:2010.05478v22020Why Diffusion Language Models Struggle with Truly Parallel (Non-Autoregressive) Decoding?
Pengxiang Li, Dilxat Muhtar, Tianlong Chen +2
cs.CLcs.AIarXiv:2602.23225v22026Look-Ahead-Bench: a Standardized Benchmark of Look-ahead Bias in Point-in-Time LLMs for Finance
Mostapha Benhenda
cs.AIcs.CLcs.LGarXiv:2601.13770v12026Sentence Embedding Alignment for Lifelong Relation Extraction
Hong Wang, Wenhan Xiong, Mo Yu +3
cs.CLarXiv:1903.02588v32019Latent Adversarial Training Improves Robustness to Persistent Harmful Behaviors in LLMs
Abhay Sheshadri, Aidan Ewart, Phillip Guo +8
cs.LGcs.AIcs.CLarXiv:2407.15549v32024RefactorPlatform: An Open-Source Harness for Controlled Evaluation of Repository-Scale Refactoring Agents
Aziz Ben Amor, Drish Mali, Mann Acharya +2
cs.CLcs.AIarXiv:2609.04898v12026Autorubric: A Unifying Framework for Rubric-Based LLM Evaluation on Non-Verifiable Tasks
Delip Rao, Chris Callison-Burch
cs.CLcs.AIarXiv:2603.00077v32026Beyond IVR: Benchmarking Customer Support LLM Agents for Business-Adherence
Sumanth Balaji, Piyush Mishra, Aashraya Sachdeva +1
cs.CLarXiv:2601.00596v12026Replaying pre-training data improves fine-tuning
Suhas Kotha, Percy Liang
cs.CLcs.LGarXiv:2603.04964v12026Efficient Inference for Large Vision-Language Models: Bottlenecks, Techniques, and Prospects
Jun Zhang, Yicheng Ji, Feiyang Ren +7
cs.CLarXiv:2604.05546v22026Mousse: Rectifying the Geometry of Muon with Curvature-Aware Preconditioning
Yechen Zhang, Shuhao Xing, Junhao Huang +5
cs.LGcs.AIcs.CLarXiv:2603.09697v22026Learning to Faithfully Rationalize by Construction
Sarthak Jain, Sarah Wiegreffe, Yuval Pinter +1
cs.CLcs.AIcs.LGarXiv:2005.00115v12020Beyond Outcome Verification: Verifiable Process Reward Models for Structured Reasoning
Massimiliano Pronesti, Anya Belz, Yufang Hou
cs.CLcs.AIarXiv:2601.17223v12026Olmix: A Framework for Data Mixing Throughout LM Development
Mayee F. Chen, Tyler Murray, David Heineman +5
cs.LGcs.AIcs.CLarXiv:2602.12237v12026