Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
5,521 to 5,580 of 11,238
A Deep Reinforcement Learning Chatbot
Iulian V. Serban, Chinnadhurai Sankar, Mathieu Germain +15
cs.CLcs.AIcs.LGarXiv:1709.02349v22017Generative vs. Encoder Models for Multilingual NER: A Comprehensive Empirical Study on Naamapadam
Jakkala Mahesh, Jatavath Shravan Kumar, Komalla Shivani +1
cs.CLarXiv:2608.29959v12026Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization
Weiyun Wang, Zhe Chen, Wenhai Wang +8
cs.CLcs.CVarXiv:2411.10442v22024Detecting Hidden Chain-of-Thought in Large Language Models with Linguistic, Behavioral, and Mechanistic Indicators
Armaan Singh, Ryan Trinh Le, Jasmine Kaur +5
cs.CLarXiv:2608.29956v12026XQDT: eXplainable and Quantitative Data-Text Alignment Metric with Feedback Signals
Kun Efimov-Zhang, Yifei Song, Claire Gardent
cs.CLarXiv:2608.29948v12026When Safety Speaks a Language: A Mechanistic Analysis of Safety-Language Identity Entanglement in LLMs
Apoorva Upadhyaya, Sandipan Sikdar
cs.CLcs.LGarXiv:2608.29936v12026Compression-Aware Abstention: Teaching LLMs to Refuse When KV-Compression Masks Remove Answer Evidence
Mohammadali Khodabandehlou, Bhaskar Krishnamachari
cs.CLcs.LGarXiv:2608.29934v12026Sleight of Word Benchmark: Can Language Models Notice If Their Own Output Was Tampered With?
Alberto Cetoli
cs.CLcs.AIarXiv:2608.29921v12026Cognitive Graph for Multi-Hop Reading Comprehension at Scale
Ming Ding, Chang Zhou, Qibin Chen +2
cs.CLarXiv:1905.05460v22019When Less is More: Understanding When Token Filtering Helps and Fails in AI-generated Text Detection
Xiaoyang Han, Lvxiaowei Xu, Ming Cai
cs.CLcs.AIarXiv:2608.29903v12026En-ViMedNER: An English-Vietnamese Parallel Biomedical Corpus with UMLS Semantic Type Annotations
Nhu Vo, Phuong Nguyen, Nu Uyen Phuong Le +4
cs.CLarXiv:2608.29890v12026Exploring wav2vec 2.0 on speaker verification and language identification
Zhiyun Fan, Meng Li, Shiyu Zhou +1
cs.SDcs.CLeess.ASarXiv:2012.06185v22020Toward Optimal Feature Selection in Naive Bayes for Text Categorization
Bo Tang, Steven Kay, Haibo He
stat.MLcs.CLcs.IRarXiv:1602.02850v12016Improving Argument Saliency Coverage in Small LLMs for Long Legal Opinion Summarization via Sequence-Level Distillation
Mohamed Elaraby, Ahmed Elhady, Diane Litman
cs.CLarXiv:2608.29884v12026mPLUG-2: A Modularized Multi-modal Foundation Model Across Text, Image and Video
Haiyang Xu, Qinghao Ye, Ming Yan +12
cs.CVcs.CLcs.MMarXiv:2302.00402v12023SkillForge: Compositional Skill Synthesis with Verification-in-the-Loop for Generating Formally Verified Dafny Programs
Yanming Liu, Xinyue Peng, Jiannan Cao +2
cs.CLcs.PLarXiv:2608.29841v12026Influence-Directed Distillation: Solving the Diversity Bottleneck in Sampled-Token On-Policy Distillation
Run Yang, Runpeng Dai, Jie Sun +5
cs.CLcs.LGarXiv:2608.29846v12026REIGN: Refurbished Embeddings with Integrated Guidance Networks for Efficient Context-Length Scaling
Devrim Çavuşoğlu, Emre Akbaş
cs.CLcs.AIcs.IRarXiv:2608.29899v12026Taskmaster-1: Toward a Realistic and Diverse Dialog Dataset
Bill Byrne, Karthik Krishnamoorthi, Chinnadhurai Sankar +7
cs.CLcs.AIcs.LGarXiv:1909.05358v12019When History Is Multimodal: Rethinking Context Management for Long-Horizon Agents
Jiaqi Su, Cong Pang, Jiawei Hong +4
cs.CLarXiv:2608.29897v12026Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization
Zhiyuan Zhao, Bin Wang, Linke Ouyang +3
cs.CVcs.CLarXiv:2311.16839v22023Code Completion with Neural Attention and Pointer Networks
Jian Li, Yue Wang, Michael R. Lyu +1
cs.CLcs.SEarXiv:1711.09573v22017Language model compression with weighted low-rank factorization
Yen-Chang Hsu, Ting Hua, Sungen Chang +3
cs.LGcs.AIcs.CLarXiv:2207.00112v12022VibeJam: A User Study Platform for Web Development with Agents
Nishant Balepur, Connor Baumler, Valerie Chen +3
cs.CLarXiv:2608.29889v12026Check The Scoreboard: An Analysis of Scoring Schemes on Multiple-Choice Evaluation
Nishant Balepur, Paiheng Xu, Wei Ai +3
cs.CLarXiv:2608.29887v12026ManGo: Manga Active Narrative Grounding Optimization
Hao Qiu, Junyan Wang, Zheyuan Liu +4
cs.CLarXiv:2608.29865v12026A Systematic Study and Comprehensive Evaluation of ChatGPT on Benchmark Datasets
Md Tahmid Rahman Laskar, M Saiful Bari, Mizanur Rahman +3
cs.CLcs.AIcs.LGarXiv:2305.18486v42023GenRubric: Self-Evolving Rubric Generation for Scalable LLM Evaluation
Yifan Chen, Haitao Li, Qingyao Ai +4
cs.CLarXiv:2608.29856v12026A Comprehensive Survey on Linguistic Steganography: Methods, Countermeasures, Evaluation, and Challenges
Ruiyi Yan, Chenhui Chu, Zhongliang Yang +1
cs.CRcs.CLarXiv:2608.29077v12026Distributional Validity and Calibration of a Korean Synthetic Persona Panel for Digital and AI Service Use: A Secondary-Data Validation Against the Korea Media Panel Survey
Howard Kim, Keun Tae Cho
cs.CYcs.CLarXiv:2608.28615v12026Zhongjing: Enhancing the Chinese Medical Capabilities of Large Language Model through Expert Feedback and Real-world Multi-turn Dialogue
Songhua Yang, Hanjie Zhao, Senbin Zhu +4
cs.CLarXiv:2308.03549v32023RegDivergence-101: An LLM Benchmark for Cross-Jurisdiction Regulatory Contradiction Detection in Life Sciences
Chuchu Wu, Zhiyin Zhou, Jingzhuo Hu +1
cs.AIcs.CLarXiv:2608.28607v12026AdaPlanner: Adaptive Planning from Feedback with Language Models
Haotian Sun, Yuchen Zhuang, Lingkai Kong +2
cs.CLcs.AIcs.LGarXiv:2305.16653v12023QuALITY: Question Answering with Long Input Texts, Yes!
Richard Yuanzhe Pang, Alicia Parrish, Nitish Joshi +8
cs.CLarXiv:2112.08608v22021A^2Agent: Action-Aware Reinforcement Learning for Repository-Level Code Localization Agents
Doyeon Kim, Suyoung Bae, Yumin Lee +1
cs.CLcs.SEarXiv:2608.29831v12026ReTrace: Rejected-Trajectory Conditioning for Speculative Decoding
Luxi Lin, Zhanpeng Zeng, Shuang Peng +2
cs.CLarXiv:2608.29748v12026Target-Speaker Voice Activity Detection: a Novel Approach for Multi-Speaker Diarization in a Dinner Party Scenario
Ivan Medennikov, Maxim Korenevsky, Tatiana Prisyach +9
eess.AScs.CLcs.SDarXiv:2005.07272v22020EVAR: Evidence-Validated Hypothesis Admission for Budget-Aware Narrative Reasoning
Peilin Liu, Zhiquan Ji, Jinglong Ping
cs.CLarXiv:2608.29835v12026Generating Fact Checking Explanations
Pepa Atanasova, Jakob Grue Simonsen, Christina Lioma +1
cs.CLcs.AIcs.LGarXiv:2004.05773v12020You Know What I Mean: A Benchmark for Agentic Conversational Reference Grounding
Karen Fuchs, Uri Katz, Yoav Goldberg
cs.CLcs.IRarXiv:2608.29834v12026Improving Factual Completeness and Consistency of Image-to-Text Radiology Report Generation
Yasuhide Miura, Yuhao Zhang, Emily Bao Tsai +2
cs.CLarXiv:2010.10042v22020Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming
Zhifei Xie, Changqiao Wu
cs.AIcs.CLcs.HCarXiv:2408.16725v32024R$^2$A: Learning Persona Policies Through Persona Representation Learning and Runtime Alignment
Mohan Zhang, Chengsong You, Xiaoyu Cao +4
cs.CLcs.AIarXiv:2608.29798v12026SentiCap: Generating Image Descriptions with Sentiments
Alexander Mathews, Lexing Xie, Xuming He
cs.CVcs.CLarXiv:1510.01431v22015Multi-Modal Hallucination Control by Visual Information Grounding
Alessandro Favero, Luca Zancato, Matthew Trager +5
cs.CVcs.CLcs.LGarXiv:2403.14003v12024Evaluating the Capabilities of LLMs for Persuasive Dialogue
Jordan Robinson, Angus R. Williams, Katie Atkinson +1
cs.CLarXiv:2608.29738v12026A Review of Large Language Models and Autonomous Agents in Chemistry
Mayk Caldas Ramos, Christopher J. Collison, Andrew D. White
cs.LGcs.AIcs.CLarXiv:2407.01603v32024HiVe: Beyond Static Prompts for Multitask Learning via Hierarchy-based Vertical Mixture-of-Experts
HyeonJik Bae, Minyeol Kim, Susik Yoon
cs.CLarXiv:2608.29790v12026DVBench: Benchmarking MLLMs for Understanding Dynamic Charts and Narratives in Data Videos
Bomiao Wang, Zekai Shao, Jiexiang Lan +3
cs.CLarXiv:2608.29711v12026The Depth Flow of Token Representations Is Nonlinear and Does Not Descend Its Own Density
Alexandre Quemy
cs.CLarXiv:2608.29706v12026A Hub of Short Rows Inflates Intrinsic Dimension Estimation of Token Embeddings
Alexandre Quemy
cs.CLarXiv:2608.29702v12026ACTD: Anchor-Based Cross-Tokenizer Distillation with Residual Regularization
Huiyi Zhang, Zijian Li, Xiaocheng Feng +4
cs.CLarXiv:2608.29662v12026MI-Distillation: Selecting from Model-Interpolated Instruct-Reasoning Data Spectrum for Chain-of-Thought Distillation
Yangsong Lan, Renkai Hu, HongKai Zheng +4
cs.CLcs.AIarXiv:2608.29623v12026How You Ask Shapes What You Get: A Theory-Seeded Measurement of Articulation in Advice-Seeking LLM Conversations
Juneha Baek, Suhyeon Lee, Donghyuk Shin
cs.CLarXiv:2608.29591v12026Gmail Smart Compose: Real-Time Assisted Writing
Mia Xu Chen, Benjamin N Lee, Gagan Bansal +9
cs.CLcs.LGarXiv:1906.00080v12019Deep Multimodal Learning for Audio-Visual Speech Recognition
Youssef Mroueh, Etienne Marcheret, Vaibhava Goel
cs.CLcs.LGarXiv:1501.05396v12015PrivBench: A Holistic and Modular Benchmarking Platform for Evaluating Text-to-Text Privatization
Stephen Meisenbacher, Andreea-Elena Bodea, Ahmet Bilal Akın +3
cs.CLarXiv:2608.29624v12026Memory-First Fact-Checking: A Knowledge-Graph-Grounded Multi-Agent System for Misinformation Detection
Amelia Petrenciuc, Alexandru Lecu, Adrian Groza
cs.CLcs.AIarXiv:2608.29617v12026Head-to-Tail: How Knowledgeable are Large Language Models (LLMs)? A.K.A. Will LLMs Replace Knowledge Graphs?
Kai Sun, Yifan Ethan Xu, Hanwen Zha +2
cs.CLarXiv:2308.10168v22023Agent Zero Memory: Provenance-Aware Long-Term Memory for LLM Agents
Ming Wu, Pengyuan Zhu
cs.CLarXiv:2608.29606v12026