Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,521 to 5,580 of 11,238

  1. A Deep Reinforcement Learning Chatbot

    Iulian V. Serban, Chinnadhurai Sankar, Mathieu Germain +15

    cs.CLcs.AIcs.LGarXiv:1709.02349v22017
  2. Generative vs. Encoder Models for Multilingual NER: A Comprehensive Empirical Study on Naamapadam

    Jakkala Mahesh, Jatavath Shravan Kumar, Komalla Shivani +1

    cs.CLarXiv:2608.29959v12026
  3. Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization

    Weiyun Wang, Zhe Chen, Wenhai Wang +8

    cs.CLcs.CVarXiv:2411.10442v22024
  4. Detecting Hidden Chain-of-Thought in Large Language Models with Linguistic, Behavioral, and Mechanistic Indicators

    Armaan Singh, Ryan Trinh Le, Jasmine Kaur +5

    cs.CLarXiv:2608.29956v12026
  5. XQDT: eXplainable and Quantitative Data-Text Alignment Metric with Feedback Signals

    Kun Efimov-Zhang, Yifei Song, Claire Gardent

    cs.CLarXiv:2608.29948v12026
  6. When Safety Speaks a Language: A Mechanistic Analysis of Safety-Language Identity Entanglement in LLMs

    Apoorva Upadhyaya, Sandipan Sikdar

    cs.CLcs.LGarXiv:2608.29936v12026
  7. Compression-Aware Abstention: Teaching LLMs to Refuse When KV-Compression Masks Remove Answer Evidence

    Mohammadali Khodabandehlou, Bhaskar Krishnamachari

    cs.CLcs.LGarXiv:2608.29934v12026
  8. Sleight of Word Benchmark: Can Language Models Notice If Their Own Output Was Tampered With?

    Alberto Cetoli

    cs.CLcs.AIarXiv:2608.29921v12026
  9. Cognitive Graph for Multi-Hop Reading Comprehension at Scale

    Ming Ding, Chang Zhou, Qibin Chen +2

    cs.CLarXiv:1905.05460v22019
  10. When Less is More: Understanding When Token Filtering Helps and Fails in AI-generated Text Detection

    Xiaoyang Han, Lvxiaowei Xu, Ming Cai

    cs.CLcs.AIarXiv:2608.29903v12026
  11. En-ViMedNER: An English-Vietnamese Parallel Biomedical Corpus with UMLS Semantic Type Annotations

    Nhu Vo, Phuong Nguyen, Nu Uyen Phuong Le +4

    cs.CLarXiv:2608.29890v12026
  12. Exploring wav2vec 2.0 on speaker verification and language identification

    Zhiyun Fan, Meng Li, Shiyu Zhou +1

    cs.SDcs.CLeess.ASarXiv:2012.06185v22020
  13. Toward Optimal Feature Selection in Naive Bayes for Text Categorization

    Bo Tang, Steven Kay, Haibo He

    stat.MLcs.CLcs.IRarXiv:1602.02850v12016
  14. Improving Argument Saliency Coverage in Small LLMs for Long Legal Opinion Summarization via Sequence-Level Distillation

    Mohamed Elaraby, Ahmed Elhady, Diane Litman

    cs.CLarXiv:2608.29884v12026
  15. mPLUG-2: A Modularized Multi-modal Foundation Model Across Text, Image and Video

    Haiyang Xu, Qinghao Ye, Ming Yan +12

    cs.CVcs.CLcs.MMarXiv:2302.00402v12023
  16. SkillForge: Compositional Skill Synthesis with Verification-in-the-Loop for Generating Formally Verified Dafny Programs

    Yanming Liu, Xinyue Peng, Jiannan Cao +2

    cs.CLcs.PLarXiv:2608.29841v12026
  17. Influence-Directed Distillation: Solving the Diversity Bottleneck in Sampled-Token On-Policy Distillation

    Run Yang, Runpeng Dai, Jie Sun +5

    cs.CLcs.LGarXiv:2608.29846v12026
  18. REIGN: Refurbished Embeddings with Integrated Guidance Networks for Efficient Context-Length Scaling

    Devrim Çavuşoğlu, Emre Akbaş

    cs.CLcs.AIcs.IRarXiv:2608.29899v12026
  19. Taskmaster-1: Toward a Realistic and Diverse Dialog Dataset

    Bill Byrne, Karthik Krishnamoorthi, Chinnadhurai Sankar +7

    cs.CLcs.AIcs.LGarXiv:1909.05358v12019
  20. When History Is Multimodal: Rethinking Context Management for Long-Horizon Agents

    Jiaqi Su, Cong Pang, Jiawei Hong +4

    cs.CLarXiv:2608.29897v12026
  21. Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

    Zhiyuan Zhao, Bin Wang, Linke Ouyang +3

    cs.CVcs.CLarXiv:2311.16839v22023
  22. Code Completion with Neural Attention and Pointer Networks

    Jian Li, Yue Wang, Michael R. Lyu +1

    cs.CLcs.SEarXiv:1711.09573v22017
  23. Language model compression with weighted low-rank factorization

    Yen-Chang Hsu, Ting Hua, Sungen Chang +3

    cs.LGcs.AIcs.CLarXiv:2207.00112v12022
  24. VibeJam: A User Study Platform for Web Development with Agents

    Nishant Balepur, Connor Baumler, Valerie Chen +3

    cs.CLarXiv:2608.29889v12026
  25. Check The Scoreboard: An Analysis of Scoring Schemes on Multiple-Choice Evaluation

    Nishant Balepur, Paiheng Xu, Wei Ai +3

    cs.CLarXiv:2608.29887v12026
  26. ManGo: Manga Active Narrative Grounding Optimization

    Hao Qiu, Junyan Wang, Zheyuan Liu +4

    cs.CLarXiv:2608.29865v12026
  27. A Systematic Study and Comprehensive Evaluation of ChatGPT on Benchmark Datasets

    Md Tahmid Rahman Laskar, M Saiful Bari, Mizanur Rahman +3

    cs.CLcs.AIcs.LGarXiv:2305.18486v42023
  28. GenRubric: Self-Evolving Rubric Generation for Scalable LLM Evaluation

    Yifan Chen, Haitao Li, Qingyao Ai +4

    cs.CLarXiv:2608.29856v12026
  29. A Comprehensive Survey on Linguistic Steganography: Methods, Countermeasures, Evaluation, and Challenges

    Ruiyi Yan, Chenhui Chu, Zhongliang Yang +1

    cs.CRcs.CLarXiv:2608.29077v12026
  30. Distributional Validity and Calibration of a Korean Synthetic Persona Panel for Digital and AI Service Use: A Secondary-Data Validation Against the Korea Media Panel Survey

    Howard Kim, Keun Tae Cho

    cs.CYcs.CLarXiv:2608.28615v12026
  31. Zhongjing: Enhancing the Chinese Medical Capabilities of Large Language Model through Expert Feedback and Real-world Multi-turn Dialogue

    Songhua Yang, Hanjie Zhao, Senbin Zhu +4

    cs.CLarXiv:2308.03549v32023
  32. RegDivergence-101: An LLM Benchmark for Cross-Jurisdiction Regulatory Contradiction Detection in Life Sciences

    Chuchu Wu, Zhiyin Zhou, Jingzhuo Hu +1

    cs.AIcs.CLarXiv:2608.28607v12026
  33. AdaPlanner: Adaptive Planning from Feedback with Language Models

    Haotian Sun, Yuchen Zhuang, Lingkai Kong +2

    cs.CLcs.AIcs.LGarXiv:2305.16653v12023
  34. QuALITY: Question Answering with Long Input Texts, Yes!

    Richard Yuanzhe Pang, Alicia Parrish, Nitish Joshi +8

    cs.CLarXiv:2112.08608v22021
  35. A^2Agent: Action-Aware Reinforcement Learning for Repository-Level Code Localization Agents

    Doyeon Kim, Suyoung Bae, Yumin Lee +1

    cs.CLcs.SEarXiv:2608.29831v12026
  36. ReTrace: Rejected-Trajectory Conditioning for Speculative Decoding

    Luxi Lin, Zhanpeng Zeng, Shuang Peng +2

    cs.CLarXiv:2608.29748v12026
  37. Target-Speaker Voice Activity Detection: a Novel Approach for Multi-Speaker Diarization in a Dinner Party Scenario

    Ivan Medennikov, Maxim Korenevsky, Tatiana Prisyach +9

    eess.AScs.CLcs.SDarXiv:2005.07272v22020
  38. EVAR: Evidence-Validated Hypothesis Admission for Budget-Aware Narrative Reasoning

    Peilin Liu, Zhiquan Ji, Jinglong Ping

    cs.CLarXiv:2608.29835v12026
  39. Generating Fact Checking Explanations

    Pepa Atanasova, Jakob Grue Simonsen, Christina Lioma +1

    cs.CLcs.AIcs.LGarXiv:2004.05773v12020
  40. You Know What I Mean: A Benchmark for Agentic Conversational Reference Grounding

    Karen Fuchs, Uri Katz, Yoav Goldberg

    cs.CLcs.IRarXiv:2608.29834v12026
  41. Improving Factual Completeness and Consistency of Image-to-Text Radiology Report Generation

    Yasuhide Miura, Yuhao Zhang, Emily Bao Tsai +2

    cs.CLarXiv:2010.10042v22020
  42. Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming

    Zhifei Xie, Changqiao Wu

    cs.AIcs.CLcs.HCarXiv:2408.16725v32024
  43. R$^2$A: Learning Persona Policies Through Persona Representation Learning and Runtime Alignment

    Mohan Zhang, Chengsong You, Xiaoyu Cao +4

    cs.CLcs.AIarXiv:2608.29798v12026
  44. SentiCap: Generating Image Descriptions with Sentiments

    Alexander Mathews, Lexing Xie, Xuming He

    cs.CVcs.CLarXiv:1510.01431v22015
  45. Multi-Modal Hallucination Control by Visual Information Grounding

    Alessandro Favero, Luca Zancato, Matthew Trager +5

    cs.CVcs.CLcs.LGarXiv:2403.14003v12024
  46. Evaluating the Capabilities of LLMs for Persuasive Dialogue

    Jordan Robinson, Angus R. Williams, Katie Atkinson +1

    cs.CLarXiv:2608.29738v12026
  47. A Review of Large Language Models and Autonomous Agents in Chemistry

    Mayk Caldas Ramos, Christopher J. Collison, Andrew D. White

    cs.LGcs.AIcs.CLarXiv:2407.01603v32024
  48. HiVe: Beyond Static Prompts for Multitask Learning via Hierarchy-based Vertical Mixture-of-Experts

    HyeonJik Bae, Minyeol Kim, Susik Yoon

    cs.CLarXiv:2608.29790v12026
  49. DVBench: Benchmarking MLLMs for Understanding Dynamic Charts and Narratives in Data Videos

    Bomiao Wang, Zekai Shao, Jiexiang Lan +3

    cs.CLarXiv:2608.29711v12026
  50. The Depth Flow of Token Representations Is Nonlinear and Does Not Descend Its Own Density

    Alexandre Quemy

    cs.CLarXiv:2608.29706v12026
  51. A Hub of Short Rows Inflates Intrinsic Dimension Estimation of Token Embeddings

    Alexandre Quemy

    cs.CLarXiv:2608.29702v12026
  52. ACTD: Anchor-Based Cross-Tokenizer Distillation with Residual Regularization

    Huiyi Zhang, Zijian Li, Xiaocheng Feng +4

    cs.CLarXiv:2608.29662v12026
  53. MI-Distillation: Selecting from Model-Interpolated Instruct-Reasoning Data Spectrum for Chain-of-Thought Distillation

    Yangsong Lan, Renkai Hu, HongKai Zheng +4

    cs.CLcs.AIarXiv:2608.29623v12026
  54. How You Ask Shapes What You Get: A Theory-Seeded Measurement of Articulation in Advice-Seeking LLM Conversations

    Juneha Baek, Suhyeon Lee, Donghyuk Shin

    cs.CLarXiv:2608.29591v12026
  55. Gmail Smart Compose: Real-Time Assisted Writing

    Mia Xu Chen, Benjamin N Lee, Gagan Bansal +9

    cs.CLcs.LGarXiv:1906.00080v12019
  56. Deep Multimodal Learning for Audio-Visual Speech Recognition

    Youssef Mroueh, Etienne Marcheret, Vaibhava Goel

    cs.CLcs.LGarXiv:1501.05396v12015
  57. PrivBench: A Holistic and Modular Benchmarking Platform for Evaluating Text-to-Text Privatization

    Stephen Meisenbacher, Andreea-Elena Bodea, Ahmet Bilal Akın +3

    cs.CLarXiv:2608.29624v12026
  58. Memory-First Fact-Checking: A Knowledge-Graph-Grounded Multi-Agent System for Misinformation Detection

    Amelia Petrenciuc, Alexandru Lecu, Adrian Groza

    cs.CLcs.AIarXiv:2608.29617v12026
  59. Head-to-Tail: How Knowledgeable are Large Language Models (LLMs)? A.K.A. Will LLMs Replace Knowledge Graphs?

    Kai Sun, Yifan Ethan Xu, Hanwen Zha +2

    cs.CLarXiv:2308.10168v22023
  60. Agent Zero Memory: Provenance-Aware Long-Term Memory for LLM Agents

    Ming Wu, Pengyuan Zhu

    cs.CLarXiv:2608.29606v12026