Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,141 to 1,200 of 11,238

  1. If It's Not Buggy, Don't Fix It: On the Dynamics of Iterative Bug-fixing with LLMs

    Xietao Wang-Lin, Anton Isopoussu, Louis Mahon

    cs.SEcs.CLarXiv:2609.10123v12026
  2. Long Time No See! Open-Domain Conversation with Long-Term Persona Memory

    Xinchao Xu, Zhibin Gou, Wenquan Wu +4

    cs.CLarXiv:2203.05797v22022
  3. Improving Knowledge Graph Embedding Using Simple Constraints

    Boyang Ding, Quan Wang, Bin Wang +1

    cs.AIcs.CLarXiv:1805.02408v22018
  4. How to Train Long-Context Language Models (Effectively)

    Tianyu Gao, Alexander Wettig, Howard Yen +1

    cs.CLcs.LGarXiv:2410.02660v42024
  5. Vague2Detect: Handling Ambiguous Prompts in Knowledge-Based Open-World Detection

    Ibrohimjon Muminov, Jihie Kim

    cs.CVcs.CLcs.LGarXiv:2609.09949v12026
  6. Who Are They to Each Other? Multi-Agent Reasoning for Speaker Relationship Inference

    Yaohan Guan, Yen-Ju Lu, Yuzhe Wang +5

    cs.MAcs.CLcs.SDarXiv:2609.09628v12026
  7. Video-3D LLM: Learning Position-Aware Video Representation for 3D Scene Understanding

    Duo Zheng, Shijia Huang, Liwei Wang

    cs.CVcs.CLarXiv:2412.00493v22024
  8. What Does MMLU Actually Measure? A Psychometric Audit of Difficulty Structure in Aggregate Benchmark Scores

    Dana Paquin, Riddhiman Jain

    math.NTcs.CLarXiv:2609.09372v12026
  9. MLLMs Hallucinate when Information Distribution Drifts in Synergy Heads

    Meng'en Qin, Junye Chen, Jucheng Liu +3

    cs.CVcs.CLarXiv:2609.09206v12026
  10. Building Multilingual Bridges: Data Mixing as the Pillar of Generalization for In-Language Reasoning

    Mehrnaz Mofakhami, Ananya Sahu, Alejandro R. Salamanca +5

    cs.CLarXiv:2609.10445v12026
  11. Do speech foundation models really learn words?

    Robin Huo, Ewan Dunbar

    cs.CLcs.SDarXiv:2609.10434v12026
  12. Deterministic Prompting for Speaker-Stable Low-Resource Greek TTS

    Georgios Syllas, Efthymios Georgiou, Kosmas Kritsis +1

    cs.SDcs.CLcs.LGarXiv:2609.10022v12026
  13. PELM: Power Efficient On-Device LLM Inference with Speculative Decoding and Dynamic Voltage Frequency Scaling

    Weisi Yang, Stephen Xia

    cs.LGcs.CLcs.OSarXiv:2609.09662v12026
  14. Cross-Target Stance Classification with Self-Attention Networks

    Chang Xu, Cecile Paris, Surya Nepal +1

    cs.CLcs.AIarXiv:1805.06593v22018
  15. An Efficient and Effective Agentic Group Shilling Attack on Recommender Systems

    Quoc Viet Nguyen, Trinh Pham, Viet Huynh +4

    cs.CRcs.CLarXiv:2609.09551v12026
  16. ScrabbleGAN: Semi-Supervised Varying Length Handwritten Text Generation

    Sharon Fogel, Hadar Averbuch-Elor, Sarel Cohen +2

    cs.CVcs.CLcs.LGarXiv:2003.10557v12020
  17. IdeaAMBIG: Benchmarking Implementation-Critical Gaps in Research-Idea Specifications

    Yiling Ma, Yilun Zhao, Sihong Wu +2

    cs.CLarXiv:2609.10539v12026
  18. EduChat: A Large-Scale Language Model-based Chatbot System for Intelligent Education

    Yuhao Dan, Zhikai Lei, Yiyang Gu +13

    cs.CLarXiv:2308.02773v12023
  19. Understanding Catastrophic Forgetting in Language Models via Implicit Inference

    Suhas Kotha, Jacob Mitchell Springer, Aditi Raghunathan

    cs.CLcs.LGarXiv:2309.10105v22023
  20. Regression Transformer: Concurrent sequence regression and generation for molecular language modeling

    Jannis Born, Matteo Manica

    cs.LGcs.AIcs.CLarXiv:2202.01338v32022
  21. Rosetta at AlexandriaX-2026: LoRA-Adapted NileChat for Context-Aware Dialectal Arabic Dialogue Translation

    Nada Esmaeil, Fathima Rena, Sibi Subhash +4

    cs.CLarXiv:2609.10395v12026
  22. On-Policy Distillation for Vision-Language Model Adaptation, an Effective Paradigm on Low-Quality Multimodal Data

    Hongyuan Zhang, Xianda Guo, Yanlun Peng +6

    cs.CLarXiv:2609.10321v12026
  23. AI Hallucinations: A Misnomer Worth Clarifying

    Negar Maleki, Balaji Padmanabhan, Kaushik Dutta

    cs.CLcs.AIarXiv:2401.06796v12024
  24. KVShareArena: KV-Cache Reuse Across Contexts and Model Checkpoints

    Xi Shi, Qian Lou

    cs.CLarXiv:2609.10266v12026
  25. Two-Token Features and Small-Large Ensembles for VLM Hallucination Detection

    Eli Schwartz

    cs.CLarXiv:2609.10244v12026
  26. From Word Models to World Models: Translating from Natural Language to the Probabilistic Language of Thought

    Lionel Wong, Gabriel Grand, Alexander K. Lew +4

    cs.CLcs.AIcs.SCarXiv:2306.12672v22023
  27. The Answer Path and the Grounding Instruction in LLM Question Answering over Knowledge Graphs

    Arquimedes Canedo

    cs.CLcs.IRarXiv:2609.10237v12026
  28. Through the Looking Glass: Directly Reading and Writing Transformers

    Mark Oskin

    cs.CLcs.LGarXiv:2609.10210v12026
  29. Politics of Feelings: Emotional Expression and Legislative Effectiveness in the U.S. Congress

    Segun Aroyehun

    cs.CLarXiv:2609.10198v12026
  30. Who Argues What? Joint Argument-Entity Detection and Classification in Political Debates

    Lucio La Cava, Stefano Francesco Monea, Sergio Greco

    cs.CLarXiv:2609.10192v12026
  31. The Semantic Bottleneck: Leveraging Semantic Representations for Non-Invasive Speech Decoding

    Gilad D. Landau, Dulhan Jayalath, Oiwi Parker Jones

    cs.CLcs.LGarXiv:2609.10296v12026
  32. VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

    Han Lin, Abhay Zala, Jaemin Cho +1

    cs.CVcs.AIcs.CLarXiv:2309.15091v22023
  33. Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization

    Navonil Majumder, Chia-Yu Hung, Deepanway Ghosal +3

    cs.SDcs.AIcs.CLarXiv:2404.09956v42024
  34. Posterior calibration and exploratory analysis for natural language processing models

    Khanh Nguyen, Brendan O'Connor

    cs.CLarXiv:1508.05154v22015
  35. One Small Step for Generative AI, One Giant Leap for AGI: A Complete Survey on ChatGPT in AIGC Era

    Chaoning Zhang, Chenshuang Zhang, Chenghao Li +13

    cs.CYcs.AIcs.CLarXiv:2304.06488v12023
  36. VLX-VR: An Agentic-Aware Video Reasoning Model

    Sheng Li, Peng Liu, Qianqian Zhang +1

    cs.CLcs.CVarXiv:2609.09985v12026
  37. Solving Math Word Problems by Combining Language Models With Symbolic Solvers

    Joy He-Yueya, Gabriel Poesia, Rose E. Wang +1

    cs.CLcs.AIarXiv:2304.09102v12023
  38. From Retrieval to Weights: Parametric Individualization of Small Language Models with Individual Text Corpora

    Christoph Wigbels, Ali Abusaleh, Markus T. Jansen +2

    cs.CLcs.IRarXiv:2609.10155v12026
  39. YallaMorph: A Benchmark for Evaluating Arabic Morphological Generation in Large Language Models

    Mahmoud Reda, Salam Khalifa, Reham Marzouk +1

    cs.CLarXiv:2609.10153v12026
  40. ProbPlug: A Plugin Uncertainty Network for Reliable Confidence in LLM Binary Classification

    Jianzong Wang, Chuhang Liu, Botao Zhao +6

    cs.CLarXiv:2609.10122v12026
  41. Data-Centric Post-Training for Financial Reasoning: Mining, Distillation, and Verifiable Learning

    Zhirayr Hayrapetyan, Andrei Kalmykov, Denis Kokosinskii +2

    cs.CLarXiv:2609.10113v12026
  42. Transfer Learning from Adult to Children for Speech Recognition: Evaluation, Analysis and Recommendations

    Prashanth Gurunath Shivakumar, Panayiotis Georgiou

    eess.AScs.CLcs.SDarXiv:1805.03322v12018
  43. MedDeID enables locally governed clinical-text de-identification from real or synthetic training data

    Stig Hellemans, Tom Stroobants, Elyne Scheurwegs +3

    cs.CLcs.LGarXiv:2609.10049v12026
  44. Stable Answers, Unfinished Reasoning: Why Self-Consensus Is Not a Safe Early-Exit Signal

    Yunxiang Mo, Donghao Zhao, Hejia Geng

    cs.CLarXiv:2609.09989v12026
  45. SalamandraTA at WMT 2026 Terminology Shared Task: Hard Examples Are Better Teachers

    Xixian Liao, Maite Melero

    cs.CLarXiv:2609.09999v12026
  46. Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing

    Ye Tian, Baolin Peng, Linfeng Song +4

    cs.CLcs.LGarXiv:2404.12253v22024
  47. Multi-Functional Embedding Models for Funder Name Disambiguation in Scientific Publication Records

    Kanyao Han, Zhiwen You, Jinseok Kim +1

    cs.CLarXiv:2609.09984v12026
  48. Towards Stress-Aware Sentence-Level Filipino G2P With Weakly-Supervised ByT5 Fine-Tuning

    Lorenz Bernard Marqueses, Paulo Grane Gabriel Silva, Chastine Cabatay +2

    cs.CLarXiv:2609.09974v12026
  49. Train Large, Then Compress: Rethinking Model Size for Efficient Training and Inference of Transformers

    Zhuohan Li, Eric Wallace, Sheng Shen +4

    cs.CLcs.LGarXiv:2002.11794v22020
  50. $S^3$-Bench: Evaluating Speech Interaction Models as Scientific Voice Assistants

    Heyang Liu, Jiayi Huang, Wenyang Xiao +8

    cs.CLarXiv:2609.09852v12026
  51. HyperTrace: Hypothesis-Based Preference Tracing for Online LLM Personalization

    Jianzhi Shen, Keyu Mao, Minghao Shao +7

    cs.CLarXiv:2609.09835v12026
  52. Towards Automated Factchecking: Developing an Annotation Schema and Benchmark for Consistent Automated Claim Detection

    Lev Konstantinovskiy, Oliver Price, Mevan Babakar +1

    cs.CLarXiv:1809.08193v22018
  53. 5-Dialects-BN: Unmasking the Impact of Transliteration on Bangla Dialectal LLMs

    Md Mahir Jawad, Galib Mahmud Jim, Rafid Ahmed +3

    cs.CLarXiv:2609.09964v12026
  54. Contrastive Projection: Reading Transformer Internals by Differencing Logit Lenses

    Olli Tuomi

    cs.CLarXiv:2609.09902v12026
  55. A Critical Evaluation of Evaluations for Long-form Question Answering

    Fangyuan Xu, Yixiao Song, Mohit Iyyer +1

    cs.CLarXiv:2305.18201v12023
  56. Explicit Sparse Transformer: Concentrated Attention Through Explicit Selection

    Guangxiang Zhao, Junyang Lin, Zhiyuan Zhang +3

    cs.CLcs.LGarXiv:1912.11637v12019
  57. Deep and shallow biases in language models

    An Vo, Vy Tuong Dang, Khai-Nguyen Nguyen +4

    cs.CLarXiv:2609.09901v12026
  58. Large Language Models can Strategically Deceive their Users when Put Under Pressure

    Jérémy Scheurer, Mikita Balesni, Marius Hobbhahn

    cs.CLcs.AIcs.LGarXiv:2311.07590v42023
  59. Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models

    Carson Denison, Monte MacDiarmid, Fazl Barez +11

    cs.AIcs.CLarXiv:2406.10162v32024
  60. Leveraging Fine-grained Error Correction in Korean Speech Recognition for Consultation Services

    Yonghyun Jun, Jimin Lee, Hwan Chang +3

    cs.CLarXiv:2609.09889v12026