Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,141 to 1,200 of 11,238
If It's Not Buggy, Don't Fix It: On the Dynamics of Iterative Bug-fixing with LLMs
Xietao Wang-Lin, Anton Isopoussu, Louis Mahon
cs.SEcs.CLarXiv:2609.10123v12026Long Time No See! Open-Domain Conversation with Long-Term Persona Memory
Xinchao Xu, Zhibin Gou, Wenquan Wu +4
cs.CLarXiv:2203.05797v22022Improving Knowledge Graph Embedding Using Simple Constraints
Boyang Ding, Quan Wang, Bin Wang +1
cs.AIcs.CLarXiv:1805.02408v22018How to Train Long-Context Language Models (Effectively)
Tianyu Gao, Alexander Wettig, Howard Yen +1
cs.CLcs.LGarXiv:2410.02660v42024Vague2Detect: Handling Ambiguous Prompts in Knowledge-Based Open-World Detection
Ibrohimjon Muminov, Jihie Kim
cs.CVcs.CLcs.LGarXiv:2609.09949v12026Who Are They to Each Other? Multi-Agent Reasoning for Speaker Relationship Inference
Yaohan Guan, Yen-Ju Lu, Yuzhe Wang +5
cs.MAcs.CLcs.SDarXiv:2609.09628v12026Video-3D LLM: Learning Position-Aware Video Representation for 3D Scene Understanding
Duo Zheng, Shijia Huang, Liwei Wang
cs.CVcs.CLarXiv:2412.00493v22024What Does MMLU Actually Measure? A Psychometric Audit of Difficulty Structure in Aggregate Benchmark Scores
Dana Paquin, Riddhiman Jain
math.NTcs.CLarXiv:2609.09372v12026MLLMs Hallucinate when Information Distribution Drifts in Synergy Heads
Meng'en Qin, Junye Chen, Jucheng Liu +3
cs.CVcs.CLarXiv:2609.09206v12026Building Multilingual Bridges: Data Mixing as the Pillar of Generalization for In-Language Reasoning
Mehrnaz Mofakhami, Ananya Sahu, Alejandro R. Salamanca +5
cs.CLarXiv:2609.10445v12026Do speech foundation models really learn words?
Robin Huo, Ewan Dunbar
cs.CLcs.SDarXiv:2609.10434v12026Deterministic Prompting for Speaker-Stable Low-Resource Greek TTS
Georgios Syllas, Efthymios Georgiou, Kosmas Kritsis +1
cs.SDcs.CLcs.LGarXiv:2609.10022v12026PELM: Power Efficient On-Device LLM Inference with Speculative Decoding and Dynamic Voltage Frequency Scaling
Weisi Yang, Stephen Xia
cs.LGcs.CLcs.OSarXiv:2609.09662v12026Cross-Target Stance Classification with Self-Attention Networks
Chang Xu, Cecile Paris, Surya Nepal +1
cs.CLcs.AIarXiv:1805.06593v22018An Efficient and Effective Agentic Group Shilling Attack on Recommender Systems
Quoc Viet Nguyen, Trinh Pham, Viet Huynh +4
cs.CRcs.CLarXiv:2609.09551v12026ScrabbleGAN: Semi-Supervised Varying Length Handwritten Text Generation
Sharon Fogel, Hadar Averbuch-Elor, Sarel Cohen +2
cs.CVcs.CLcs.LGarXiv:2003.10557v12020IdeaAMBIG: Benchmarking Implementation-Critical Gaps in Research-Idea Specifications
Yiling Ma, Yilun Zhao, Sihong Wu +2
cs.CLarXiv:2609.10539v12026EduChat: A Large-Scale Language Model-based Chatbot System for Intelligent Education
Yuhao Dan, Zhikai Lei, Yiyang Gu +13
cs.CLarXiv:2308.02773v12023Understanding Catastrophic Forgetting in Language Models via Implicit Inference
Suhas Kotha, Jacob Mitchell Springer, Aditi Raghunathan
cs.CLcs.LGarXiv:2309.10105v22023Regression Transformer: Concurrent sequence regression and generation for molecular language modeling
Jannis Born, Matteo Manica
cs.LGcs.AIcs.CLarXiv:2202.01338v32022Rosetta at AlexandriaX-2026: LoRA-Adapted NileChat for Context-Aware Dialectal Arabic Dialogue Translation
Nada Esmaeil, Fathima Rena, Sibi Subhash +4
cs.CLarXiv:2609.10395v12026On-Policy Distillation for Vision-Language Model Adaptation, an Effective Paradigm on Low-Quality Multimodal Data
Hongyuan Zhang, Xianda Guo, Yanlun Peng +6
cs.CLarXiv:2609.10321v12026AI Hallucinations: A Misnomer Worth Clarifying
Negar Maleki, Balaji Padmanabhan, Kaushik Dutta
cs.CLcs.AIarXiv:2401.06796v12024KVShareArena: KV-Cache Reuse Across Contexts and Model Checkpoints
Xi Shi, Qian Lou
cs.CLarXiv:2609.10266v12026Two-Token Features and Small-Large Ensembles for VLM Hallucination Detection
Eli Schwartz
cs.CLarXiv:2609.10244v12026From Word Models to World Models: Translating from Natural Language to the Probabilistic Language of Thought
Lionel Wong, Gabriel Grand, Alexander K. Lew +4
cs.CLcs.AIcs.SCarXiv:2306.12672v22023The Answer Path and the Grounding Instruction in LLM Question Answering over Knowledge Graphs
Arquimedes Canedo
cs.CLcs.IRarXiv:2609.10237v12026Through the Looking Glass: Directly Reading and Writing Transformers
Mark Oskin
cs.CLcs.LGarXiv:2609.10210v12026Politics of Feelings: Emotional Expression and Legislative Effectiveness in the U.S. Congress
Segun Aroyehun
cs.CLarXiv:2609.10198v12026Who Argues What? Joint Argument-Entity Detection and Classification in Political Debates
Lucio La Cava, Stefano Francesco Monea, Sergio Greco
cs.CLarXiv:2609.10192v12026The Semantic Bottleneck: Leveraging Semantic Representations for Non-Invasive Speech Decoding
Gilad D. Landau, Dulhan Jayalath, Oiwi Parker Jones
cs.CLcs.LGarXiv:2609.10296v12026VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning
Han Lin, Abhay Zala, Jaemin Cho +1
cs.CVcs.AIcs.CLarXiv:2309.15091v22023Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization
Navonil Majumder, Chia-Yu Hung, Deepanway Ghosal +3
cs.SDcs.AIcs.CLarXiv:2404.09956v42024Posterior calibration and exploratory analysis for natural language processing models
Khanh Nguyen, Brendan O'Connor
cs.CLarXiv:1508.05154v22015One Small Step for Generative AI, One Giant Leap for AGI: A Complete Survey on ChatGPT in AIGC Era
Chaoning Zhang, Chenshuang Zhang, Chenghao Li +13
cs.CYcs.AIcs.CLarXiv:2304.06488v12023VLX-VR: An Agentic-Aware Video Reasoning Model
Sheng Li, Peng Liu, Qianqian Zhang +1
cs.CLcs.CVarXiv:2609.09985v12026Solving Math Word Problems by Combining Language Models With Symbolic Solvers
Joy He-Yueya, Gabriel Poesia, Rose E. Wang +1
cs.CLcs.AIarXiv:2304.09102v12023From Retrieval to Weights: Parametric Individualization of Small Language Models with Individual Text Corpora
Christoph Wigbels, Ali Abusaleh, Markus T. Jansen +2
cs.CLcs.IRarXiv:2609.10155v12026YallaMorph: A Benchmark for Evaluating Arabic Morphological Generation in Large Language Models
Mahmoud Reda, Salam Khalifa, Reham Marzouk +1
cs.CLarXiv:2609.10153v12026ProbPlug: A Plugin Uncertainty Network for Reliable Confidence in LLM Binary Classification
Jianzong Wang, Chuhang Liu, Botao Zhao +6
cs.CLarXiv:2609.10122v12026Data-Centric Post-Training for Financial Reasoning: Mining, Distillation, and Verifiable Learning
Zhirayr Hayrapetyan, Andrei Kalmykov, Denis Kokosinskii +2
cs.CLarXiv:2609.10113v12026Transfer Learning from Adult to Children for Speech Recognition: Evaluation, Analysis and Recommendations
Prashanth Gurunath Shivakumar, Panayiotis Georgiou
eess.AScs.CLcs.SDarXiv:1805.03322v12018MedDeID enables locally governed clinical-text de-identification from real or synthetic training data
Stig Hellemans, Tom Stroobants, Elyne Scheurwegs +3
cs.CLcs.LGarXiv:2609.10049v12026Stable Answers, Unfinished Reasoning: Why Self-Consensus Is Not a Safe Early-Exit Signal
Yunxiang Mo, Donghao Zhao, Hejia Geng
cs.CLarXiv:2609.09989v12026SalamandraTA at WMT 2026 Terminology Shared Task: Hard Examples Are Better Teachers
Xixian Liao, Maite Melero
cs.CLarXiv:2609.09999v12026Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing
Ye Tian, Baolin Peng, Linfeng Song +4
cs.CLcs.LGarXiv:2404.12253v22024Multi-Functional Embedding Models for Funder Name Disambiguation in Scientific Publication Records
Kanyao Han, Zhiwen You, Jinseok Kim +1
cs.CLarXiv:2609.09984v12026Towards Stress-Aware Sentence-Level Filipino G2P With Weakly-Supervised ByT5 Fine-Tuning
Lorenz Bernard Marqueses, Paulo Grane Gabriel Silva, Chastine Cabatay +2
cs.CLarXiv:2609.09974v12026Train Large, Then Compress: Rethinking Model Size for Efficient Training and Inference of Transformers
Zhuohan Li, Eric Wallace, Sheng Shen +4
cs.CLcs.LGarXiv:2002.11794v22020$S^3$-Bench: Evaluating Speech Interaction Models as Scientific Voice Assistants
Heyang Liu, Jiayi Huang, Wenyang Xiao +8
cs.CLarXiv:2609.09852v12026HyperTrace: Hypothesis-Based Preference Tracing for Online LLM Personalization
Jianzhi Shen, Keyu Mao, Minghao Shao +7
cs.CLarXiv:2609.09835v12026Towards Automated Factchecking: Developing an Annotation Schema and Benchmark for Consistent Automated Claim Detection
Lev Konstantinovskiy, Oliver Price, Mevan Babakar +1
cs.CLarXiv:1809.08193v220185-Dialects-BN: Unmasking the Impact of Transliteration on Bangla Dialectal LLMs
Md Mahir Jawad, Galib Mahmud Jim, Rafid Ahmed +3
cs.CLarXiv:2609.09964v12026Contrastive Projection: Reading Transformer Internals by Differencing Logit Lenses
Olli Tuomi
cs.CLarXiv:2609.09902v12026A Critical Evaluation of Evaluations for Long-form Question Answering
Fangyuan Xu, Yixiao Song, Mohit Iyyer +1
cs.CLarXiv:2305.18201v12023Explicit Sparse Transformer: Concentrated Attention Through Explicit Selection
Guangxiang Zhao, Junyang Lin, Zhiyuan Zhang +3
cs.CLcs.LGarXiv:1912.11637v12019Deep and shallow biases in language models
An Vo, Vy Tuong Dang, Khai-Nguyen Nguyen +4
cs.CLarXiv:2609.09901v12026Large Language Models can Strategically Deceive their Users when Put Under Pressure
Jérémy Scheurer, Mikita Balesni, Marius Hobbhahn
cs.CLcs.AIcs.LGarXiv:2311.07590v42023Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models
Carson Denison, Monte MacDiarmid, Fazl Barez +11
cs.AIcs.CLarXiv:2406.10162v32024Leveraging Fine-grained Error Correction in Korean Speech Recognition for Consultation Services
Yonghyun Jun, Jimin Lee, Hwan Chang +3
cs.CLarXiv:2609.09889v12026