Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,201 to 1,260 of 11,259

  1. ProbPlug: A Plugin Uncertainty Network for Reliable Confidence in LLM Binary Classification

    Jianzong Wang, Chuhang Liu, Botao Zhao +6

    cs.CLarXiv:2609.10122v12026
  2. Data-Centric Post-Training for Financial Reasoning: Mining, Distillation, and Verifiable Learning

    Zhirayr Hayrapetyan, Andrei Kalmykov, Denis Kokosinskii +2

    cs.CLarXiv:2609.10113v12026
  3. Transfer Learning from Adult to Children for Speech Recognition: Evaluation, Analysis and Recommendations

    Prashanth Gurunath Shivakumar, Panayiotis Georgiou

    eess.AScs.CLcs.SDarXiv:1805.03322v12018
  4. MedDeID enables locally governed clinical-text de-identification from real or synthetic training data

    Stig Hellemans, Tom Stroobants, Elyne Scheurwegs +3

    cs.CLcs.LGarXiv:2609.10049v12026
  5. Stable Answers, Unfinished Reasoning: Why Self-Consensus Is Not a Safe Early-Exit Signal

    Yunxiang Mo, Donghao Zhao, Hejia Geng

    cs.CLarXiv:2609.09989v12026
  6. SalamandraTA at WMT 2026 Terminology Shared Task: Hard Examples Are Better Teachers

    Xixian Liao, Maite Melero

    cs.CLarXiv:2609.09999v12026
  7. Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing

    Ye Tian, Baolin Peng, Linfeng Song +4

    cs.CLcs.LGarXiv:2404.12253v22024
  8. Multi-Functional Embedding Models for Funder Name Disambiguation in Scientific Publication Records

    Kanyao Han, Zhiwen You, Jinseok Kim +1

    cs.CLarXiv:2609.09984v12026
  9. Towards Stress-Aware Sentence-Level Filipino G2P With Weakly-Supervised ByT5 Fine-Tuning

    Lorenz Bernard Marqueses, Paulo Grane Gabriel Silva, Chastine Cabatay +2

    cs.CLarXiv:2609.09974v12026
  10. Train Large, Then Compress: Rethinking Model Size for Efficient Training and Inference of Transformers

    Zhuohan Li, Eric Wallace, Sheng Shen +4

    cs.CLcs.LGarXiv:2002.11794v22020
  11. $S^3$-Bench: Evaluating Speech Interaction Models as Scientific Voice Assistants

    Heyang Liu, Jiayi Huang, Wenyang Xiao +8

    cs.CLarXiv:2609.09852v12026
  12. HyperTrace: Hypothesis-Based Preference Tracing for Online LLM Personalization

    Jianzhi Shen, Keyu Mao, Minghao Shao +7

    cs.CLarXiv:2609.09835v12026
  13. Towards Automated Factchecking: Developing an Annotation Schema and Benchmark for Consistent Automated Claim Detection

    Lev Konstantinovskiy, Oliver Price, Mevan Babakar +1

    cs.CLarXiv:1809.08193v22018
  14. 5-Dialects-BN: Unmasking the Impact of Transliteration on Bangla Dialectal LLMs

    Md Mahir Jawad, Galib Mahmud Jim, Rafid Ahmed +3

    cs.CLarXiv:2609.09964v12026
  15. Contrastive Projection: Reading Transformer Internals by Differencing Logit Lenses

    Olli Tuomi

    cs.CLarXiv:2609.09902v12026
  16. A Critical Evaluation of Evaluations for Long-form Question Answering

    Fangyuan Xu, Yixiao Song, Mohit Iyyer +1

    cs.CLarXiv:2305.18201v12023
  17. Explicit Sparse Transformer: Concentrated Attention Through Explicit Selection

    Guangxiang Zhao, Junyang Lin, Zhiyuan Zhang +3

    cs.CLcs.LGarXiv:1912.11637v12019
  18. Deep and shallow biases in language models

    An Vo, Vy Tuong Dang, Khai-Nguyen Nguyen +4

    cs.CLarXiv:2609.09901v12026
  19. Large Language Models can Strategically Deceive their Users when Put Under Pressure

    Jérémy Scheurer, Mikita Balesni, Marius Hobbhahn

    cs.CLcs.AIcs.LGarXiv:2311.07590v42023
  20. Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models

    Carson Denison, Monte MacDiarmid, Fazl Barez +11

    cs.AIcs.CLarXiv:2406.10162v32024
  21. Leveraging Fine-grained Error Correction in Korean Speech Recognition for Consultation Services

    Yonghyun Jun, Jimin Lee, Hwan Chang +3

    cs.CLarXiv:2609.09889v12026
  22. When Does Defendant Statement Matter? A Study of Bias and Persuasion in LLM-Simulated Jurors

    Cho-Ying Wu

    cs.CLcs.CYarXiv:2609.09887v12026
  23. The Effect of Natural Distribution Shift on Question Answering Models

    John Miller, Karl Krauth, Benjamin Recht +1

    cs.LGcs.CLstat.MLarXiv:2004.14444v12020
  24. MUCnoHARM@GermEval Shared Task 2026: Retrieval-based In-Context Learning for Defamatory Offences, and Where It Falls Short

    Kristin Gnadt, Maximilian Meidinger, Matthias Aßenmacher

    cs.CLarXiv:2609.09791v12026
  25. Iterative Pseudo-Labeling for Speech Recognition

    Qiantong Xu, Tatiana Likhomanenko, Jacob Kahn +3

    cs.CLcs.SDeess.ASarXiv:2005.09267v22020
  26. ROAM: Robust Organization of Atomic Memories for Agents through Semantic Relations

    Jianjie Zheng, Peng Lai, Sijie Cheng +3

    cs.CLarXiv:2609.09778v12026
  27. Is Model Collapse Inevitable? Breaking the Curse of Recursion by Accumulating Real and Synthetic Data

    Matthias Gerstgrasser, Rylan Schaeffer, Apratim Dey +11

    cs.LGcs.AIcs.CLarXiv:2404.01413v22024
  28. SymbolicLight V2: Hybrid Neuromorphic Architecture and Sparse Execution for Low-Energy Language Inference

    Ting Liu

    cs.CLarXiv:2609.09772v12026
  29. The Self-Perception and Political Biases of ChatGPT

    Jérôme Rutinowski, Sven Franke, Jan Endendyk +2

    cs.CYcs.AIcs.CLarXiv:2304.07333v12023
  30. Beyond Top Words: MonoTM for Topic Modeling with Interpretable Monosemantic Features

    Una Joh, Bei Yu

    cs.CLcs.HCarXiv:2609.09575v12026
  31. Intrinsic Dimension Estimation for Robust Detection of AI-Generated Texts

    Eduard Tulchinskii, Kristian Kuznetsov, Laida Kushnareva +5

    cs.CLcs.AIcs.ITarXiv:2306.04723v22023
  32. Eight Methods to Evaluate Robust Unlearning in LLMs

    Aengus Lynch, Phillip Guo, Aidan Ewart +2

    cs.CLarXiv:2402.16835v12024
  33. CARRE: Counterfactual Action Retrieval and Reason Evaluation for Explainable Churn Prescription

    MinJoo Kim, SanJin Park, SeungHwan Cho

    cs.CLarXiv:2609.09766v12026
  34. Towards String-to-Tree Neural Machine Translation

    Roee Aharoni, Yoav Goldberg

    cs.CLarXiv:1704.04743v32017
  35. MISC: A MIxed Strategy-Aware Model Integrating COMET for Emotional Support Conversation

    Quan Tu, Yanran Li, Jianwei Cui +3

    cs.CLarXiv:2203.13560v22022
  36. StreamAlign: Streaming Text-Aligned Speech Tokenization

    Kang-wook Kim, Jinyoung Park, Jinsoo Kim +3

    cs.CLcs.SDeess.ASarXiv:2609.09719v12026
  37. Semi-Supervised QA with Generative Domain-Adaptive Nets

    Zhilin Yang, Junjie Hu, Ruslan Salakhutdinov +1

    cs.CLcs.LGarXiv:1702.02206v22017
  38. SocialRL: Refining LLMs' Social Intelligence through Multi-turn Reinforcement Learning and Reward Design

    Jianing Wang, Xintao Wang, Aili Chen +7

    cs.CLarXiv:2609.09764v12026
  39. Scaling E-Commerce Attribute Extraction with Parallel Decoding

    Nikhita Vedula, Dushyanta Dhyani, Bryan Wang +1

    cs.CLarXiv:2609.09716v12026
  40. Knowledgeable or Educated Guess? Revisiting Language Models as Knowledge Bases

    Boxi Cao, Hongyu Lin, Xianpei Han +5

    cs.CLcs.AIarXiv:2106.09231v12021
  41. X2-NativeCursor: Native-Token Text Progress Tracking for Incremental-Text Streaming Codec TTS

    Zehan Liu, Carl Chen, Rime Wen +7

    cs.CLarXiv:2609.09677v12026
  42. SEA-SpeechBench: A Large-Scale Multitask Benchmark for Speech Understanding Across Southeast Asia

    Jingyi Liao, Wenyu Zhang, Zhuohan Liu +6

    cs.CLarXiv:2609.09672v12026
  43. Probabilistic FastText for Multi-Sense Word Embeddings

    Ben Athiwaratkun, Andrew Gordon Wilson, Anima Anandkumar

    cs.CLcs.AIcs.LGarXiv:1806.02901v12018
  44. AraGPT2: Pre-Trained Transformer for Arabic Language Generation

    Wissam Antoun, Fady Baly, Hazem Hajj

    cs.CLarXiv:2012.15520v22020
  45. Reproducing Omitted Temporal Expressions in Japanese News for Retrieval-Augmented Applications

    Tomoaki Yasuda, Shotaro Ishihara

    cs.CLarXiv:2609.09569v12026
  46. A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts

    Kuang-Huei Lee, Xinyun Chen, Hiroki Furuta +2

    cs.CLcs.AIcs.IRarXiv:2402.09727v32024
  47. Towards Automatic Evolution Tree Generation from Citation Graphs

    Zexing Zhao, Yuntong Hu, Liang Zhao

    cs.CLarXiv:2609.09561v12026
  48. BuzzASR: A Swarm of 100+ Monolingual Speech Recognition Models

    Shivam Singh, Aditya Yadavalli, Catherine Arnett +1

    cs.CLarXiv:2609.09554v12026
  49. Instruction-driven history-aware policies for robotic manipulations

    Pierre-Louis Guhur, Shizhe Chen, Ricardo Garcia +3

    cs.ROcs.AIcs.CLarXiv:2209.04899v32022
  50. The Mutations of Machine Speech

    Mauricio Figueroa

    cs.CLcs.CYcs.HCarXiv:2609.09496v12026
  51. Do LLMs Make More Mistakes If They Do Not Believe the Input Data?

    Peter Kochelka, Aleš Manuel Papáček, Vojtěch Dvořák +1

    cs.CLarXiv:2609.09363v12026
  52. TEFM: Token-Efficient Faithful Modeling for Structured Data

    Zhichao Hou, Lingdao Sha, Xueyu Mao +3

    cs.CLcs.LGarXiv:2609.09552v12026
  53. NL4Opt Competition: Formulating Optimization Problems Based on Their Natural Language Descriptions

    Rindranirina Ramamonjison, Timothy T. Yu, Raymond Li +8

    cs.CLcs.AIarXiv:2303.08233v22023
  54. SciCode: A Research Coding Benchmark Curated by Scientists

    Minyang Tian, Luyu Gao, Shizhuo Dylan Zhang +27

    cs.AIcs.CLarXiv:2407.13168v12024
  55. Benchmarking Hybrid Deep Research Across Database Querying and Web Search

    Ruofan Wu, Peiran Xu, Xiaolong Li +9

    cs.CLarXiv:2609.09410v12026
  56. SpecTr: Fast Speculative Decoding via Optimal Transport

    Ziteng Sun, Ananda Theertha Suresh, Jae Hun Ro +3

    cs.LGcs.CLcs.DSarXiv:2310.15141v22023
  57. SWORD: Wikidata-based Distortions Reveal Hidden Cross-Lingual Inconsistencies in LLM Factual Error Rejection

    Sanghyeok Park, Minji Kang, Hosung Kwak +1

    cs.CLarXiv:2609.09349v12026
  58. Toward a realistic model of speech processing in the brain with self-supervised learning

    Juliette Millet, Charlotte Caucheteux, Pierre Orhan +5

    q-bio.NCcs.AIcs.CLarXiv:2206.01685v22022
  59. AMR Parsing as Sequence-to-Graph Transduction

    Sheng Zhang, Xutai Ma, Kevin Duh +1

    cs.CLarXiv:1905.08704v22019
  60. Osprey: Target-agnostic Pre-training Makes Stronger Drafters in Speculative Decoding

    Fengxiang Bie, Yuqing Jian, Yifan Yu +7

    cs.CLarXiv:2609.09338v12026