Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,281 to 8,340 of 11,247

  1. PARTAB: Partition-Aware Reasoning with Structured Evidence for Scalable Table Understanding

    Md Mahadi Hasan Nahid, Davood Rafiei

    cs.CLcs.AIcs.IRarXiv:2608.24082v12026
  2. MortalMATH: Evaluating the Conflict Between Reasoning Objectives and Emergency Contexts

    Etienne Lanzeray, Stephane Meilliez, Malo Ruelle +1

    cs.CLarXiv:2601.18790v12026
  3. Influence Guided Sampling for Domain Adaptation of Text Retrievers

    Meet Doshi, Vishwajeet Kumar, Yulong Li +1

    cs.IRcs.CLarXiv:2601.21759v12026
  4. Automata from Agent Traces: Failure and Next-Step Prediction

    Seonglae Cho, Franklin Cardenoso Fernandez, Umar Mohammed +4

    cs.AIcs.CLcs.LGarXiv:2608.23670v12026
  5. nocaps: novel object captioning at scale

    Harsh Agrawal, Karan Desai, Yufei Wang +7

    cs.CVcs.AIcs.CLarXiv:1812.08658v32018
  6. SSL: Sweet Spot Learning for Differentiated Guidance in Agentic Optimization

    Jinyang Wu, Changpeng Yang, Yuhao Shen +9

    cs.CLarXiv:2601.22491v12026
  7. ORPO: Monolithic Preference Optimization without Reference Model

    Jiwoo Hong, Noah Lee, James Thorne

    cs.CLcs.AIarXiv:2403.07691v22024
  8. CoDiQ: Test-Time Scaling for Controllable Difficult Question Generation

    Zhongyuan Peng, Caijun Xu, Changyi Xiao +4

    cs.CLcs.AIarXiv:2602.01660v12026
  9. Discovering Hidden Gems in Model Repositories

    Jonathan Kahana, Eliahu Horwitz, Yedid Hoshen

    cs.LGcs.CLarXiv:2601.22157v12026
  10. The Linear Representation Hypothesis and the Geometry of Large Language Models

    Kiho Park, Yo Joong Choe, Victor Veitch

    cs.CLcs.AIcs.LGarXiv:2311.03658v22023
  11. Universal Dependencies v2: An Evergrowing Multilingual Treebank Collection

    Joakim Nivre, Marie-Catherine de Marneffe, Filip Ginter +6

    cs.CLarXiv:2004.10643v12020
  12. FourierSampler: Unlocking Non-Autoregressive Potential in Diffusion Language Models via Frequency-Guided Generation

    Siyang He, Qiqi Wang, Xiaoran Liu +8

    cs.CLarXiv:2601.23182v22026
  13. PICARD: Parsing Incrementally for Constrained Auto-Regressive Decoding from Language Models

    Torsten Scholak, Nathan Schucher, Dzmitry Bahdanau

    cs.CLcs.PLarXiv:2109.05093v12021
  14. SQLite is Enough. Lexical, Semantic, and Hybrid Search with scrydb

    Timo Breuer

    cs.IRcs.CLcs.DBarXiv:2608.24060v12026
  15. TrustDABench: Benchmarking Reliability and Robustness of LLMs for Structured Data Analysis

    Boshen Shi, Yize Liu, Chen Zhao +4

    cs.CLcs.SEarXiv:2608.24145v12026
  16. LLM-Blender: Ensembling Large Language Models with Pairwise Ranking and Generative Fusion

    Dongfu Jiang, Xiang Ren, Bill Yuchen Lin

    cs.CLcs.AIcs.LGarXiv:2306.02561v32023
  17. Textbooks Are All You Need II: phi-1.5 technical report

    Yuanzhi Li, Sébastien Bubeck, Ronen Eldan +3

    cs.CLcs.AIarXiv:2309.05463v12023
  18. Closing the Loop: Universal Repository Representation with RPG-Encoder

    Jane Luo, Chengyu Yin, Xin Zhang +10

    cs.CLcs.SEarXiv:2602.02084v22026
  19. A Survey on Recent Advances in Named Entity Recognition from Deep Learning models

    Vikas Yadav, Steven Bethard

    cs.CLcs.LGarXiv:1910.11470v12019
  20. $C$-$ΔΘ$: Circuit-Restricted Weight Arithmetic for Selective Refusal

    Aditya Kasliwal, Pratinav Seth, Vinay Kumar Sankarapu

    cs.CLcs.ETarXiv:2602.04521v22026
  21. Learning the Difference that Makes a Difference with Counterfactually-Augmented Data

    Divyansh Kaushik, Eduard Hovy, Zachary C. Lipton

    cs.CLcs.AIcs.LGarXiv:1909.12434v22019
  22. RexBERT: Context Specialized Bidirectional Encoders for E-commerce

    Rahul Bajaj, Anuj Garg

    cs.CLcs.AIarXiv:2602.04605v12026
  23. Entropy Aware Reward Guidance for Diffusion Language Model Alignment

    Atula Tejaswi, Litu Rout, Constantine Caramanis +2

    cs.LGcs.AIcs.CLarXiv:2602.05000v22026
  24. Back to Basics: Revisiting Exploration in Reinforcement Learning for LLM Reasoning via Generative Probabilities

    Pengyi Li, Elizaveta Goncharova, Andrey Kuznetsov +1

    cs.LGcs.CLarXiv:2602.05281v22026
  25. CoPE: Clipped RoPE as A Scalable Free Lunch for Long Context LLMs

    Haoran Li, Sucheng Ren, Alan Yuille +1

    cs.CLcs.AIcs.LGarXiv:2602.05258v12026
  26. Anchored Decoding: Provably Reducing Copyright Risk for Any Language Model

    Jacqueline He, Jonathan Hayase, Wen-tau Yih +3

    cs.CLarXiv:2602.07120v22026
  27. OmniMoE: An Efficient MoE by Orchestrating Atomic Experts at Scale

    Jingze Shi, Zhangyang Peng, Yizhang Zhu +3

    cs.CLcs.AIarXiv:2602.05711v22026
  28. Revisiting the Shape Convention of Transformer Language Models

    Feng-Ting Liao, Meng-Hsi Chen, Guan-Ting Yi +1

    cs.CLcs.AIcs.LGarXiv:2602.06471v12026
  29. TodoEvolve: Learning to Architect Agent Planning Systems

    Jiaxi Liu, Yanzuo Jiang, Guibin Zhang +5

    cs.CLcs.AIcs.LGarXiv:2602.07839v12026
  30. BAE: BERT-based Adversarial Examples for Text Classification

    Siddhant Garg, Goutham Ramakrishnan

    cs.CLarXiv:2004.01970v32020
  31. AgentCPM-Report: Interleaving Drafting and Deepening for Open-Ended Deep Research

    Yishan Li, Wentong Chen, Yukun Yan +12

    cs.AIcs.CLarXiv:2602.06540v12026
  32. Method, Mind, and Morality: How People Make Sense of Artificial Intelligence

    Jacy Reese Anthis, Erik Brynjolfsson, James Evans

    cs.CYcs.AIcs.CLarXiv:2608.24748v12026
  33. Streaming End-to-end Speech Recognition For Mobile Devices

    Yanzhang He, Tara N. Sainath, Rohit Prabhavalkar +17

    cs.CLarXiv:1811.06621v12018
  34. Blind to the Human Touch: Overlap Bias in LLM-Based Summary Evaluation

    Jiangnan Fang, Cheng-Tse Liu, Hanieh Deilamsalehy +5

    cs.CLarXiv:2602.07673v12026
  35. QP-OneModel: A Unified Generative LLM for Multi-Task Query Understanding in Xiaohongshu Search

    Jianzhao Huang, Xiaorui Huang, Fei Zhao +9

    cs.IRcs.CLarXiv:2602.09901v12026
  36. SAGE: From Direct Answering to Evidence-Grounded Inference for Chinese Ancient Document Understanding

    Yuchuan Wu, Xuan Luo, Yinglian Zhu +3

    cs.CLcs.AIarXiv:2608.24011v12026
  37. Faith and Fate: Limits of Transformers on Compositionality

    Nouha Dziri, Ximing Lu, Melanie Sclar +13

    cs.CLcs.AIcs.LGarXiv:2305.18654v32023
  38. LLM Evaluators Recognize and Favor Their Own Generations

    Arjun Panickssery, Samuel R. Bowman, Shi Feng

    cs.CLcs.AIarXiv:2404.13076v12024
  39. RecurSE: Bounded Recursive Self-Evaluation for LLM Rubric Judges

    Kaiyuan Liu, Ziyuan Zhuang, Rongxiang Weng +1

    cs.CLarXiv:2608.24231v12026
  40. Pretraining A Large Language Model using Distributed GPUs: A Memory-Efficient Decentralized Paradigm

    Jinrui Zhang, Chaodong Xiao, Aoqi Wu +2

    cs.CLarXiv:2602.11543v32026
  41. Thinking with Drafting: Optical Decompression via Logical Reconstruction

    Jingxuan Wei, Honghao He, Caijun Jia +9

    cs.CLarXiv:2602.11731v22026
  42. DeepSight: An All-in-One LM Safety Toolkit

    Bo Zhang, Jiaxuan Guo, Lijun Li +17

    cs.CLcs.AIcs.CRarXiv:2602.12092v12026
  43. Detecting Overflow in Compressed Token Representations for Retrieval-Augmented Generation

    Julia Belikova, Danila Rozhevskii, Dennis Svirin +2

    cs.CLarXiv:2602.12235v22026
  44. LM-Lexicon: Improving Definition Modeling via Harmonizing Semantic Experts

    Yang Liu, Jiaye Yang, Weikang Li +3

    cs.CLarXiv:2602.14060v12026
  45. BrowserForge: Scaling Web Episode via Parallel Browser Sandboxes

    Fei Tang, Huawen Shen, Zhiqiong Lu +7

    cs.CLarXiv:2608.24848v12026
  46. Neural Variational Inference for Text Processing

    Yishu Miao, Lei Yu, Phil Blunsom

    cs.CLcs.LGstat.MLarXiv:1511.06038v42015
  47. Avey-B

    Devang Acharya, Mohammad Hammoud

    cs.CLcs.AIarXiv:2602.15814v12026
  48. ArXiv-to-Model: A Practical Study of Scientific LM Training

    Anuj Gupta

    cs.AIcs.CLarXiv:2602.17288v12026
  49. STATe-of-Thoughts: Structured Action Templates for Tree-of-Thoughts

    Zachary Bamberger, Till R. Saenger, Gilad Morad +3

    cs.CLcs.LGarXiv:2602.14265v32026
  50. Whom to Query for What: Adaptive Group Elicitation via Multi-Turn LLM Interactions

    Ruomeng Ding, Tianwei Gao, Thomas P. Zollo +3

    cs.LGcs.AIcs.CLarXiv:2602.14279v22026
  51. BEATs: Audio Pre-Training with Acoustic Tokenizers

    Sanyuan Chen, Yu Wu, Chengyi Wang +4

    eess.AScs.AIcs.CLarXiv:2212.09058v12022
  52. Reinforced Fast Weights with Next-Sequence Prediction

    Hee Seung Hwang, Xindi Wu, Sanghyuk Chun +1

    cs.CLarXiv:2602.16704v12026
  53. Compacter: Efficient Low-Rank Hypercomplex Adapter Layers

    Rabeeh Karimi Mahabadi, James Henderson, Sebastian Ruder

    cs.CLarXiv:2106.04647v22021
  54. No One Size Fits All: QueryBandits for Hallucination Mitigation

    Nicole Cho, William Watson, Alec Koppel +2

    cs.CLcs.AIcs.LGarXiv:2602.20332v12026
  55. What BERT is not: Lessons from a new suite of psycholinguistic diagnostics for language models

    Allyson Ettinger

    cs.CLcs.AIarXiv:1907.13528v22019
  56. Learning to Detect Language Model Training Data via Active Reconstruction

    Junjie Oscar Yin, John X. Morris, Vitaly Shmatikov +2

    cs.LGcs.AIcs.CLarXiv:2602.19020v12026
  57. Adaptive Text Anonymization: Learning Privacy-Utility Trade-offs via Prompt Optimization

    Gabriel Loiseau, Damien Sileo, Damien Riquet +2

    cs.CLarXiv:2602.20743v22026
  58. SLAKE: A Semantically-Labeled Knowledge-Enhanced Dataset for Medical Visual Question Answering

    Bo Liu, Li-Ming Zhan, Li Xu +3

    cs.CVcs.AIcs.CLarXiv:2102.09542v12021
  59. DLT-Corpus: A Large-Scale Text Collection for the Distributed Ledger Technology Domain

    Walter Hernandez Cruz, Peter Devine, Nikhil Vadgama +2

    cs.CLarXiv:2602.22045v22026
  60. Contextual Augmentation: Data Augmentation by Words with Paradigmatic Relations

    Sosuke Kobayashi

    cs.CLcs.LGarXiv:1805.06201v12018