Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,281 to 8,340 of 11,247
PARTAB: Partition-Aware Reasoning with Structured Evidence for Scalable Table Understanding
Md Mahadi Hasan Nahid, Davood Rafiei
cs.CLcs.AIcs.IRarXiv:2608.24082v12026MortalMATH: Evaluating the Conflict Between Reasoning Objectives and Emergency Contexts
Etienne Lanzeray, Stephane Meilliez, Malo Ruelle +1
cs.CLarXiv:2601.18790v12026Influence Guided Sampling for Domain Adaptation of Text Retrievers
Meet Doshi, Vishwajeet Kumar, Yulong Li +1
cs.IRcs.CLarXiv:2601.21759v12026Automata from Agent Traces: Failure and Next-Step Prediction
Seonglae Cho, Franklin Cardenoso Fernandez, Umar Mohammed +4
cs.AIcs.CLcs.LGarXiv:2608.23670v12026nocaps: novel object captioning at scale
Harsh Agrawal, Karan Desai, Yufei Wang +7
cs.CVcs.AIcs.CLarXiv:1812.08658v32018SSL: Sweet Spot Learning for Differentiated Guidance in Agentic Optimization
Jinyang Wu, Changpeng Yang, Yuhao Shen +9
cs.CLarXiv:2601.22491v12026ORPO: Monolithic Preference Optimization without Reference Model
Jiwoo Hong, Noah Lee, James Thorne
cs.CLcs.AIarXiv:2403.07691v22024CoDiQ: Test-Time Scaling for Controllable Difficult Question Generation
Zhongyuan Peng, Caijun Xu, Changyi Xiao +4
cs.CLcs.AIarXiv:2602.01660v12026Discovering Hidden Gems in Model Repositories
Jonathan Kahana, Eliahu Horwitz, Yedid Hoshen
cs.LGcs.CLarXiv:2601.22157v12026The Linear Representation Hypothesis and the Geometry of Large Language Models
Kiho Park, Yo Joong Choe, Victor Veitch
cs.CLcs.AIcs.LGarXiv:2311.03658v22023Universal Dependencies v2: An Evergrowing Multilingual Treebank Collection
Joakim Nivre, Marie-Catherine de Marneffe, Filip Ginter +6
cs.CLarXiv:2004.10643v12020FourierSampler: Unlocking Non-Autoregressive Potential in Diffusion Language Models via Frequency-Guided Generation
Siyang He, Qiqi Wang, Xiaoran Liu +8
cs.CLarXiv:2601.23182v22026PICARD: Parsing Incrementally for Constrained Auto-Regressive Decoding from Language Models
Torsten Scholak, Nathan Schucher, Dzmitry Bahdanau
cs.CLcs.PLarXiv:2109.05093v12021SQLite is Enough. Lexical, Semantic, and Hybrid Search with scrydb
Timo Breuer
cs.IRcs.CLcs.DBarXiv:2608.24060v12026TrustDABench: Benchmarking Reliability and Robustness of LLMs for Structured Data Analysis
Boshen Shi, Yize Liu, Chen Zhao +4
cs.CLcs.SEarXiv:2608.24145v12026LLM-Blender: Ensembling Large Language Models with Pairwise Ranking and Generative Fusion
Dongfu Jiang, Xiang Ren, Bill Yuchen Lin
cs.CLcs.AIcs.LGarXiv:2306.02561v32023Textbooks Are All You Need II: phi-1.5 technical report
Yuanzhi Li, Sébastien Bubeck, Ronen Eldan +3
cs.CLcs.AIarXiv:2309.05463v12023Closing the Loop: Universal Repository Representation with RPG-Encoder
Jane Luo, Chengyu Yin, Xin Zhang +10
cs.CLcs.SEarXiv:2602.02084v22026A Survey on Recent Advances in Named Entity Recognition from Deep Learning models
Vikas Yadav, Steven Bethard
cs.CLcs.LGarXiv:1910.11470v12019$C$-$ΔΘ$: Circuit-Restricted Weight Arithmetic for Selective Refusal
Aditya Kasliwal, Pratinav Seth, Vinay Kumar Sankarapu
cs.CLcs.ETarXiv:2602.04521v22026Learning the Difference that Makes a Difference with Counterfactually-Augmented Data
Divyansh Kaushik, Eduard Hovy, Zachary C. Lipton
cs.CLcs.AIcs.LGarXiv:1909.12434v22019RexBERT: Context Specialized Bidirectional Encoders for E-commerce
Rahul Bajaj, Anuj Garg
cs.CLcs.AIarXiv:2602.04605v12026Entropy Aware Reward Guidance for Diffusion Language Model Alignment
Atula Tejaswi, Litu Rout, Constantine Caramanis +2
cs.LGcs.AIcs.CLarXiv:2602.05000v22026Back to Basics: Revisiting Exploration in Reinforcement Learning for LLM Reasoning via Generative Probabilities
Pengyi Li, Elizaveta Goncharova, Andrey Kuznetsov +1
cs.LGcs.CLarXiv:2602.05281v22026CoPE: Clipped RoPE as A Scalable Free Lunch for Long Context LLMs
Haoran Li, Sucheng Ren, Alan Yuille +1
cs.CLcs.AIcs.LGarXiv:2602.05258v12026Anchored Decoding: Provably Reducing Copyright Risk for Any Language Model
Jacqueline He, Jonathan Hayase, Wen-tau Yih +3
cs.CLarXiv:2602.07120v22026OmniMoE: An Efficient MoE by Orchestrating Atomic Experts at Scale
Jingze Shi, Zhangyang Peng, Yizhang Zhu +3
cs.CLcs.AIarXiv:2602.05711v22026Revisiting the Shape Convention of Transformer Language Models
Feng-Ting Liao, Meng-Hsi Chen, Guan-Ting Yi +1
cs.CLcs.AIcs.LGarXiv:2602.06471v12026TodoEvolve: Learning to Architect Agent Planning Systems
Jiaxi Liu, Yanzuo Jiang, Guibin Zhang +5
cs.CLcs.AIcs.LGarXiv:2602.07839v12026BAE: BERT-based Adversarial Examples for Text Classification
Siddhant Garg, Goutham Ramakrishnan
cs.CLarXiv:2004.01970v32020AgentCPM-Report: Interleaving Drafting and Deepening for Open-Ended Deep Research
Yishan Li, Wentong Chen, Yukun Yan +12
cs.AIcs.CLarXiv:2602.06540v12026Method, Mind, and Morality: How People Make Sense of Artificial Intelligence
Jacy Reese Anthis, Erik Brynjolfsson, James Evans
cs.CYcs.AIcs.CLarXiv:2608.24748v12026Streaming End-to-end Speech Recognition For Mobile Devices
Yanzhang He, Tara N. Sainath, Rohit Prabhavalkar +17
cs.CLarXiv:1811.06621v12018Blind to the Human Touch: Overlap Bias in LLM-Based Summary Evaluation
Jiangnan Fang, Cheng-Tse Liu, Hanieh Deilamsalehy +5
cs.CLarXiv:2602.07673v12026QP-OneModel: A Unified Generative LLM for Multi-Task Query Understanding in Xiaohongshu Search
Jianzhao Huang, Xiaorui Huang, Fei Zhao +9
cs.IRcs.CLarXiv:2602.09901v12026SAGE: From Direct Answering to Evidence-Grounded Inference for Chinese Ancient Document Understanding
Yuchuan Wu, Xuan Luo, Yinglian Zhu +3
cs.CLcs.AIarXiv:2608.24011v12026Faith and Fate: Limits of Transformers on Compositionality
Nouha Dziri, Ximing Lu, Melanie Sclar +13
cs.CLcs.AIcs.LGarXiv:2305.18654v32023LLM Evaluators Recognize and Favor Their Own Generations
Arjun Panickssery, Samuel R. Bowman, Shi Feng
cs.CLcs.AIarXiv:2404.13076v12024RecurSE: Bounded Recursive Self-Evaluation for LLM Rubric Judges
Kaiyuan Liu, Ziyuan Zhuang, Rongxiang Weng +1
cs.CLarXiv:2608.24231v12026Pretraining A Large Language Model using Distributed GPUs: A Memory-Efficient Decentralized Paradigm
Jinrui Zhang, Chaodong Xiao, Aoqi Wu +2
cs.CLarXiv:2602.11543v32026Thinking with Drafting: Optical Decompression via Logical Reconstruction
Jingxuan Wei, Honghao He, Caijun Jia +9
cs.CLarXiv:2602.11731v22026DeepSight: An All-in-One LM Safety Toolkit
Bo Zhang, Jiaxuan Guo, Lijun Li +17
cs.CLcs.AIcs.CRarXiv:2602.12092v12026Detecting Overflow in Compressed Token Representations for Retrieval-Augmented Generation
Julia Belikova, Danila Rozhevskii, Dennis Svirin +2
cs.CLarXiv:2602.12235v22026LM-Lexicon: Improving Definition Modeling via Harmonizing Semantic Experts
Yang Liu, Jiaye Yang, Weikang Li +3
cs.CLarXiv:2602.14060v12026BrowserForge: Scaling Web Episode via Parallel Browser Sandboxes
Fei Tang, Huawen Shen, Zhiqiong Lu +7
cs.CLarXiv:2608.24848v12026Neural Variational Inference for Text Processing
Yishu Miao, Lei Yu, Phil Blunsom
cs.CLcs.LGstat.MLarXiv:1511.06038v42015Avey-B
Devang Acharya, Mohammad Hammoud
cs.CLcs.AIarXiv:2602.15814v12026ArXiv-to-Model: A Practical Study of Scientific LM Training
Anuj Gupta
cs.AIcs.CLarXiv:2602.17288v12026STATe-of-Thoughts: Structured Action Templates for Tree-of-Thoughts
Zachary Bamberger, Till R. Saenger, Gilad Morad +3
cs.CLcs.LGarXiv:2602.14265v32026Whom to Query for What: Adaptive Group Elicitation via Multi-Turn LLM Interactions
Ruomeng Ding, Tianwei Gao, Thomas P. Zollo +3
cs.LGcs.AIcs.CLarXiv:2602.14279v22026BEATs: Audio Pre-Training with Acoustic Tokenizers
Sanyuan Chen, Yu Wu, Chengyi Wang +4
eess.AScs.AIcs.CLarXiv:2212.09058v12022Reinforced Fast Weights with Next-Sequence Prediction
Hee Seung Hwang, Xindi Wu, Sanghyuk Chun +1
cs.CLarXiv:2602.16704v12026Compacter: Efficient Low-Rank Hypercomplex Adapter Layers
Rabeeh Karimi Mahabadi, James Henderson, Sebastian Ruder
cs.CLarXiv:2106.04647v22021No One Size Fits All: QueryBandits for Hallucination Mitigation
Nicole Cho, William Watson, Alec Koppel +2
cs.CLcs.AIcs.LGarXiv:2602.20332v12026What BERT is not: Lessons from a new suite of psycholinguistic diagnostics for language models
Allyson Ettinger
cs.CLcs.AIarXiv:1907.13528v22019Learning to Detect Language Model Training Data via Active Reconstruction
Junjie Oscar Yin, John X. Morris, Vitaly Shmatikov +2
cs.LGcs.AIcs.CLarXiv:2602.19020v12026Adaptive Text Anonymization: Learning Privacy-Utility Trade-offs via Prompt Optimization
Gabriel Loiseau, Damien Sileo, Damien Riquet +2
cs.CLarXiv:2602.20743v22026SLAKE: A Semantically-Labeled Knowledge-Enhanced Dataset for Medical Visual Question Answering
Bo Liu, Li-Ming Zhan, Li Xu +3
cs.CVcs.AIcs.CLarXiv:2102.09542v12021DLT-Corpus: A Large-Scale Text Collection for the Distributed Ledger Technology Domain
Walter Hernandez Cruz, Peter Devine, Nikhil Vadgama +2
cs.CLarXiv:2602.22045v22026Contextual Augmentation: Data Augmentation by Words with Paradigmatic Relations
Sosuke Kobayashi
cs.CLcs.LGarXiv:1805.06201v12018