Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
3,781 to 3,840 of 11,295
GRPO Beyond English: A Large-Scale Study of GRPO in Non-English and Multilingual Settings
Konstantin Dobler, Federico Scozzafava, Jonathan Janke +2
cs.CLcs.LGarXiv:2608.13698v12026Does a Language Server Save Tokens for Coding Agents? A Measurement Methodology and Preliminary Study
Pengcheng Xu
cs.CLcs.AIarXiv:2608.13568v12026NVIDIA-labs OO Agents: Native Python Object-Oriented Agents
Paul Furgale, Severin Klingler, James Nolan +12
cs.AIcs.CLarXiv:2607.20709v12026Constitutional Midtraining: Content Presence Drives Alignment Gains
Desiree Cho, Cameron Tice, Bernie Hogan +4
cs.CLcs.AIcs.CYarXiv:2607.26654v22026Agent Retrieval Bench: Evaluating Repository Context Retrieval for Coding Agents
Bowen Qin, Yi Xie
cs.IRcs.AIcs.CLarXiv:2607.24882v12026Discrete Diffusion Models: A Unified Framework from Tokenization to Generation
Ye Yuan, Weien Li, Rui Song +19
cs.LGcs.AIcs.CLarXiv:2607.13431v12026Subliminal Clocks: Latent Time Modelling in Diffusion Language Models
Maximo Eduardo Rulli, Thomas Vaitses Fontanari, Simone Petruzzi +9
cs.AIcs.CLarXiv:2607.01774v22026Seeing Is Not Sharing: Some Vision-Language Models Overestimate Common Ground in Asymmetric Dialogue
Nan Li, Albert Gatt, Massimo Poesio
cs.CLcs.AIarXiv:2606.31719v12026HAKARI-Bench: A Lightweight Benchmark for Comparing Retrieval Architectures and Efficiency Settings under Unified Conditions
Yuichi Tateno
cs.IRcs.CLarXiv:2606.22778v12026iOSWorld: A Benchmark for Personally Intelligent Phone Agents
Lawrence Keunho Jang, Mareks Woodside, Geronimo Carom +3
cs.LGcs.CLarXiv:2606.09764v12026SCOPE: Self-Play via Co-Evolving Policies for Open-Ended Tasks
Wai-Chung Kwan, Aryo Pradipta Gema, Joshua Ong Jun Leang +1
cs.CLarXiv:2605.31433v12026Crafter: A Multi-Agent Harness for Editable Scientific Figure Generation from Diverse Inputs
Haozhe Zhao, Shuzheng Si, Zhenhailong Wang +6
cs.CVcs.AIcs.CLarXiv:2605.30611v12026dMoE: dLLMs with Learnable Block Experts
Sicheng Feng, Zigeng Chen, Gongfan Fang +2
cs.CLarXiv:2605.30876v22026Token-Level Generalization in LoRA Adapter Backdoors: Attack Characterization and Behavioral Detection
Travis Lelle
cs.CRcs.AIcs.CLarXiv:2605.30189v12026Learning to Foresee: Unveiling the Unlocking Efficiency of On-Policy Distillation
Yuchen Cai, Ding Cao, Liang Lin +9
cs.CLarXiv:2605.11739v32026"I Didn't Make the Micro Decisions": Measuring, Inducing, and Exposing Goal-Level AI Contributions in Collaboration
Eunsu Kim, Jessica R. Mindel, Kyungjin Kim +1
cs.CLarXiv:2605.21363v22026Language-Switching Triggers Take a Latent Detour Through Language Models
Francis Kulumba, Wissam Antoun, Théo Lasnier +2
cs.CLarXiv:2605.18646v22026CASCADE: Case-Based Continual Adaptation for Large Language Models During Deployment
Siyuan Guo, Yali Du, Hechang Chen +2
cs.AIcs.CLcs.LGarXiv:2605.06702v12026MEME: Multi-entity & Evolving Memory Evaluation
Seokwon Jung, Alexander Rubinstein, Arnas Uselis +2
cs.LGcs.CLarXiv:2605.12477v12026Rethinking Reasoning-Intensive Retrieval: Evaluating and Advancing Retrievers in Agentic Search Systems
Yilun Zhao, Jinbiao Wei, Tingyu Song +3
cs.CLcs.IRarXiv:2605.04018v12026Micro Language Models Enable Instant Responses
Wen Cheng, Tuochao Chen, Karim Helwani +3
cs.CLarXiv:2604.19642v12026OptiMer: Optimal Distribution Vector Merging Is Better than Data Mixing for Continual Pre-Training
Haiyue Song, Masao Utiyama
cs.CLcs.AIcs.LGarXiv:2603.28858v22026The Cognitive Penalty: Ablating System 1 and System 2 Reasoning in Edge-Native SLMs for Decentralized Consensus
Syed Muhammad Aqdas Rizvi
cs.AIcs.CLcs.CRarXiv:2604.16913v12026CREATE: Testing LLMs for Associative Creativity
Manya Wadhwa, Tiasa Singha Roy, Harvey Lederman +2
cs.CLarXiv:2603.09970v22026Human Psychometric Questionnaires Mischaracterize LLM Behavior
Woojung Song, Dongmin Choi, Yoonah Park +3
cs.CLcs.AIarXiv:2509.10078v42025Natural Language Processing in Electronic Health Records in Relation to Healthcare Decision-making: A Systematic Review
Elias Hossain, Rajib Rana, Niall Higgins +5
cs.CLcs.CYarXiv:2306.12834v12023AnyGPT: Unified Multimodal LLM with Discrete Sequence Modeling
Jun Zhan, Junqi Dai, Jiasheng Ye +13
cs.CLcs.AIcs.CVarXiv:2402.12226v52024AlpacaFarm: A Simulation Framework for Methods that Learn from Human Feedback
Yann Dubois, Xuechen Li, Rohan Taori +6
cs.LGcs.AIcs.CLarXiv:2305.14387v42023OmniSQL: Synthesizing High-quality Text-to-SQL Data at Scale
Haoyang Li, Shang Wu, Xiaokang Zhang +9
cs.CLcs.DBarXiv:2503.02240v22025Sign Language Recognition, Generation, and Translation: An Interdisciplinary Perspective
Danielle Bragg, Oscar Koller, Mary Bellard +9
cs.CVcs.CLcs.CYarXiv:1908.08597v12019Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory
Tianxin Wei, Noveen Sachdeva, Benjamin Coleman +12
cs.CLcs.AIarXiv:2511.20857v22025Do LLMs Recognize Your Preferences? Evaluating Personalized Preference Following in LLMs
Siyan Zhao, Mingyi Hong, Yang Liu +2
cs.LGcs.CLarXiv:2502.09597v12025Locked at the Entrance, Open Inside: Where RLVR Narrows the Solution Space
Qiancheng Zhou, Ruizhe Li
cs.LGcs.AIcs.CLarXiv:2608.29188v12026dKV-Cache: The Cache for Diffusion Language Models
Xinyin Ma, Runpeng Yu, Gongfan Fang +1
cs.CLarXiv:2505.15781v12025A Survey of Context Engineering for Large Language Models
Lingrui Mei, Jiayu Yao, Yuyao Ge +12
cs.CLarXiv:2507.13334v22025MMBERT: Multimodal BERT Pretraining for Improved Medical VQA
Yash Khare, Viraj Bagal, Minesh Mathew +3
cs.CVcs.CLcs.LGarXiv:2104.01394v12021Mixture-of-Recursions: Learning Dynamic Recursive Depths for Adaptive Token-Level Computation
Sangmin Bae, Yujin Kim, Reza Bayat +8
cs.CLcs.LGarXiv:2507.10524v32025LocAgent: Graph-Guided LLM Agents for Code Localization
Zhaoling Chen, Xiangru Tang, Gangda Deng +6
cs.SEcs.AIcs.CLarXiv:2503.09089v22025R2E-Gym: Procedural Environments and Hybrid Verifiers for Scaling Open-Weights SWE Agents
Naman Jain, Jaskirat Singh, Manish Shetty +3
cs.SEcs.CLcs.LGarXiv:2504.07164v12025RuleMem: Active Rule Memory for Long-Term Conversational Agents
Xingyuan Zeng, Zuohan Wu, Quanming Yao +5
cs.CLcs.IRarXiv:2609.03915v12026Agentic Reinforced Policy Optimization
Guanting Dong, Hangyu Mao, Kai Ma +11
cs.LGcs.AIcs.CLarXiv:2507.19849v12025AdaptThink: Reasoning Models Can Learn When to Think
Jiajie Zhang, Nianyi Lin, Lei Hou +2
cs.CLcs.AIcs.LGarXiv:2505.13417v12025Transfiver: Human-AI Co-Inference through a Shared Editable State
Minji Park, Seunghyun Yoon, Hyuk Lim
cs.AIcs.CLcs.HCarXiv:2609.03797v12026Self-Distilled RLVR
Chenxu Yang, Chuanyu Qin, Qingyi Si +7
cs.LGcs.CLarXiv:2604.03128v22026Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling
Runze Liu, Junqi Gao, Jian Zhao +5
cs.CLarXiv:2502.06703v12025MCP-Universe: Benchmarking Large Language Models with Real-World Model Context Protocol Servers
Ziyang Luo, Zhiqi Shen, Wenzhuo Yang +7
cs.AIcs.CLarXiv:2508.14704v12025LLaVA-Mini: Efficient Image and Video Large Multimodal Models with One Vision Token
Shaolei Zhang, Qingkai Fang, Zhe Yang +1
cs.CVcs.AIcs.CLarXiv:2501.03895v22025Two-Stage Reinforcement Learning for Sound and Adversarial Test Generation in Code LLMs
Jiacheng Xu, Wentao Zhang, Zhiyi Lyu +4
cs.CLcs.LGarXiv:2609.03955v12026Benchmarking Cognitive Biases in Large Language Models as Evaluators
Ryan Koo, Minhwa Lee, Vipul Raheja +3
cs.CLcs.AIcs.LGarXiv:2309.17012v32023PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding
Wei Chow, Jiageng Mao, Boyi Li +3
cs.CVcs.AIcs.CLarXiv:2501.16411v22025Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification
Anqi Zhang, Yulin Chen, Jane Pan +4
cs.AIcs.CLarXiv:2504.05419v12025Lost but not erased: Finding traces of a forgotten language in neural speech models
Peter Plantinga, Charlotte Moore, Peter W. Donhauser +2
cs.CLcs.LGarXiv:2608.25976v12026Which Economic Tasks are Performed with AI? Evidence from Millions of Claude Conversations
Kunal Handa, Alex Tamkin, Miles McCain +12
cs.CYcs.AIcs.CLarXiv:2503.04761v12025Entity-level Factual Consistency of Abstractive Text Summarization
Feng Nan, Ramesh Nallapati, Zhiguo Wang +5
cs.CLcs.AIarXiv:2102.09130v12021GRIT: Teaching MLLMs to Think with Images
Yue Fan, Xuehai He, Diji Yang +6
cs.CVcs.AIcs.CLarXiv:2505.15879v22025NeuroNER: an easy-to-use program for named-entity recognition based on neural networks
Franck Dernoncourt, Ji Young Lee, Peter Szolovits
cs.CLcs.NEstat.MLarXiv:1705.05487v12017When Retrieval Helps: Selective Retrieval for Single-Turn Mental-Health QA
Hyunseo Oh, Chong-Kwon Kim, Yoonhyuk Choi
cs.CLcs.IRarXiv:2609.03454v12026Time-R1: Post-Training Large Vision Language Model for Temporal Video Grounding
Ye Wang, Ziheng Wang, Boshen Xu +14
cs.CVcs.AIcs.CLarXiv:2503.13377v32025FlexPrefill: A Context-Aware Sparse Attention Mechanism for Efficient Long-Sequence Inference
Xunhao Lai, Jianqiao Lu, Yao Luo +2
cs.LGcs.CLarXiv:2502.20766v12025MemoryArena: Benchmarking Agent Memory in Interdependent Multi-Session Agentic Tasks
Zexue He, Yu Wang, Churan Zhi +11
cs.CLarXiv:2602.16313v12026