Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

3,781 to 3,840 of 11,295

  1. GRPO Beyond English: A Large-Scale Study of GRPO in Non-English and Multilingual Settings

    Konstantin Dobler, Federico Scozzafava, Jonathan Janke +2

    cs.CLcs.LGarXiv:2608.13698v12026
  2. Does a Language Server Save Tokens for Coding Agents? A Measurement Methodology and Preliminary Study

    Pengcheng Xu

    cs.CLcs.AIarXiv:2608.13568v12026
  3. NVIDIA-labs OO Agents: Native Python Object-Oriented Agents

    Paul Furgale, Severin Klingler, James Nolan +12

    cs.AIcs.CLarXiv:2607.20709v12026
  4. Constitutional Midtraining: Content Presence Drives Alignment Gains

    Desiree Cho, Cameron Tice, Bernie Hogan +4

    cs.CLcs.AIcs.CYarXiv:2607.26654v22026
  5. Agent Retrieval Bench: Evaluating Repository Context Retrieval for Coding Agents

    Bowen Qin, Yi Xie

    cs.IRcs.AIcs.CLarXiv:2607.24882v12026
  6. Discrete Diffusion Models: A Unified Framework from Tokenization to Generation

    Ye Yuan, Weien Li, Rui Song +19

    cs.LGcs.AIcs.CLarXiv:2607.13431v12026
  7. Subliminal Clocks: Latent Time Modelling in Diffusion Language Models

    Maximo Eduardo Rulli, Thomas Vaitses Fontanari, Simone Petruzzi +9

    cs.AIcs.CLarXiv:2607.01774v22026
  8. Seeing Is Not Sharing: Some Vision-Language Models Overestimate Common Ground in Asymmetric Dialogue

    Nan Li, Albert Gatt, Massimo Poesio

    cs.CLcs.AIarXiv:2606.31719v12026
  9. HAKARI-Bench: A Lightweight Benchmark for Comparing Retrieval Architectures and Efficiency Settings under Unified Conditions

    Yuichi Tateno

    cs.IRcs.CLarXiv:2606.22778v12026
  10. iOSWorld: A Benchmark for Personally Intelligent Phone Agents

    Lawrence Keunho Jang, Mareks Woodside, Geronimo Carom +3

    cs.LGcs.CLarXiv:2606.09764v12026
  11. SCOPE: Self-Play via Co-Evolving Policies for Open-Ended Tasks

    Wai-Chung Kwan, Aryo Pradipta Gema, Joshua Ong Jun Leang +1

    cs.CLarXiv:2605.31433v12026
  12. Crafter: A Multi-Agent Harness for Editable Scientific Figure Generation from Diverse Inputs

    Haozhe Zhao, Shuzheng Si, Zhenhailong Wang +6

    cs.CVcs.AIcs.CLarXiv:2605.30611v12026
  13. dMoE: dLLMs with Learnable Block Experts

    Sicheng Feng, Zigeng Chen, Gongfan Fang +2

    cs.CLarXiv:2605.30876v22026
  14. Token-Level Generalization in LoRA Adapter Backdoors: Attack Characterization and Behavioral Detection

    Travis Lelle

    cs.CRcs.AIcs.CLarXiv:2605.30189v12026
  15. Learning to Foresee: Unveiling the Unlocking Efficiency of On-Policy Distillation

    Yuchen Cai, Ding Cao, Liang Lin +9

    cs.CLarXiv:2605.11739v32026
  16. "I Didn't Make the Micro Decisions": Measuring, Inducing, and Exposing Goal-Level AI Contributions in Collaboration

    Eunsu Kim, Jessica R. Mindel, Kyungjin Kim +1

    cs.CLarXiv:2605.21363v22026
  17. Language-Switching Triggers Take a Latent Detour Through Language Models

    Francis Kulumba, Wissam Antoun, Théo Lasnier +2

    cs.CLarXiv:2605.18646v22026
  18. CASCADE: Case-Based Continual Adaptation for Large Language Models During Deployment

    Siyuan Guo, Yali Du, Hechang Chen +2

    cs.AIcs.CLcs.LGarXiv:2605.06702v12026
  19. MEME: Multi-entity & Evolving Memory Evaluation

    Seokwon Jung, Alexander Rubinstein, Arnas Uselis +2

    cs.LGcs.CLarXiv:2605.12477v12026
  20. Rethinking Reasoning-Intensive Retrieval: Evaluating and Advancing Retrievers in Agentic Search Systems

    Yilun Zhao, Jinbiao Wei, Tingyu Song +3

    cs.CLcs.IRarXiv:2605.04018v12026
  21. Micro Language Models Enable Instant Responses

    Wen Cheng, Tuochao Chen, Karim Helwani +3

    cs.CLarXiv:2604.19642v12026
  22. OptiMer: Optimal Distribution Vector Merging Is Better than Data Mixing for Continual Pre-Training

    Haiyue Song, Masao Utiyama

    cs.CLcs.AIcs.LGarXiv:2603.28858v22026
  23. The Cognitive Penalty: Ablating System 1 and System 2 Reasoning in Edge-Native SLMs for Decentralized Consensus

    Syed Muhammad Aqdas Rizvi

    cs.AIcs.CLcs.CRarXiv:2604.16913v12026
  24. CREATE: Testing LLMs for Associative Creativity

    Manya Wadhwa, Tiasa Singha Roy, Harvey Lederman +2

    cs.CLarXiv:2603.09970v22026
  25. Human Psychometric Questionnaires Mischaracterize LLM Behavior

    Woojung Song, Dongmin Choi, Yoonah Park +3

    cs.CLcs.AIarXiv:2509.10078v42025
  26. Natural Language Processing in Electronic Health Records in Relation to Healthcare Decision-making: A Systematic Review

    Elias Hossain, Rajib Rana, Niall Higgins +5

    cs.CLcs.CYarXiv:2306.12834v12023
  27. AnyGPT: Unified Multimodal LLM with Discrete Sequence Modeling

    Jun Zhan, Junqi Dai, Jiasheng Ye +13

    cs.CLcs.AIcs.CVarXiv:2402.12226v52024
  28. AlpacaFarm: A Simulation Framework for Methods that Learn from Human Feedback

    Yann Dubois, Xuechen Li, Rohan Taori +6

    cs.LGcs.AIcs.CLarXiv:2305.14387v42023
  29. OmniSQL: Synthesizing High-quality Text-to-SQL Data at Scale

    Haoyang Li, Shang Wu, Xiaokang Zhang +9

    cs.CLcs.DBarXiv:2503.02240v22025
  30. Sign Language Recognition, Generation, and Translation: An Interdisciplinary Perspective

    Danielle Bragg, Oscar Koller, Mary Bellard +9

    cs.CVcs.CLcs.CYarXiv:1908.08597v12019
  31. Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

    Tianxin Wei, Noveen Sachdeva, Benjamin Coleman +12

    cs.CLcs.AIarXiv:2511.20857v22025
  32. Do LLMs Recognize Your Preferences? Evaluating Personalized Preference Following in LLMs

    Siyan Zhao, Mingyi Hong, Yang Liu +2

    cs.LGcs.CLarXiv:2502.09597v12025
  33. Locked at the Entrance, Open Inside: Where RLVR Narrows the Solution Space

    Qiancheng Zhou, Ruizhe Li

    cs.LGcs.AIcs.CLarXiv:2608.29188v12026
  34. dKV-Cache: The Cache for Diffusion Language Models

    Xinyin Ma, Runpeng Yu, Gongfan Fang +1

    cs.CLarXiv:2505.15781v12025
  35. A Survey of Context Engineering for Large Language Models

    Lingrui Mei, Jiayu Yao, Yuyao Ge +12

    cs.CLarXiv:2507.13334v22025
  36. MMBERT: Multimodal BERT Pretraining for Improved Medical VQA

    Yash Khare, Viraj Bagal, Minesh Mathew +3

    cs.CVcs.CLcs.LGarXiv:2104.01394v12021
  37. Mixture-of-Recursions: Learning Dynamic Recursive Depths for Adaptive Token-Level Computation

    Sangmin Bae, Yujin Kim, Reza Bayat +8

    cs.CLcs.LGarXiv:2507.10524v32025
  38. LocAgent: Graph-Guided LLM Agents for Code Localization

    Zhaoling Chen, Xiangru Tang, Gangda Deng +6

    cs.SEcs.AIcs.CLarXiv:2503.09089v22025
  39. R2E-Gym: Procedural Environments and Hybrid Verifiers for Scaling Open-Weights SWE Agents

    Naman Jain, Jaskirat Singh, Manish Shetty +3

    cs.SEcs.CLcs.LGarXiv:2504.07164v12025
  40. RuleMem: Active Rule Memory for Long-Term Conversational Agents

    Xingyuan Zeng, Zuohan Wu, Quanming Yao +5

    cs.CLcs.IRarXiv:2609.03915v12026
  41. Agentic Reinforced Policy Optimization

    Guanting Dong, Hangyu Mao, Kai Ma +11

    cs.LGcs.AIcs.CLarXiv:2507.19849v12025
  42. AdaptThink: Reasoning Models Can Learn When to Think

    Jiajie Zhang, Nianyi Lin, Lei Hou +2

    cs.CLcs.AIcs.LGarXiv:2505.13417v12025
  43. Transfiver: Human-AI Co-Inference through a Shared Editable State

    Minji Park, Seunghyun Yoon, Hyuk Lim

    cs.AIcs.CLcs.HCarXiv:2609.03797v12026
  44. Self-Distilled RLVR

    Chenxu Yang, Chuanyu Qin, Qingyi Si +7

    cs.LGcs.CLarXiv:2604.03128v22026
  45. Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling

    Runze Liu, Junqi Gao, Jian Zhao +5

    cs.CLarXiv:2502.06703v12025
  46. MCP-Universe: Benchmarking Large Language Models with Real-World Model Context Protocol Servers

    Ziyang Luo, Zhiqi Shen, Wenzhuo Yang +7

    cs.AIcs.CLarXiv:2508.14704v12025
  47. LLaVA-Mini: Efficient Image and Video Large Multimodal Models with One Vision Token

    Shaolei Zhang, Qingkai Fang, Zhe Yang +1

    cs.CVcs.AIcs.CLarXiv:2501.03895v22025
  48. Two-Stage Reinforcement Learning for Sound and Adversarial Test Generation in Code LLMs

    Jiacheng Xu, Wentao Zhang, Zhiyi Lyu +4

    cs.CLcs.LGarXiv:2609.03955v12026
  49. Benchmarking Cognitive Biases in Large Language Models as Evaluators

    Ryan Koo, Minhwa Lee, Vipul Raheja +3

    cs.CLcs.AIcs.LGarXiv:2309.17012v32023
  50. PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

    Wei Chow, Jiageng Mao, Boyi Li +3

    cs.CVcs.AIcs.CLarXiv:2501.16411v22025
  51. Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

    Anqi Zhang, Yulin Chen, Jane Pan +4

    cs.AIcs.CLarXiv:2504.05419v12025
  52. Lost but not erased: Finding traces of a forgotten language in neural speech models

    Peter Plantinga, Charlotte Moore, Peter W. Donhauser +2

    cs.CLcs.LGarXiv:2608.25976v12026
  53. Which Economic Tasks are Performed with AI? Evidence from Millions of Claude Conversations

    Kunal Handa, Alex Tamkin, Miles McCain +12

    cs.CYcs.AIcs.CLarXiv:2503.04761v12025
  54. Entity-level Factual Consistency of Abstractive Text Summarization

    Feng Nan, Ramesh Nallapati, Zhiguo Wang +5

    cs.CLcs.AIarXiv:2102.09130v12021
  55. GRIT: Teaching MLLMs to Think with Images

    Yue Fan, Xuehai He, Diji Yang +6

    cs.CVcs.AIcs.CLarXiv:2505.15879v22025
  56. NeuroNER: an easy-to-use program for named-entity recognition based on neural networks

    Franck Dernoncourt, Ji Young Lee, Peter Szolovits

    cs.CLcs.NEstat.MLarXiv:1705.05487v12017
  57. When Retrieval Helps: Selective Retrieval for Single-Turn Mental-Health QA

    Hyunseo Oh, Chong-Kwon Kim, Yoonhyuk Choi

    cs.CLcs.IRarXiv:2609.03454v12026
  58. Time-R1: Post-Training Large Vision Language Model for Temporal Video Grounding

    Ye Wang, Ziheng Wang, Boshen Xu +14

    cs.CVcs.AIcs.CLarXiv:2503.13377v32025
  59. FlexPrefill: A Context-Aware Sparse Attention Mechanism for Efficient Long-Sequence Inference

    Xunhao Lai, Jianqiao Lu, Yao Luo +2

    cs.LGcs.CLarXiv:2502.20766v12025
  60. MemoryArena: Benchmarking Agent Memory in Interdependent Multi-Session Agentic Tasks

    Zexue He, Yu Wang, Churan Zhi +11

    cs.CLarXiv:2602.16313v12026