Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,081 to 4,140 of 11,252

  1. Revisiting Lossy Verification in Speculative Decoding: Mechanisms, Trade-offs, and Failure Modes

    Tianyu Wang, Yuxuan Zhou, Wenbin Wang +3

    cs.CLarXiv:2607.26627v12026
  2. Kimi Linear: An Expressive, Efficient Attention Architecture

    Kimi Team, Yu Zhang, Zongyu Lin +57

    cs.CLcs.LGarXiv:2510.26692v22025
  3. Token Time Continuous Diffusion for Language Modeling

    Parikshit Bansal, Sujay Sanghavi

    cs.CLcs.AIarXiv:2607.14106v12026
  4. Group Entropy-Controlled Policy Optimization

    Guangran Cheng, Chengqi Lyu, Songyang Gao +2

    cs.CLarXiv:2607.16850v12026
  5. UI-MOPD: Multi-Platform On-Policy Distillation for Unified GUI Agents

    Niu Lian, Tongbo Chen, Zhehao Yu +8

    cs.CLcs.AIcs.CVarXiv:2607.04425v22026
  6. SeKV: Resolution-Adaptive KV Cache with Hierarchical Semantic Memory for Long-Context LLM Inference

    Amirhossein Abaskohi, Giuseppe Carenini, Peter West +1

    cs.CLarXiv:2606.31145v12026
  7. Capable but Careless: Do Computer-Use Agents Follow Contextual Integrity?

    Anmol Goel, Iryna Gurevych

    cs.AIcs.CLarXiv:2606.23189v12026
  8. Orchestra-o1: Omnimodal Agent Orchestration

    Fan Zhang, Vireo Zhang, Shengju Qian +8

    cs.AIcs.CLcs.CVarXiv:2606.13707v12026
  9. Redesign Mixture-of-Experts Routers with Manifold Power Iteration

    Songhao Wu, Ang Lv, Ruobing Xie +1

    cs.LGcs.AIcs.CLarXiv:2606.12397v12026
  10. UnpredictaBench: A Benchmark for Evaluating Distributional Randomness in LLMs

    Amirhossein Abaskohi, Amirhossein Dabiriaghdam, Liang Luo +4

    cs.CLarXiv:2606.06622v32026
  11. Do Coding Agents Deceive Us? Detecting and Preventing Cheating via Capped Evaluation with Randomized Tests

    Thanawat Lodkaew, Johannes Ackermann, Soichiro Nishimori +3

    cs.LGcs.AIcs.CLarXiv:2606.07379v22026
  12. AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints

    Jiayu Liu, Cheng Qian, Zhenhailong Wang +10

    cs.CLarXiv:2606.05622v22026
  13. GRAIL: Gradient-Reweighted Advantages for Reinforcement Learning with Verifiable Rewards

    Tej Deep Pala, Vernon Toh, Soujanya Poria

    cs.CLarXiv:2606.04889v12026
  14. OmniOPD: Logit-Free On-Policy Distillation via Speculative Verification

    Yuhang Zhou, Lizhu Zhang, Yifan Wu +5

    cs.LGcs.CLarXiv:2606.01476v22026
  15. MemTrace: Tracing and Attributing Errors in Large Language Model Memory Systems

    Xinle Deng, Ruobin Zhong, Hujin Peng +15

    cs.CLcs.AIcs.LGarXiv:2605.28732v32026
  16. Parallax: Parameterized Local Linear Attention for Language Modeling

    Yifei Zuo, Dhruv Pai, Zhichen Zeng +3

    cs.LGcs.AIcs.CLarXiv:2605.29157v12026
  17. RUBRIC-ARROW: Alternating Pointwise Rubric Reward Modeling for LLM Post-training in Non-verifiable Domains

    Haoxiang Jiang, Zihan Dong, Tianci Liu +5

    cs.LGcs.CLarXiv:2605.29156v12026
  18. Pruning and Distilling Mixture-of-Experts into Dense Language Models

    Junhyuck Kim, Jihun Yun, Haechan Kim +3

    cs.CLcs.AIcs.LGarXiv:2605.28207v22026
  19. Monitoring the Internal Monologue: Probe Trajectories Reveal Reasoning Dynamics

    Maciej Chrabąszcz, Aleksander Szymczyk, Marcin Sendera +2

    cs.CLcs.CRarXiv:2605.18549v12026
  20. optimize_anything: A Universal API for Optimizing any Text Parameter

    Lakshya A Agrawal, Donghyun Lee, Shangyin Tan +11

    cs.CLcs.AIcs.LGarXiv:2605.19633v12026
  21. Nudging Beyond the Comfort Zone: Efficient Strategy-Guided Exploration for RLVR

    Chanuk Lee, Sangwoo Park, Minki Kang +1

    cs.AIcs.CLarXiv:2605.15726v12026
  22. Forgetting That Sticks: Quantization-Permanent Unlearning via Circuit Attribution

    Saisab Sadhu, Pratinav Seth, Vinay Kumar Sankarapu

    cs.LGcs.CLcs.ETarXiv:2605.15138v12026
  23. HodgeCover: Higher-Order Topological Coverage Drives Compression of Sparse Mixture-of-Experts

    Tao Zhong, Dongzhe Zheng, Christine Allen-Blanchette

    cs.LGcs.AIcs.CLarXiv:2605.13997v12026
  24. Crosslingual On-Policy Self-Distillation for Multilingual Reasoning

    Yihong Liu, Raoyuan Zhao, Michael A. Hedderich +1

    cs.CLarXiv:2605.09548v12026
    Summaries:한국어
  25. Reliable Chain-of-Thought via Prefix Consistency

    Naoto Iwase, Yuki Ichihara, Mohammad Atif Quamar +1

    stat.MLcs.CLcs.LGarXiv:2605.07654v12026
  26. G-Zero: Self-Play for Open-Ended Generation from Zero Data

    Chengsong Huang, Haolin Liu, Tong Zheng +7

    cs.LGcs.AIcs.CLarXiv:2605.09959v12026
  27. Rethinking State Tracking in Recurrent Models Through Error Control Dynamics

    Jiwan Chung, Heechan Choi, Seon Joo Kim

    cs.LGcs.CLarXiv:2605.07755v12026
  28. Addressing Performance Saturation for LLM RL via Precise Entropy Curve Control

    Bolian Li, Yifan Wang, Yi Ding +3

    cs.LGcs.CLstat.MLarXiv:2604.26326v22026
  29. Rewarding the Scientific Process: Process-Level Reward Modeling for Agentic Data Analysis

    Zhisong Qiu, Shuofei Qiao, Kewei Xu +4

    cs.CLcs.AIcs.CEarXiv:2604.24198v22026
  30. R$^3$-SQL: Ranking Reward and Resampling for Text-to-SQL

    Hojae Han, Yeonseok Jeong, Seung-won Hwang +2

    cs.SEcs.AIcs.CLarXiv:2604.25325v12026
  31. Convergent Evolution: How Different Language Models Learn Similar Number Representations

    Deqing Fu, Tianyi Zhou, Mikhail Belkin +2

    cs.CLcs.AIcs.LGarXiv:2604.20817v22026
  32. On the Robustness of LLM-Based Dense Retrievers: A Systematic Analysis of Generalizability and Stability

    Yongkang Li, Panagiotis Eustratiadis, Yixing Fan +1

    cs.IRcs.CLarXiv:2604.16576v12026
  33. $p1$: Better Prompt Optimization with Fewer Prompts

    Zhaolin Gao, Yu, Wang +4

    cs.LGcs.CLarXiv:2604.08801v12026
  34. Learning to Hint for Reinforcement Learning

    Yu Xia, Canwen Xu, Zhewei Yao +2

    cs.LGcs.AIcs.CLarXiv:2604.00698v12026
  35. Locally Confident, Globally Stuck: The Quality-Exploration Dilemma in Diffusion Language Models

    Liancheng Fang, Aiwei Liu, Henry Peng Zou +7

    cs.CLarXiv:2604.00375v12026
  36. Parametric Social Identity Injection and Diversification in Public Opinion Simulation

    Hexi Wang, Yujia Zhou, Bangde Du +2

    cs.CLarXiv:2603.16142v22026
  37. Structural Abstraction as an Inductive Bias for Non-Stationary Language Model Training

    Elnaz Rahmati, Nona Ghazizadeh, Zhivar Sourati +2

    cs.LGcs.CLarXiv:2603.17198v22026
  38. Code-A1: Adversarial Evolving of Code LLM and Test LLM via Reinforcement Learning

    Aozhe Wang, Yuchen Yan, Nan Zhou +5

    cs.CLarXiv:2603.15611v12026
  39. ReMix: Reinforcement routing for mixtures of LoRAs in LLM finetuning

    Ruizhong Qiu, Hanqing Zeng, Yinglong Xia +15

    cs.LGcs.CLarXiv:2603.10160v12026
  40. SciDER: Scientific Data-centric End-to-end Researcher

    Ke Lin, Owais Aijaz, Yilin Lu +3

    cs.AIcs.CLarXiv:2603.01421v32026
  41. Truncated Step-Level Sampling with Process Rewards for Retrieval-Augmented Reasoning

    Chris Samarinas, Haw-Shiuan Chang, Hamed Zamani

    cs.CLcs.IRarXiv:2602.23440v42026
  42. Learning Personalized Agents from Human Feedback

    Kaiqu Liang, Julia Kruk, Shengyi Qian +9

    cs.AIcs.CLcs.LGarXiv:2602.16173v12026
  43. STAPO: Stabilizing Reinforcement Learning for LLMs by Silencing Rare Spurious Tokens

    Shiqi Liu, Zeyu He, Guojian Zhan +10

    cs.CLcs.AIarXiv:2602.15620v52026
  44. Group Distributionally Robust Optimization-Driven Reinforcement Learning for LLM Reasoning

    Kishan Panaganti, Zhenwen Liang, Wenhao Yu +2

    cs.LGcs.AIcs.CLarXiv:2601.19280v12026
  45. CooperBench: Why Coding Agents Cannot be Your Teammates Yet

    Arpandeep Khatua, Hao Zhu, Peter Tran +8

    cs.LGcs.AIcs.CLarXiv:2601.13295v22026
  46. Harder Is Better: Boosting Mathematical Reasoning via Difficulty-Aware GRPO and Multi-Aspect Question Reformulation

    Yanqi Dai, Yuxiang Ji, Xiao Zhang +3

    cs.AIcs.CLarXiv:2601.20614v12026
  47. AgentIF-OneDay: A Task-level Instruction-Following Benchmark for General AI Agents in Daily Scenarios

    Kaiyuan Chen, Qimin Wu, Taiyu Hou +42

    cs.CLarXiv:2601.20613v22026
  48. ProFit: Leveraging High-Value Signals in SFT via Probability-Guided Token Selection

    Tao Liu, Taiqiang Wu, Runming Yang +3

    cs.CLcs.AIarXiv:2601.09195v32026
  49. Imagine-then-Plan: Agent Learning from Adaptive Lookahead with World Models

    Youwei Liu, Jian Wang, Hanlin Wang +2

    cs.CLcs.AIcs.LGarXiv:2601.08955v22026
  50. $τ^2$-Bench: Evaluating Conversational Agents in a Dual-Control Environment

    Victor Barres, Honghua Dong, Soham Ray +2

    cs.AIcs.CLarXiv:2506.07982v12025
  51. Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

    An Yang, Beichen Zhang, Binyuan Hui +13

    cs.CLcs.AIcs.LGarXiv:2409.12122v12024
  52. End-to-End Bias Mitigation by Modelling Biases in Corpora

    Rabeeh Karimi Mahabadi, Yonatan Belinkov, James Henderson

    cs.CLarXiv:1909.06321v32019
  53. Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

    Gemini Team, Petko Georgiev, Ving Ian Lei +1134

    cs.CLcs.AIarXiv:2403.05530v52024
  54. RankRAG: Unifying Context Ranking with Retrieval-Augmented Generation in LLMs

    Yue Yu, Wei Ping, Zihan Liu +5

    cs.CLcs.AIcs.IRarXiv:2407.02485v12024
  55. OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments

    Tianbao Xie, Danyang Zhang, Jixuan Chen +14

    cs.AIcs.CLarXiv:2404.07972v22024
  56. WildBench: Benchmarking LLMs with Challenging Tasks from Real Users in the Wild

    Bill Yuchen Lin, Yuntian Deng, Khyathi Chandu +6

    cs.CLcs.AIarXiv:2406.04770v22024
  57. The Dawn of LMMs: Preliminary Explorations with GPT-4V(ision)

    Zhengyuan Yang, Linjie Li, Kevin Lin +4

    cs.CVcs.CLarXiv:2309.17421v22023
  58. The opportunities and risks of large language models in mental health

    Hannah R. Lawrence, Renee A. Schneider, Susan B. Rubin +3

    cs.CLcs.AIcs.CYarXiv:2403.14814v32024
  59. We Can Detect Your Bias: Predicting the Political Ideology of News Articles

    Ramy Baly, Giovanni Da San Martino, James Glass +1

    cs.CLarXiv:2010.05338v12020
  60. Griffin: Mixing Gated Linear Recurrences with Local Attention for Efficient Language Models

    Soham De, Samuel L. Smith, Anushan Fernando +14

    cs.LGcs.CLarXiv:2402.19427v12024