Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
61 to 120 of 11,225
The Fragility of Chain-of-Thought Monitoring Across Typologically Diverse Languages
Eric Onyame, Runtao Zhou, Kowshik Thopalli +2
cs.CLcs.AIarXiv:2605.27901v12026The MiniMax-M2 Series: Mini Activations Unleashing Max Real-World Intelligence
Aili Chen, Aonian Li, Baichuan Zhou +215
cs.AIcs.CLcs.LGarXiv:2605.26494v22026Verus-SpecGym: An Agentic Environment for Evaluating Specification Autoformalization
Anmol Agarwal, Natalie Neamtu, Pranjal Aggarwal +6
cs.SEcs.AIcs.CLarXiv:2605.26457v12026Forecasting Downstream Performance of LLMs With Proxy Metrics
Arkil Patel, Siva Reddy, Marius Mosbach +1
cs.CLcs.LGarXiv:2605.18607v12026Dynamic Skill Lifecycle Management for Agentic Reinforcement Learning
Junhao Shen, Teng Zhang, Xiaoyan Zhao +1
cs.LGcs.CLarXiv:2605.10923v22026MemPrivacy: Privacy-Preserving Personalized Memory Management for Edge-Cloud Agents
Yining Chen, Jihao Zhao, Bo Tang +5
cs.CRcs.CLarXiv:2605.09530v32026Language models fail at extended rule following
Tianxiang Dai, Jonathan Fan
cs.CLarXiv:2605.02028v22026Red Teaming Language Models with Language Models
Ethan Perez, Saffron Huang, Francis Song +6
cs.CLcs.AIcs.CRarXiv:2202.03286v12022Beyond SFT-to-RL: Pre-alignment via Black-Box On-Policy Distillation for Multimodal RL
Sudong Wang, Weiquan Huang, Xiaomin Yu +9
cs.CVcs.AIcs.CLarXiv:2604.28123v32026WindowsWorld: A Process-Centric Benchmark of Autonomous GUI Agents in Professional Cross-Application Environments
Jinchao Li, Yunxin Li, Chenrui Zhao +3
cs.AIcs.CLarXiv:2604.27776v12026Register Tokens for Bounded-State Reasoning in Diffusion Language Models
Albert Ge, Chandan Singh, Yufan Zhuang +3
cs.CLarXiv:2609.16372v12026Few-Shot NLG with Pre-Trained Language Model
Zhiyu Chen, Harini Eavani, Wenhu Chen +2
cs.CLarXiv:1904.09521v32019VLAA-GUI: Knowing When to Stop, Recover, and Search, A Modular Framework for GUI Automation
Qijun Han, Haoqin Tu, Zijun Wang +11
cs.CLcs.AIcs.SEarXiv:2604.21375v22026End-to-End Knowledge-Routed Relational Dialogue System for Automatic Diagnosis
Lin Xu, Qixian Zhou, Ke Gong +3
cs.CLarXiv:1901.10623v22019Scaling Test-Time Compute for Agentic Coding
Joongwon Kim, Wannan Yang, Kelvin Niu +13
cs.SEcs.AIcs.CLarXiv:2604.16529v12026NovBench: Evaluating Large Language Models on Academic Paper Novelty Assessment
Wenqing Wu, Yi Zhao, Yuzhuo Wang +4
cs.CLcs.AIcs.DLarXiv:2604.11543v12026CPT: A Pre-Trained Unbalanced Transformer for Both Chinese Language Understanding and Generation
Yunfan Shao, Zhichao Geng, Yitao Liu +6
cs.CLarXiv:2109.05729v42021Combee: Scaling Prompt Learning for Self-Improving Language Model Agents
Hanchen Li, Runyuan He, Qizheng Zhang +11
cs.AIcs.CLcs.LGarXiv:2604.04247v12026Sparse Readout Prism: Explaining Logit-Lens Scores in Features Instead of Tokens
Matteo He, William F. Shen, Xinchi Qiu +1
cs.CLcs.AIcs.LGarXiv:2609.01936v12026Probing Factual Knowledge Transfer with Training Data Interventions
Romina Oji, Marc Braun, Marcel Bollmann +2
cs.CLcs.AIarXiv:2609.01341v12026PhyCRNet: Physics-informed Convolutional-Recurrent Network for Solving Spatiotemporal PDEs
Pu Ren, Chengping Rao, Yang Liu +2
cs.LGcs.CLmath.NAarXiv:2106.14103v12021Post-Training Science for Supervised Fine-Tuning
Charles O'Neill, Mudith Jayasekara, Harry Partridge
cs.LGcs.CLarXiv:2609.01244v12026Seq2Seq-Vis: A Visual Debugging Tool for Sequence-to-Sequence Models
Hendrik Strobelt, Sebastian Gehrmann, Michael Behrisch +3
cs.CLcs.AIcs.NEarXiv:1804.09299v22018Fast Lexically Constrained Decoding with Dynamic Beam Allocation for Neural Machine Translation
Matt Post, David Vilar
cs.CLarXiv:1804.06609v22018Embarrassingly Simple Self-Distillation Improves Code Generation
Ruixiang Zhang, Richard He Bai, Huangjie Zheng +3
cs.CLarXiv:2604.01193v22026Demystifying When Pruning Works via Representation Hierarchies
Shwai He, Guoheng Sun, Haichao Zhang +2
cs.CLcs.LGarXiv:2603.24652v32026CAST: Critique-Aware Supervision for Training Reliable Long-Horizon Tool-Calling Agents
Amir Saeidi, Zehua Zhang, Rishitosh Singh +6
cs.CLarXiv:2608.30147v12026AgentHER: Hindsight Experience Replay for LLM Agent Trajectory Relabeling
Liang Ding
cs.AIcs.CLarXiv:2603.21357v42026AI Historian: Helping historians organize and verify person-centred temporal clues from dispersed historical narratives
Yifeng Lu, Zijie Yang, Jie Li +2
cs.CLarXiv:2608.29133v12026BERTology of Molecular Property Prediction
Mohammad Mostafanejad, Paul Saxe, T. Daniel Crawford
cs.LGcs.CLarXiv:2603.13627v12026ContextPilot: Teaching Agents for Proactive Context Management via Fine-grained RL
Zhuoshi Pan, Qizhi Pei, Junru Lu +4
cs.CLarXiv:2608.28476v12026Stranger, Fan, or Peer? A Systematic Study on the Role of Interlocutor in Persona-Based Dialogue Generation
Daniela Occhipinti, Malvina Nissim, Marco Guerini
cs.CLarXiv:2608.28467v12026Safe and Scalable Web Agent Learning via Recreated Websites
Hyungjoo Chae, Jungsoo Park, Alan Ritter
cs.CLarXiv:2603.10505v12026VideoSET: Video Summary Evaluation through Text
Serena Yeung, Alireza Fathi, Li Fei-Fei
cs.CVcs.CLcs.IRarXiv:1406.5824v12014Automated Phrase Mining from Massive Text Corpora
Jingbo Shang, Jialu Liu, Meng Jiang +3
cs.CLarXiv:1702.04457v22017DocTalkBN: A Novel Dataset of Expert Telemedicine Conversations in Bengali
Anik Saha, Fahmida Sultana Naznin, Sadatul Islam Sadi +3
cs.CLarXiv:2608.27110v12026Equal Ranking Quality, Different Decisions: Training Order-Consistent LLM Scorers
Markus Frohmann, Mahdiyar Alavi, Elizabeth Lingg +1
cs.CLcs.IRcs.LGarXiv:2608.26762v12026Not Just Reason, Not Just Scan: Reinforcement Learning for Proactive Scientific Error Verification over Academic Paper
Rongjin Li, Yuanxin Liu, Hao Zhou +3
cs.CLarXiv:2608.26596v12026TreeGraft: Adaptive Multi-Drafter Grafting for Tree-Based Speculative Decoding
Jiaming Fan, Daming Cao, Canchen Huang +6
cs.CLarXiv:2608.26112v12026ReIn: Conversational Error Recovery with Reasoning Inception
Takyoung Kim, Jinseok Nam, Chandrayee Basu +5
cs.CLcs.AIarXiv:2602.17022v12026Evaluating Memory Structure in LLM Agents
Alina Shutova, Alexandra Olenina, Ivan Vinogradov +1
cs.LGcs.CLarXiv:2602.11243v22026What is Wrong with Topic Modeling? (and How to Fix it Using Search-based Software Engineering)
Amritanshu Agrawal, Wei Fu, Tim Menzies
cs.SEcs.AIcs.CLarXiv:1608.08176v42016Continual Visual Learning under Evolving Semantic Concept Shift
Ismail Lamaakal, Chaymae Yahyati, Yassine Maleh +2
cs.CVcs.CLarXiv:2608.23903v12026Variational Neural Machine Translation
Biao Zhang, Deyi Xiong, Jinsong Su +2
cs.CLarXiv:1605.07869v22016Less Noise, More Voice: Reinforcement Learning for Reasoning via Instruction Purification
Yiju Guo, Tianyi Hu, Zexu Sun +1
cs.LGcs.AIcs.CLarXiv:2601.21244v32026Mem2ActBench: A Benchmark for Evaluating Long-Term Memory Utilization in Task-Oriented Autonomous Agents
Yiting Shen, Kun Li, Wei Zhou +1
cs.CLcs.AIarXiv:2601.19935v12026PropUQ-MAS: Propagation-Aware Uncertainty Quantification for LLM Multi-Agent Systems
Yaokun Liu, Yifan Liu, Daniel Yue Zhang +3
cs.MAcs.CLarXiv:2608.22130v12026On the Role of Citations in Preference Data
Yu Hou, Hal Daumé, Rachel Rudinger +1
cs.CLcs.AIarXiv:2608.21376v12026Beyond Hard Masks: Progressive Token Evolution for Diffusion Language Models
Linhao Zhong, Linyu Wu, Bozhen Fang +6
cs.CLcs.AIarXiv:2601.07351v22026Beyond Static Summarization: Proactive Memory Extraction for LLM Agents
Chengyuan Yang, Zequn Sun, Wei Wei +1
cs.CLcs.AIarXiv:2601.04463v22026MentorPulse: Refreshing Cross-Model Latent Guidance for Long-Form Generation
Ziwu Liu, Guozhong Li, Chen Qiu +2
cs.CLcs.AIarXiv:2608.20927v12026MemEvolve: Meta-Evolution of Agent Memory Systems
Guibin Zhang, Haotian Ren, Chong Zhan +5
cs.CLcs.MAarXiv:2512.18746v12025Toward Auto-Research: Mining Falsifiable Research Ideas from Paper Knowledge Graphs with Categorical Structure
Yuchen Wang, Zhongzhi Luan
cs.CLcs.AIarXiv:2608.20361v12026How to Train a Real-World Silicon Concierge? Internalizing Complex Business Workflow to Only OneModel
Chang Liu, Chaoyang Ning, Dayi Jiang +32
cs.CLcs.AIarXiv:2608.20350v12026What is Missing from AI Post-Training AI: An Empirical Analysis
Joy Jia Yin Lim, Xin Huang, Hao Peng +5
cs.AIcs.CLcs.LGarXiv:2608.19072v12026The Plot Thins: Uniformity and Linearity in Literary Summaries
Rebecca M. M. Hicke, Sil Hamilton, David Mimno +1
cs.CLarXiv:2608.17218v12026A Survey of Reinforcement Learning for Large Reasoning Models
Kaiyan Zhang, Yuxin Zuo, Bingxiang He +36
cs.CLcs.AIcs.LGarXiv:2509.08827v32025Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
Gheorghe Comanici, Eric Bieber, Mike Schaekermann +3432
cs.CLcs.AIarXiv:2507.06261v62025Pre-Training BERT on Arabic Tweets: Practical Considerations
Ahmed Abdelali, Sabit Hassan, Hamdy Mubarak +2
cs.CLcs.AIarXiv:2102.10684v12021Moral Stories: Situated Reasoning about Norms, Intents, Actions, and their Consequences
Denis Emelin, Ronan Le Bras, Jena D. Hwang +2
cs.CLcs.AIarXiv:2012.15738v12020