Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,441 to 4,500 of 11,259
The Lessons of Developing Process Reward Models in Mathematical Reasoning
Zhenru Zhang, Chujie Zheng, Yangzhen Wu +6
cs.CLcs.AIcs.LGarXiv:2501.07301v22025The Right Tool for the Job: Matching Model and Instance Complexities
Roy Schwartz, Gabriel Stanovsky, Swabha Swayamdipta +2
cs.CLcs.LGarXiv:2004.07453v22020Search-o1: Agentic Search-Enhanced Large Reasoning Models
Xiaoxi Li, Guanting Dong, Jiajie Jin +5
cs.AIcs.CLcs.IRarXiv:2501.05366v12025Reasoning with Exploration: An Entropy Perspective
Daixuan Cheng, Shaohan Huang, Xuekai Zhu +4
cs.CLarXiv:2506.14758v42025Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs
Kanishk Gandhi, Ayush Chakravarthy, Anikait Singh +2
cs.CLcs.LGarXiv:2503.01307v22025Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models
Qizheng Zhang, Changran Hu, Shubhangi Upasani +10
cs.LGcs.AIcs.CLarXiv:2510.04618v32025Process Reinforcement through Implicit Rewards
Ganqu Cui, Lifan Yuan, Zefan Wang +22
cs.LGcs.AIcs.CLarXiv:2502.01456v22025From Reweighting to Rewriting: Unlocking the Intervention Effects of Influential Samples in Training Data Attribution
Yuzhang Luo, Chenpeng Wang, Jianhui Chen +1
cs.CLcs.AIcs.LGarXiv:2609.02771v12026Political Compass or Spinning Arrow? Towards More Meaningful Evaluations for Values and Opinions in Large Language Models
Paul Röttger, Valentin Hofmann, Valentina Pyatkin +4
cs.CLcs.AIarXiv:2402.16786v22024DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
Huajian Xin, Z. Z. Ren, Junxiao Song +14
cs.CLcs.AIcs.LGarXiv:2408.08152v12024ELEVATER: A Benchmark and Toolkit for Evaluating Language-Augmented Visual Models
Chunyuan Li, Haotian Liu, Liunian Harold Li +8
cs.CVcs.CLcs.LGarXiv:2204.08790v62022SMHD: A Large-Scale Resource for Exploring Online Language Usage for Multiple Mental Health Conditions
Arman Cohan, Bart Desmet, Andrew Yates +3
cs.CLarXiv:1806.05258v22018Multi-Source Domain Adaptation with Mixture of Experts
Jiang Guo, Darsh J Shah, Regina Barzilay
cs.CLarXiv:1809.02256v22018C3oT: Generating Shorter Chain-of-Thought without Compromising Effectiveness
Yu Kang, Xianghui Sun, Liangyu Chen +1
cs.CLcs.LGarXiv:2412.11664v12024VakyArth: Evaluating Pragmatic Competence in LLMs across Indic Languages
Usneek Singh, Poorvaja Veera Balaji Kumar, Parth Nanda +4
cs.CLcs.AIarXiv:2609.01788v12026Jointly embedding the local and global relations of heterogeneous graph for rumor detection
Chunyuan Yuan, Qianwen Ma, Wei Zhou +2
cs.CLcs.IRcs.SIarXiv:1909.04465v22019Stochastic Multiple Choice Learning for Training Diverse Deep Ensembles
Stefan Lee, Senthil Purushwalkam, Michael Cogswell +3
cs.CVcs.CLarXiv:1606.07839v32016FlowSeq: Non-Autoregressive Conditional Sequence Generation with Generative Flow
Xuezhe Ma, Chunting Zhou, Xian Li +2
cs.CLcs.LGarXiv:1909.02480v32019NE-R1: Enhancing Named Entity Recognition Model via Reinforcement Learning
Meixuan Chen, Hehan Li, Ruizhi Zhao +8
cs.CLcs.AIarXiv:2609.02366v12026Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention
Jingyang Yuan, Huazuo Gao, Damai Dai +12
cs.CLcs.AIcs.LGarXiv:2502.11089v22025ColPali: Efficient Document Retrieval with Vision Language Models
Manuel Faysse, Hugues Sibille, Tony Wu +4
cs.IRcs.CLcs.CVarXiv:2407.01449v62024Automatic Prompt Augmentation and Selection with Chain-of-Thought from Labeled Data
KaShun Shum, Shizhe Diao, Tong Zhang
cs.CLarXiv:2302.12822v32023VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model
Haozhan Shen, Peng Liu, Jingcheng Li +9
cs.CVcs.CLarXiv:2504.07615v22025Interactive and Visual Prompt Engineering for Ad-hoc Task Adaptation with Large Language Models
Hendrik Strobelt, Albert Webson, Victor Sanh +4
cs.CLcs.HCcs.LGarXiv:2208.07852v12022Automated Concatenation of Embeddings for Structured Prediction
Xinyu Wang, Yong Jiang, Nguyen Bach +4
cs.CLcs.AIcs.LGarXiv:2010.05006v42020Learning to Understand Phrases by Embedding the Dictionary
Felix Hill, Kyunghyun Cho, Anna Korhonen +1
cs.CLarXiv:1504.00548v42015Select, Compress, Reinvest: A Controlled Study of Visual-Token Allocation in Long-Video MLLMs
Prakhar Khatri
cs.CVcs.CLarXiv:2609.03820v12026ReTool: Reinforcement Learning for Strategic Tool Use in LLMs
Jiazhan Feng, Shijue Huang, Xingwei Qu +6
cs.CLcs.AIarXiv:2504.11536v22025Reasoning Models Don't Always Say What They Think
Yanda Chen, Joe Benton, Ansh Radhakrishnan +12
cs.CLcs.AIcs.LGarXiv:2505.05410v12025HyperStyler: Low-resource Authorship Style Transfer via Context-aware Style Navigation and Hypernetworks
Jongkyung Shin, Minguk Jeon, Chanwoo Park +1
cs.CLarXiv:2609.02772v12026Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking
Mingxin Li, Yanzhao Zhang, Dingkun Long +9
cs.CLarXiv:2601.04720v22026Kimi K2: Open Agentic Intelligence
Kimi Team, Yifan Bai, Yiping Bao +197
cs.LGcs.AIcs.CLarXiv:2507.20534v22025Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model
Jingcheng Hu, Yinmin Zhang, Qi Han +3
cs.LGcs.CLarXiv:2503.24290v22025GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models
GLM-4. 5 Team, :, Aohan Zeng +169
cs.CLarXiv:2508.06471v12025Scalable Kronecker-Fisher Approximation: Efficient Hessian Analysis for Billion-Parameter Language Models Compression
Viacheslav Yusupov, Daria Cherniuk, Evgeny Frolov
cs.LGcs.AIcs.CLarXiv:2609.02451v12026Bounded Personas Match Retrieval on Classification but Not Regression for a Frozen Agent
JaeHa Yoon, Minjun Park, Seoyeon Kim +3
cs.CLarXiv:2609.02890v12026Where Does Harness-Optimization Value Live? Localized Gains and the Budget-Splitting Trap in Self-Evolving LLM Agents
Michael Nguyen, Wei Chen Tan, Nurul Aisyah Hassan +3
cs.CLarXiv:2609.02889v12026Learning to Fuse LLMs with Ontology Rankers for Rare-Disease Diagnosis
Zhaoyang Jiang, Zhizhong Fu, Yunsoo Kim +5
cs.CLarXiv:2609.02473v12026CORAL: An LLM-Native Harness for Production Recommender Systems
Muhammad Rafay Azhar, Yuhang Zhou, Gilbert Jiang +7
cs.CLarXiv:2609.02730v12026AI agents reshape consensus formation in human groups
Lin Chen, Ziyi Liu, Xia Hu +1
cs.CLcs.CYcs.SIarXiv:2609.02122v12026Speculative Decoding: Exploiting Speculative Execution for Accelerating Seq2seq Generation
Heming Xia, Tao Ge, Peiyi Wang +3
cs.CLcs.LGarXiv:2203.16487v62022R$^{2}$Adapter: A Routing and Rewriting Adapter for Efficient Hybrid RAG
Yucan Guo, Miao Su, Saiping Guan +6
cs.CLcs.IRarXiv:2609.02894v12026Unsupervised Word and Dependency Path Embeddings for Aspect Term Extraction
Yichun Yin, Furu Wei, Li Dong +3
cs.CLarXiv:1605.07843v12016End-to-end Generative Pretraining for Multimodal Video Captioning
Paul Hongsuck Seo, Arsha Nagrani, Anurag Arnab +1
cs.CVcs.AIcs.CLarXiv:2201.08264v22022SituatedQA: Incorporating Extra-Linguistic Contexts into QA
Michael J. Q. Zhang, Eunsol Choi
cs.CLarXiv:2109.06157v12021Margins, Not Windows: Training-Free Per-Step Lossy Speculative Decoding
Oszkár Urbán, Young D. Kwon, Stylianos I. Venieris +1
cs.CLarXiv:2609.02897v12026Evaluating Prerequisite Qualities for Learning End-to-End Dialog Systems
Jesse Dodge, Andreea Gane, Xiang Zhang +5
cs.CLcs.LGarXiv:1511.06931v62015Probe Generalization as Subspace Selection for OOD Deception Detection
Daniel Yoo, Adrians Skapars
cs.CLarXiv:2609.02893v12026Counterexamples as Feedback for Agent Self-Correction
Sidhesh Badrinarayan, Adithya Parthasarathy
cs.CLcs.AIarXiv:2609.02892v12026SenseBERT: Driving Some Sense into BERT
Yoav Levine, Barak Lenz, Or Dagan +6
cs.CLcs.LGarXiv:1908.05646v22019Before the Script, Set the Stage: How Worldview Simulation Amplifies Psychologically Grounded Persuasion in Multi-Turn Jailbreaking
Siyu Chen, Haoran Wang, Xiaojian Li +3
cs.CLcs.AIarXiv:2609.02414v12026BharatGather: A Culturally-Informed Benchmark Dataset for Misinformation and Fake News Detection in Indian Public Events
Parth Bramhecha, Smit Deshmukh, Sairaj Bodhale +2
cs.CLcs.LGarXiv:2609.02895v12026Personalizing Dialogue Agents via Meta-Learning
Zhaojiang Lin, Andrea Madotto, Chien-Sheng Wu +1
cs.CLcs.AIarXiv:1905.10033v12019Text Segmentation as a Supervised Learning Task
Omri Koshorek, Adir Cohen, Noam Mor +2
cs.CLarXiv:1803.09337v12018SFAD: Speculative Factuality-Aware Decoding
Guanqiao Chen, Di Wang, Lijie Hu
cs.CLarXiv:2609.00796v22026SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild
Weihao Zeng, Yuzhen Huang, Qian Liu +4
cs.LGcs.AIcs.CLarXiv:2503.18892v32025Scaling Near-Optimal SFT-RL Annotation Budget Allocation from Small to Large LLMs
Jingtan Wang, Arun Verma, Xiaoqiang Lin +4
cs.CLcs.AIcs.LGarXiv:2609.01573v12026Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step
Li Zhong, Zilong Wang, Jingbo Shang
cs.SEcs.AIcs.CLarXiv:2402.16906v62024GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning
Lakshya A Agrawal, Shangyin Tan, Dilara Soylu +14
cs.CLcs.AIcs.LGarXiv:2507.19457v22025Sequence-to-Sequence Knowledge Graph Completion and Question Answering
Apoorv Saxena, Adrian Kochsiek, Rainer Gemulla
cs.CLcs.LGarXiv:2203.10321v12022