Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,441 to 4,500 of 11,259

  1. The Lessons of Developing Process Reward Models in Mathematical Reasoning

    Zhenru Zhang, Chujie Zheng, Yangzhen Wu +6

    cs.CLcs.AIcs.LGarXiv:2501.07301v22025
  2. The Right Tool for the Job: Matching Model and Instance Complexities

    Roy Schwartz, Gabriel Stanovsky, Swabha Swayamdipta +2

    cs.CLcs.LGarXiv:2004.07453v22020
  3. Search-o1: Agentic Search-Enhanced Large Reasoning Models

    Xiaoxi Li, Guanting Dong, Jiajie Jin +5

    cs.AIcs.CLcs.IRarXiv:2501.05366v12025
  4. Reasoning with Exploration: An Entropy Perspective

    Daixuan Cheng, Shaohan Huang, Xuekai Zhu +4

    cs.CLarXiv:2506.14758v42025
  5. Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs

    Kanishk Gandhi, Ayush Chakravarthy, Anikait Singh +2

    cs.CLcs.LGarXiv:2503.01307v22025
  6. Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models

    Qizheng Zhang, Changran Hu, Shubhangi Upasani +10

    cs.LGcs.AIcs.CLarXiv:2510.04618v32025
  7. Process Reinforcement through Implicit Rewards

    Ganqu Cui, Lifan Yuan, Zefan Wang +22

    cs.LGcs.AIcs.CLarXiv:2502.01456v22025
  8. From Reweighting to Rewriting: Unlocking the Intervention Effects of Influential Samples in Training Data Attribution

    Yuzhang Luo, Chenpeng Wang, Jianhui Chen +1

    cs.CLcs.AIcs.LGarXiv:2609.02771v12026
  9. Political Compass or Spinning Arrow? Towards More Meaningful Evaluations for Values and Opinions in Large Language Models

    Paul Röttger, Valentin Hofmann, Valentina Pyatkin +4

    cs.CLcs.AIarXiv:2402.16786v22024
  10. DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

    Huajian Xin, Z. Z. Ren, Junxiao Song +14

    cs.CLcs.AIcs.LGarXiv:2408.08152v12024
  11. ELEVATER: A Benchmark and Toolkit for Evaluating Language-Augmented Visual Models

    Chunyuan Li, Haotian Liu, Liunian Harold Li +8

    cs.CVcs.CLcs.LGarXiv:2204.08790v62022
  12. SMHD: A Large-Scale Resource for Exploring Online Language Usage for Multiple Mental Health Conditions

    Arman Cohan, Bart Desmet, Andrew Yates +3

    cs.CLarXiv:1806.05258v22018
  13. Multi-Source Domain Adaptation with Mixture of Experts

    Jiang Guo, Darsh J Shah, Regina Barzilay

    cs.CLarXiv:1809.02256v22018
  14. C3oT: Generating Shorter Chain-of-Thought without Compromising Effectiveness

    Yu Kang, Xianghui Sun, Liangyu Chen +1

    cs.CLcs.LGarXiv:2412.11664v12024
  15. VakyArth: Evaluating Pragmatic Competence in LLMs across Indic Languages

    Usneek Singh, Poorvaja Veera Balaji Kumar, Parth Nanda +4

    cs.CLcs.AIarXiv:2609.01788v12026
  16. Jointly embedding the local and global relations of heterogeneous graph for rumor detection

    Chunyuan Yuan, Qianwen Ma, Wei Zhou +2

    cs.CLcs.IRcs.SIarXiv:1909.04465v22019
  17. Stochastic Multiple Choice Learning for Training Diverse Deep Ensembles

    Stefan Lee, Senthil Purushwalkam, Michael Cogswell +3

    cs.CVcs.CLarXiv:1606.07839v32016
  18. FlowSeq: Non-Autoregressive Conditional Sequence Generation with Generative Flow

    Xuezhe Ma, Chunting Zhou, Xian Li +2

    cs.CLcs.LGarXiv:1909.02480v32019
  19. NE-R1: Enhancing Named Entity Recognition Model via Reinforcement Learning

    Meixuan Chen, Hehan Li, Ruizhi Zhao +8

    cs.CLcs.AIarXiv:2609.02366v12026
  20. Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention

    Jingyang Yuan, Huazuo Gao, Damai Dai +12

    cs.CLcs.AIcs.LGarXiv:2502.11089v22025
  21. ColPali: Efficient Document Retrieval with Vision Language Models

    Manuel Faysse, Hugues Sibille, Tony Wu +4

    cs.IRcs.CLcs.CVarXiv:2407.01449v62024
  22. Automatic Prompt Augmentation and Selection with Chain-of-Thought from Labeled Data

    KaShun Shum, Shizhe Diao, Tong Zhang

    cs.CLarXiv:2302.12822v32023
  23. VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model

    Haozhan Shen, Peng Liu, Jingcheng Li +9

    cs.CVcs.CLarXiv:2504.07615v22025
  24. Interactive and Visual Prompt Engineering for Ad-hoc Task Adaptation with Large Language Models

    Hendrik Strobelt, Albert Webson, Victor Sanh +4

    cs.CLcs.HCcs.LGarXiv:2208.07852v12022
  25. Automated Concatenation of Embeddings for Structured Prediction

    Xinyu Wang, Yong Jiang, Nguyen Bach +4

    cs.CLcs.AIcs.LGarXiv:2010.05006v42020
  26. Learning to Understand Phrases by Embedding the Dictionary

    Felix Hill, Kyunghyun Cho, Anna Korhonen +1

    cs.CLarXiv:1504.00548v42015
  27. Select, Compress, Reinvest: A Controlled Study of Visual-Token Allocation in Long-Video MLLMs

    Prakhar Khatri

    cs.CVcs.CLarXiv:2609.03820v12026
  28. ReTool: Reinforcement Learning for Strategic Tool Use in LLMs

    Jiazhan Feng, Shijue Huang, Xingwei Qu +6

    cs.CLcs.AIarXiv:2504.11536v22025
  29. Reasoning Models Don't Always Say What They Think

    Yanda Chen, Joe Benton, Ansh Radhakrishnan +12

    cs.CLcs.AIcs.LGarXiv:2505.05410v12025
  30. HyperStyler: Low-resource Authorship Style Transfer via Context-aware Style Navigation and Hypernetworks

    Jongkyung Shin, Minguk Jeon, Chanwoo Park +1

    cs.CLarXiv:2609.02772v12026
  31. Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking

    Mingxin Li, Yanzhao Zhang, Dingkun Long +9

    cs.CLarXiv:2601.04720v22026
  32. Kimi K2: Open Agentic Intelligence

    Kimi Team, Yifan Bai, Yiping Bao +197

    cs.LGcs.AIcs.CLarXiv:2507.20534v22025
  33. Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model

    Jingcheng Hu, Yinmin Zhang, Qi Han +3

    cs.LGcs.CLarXiv:2503.24290v22025
  34. GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models

    GLM-4. 5 Team, :, Aohan Zeng +169

    cs.CLarXiv:2508.06471v12025
  35. Scalable Kronecker-Fisher Approximation: Efficient Hessian Analysis for Billion-Parameter Language Models Compression

    Viacheslav Yusupov, Daria Cherniuk, Evgeny Frolov

    cs.LGcs.AIcs.CLarXiv:2609.02451v12026
  36. Bounded Personas Match Retrieval on Classification but Not Regression for a Frozen Agent

    JaeHa Yoon, Minjun Park, Seoyeon Kim +3

    cs.CLarXiv:2609.02890v12026
  37. Where Does Harness-Optimization Value Live? Localized Gains and the Budget-Splitting Trap in Self-Evolving LLM Agents

    Michael Nguyen, Wei Chen Tan, Nurul Aisyah Hassan +3

    cs.CLarXiv:2609.02889v12026
  38. Learning to Fuse LLMs with Ontology Rankers for Rare-Disease Diagnosis

    Zhaoyang Jiang, Zhizhong Fu, Yunsoo Kim +5

    cs.CLarXiv:2609.02473v12026
  39. CORAL: An LLM-Native Harness for Production Recommender Systems

    Muhammad Rafay Azhar, Yuhang Zhou, Gilbert Jiang +7

    cs.CLarXiv:2609.02730v12026
  40. AI agents reshape consensus formation in human groups

    Lin Chen, Ziyi Liu, Xia Hu +1

    cs.CLcs.CYcs.SIarXiv:2609.02122v12026
  41. Speculative Decoding: Exploiting Speculative Execution for Accelerating Seq2seq Generation

    Heming Xia, Tao Ge, Peiyi Wang +3

    cs.CLcs.LGarXiv:2203.16487v62022
  42. R$^{2}$Adapter: A Routing and Rewriting Adapter for Efficient Hybrid RAG

    Yucan Guo, Miao Su, Saiping Guan +6

    cs.CLcs.IRarXiv:2609.02894v12026
  43. Unsupervised Word and Dependency Path Embeddings for Aspect Term Extraction

    Yichun Yin, Furu Wei, Li Dong +3

    cs.CLarXiv:1605.07843v12016
  44. End-to-end Generative Pretraining for Multimodal Video Captioning

    Paul Hongsuck Seo, Arsha Nagrani, Anurag Arnab +1

    cs.CVcs.AIcs.CLarXiv:2201.08264v22022
  45. SituatedQA: Incorporating Extra-Linguistic Contexts into QA

    Michael J. Q. Zhang, Eunsol Choi

    cs.CLarXiv:2109.06157v12021
  46. Margins, Not Windows: Training-Free Per-Step Lossy Speculative Decoding

    Oszkár Urbán, Young D. Kwon, Stylianos I. Venieris +1

    cs.CLarXiv:2609.02897v12026
  47. Evaluating Prerequisite Qualities for Learning End-to-End Dialog Systems

    Jesse Dodge, Andreea Gane, Xiang Zhang +5

    cs.CLcs.LGarXiv:1511.06931v62015
  48. Probe Generalization as Subspace Selection for OOD Deception Detection

    Daniel Yoo, Adrians Skapars

    cs.CLarXiv:2609.02893v12026
  49. Counterexamples as Feedback for Agent Self-Correction

    Sidhesh Badrinarayan, Adithya Parthasarathy

    cs.CLcs.AIarXiv:2609.02892v12026
  50. SenseBERT: Driving Some Sense into BERT

    Yoav Levine, Barak Lenz, Or Dagan +6

    cs.CLcs.LGarXiv:1908.05646v22019
  51. Before the Script, Set the Stage: How Worldview Simulation Amplifies Psychologically Grounded Persuasion in Multi-Turn Jailbreaking

    Siyu Chen, Haoran Wang, Xiaojian Li +3

    cs.CLcs.AIarXiv:2609.02414v12026
  52. BharatGather: A Culturally-Informed Benchmark Dataset for Misinformation and Fake News Detection in Indian Public Events

    Parth Bramhecha, Smit Deshmukh, Sairaj Bodhale +2

    cs.CLcs.LGarXiv:2609.02895v12026
  53. Personalizing Dialogue Agents via Meta-Learning

    Zhaojiang Lin, Andrea Madotto, Chien-Sheng Wu +1

    cs.CLcs.AIarXiv:1905.10033v12019
  54. Text Segmentation as a Supervised Learning Task

    Omri Koshorek, Adir Cohen, Noam Mor +2

    cs.CLarXiv:1803.09337v12018
  55. SFAD: Speculative Factuality-Aware Decoding

    Guanqiao Chen, Di Wang, Lijie Hu

    cs.CLarXiv:2609.00796v22026
  56. SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

    Weihao Zeng, Yuzhen Huang, Qian Liu +4

    cs.LGcs.AIcs.CLarXiv:2503.18892v32025
  57. Scaling Near-Optimal SFT-RL Annotation Budget Allocation from Small to Large LLMs

    Jingtan Wang, Arun Verma, Xiaoqiang Lin +4

    cs.CLcs.AIcs.LGarXiv:2609.01573v12026
  58. Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

    Li Zhong, Zilong Wang, Jingbo Shang

    cs.SEcs.AIcs.CLarXiv:2402.16906v62024
  59. GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning

    Lakshya A Agrawal, Shangyin Tan, Dilara Soylu +14

    cs.CLcs.AIcs.LGarXiv:2507.19457v22025
  60. Sequence-to-Sequence Knowledge Graph Completion and Question Answering

    Apoorv Saxena, Adrian Kochsiek, Rainer Gemulla

    cs.CLcs.LGarXiv:2203.10321v12022