Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,381 to 4,440 of 11,233

  1. Zep: A Temporal Knowledge Graph Architecture for Agent Memory

    Preston Rasmussen, Pavlo Paliychuk, Travis Beauvais +2

    cs.CLcs.AIcs.IRarXiv:2501.13956v12025
  2. SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model

    Loubna Ben Allal, Anton Lozhkov, Elie Bakouch +19

    cs.CLarXiv:2502.02737v12025
  3. O1-Pruner: Length-Harmonizing Fine-Tuning for O1-Like Reasoning Pruning

    Haotian Luo, Li Shen, Haiying He +6

    cs.CLarXiv:2501.12570v22025
  4. DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents

    Mingxuan Du, Benfeng Xu, Chiwei Zhu +2

    cs.CLcs.IRarXiv:2506.11763v12025
  5. Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning

    Haozhe Wang, Alex Su, Weiming Ren +2

    cs.CVcs.AIcs.CLarXiv:2505.15966v32025
  6. Learning a Recurrent Visual Representation for Image Caption Generation

    Xinlei Chen, C. Lawrence Zitnick

    cs.CVcs.AIcs.CLarXiv:1411.5654v12014
  7. WebThinker: Empowering Large Reasoning Models with Deep Research Capability

    Xiaoxi Li, Jiajie Jin, Guanting Dong +5

    cs.CLcs.AIcs.IRarXiv:2504.21776v22025
  8. FinBen: A Holistic Financial Benchmark for Large Language Models

    Qianqian Xie, Weiguang Han, Zhengyu Chen +31

    cs.CLcs.AIcs.CEarXiv:2402.12659v22024
  9. Learning to Reason under Off-Policy Guidance

    Jianhao Yan, Yafu Li, Zican Hu +5

    cs.LGcs.AIcs.CLarXiv:2504.14945v52025
  10. Beyond Magnitude: Contrastive Routing for Modular Mixture-of-Experts

    Nikolaos Xiros, Dimitrios Damianos, Maria-Eleni Zoumpoulidi +3

    cs.CLarXiv:2609.01100v12026
  11. Learning Mixtures of Submodular Shells with Application to Document Summarization

    Hui Lin, Jeff A. Bilmes

    cs.LGcs.CLcs.IRarXiv:1210.4871v12012
  12. Towards AI-Assisted Clinical Trial Matching: Practical Considerations, Multicenter Evaluation, and Real-World Deployment

    Yin Fang, Qiao Jin, Shubo Tian +24

    cs.CLcs.AIcs.CYarXiv:2609.01202v12026
  13. Accelerating scientific discovery with Co-Scientist

    Juraj Gottweis, Wei-Hung Weng, Alexander Daryin +48

    cs.AIcs.CLcs.HCarXiv:2502.18864v22025
  14. ToolRL: Reward is All Tool Learning Needs

    Cheng Qian, Emre Can Acikgoz, Qi He +5

    cs.LGcs.AIcs.CLarXiv:2504.13958v12025
  15. Barack's Wife Hillary: Using Knowledge-Graphs for Fact-Aware Language Modeling

    Robert L. Logan, Nelson F. Liu, Matthew E. Peters +2

    cs.CLarXiv:1906.07241v22019
  16. Humanity's Last Exam

    Long Phan, Alice Gatti, Ziwen Han +1155

    cs.LGcs.AIcs.CLarXiv:2501.14249v112025
  17. Designing Proactive Thought Partners for Writing

    Chao Zhang, Abe Davis, Chih-Wei Chen +1

    cs.HCcs.AIcs.CLarXiv:2609.01588v12026
  18. The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity

    Parshin Shojaee, Iman Mirzadeh, Keivan Alizadeh +3

    cs.AIcs.CLcs.LGarXiv:2506.06941v32025
  19. MathArena: Evaluating LLMs on Uncontaminated Math Competitions

    Mislav Balunović, Jasper Dekoninck, Ivo Petrov +2

    cs.AIcs.CLarXiv:2505.23281v32025
  20. On the Effect of Dropping Layers of Pre-trained Transformer Models

    Hassan Sajjad, Fahim Dalvi, Nadir Durrani +1

    cs.CLcs.LGarXiv:2004.03844v32020
  21. Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models

    Yang Sui, Yu-Neng Chuang, Guanchu Wang +9

    cs.CLarXiv:2503.16419v42025
  22. L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning

    Pranjal Aggarwal, Sean Welleck

    cs.CLcs.AIcs.LGarXiv:2503.04697v22025
  23. Memory in the Age of AI Agents

    Yuyang Hu, Shichun Liu, Yanwei Yue +44

    cs.CLcs.AIarXiv:2512.13564v22025
  24. EAGLE-3: Scaling up Inference Acceleration of Large Language Models via Training-Time Test

    Yuhui Li, Fangyun Wei, Chao Zhang +1

    cs.CLarXiv:2503.01840v32025
  25. DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition

    Z. Z. Ren, Zhihong Shao, Junxiao Song +15

    cs.CLcs.AIarXiv:2504.21801v22025
  26. Faithful Logical Reasoning via Symbolic Chain-of-Thought

    Jundong Xu, Hao Fei, Liangming Pan +3

    cs.CLarXiv:2405.18357v22024
  27. Agent Laboratory: Using LLM Agents as Research Assistants

    Samuel Schmidgall, Yusheng Su, Ze Wang +7

    cs.HCcs.AIcs.CLarXiv:2501.04227v22025
  28. Last Translation Benchmark

    Vilém Zouhar, Niyati Bafna, Mukund Choudhary +241

    cs.CLarXiv:2609.04173v12026
  29. Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach

    Jonas Geiping, Sean McLeish, Neel Jain +6

    cs.LGcs.CLarXiv:2502.05171v22025
  30. Persona Vectors: Monitoring and Controlling Character Traits in Language Models

    Runjin Chen, Andy Arditi, Henry Sleight +2

    cs.CLcs.LGarXiv:2507.21509v32025
  31. LIBERO-Plus: In-depth Robustness Analysis of Vision-Language-Action Models

    Senyu Fei, Siyin Wang, Junhao Shi +10

    cs.ROcs.CLcs.CVarXiv:2510.13626v32025
  32. DepecheMood: a Lexicon for Emotion Analysis from Crowd-Annotated News

    Jacopo Staiano, Marco Guerini

    cs.CLcs.CYarXiv:1405.1605v12014
  33. Diverse Demonstrations Improve In-context Compositional Generalization

    Itay Levy, Ben Bogin, Jonathan Berant

    cs.CLarXiv:2212.06800v32022
  34. Muon is Scalable for LLM Training

    Jingyuan Liu, Jianlin Su, Xingcheng Yao +25

    cs.LGcs.AIcs.CLarXiv:2502.16982v12025
  35. The Lessons of Developing Process Reward Models in Mathematical Reasoning

    Zhenru Zhang, Chujie Zheng, Yangzhen Wu +6

    cs.CLcs.AIcs.LGarXiv:2501.07301v22025
  36. The Right Tool for the Job: Matching Model and Instance Complexities

    Roy Schwartz, Gabriel Stanovsky, Swabha Swayamdipta +2

    cs.CLcs.LGarXiv:2004.07453v22020
  37. Search-o1: Agentic Search-Enhanced Large Reasoning Models

    Xiaoxi Li, Guanting Dong, Jiajie Jin +5

    cs.AIcs.CLcs.IRarXiv:2501.05366v12025
  38. Reasoning with Exploration: An Entropy Perspective

    Daixuan Cheng, Shaohan Huang, Xuekai Zhu +4

    cs.CLarXiv:2506.14758v42025
  39. Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs

    Kanishk Gandhi, Ayush Chakravarthy, Anikait Singh +2

    cs.CLcs.LGarXiv:2503.01307v22025
  40. Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models

    Qizheng Zhang, Changran Hu, Shubhangi Upasani +10

    cs.LGcs.AIcs.CLarXiv:2510.04618v32025
  41. Process Reinforcement through Implicit Rewards

    Ganqu Cui, Lifan Yuan, Zefan Wang +22

    cs.LGcs.AIcs.CLarXiv:2502.01456v22025
  42. From Reweighting to Rewriting: Unlocking the Intervention Effects of Influential Samples in Training Data Attribution

    Yuzhang Luo, Chenpeng Wang, Jianhui Chen +1

    cs.CLcs.AIcs.LGarXiv:2609.02771v12026
  43. Political Compass or Spinning Arrow? Towards More Meaningful Evaluations for Values and Opinions in Large Language Models

    Paul Röttger, Valentin Hofmann, Valentina Pyatkin +4

    cs.CLcs.AIarXiv:2402.16786v22024
  44. DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

    Huajian Xin, Z. Z. Ren, Junxiao Song +14

    cs.CLcs.AIcs.LGarXiv:2408.08152v12024
  45. ELEVATER: A Benchmark and Toolkit for Evaluating Language-Augmented Visual Models

    Chunyuan Li, Haotian Liu, Liunian Harold Li +8

    cs.CVcs.CLcs.LGarXiv:2204.08790v62022
  46. SMHD: A Large-Scale Resource for Exploring Online Language Usage for Multiple Mental Health Conditions

    Arman Cohan, Bart Desmet, Andrew Yates +3

    cs.CLarXiv:1806.05258v22018
  47. Multi-Source Domain Adaptation with Mixture of Experts

    Jiang Guo, Darsh J Shah, Regina Barzilay

    cs.CLarXiv:1809.02256v22018
  48. C3oT: Generating Shorter Chain-of-Thought without Compromising Effectiveness

    Yu Kang, Xianghui Sun, Liangyu Chen +1

    cs.CLcs.LGarXiv:2412.11664v12024
  49. VakyArth: Evaluating Pragmatic Competence in LLMs across Indic Languages

    Usneek Singh, Poorvaja Veera Balaji Kumar, Parth Nanda +4

    cs.CLcs.AIarXiv:2609.01788v12026
  50. Jointly embedding the local and global relations of heterogeneous graph for rumor detection

    Chunyuan Yuan, Qianwen Ma, Wei Zhou +2

    cs.CLcs.IRcs.SIarXiv:1909.04465v22019
  51. Stochastic Multiple Choice Learning for Training Diverse Deep Ensembles

    Stefan Lee, Senthil Purushwalkam, Michael Cogswell +3

    cs.CVcs.CLarXiv:1606.07839v32016
  52. FlowSeq: Non-Autoregressive Conditional Sequence Generation with Generative Flow

    Xuezhe Ma, Chunting Zhou, Xian Li +2

    cs.CLcs.LGarXiv:1909.02480v32019
  53. NE-R1: Enhancing Named Entity Recognition Model via Reinforcement Learning

    Meixuan Chen, Hehan Li, Ruizhi Zhao +8

    cs.CLcs.AIarXiv:2609.02366v12026
  54. Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention

    Jingyang Yuan, Huazuo Gao, Damai Dai +12

    cs.CLcs.AIcs.LGarXiv:2502.11089v22025
  55. ColPali: Efficient Document Retrieval with Vision Language Models

    Manuel Faysse, Hugues Sibille, Tony Wu +4

    cs.IRcs.CLcs.CVarXiv:2407.01449v62024
  56. Automatic Prompt Augmentation and Selection with Chain-of-Thought from Labeled Data

    KaShun Shum, Shizhe Diao, Tong Zhang

    cs.CLarXiv:2302.12822v32023
  57. VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model

    Haozhan Shen, Peng Liu, Jingcheng Li +9

    cs.CVcs.CLarXiv:2504.07615v22025
  58. Interactive and Visual Prompt Engineering for Ad-hoc Task Adaptation with Large Language Models

    Hendrik Strobelt, Albert Webson, Victor Sanh +4

    cs.CLcs.HCcs.LGarXiv:2208.07852v12022
  59. Automated Concatenation of Embeddings for Structured Prediction

    Xinyu Wang, Yong Jiang, Nguyen Bach +4

    cs.CLcs.AIcs.LGarXiv:2010.05006v42020
  60. Learning to Understand Phrases by Embedding the Dictionary

    Felix Hill, Kyunghyun Cho, Anna Korhonen +1

    cs.CLarXiv:1504.00548v42015