Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,381 to 4,440 of 11,233
Zep: A Temporal Knowledge Graph Architecture for Agent Memory
Preston Rasmussen, Pavlo Paliychuk, Travis Beauvais +2
cs.CLcs.AIcs.IRarXiv:2501.13956v12025SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model
Loubna Ben Allal, Anton Lozhkov, Elie Bakouch +19
cs.CLarXiv:2502.02737v12025O1-Pruner: Length-Harmonizing Fine-Tuning for O1-Like Reasoning Pruning
Haotian Luo, Li Shen, Haiying He +6
cs.CLarXiv:2501.12570v22025DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents
Mingxuan Du, Benfeng Xu, Chiwei Zhu +2
cs.CLcs.IRarXiv:2506.11763v12025Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning
Haozhe Wang, Alex Su, Weiming Ren +2
cs.CVcs.AIcs.CLarXiv:2505.15966v32025Learning a Recurrent Visual Representation for Image Caption Generation
Xinlei Chen, C. Lawrence Zitnick
cs.CVcs.AIcs.CLarXiv:1411.5654v12014WebThinker: Empowering Large Reasoning Models with Deep Research Capability
Xiaoxi Li, Jiajie Jin, Guanting Dong +5
cs.CLcs.AIcs.IRarXiv:2504.21776v22025FinBen: A Holistic Financial Benchmark for Large Language Models
Qianqian Xie, Weiguang Han, Zhengyu Chen +31
cs.CLcs.AIcs.CEarXiv:2402.12659v22024Learning to Reason under Off-Policy Guidance
Jianhao Yan, Yafu Li, Zican Hu +5
cs.LGcs.AIcs.CLarXiv:2504.14945v52025Beyond Magnitude: Contrastive Routing for Modular Mixture-of-Experts
Nikolaos Xiros, Dimitrios Damianos, Maria-Eleni Zoumpoulidi +3
cs.CLarXiv:2609.01100v12026Learning Mixtures of Submodular Shells with Application to Document Summarization
Hui Lin, Jeff A. Bilmes
cs.LGcs.CLcs.IRarXiv:1210.4871v12012Towards AI-Assisted Clinical Trial Matching: Practical Considerations, Multicenter Evaluation, and Real-World Deployment
Yin Fang, Qiao Jin, Shubo Tian +24
cs.CLcs.AIcs.CYarXiv:2609.01202v12026Accelerating scientific discovery with Co-Scientist
Juraj Gottweis, Wei-Hung Weng, Alexander Daryin +48
cs.AIcs.CLcs.HCarXiv:2502.18864v22025ToolRL: Reward is All Tool Learning Needs
Cheng Qian, Emre Can Acikgoz, Qi He +5
cs.LGcs.AIcs.CLarXiv:2504.13958v12025Barack's Wife Hillary: Using Knowledge-Graphs for Fact-Aware Language Modeling
Robert L. Logan, Nelson F. Liu, Matthew E. Peters +2
cs.CLarXiv:1906.07241v22019Humanity's Last Exam
Long Phan, Alice Gatti, Ziwen Han +1155
cs.LGcs.AIcs.CLarXiv:2501.14249v112025Designing Proactive Thought Partners for Writing
Chao Zhang, Abe Davis, Chih-Wei Chen +1
cs.HCcs.AIcs.CLarXiv:2609.01588v12026The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity
Parshin Shojaee, Iman Mirzadeh, Keivan Alizadeh +3
cs.AIcs.CLcs.LGarXiv:2506.06941v32025MathArena: Evaluating LLMs on Uncontaminated Math Competitions
Mislav Balunović, Jasper Dekoninck, Ivo Petrov +2
cs.AIcs.CLarXiv:2505.23281v32025On the Effect of Dropping Layers of Pre-trained Transformer Models
Hassan Sajjad, Fahim Dalvi, Nadir Durrani +1
cs.CLcs.LGarXiv:2004.03844v32020Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models
Yang Sui, Yu-Neng Chuang, Guanchu Wang +9
cs.CLarXiv:2503.16419v42025L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
Pranjal Aggarwal, Sean Welleck
cs.CLcs.AIcs.LGarXiv:2503.04697v22025Memory in the Age of AI Agents
Yuyang Hu, Shichun Liu, Yanwei Yue +44
cs.CLcs.AIarXiv:2512.13564v22025EAGLE-3: Scaling up Inference Acceleration of Large Language Models via Training-Time Test
Yuhui Li, Fangyun Wei, Chao Zhang +1
cs.CLarXiv:2503.01840v32025DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition
Z. Z. Ren, Zhihong Shao, Junxiao Song +15
cs.CLcs.AIarXiv:2504.21801v22025Faithful Logical Reasoning via Symbolic Chain-of-Thought
Jundong Xu, Hao Fei, Liangming Pan +3
cs.CLarXiv:2405.18357v22024Agent Laboratory: Using LLM Agents as Research Assistants
Samuel Schmidgall, Yusheng Su, Ze Wang +7
cs.HCcs.AIcs.CLarXiv:2501.04227v22025Last Translation Benchmark
Vilém Zouhar, Niyati Bafna, Mukund Choudhary +241
cs.CLarXiv:2609.04173v12026Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
Jonas Geiping, Sean McLeish, Neel Jain +6
cs.LGcs.CLarXiv:2502.05171v22025Persona Vectors: Monitoring and Controlling Character Traits in Language Models
Runjin Chen, Andy Arditi, Henry Sleight +2
cs.CLcs.LGarXiv:2507.21509v32025LIBERO-Plus: In-depth Robustness Analysis of Vision-Language-Action Models
Senyu Fei, Siyin Wang, Junhao Shi +10
cs.ROcs.CLcs.CVarXiv:2510.13626v32025DepecheMood: a Lexicon for Emotion Analysis from Crowd-Annotated News
Jacopo Staiano, Marco Guerini
cs.CLcs.CYarXiv:1405.1605v12014Diverse Demonstrations Improve In-context Compositional Generalization
Itay Levy, Ben Bogin, Jonathan Berant
cs.CLarXiv:2212.06800v32022Muon is Scalable for LLM Training
Jingyuan Liu, Jianlin Su, Xingcheng Yao +25
cs.LGcs.AIcs.CLarXiv:2502.16982v12025The Lessons of Developing Process Reward Models in Mathematical Reasoning
Zhenru Zhang, Chujie Zheng, Yangzhen Wu +6
cs.CLcs.AIcs.LGarXiv:2501.07301v22025The Right Tool for the Job: Matching Model and Instance Complexities
Roy Schwartz, Gabriel Stanovsky, Swabha Swayamdipta +2
cs.CLcs.LGarXiv:2004.07453v22020Search-o1: Agentic Search-Enhanced Large Reasoning Models
Xiaoxi Li, Guanting Dong, Jiajie Jin +5
cs.AIcs.CLcs.IRarXiv:2501.05366v12025Reasoning with Exploration: An Entropy Perspective
Daixuan Cheng, Shaohan Huang, Xuekai Zhu +4
cs.CLarXiv:2506.14758v42025Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs
Kanishk Gandhi, Ayush Chakravarthy, Anikait Singh +2
cs.CLcs.LGarXiv:2503.01307v22025Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models
Qizheng Zhang, Changran Hu, Shubhangi Upasani +10
cs.LGcs.AIcs.CLarXiv:2510.04618v32025Process Reinforcement through Implicit Rewards
Ganqu Cui, Lifan Yuan, Zefan Wang +22
cs.LGcs.AIcs.CLarXiv:2502.01456v22025From Reweighting to Rewriting: Unlocking the Intervention Effects of Influential Samples in Training Data Attribution
Yuzhang Luo, Chenpeng Wang, Jianhui Chen +1
cs.CLcs.AIcs.LGarXiv:2609.02771v12026Political Compass or Spinning Arrow? Towards More Meaningful Evaluations for Values and Opinions in Large Language Models
Paul Röttger, Valentin Hofmann, Valentina Pyatkin +4
cs.CLcs.AIarXiv:2402.16786v22024DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
Huajian Xin, Z. Z. Ren, Junxiao Song +14
cs.CLcs.AIcs.LGarXiv:2408.08152v12024ELEVATER: A Benchmark and Toolkit for Evaluating Language-Augmented Visual Models
Chunyuan Li, Haotian Liu, Liunian Harold Li +8
cs.CVcs.CLcs.LGarXiv:2204.08790v62022SMHD: A Large-Scale Resource for Exploring Online Language Usage for Multiple Mental Health Conditions
Arman Cohan, Bart Desmet, Andrew Yates +3
cs.CLarXiv:1806.05258v22018Multi-Source Domain Adaptation with Mixture of Experts
Jiang Guo, Darsh J Shah, Regina Barzilay
cs.CLarXiv:1809.02256v22018C3oT: Generating Shorter Chain-of-Thought without Compromising Effectiveness
Yu Kang, Xianghui Sun, Liangyu Chen +1
cs.CLcs.LGarXiv:2412.11664v12024VakyArth: Evaluating Pragmatic Competence in LLMs across Indic Languages
Usneek Singh, Poorvaja Veera Balaji Kumar, Parth Nanda +4
cs.CLcs.AIarXiv:2609.01788v12026Jointly embedding the local and global relations of heterogeneous graph for rumor detection
Chunyuan Yuan, Qianwen Ma, Wei Zhou +2
cs.CLcs.IRcs.SIarXiv:1909.04465v22019Stochastic Multiple Choice Learning for Training Diverse Deep Ensembles
Stefan Lee, Senthil Purushwalkam, Michael Cogswell +3
cs.CVcs.CLarXiv:1606.07839v32016FlowSeq: Non-Autoregressive Conditional Sequence Generation with Generative Flow
Xuezhe Ma, Chunting Zhou, Xian Li +2
cs.CLcs.LGarXiv:1909.02480v32019NE-R1: Enhancing Named Entity Recognition Model via Reinforcement Learning
Meixuan Chen, Hehan Li, Ruizhi Zhao +8
cs.CLcs.AIarXiv:2609.02366v12026Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention
Jingyang Yuan, Huazuo Gao, Damai Dai +12
cs.CLcs.AIcs.LGarXiv:2502.11089v22025ColPali: Efficient Document Retrieval with Vision Language Models
Manuel Faysse, Hugues Sibille, Tony Wu +4
cs.IRcs.CLcs.CVarXiv:2407.01449v62024Automatic Prompt Augmentation and Selection with Chain-of-Thought from Labeled Data
KaShun Shum, Shizhe Diao, Tong Zhang
cs.CLarXiv:2302.12822v32023VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model
Haozhan Shen, Peng Liu, Jingcheng Li +9
cs.CVcs.CLarXiv:2504.07615v22025Interactive and Visual Prompt Engineering for Ad-hoc Task Adaptation with Large Language Models
Hendrik Strobelt, Albert Webson, Victor Sanh +4
cs.CLcs.HCcs.LGarXiv:2208.07852v12022Automated Concatenation of Embeddings for Structured Prediction
Xinyu Wang, Yong Jiang, Nguyen Bach +4
cs.CLcs.AIcs.LGarXiv:2010.05006v42020Learning to Understand Phrases by Embedding the Dictionary
Felix Hill, Kyunghyun Cho, Anna Korhonen +1
cs.CLarXiv:1504.00548v42015