Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,641 to 8,700 of 11,247
CodeT5+: Open Code Large Language Models for Code Understanding and Generation
Yue Wang, Hung Le, Akhilesh Deepak Gotmare +3
cs.CLcs.LGcs.PLarXiv:2305.07922v22023An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning
Yun Luo, Zhen Yang, Fandong Meng +3
cs.CLarXiv:2308.08747v52023Visualizing and Understanding Neural Models in NLP
Jiwei Li, Xinlei Chen, Eduard Hovy +1
cs.CLarXiv:1506.01066v22015Recursive Experiential-Working Memory Evolution for Long-Horizon Agent Harnesses
Zhaochen Yu, Yingcheng Wu, Zhenfei Yin +5
cs.AIcs.CLarXiv:2608.24876v12026Privacy Collapse: Benign Fine-Tuning Can Break Contextual Privacy in Language Models
Anmol Goel, Cornelius Emde, Sangdoo Yun +2
cs.CLarXiv:2601.15220v22026Memorization Dynamics in Knowledge Distillation for Language Models
Jaydeep Borkar, Karan Chadha, Niloofar Mireshghallah +6
cs.CLarXiv:2601.15394v22026All four leading LLMs talk more than they listen to personality-verified synthetic help-seekers
Pablo A. Fonseca, Raquel Rodríguez-Carvajal, Rafael A. Calvo
cs.HCcs.CLcs.CYarXiv:2608.22425v12026JudgeRLVR: Judge First, Generate Second for Efficient Reasoning
Jiangshan Duo, Hanyu Li, Hailin Zhang +3
cs.CLcs.AIcs.LGarXiv:2601.08468v12026Whitewashing Hate, Smearing Harmless Content: Annotator-Style Rebuttal Attacks on LLM-Based Moderation
Junyu Lu, Kaiyuan Liu, Jingyi Kang +7
cs.CLarXiv:2608.22230v12026LM-Nav: Robotic Navigation with Large Pre-Trained Models of Language, Vision, and Action
Dhruv Shah, Blazej Osinski, Brian Ichter +1
cs.ROcs.AIcs.CLarXiv:2207.04429v22022Meta$^n$: Recursive Self-Improvement through Emergent Depth
Zae Myung Kim, Young-Jun Lee, Seungyeon Jwa +1
cs.AIcs.CLeess.SYarXiv:2608.24735v12026Lost in the Prompt Order: Revealing the Limitations of Causal Attention in Language Models
Hyunjong Ok, Jaeho Lee
cs.CLcs.AIcs.LGarXiv:2601.14152v22026Training Reasoning Models on Saturated Problems via Failure-Prefix Conditioning
Minwu Kim, Safal Shrestha, Anubhav Shrestha +1
cs.LGcs.AIcs.CLarXiv:2601.20829v22026GDCNet: Generative Discrepancy Comparison Network for Multimodal Sarcasm Detection
Shuguang Zhang, Junhong Lian, Guoxin Yu +2
cs.CVcs.AIcs.CLarXiv:2601.20618v12026ECO: Quantized Training without Full-Precision Master Weights
Mahdi Nikdan, Amir Zandieh, Dan Alistarh +1
cs.CLcs.AIcs.LGarXiv:2601.22101v12026DIFFA-2: A Practical Diffusion Large Language Model for General Audio Understanding
Jiaming Zhou, Xuxin Cheng, Shiwan Zhao +5
cs.SDcs.CLarXiv:2601.23161v12026One Adapts to Any: Meta Reward Modeling for Personalized LLM Alignment
Hongru Cai, Yongqi Li, Tiezheng Yu +4
cs.CLcs.AIarXiv:2601.18731v22026BMAM: Brain-inspired Multi-Agent Memory Framework
Yang Li, Jiaxiang Liu, Yusong Wang +2
cs.CLarXiv:2601.20465v12026Grad-TTS: A Diffusion Probabilistic Model for Text-to-Speech
Vadim Popov, Ivan Vovk, Vladimir Gogoryan +2
cs.LGcs.CLstat.MLarXiv:2105.06337v22021Iteration Without Elaboration: A Simple ReAct Architecture Suffices for Text-to-SQL Generation
Jian Lu, Haiwei Yu, Raymond M Xiong +2
cs.CLarXiv:2608.22651v12026Why Attention Patterns Exist: A Unifying Temporal Perspective Analysis
Qingyue Yang, Jie Wang, Xing Li +6
cs.CLarXiv:2601.21709v12026RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment
Hanze Dong, Wei Xiong, Deepanshu Goyal +7
cs.LGcs.AIcs.CLarXiv:2304.06767v42023Automatic Prompt Optimization with "Gradient Descent" and Beam Search
Reid Pryzant, Dan Iter, Jerry Li +3
cs.CLcs.AIcs.LGarXiv:2305.03495v22023RocketQA: An Optimized Training Approach to Dense Passage Retrieval for Open-Domain Question Answering
Yingqi Qu, Yuchen Ding, Jing Liu +6
cs.CLcs.IRarXiv:2010.08191v22020Where Cognition Lives: Dissecting Emergent from Computed Function in a Minimal Complete Cognitive Architecture
Francisco M. Arrabal-Campos, Francisco G. Montoya, Alfredo Alcayde +1
cs.AIcs.CLcs.LGarXiv:2608.22347v12026Automatic detection of Gen-AI texts: A comparative framework of neural models
Cristian Buttaro, Irene Amerini
cs.CLarXiv:2603.18750v12026TyDi QA: A Benchmark for Information-Seeking Question Answering in Typologically Diverse Languages
Jonathan H. Clark, Eunsol Choi, Michael Collins +4
cs.CLcs.LGarXiv:2003.05002v12020Summaries:한국어Decomposed Prompting: A Modular Approach for Solving Complex Tasks
Tushar Khot, Harsh Trivedi, Matthew Finlayson +4
cs.CLarXiv:2210.02406v22022Summaries:한국어Scaling Small Agents Through Strategy Auctions
Lisa Alazraki, William F. Shen, Yoram Bachrach +1
cs.MAcs.AIcs.CLarXiv:2602.02751v32026Language Models are Super Mario: Absorbing Abilities from Homologous Models as a Free Lunch
Le Yu, Bowen Yu, Haiyang Yu +2
cs.CLcs.LGarXiv:2311.03099v32023SimpleGPT: Improving GPT via A Simple Normalization Strategy
Marco Chen, Xianbiao Qi, Yelin He +2
cs.LGcs.CLcs.CVarXiv:2602.01212v12026Dynamic Memory Networks for Visual and Textual Question Answering
Caiming Xiong, Stephen Merity, Richard Socher
cs.NEcs.CLcs.CVarXiv:1603.01417v12016WideSeek: Advancing Wide Research via Multi-Agent Scaling
Ziyang Huang, Haolin Ren, Xiaowei Yuan +6
cs.CLcs.AIcs.IRarXiv:2602.02636v12026What learning algorithm is in-context learning? Investigations with linear models
Ekin Akyürek, Dale Schuurmans, Jacob Andreas +2
cs.LGcs.CLarXiv:2211.15661v32022Adaptive Ability Decomposing for Unlocking Large Reasoning Model Effective Reinforcement Learning
Zhipeng Chen, Xiaobo Qin, Wayne Xin Zhao +2
cs.CLcs.AIarXiv:2602.00759v12026Rethinking Selective Knowledge Distillation
Almog Tavor, Itay Ebenspanger, Neil Cnaan +1
cs.CLarXiv:2602.01395v12026Making Avatars Interact: Towards Text-Driven Human-Object Interaction for Controllable Talking Avatars
Youliang Zhang, Zhengguang Zhou, Zhentao Yu +11
cs.CVcs.AIcs.CLarXiv:2602.01538v12026SEA-Guard: Culturally Grounded Multilingual Safeguard for Southeast Asia
Panuthep Tasawong, Jian Gang Ngui, Alham Fikri Aji +2
cs.CLarXiv:2602.01618v12026Hunt Instead of Wait: Evaluating Deep Data Research on Large Language Models
Wei Liu, Peijie Yu, Michele Orini +2
cs.AIcs.CLcs.DBarXiv:2602.02039v22026WildGraphBench: Benchmarking GraphRAG with Wild-Source Corpora
Pengyu Wang, Benfeng Xu, Licheng Zhang +4
cs.CLarXiv:2602.02053v22026LiT: Zero-Shot Transfer with Locked-image text Tuning
Xiaohua Zhai, Xiao Wang, Basil Mustafa +4
cs.CVcs.CLcs.LGarXiv:2111.07991v32021Echoes as Anchors: Probabilistic Costs and Attention Refocusing in LLM Reasoning
Zhuoyuan Hao, Zhuo Li, Wu Li +3
cs.CLarXiv:2602.06600v12026ReMiT: RL-Guided Mid-Training for Iterative LLM Evolution
Junjie Huang, Jiarui Qin, Di Yin +4
cs.CLarXiv:2602.03075v12026EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Karttikeya Mangalam, Raiymbek Akshulakov, Jitendra Malik
cs.CVcs.AIcs.CLarXiv:2308.09126v12023Fix the Structural Bottleneck: Context Compression via Explicit Information Transmission
Jiangnan Ye, Hanqi Yan, Zhenyi Shen +3
cs.CLarXiv:2602.03784v42026Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference
Benjamin Warner, Antoine Chaffin, Benjamin Clavié +11
cs.CLcs.AIarXiv:2412.13663v22024Language Chain in Alignment: Cross-Lingual Ranking Preference Optimization
Seungyoon Lee, Minhyuk Kim, Jungseob Lee +1
cs.CLcs.AIarXiv:2608.23149v12026FNet: Mixing Tokens with Fourier Transforms
James Lee-Thorp, Joshua Ainslie, Ilya Eckstein +1
cs.CLcs.LGarXiv:2105.03824v42021Fundamental Reasoning Paradigms Induce Out-of-Domain Generalization in Language Models
Mingzi Cao, Xingwei Tan, Mahmud Elahi Akhter +4
cs.CLarXiv:2602.08658v22026SocialVeil: Probing Social Intelligence of Language Agents under Communication Barriers
Keyang Xuan, Pengda Wang, Chongrui Ye +3
cs.AIcs.CLarXiv:2602.05115v12026Uncovering Cross-Objective Interference in Multi-Objective Alignment
Yining Lu, Meng Jiang
cs.CLcs.LGarXiv:2602.06869v22026Secure Code Generation via Online Reinforcement Learning with Vulnerability Reward Model
Tianyi Wu, Mingzhe Du, Yue Liu +4
cs.CRcs.AIcs.CLarXiv:2602.07422v12026A Survey on Automated Fact-Checking
Zhijiang Guo, Michael Schlichtkrull, Andreas Vlachos
cs.CLarXiv:2108.11896v32021compar:IA: The French Government's LLM arena to collect French-language human prompts and preference data
Lucie Termignon, Simonas Zilinskas, Hadrien Pélissier +3
cs.CLcs.AIarXiv:2602.06669v12026Improving Data and Reward Design for Scientific Reasoning in Large Language Models
Zijie Chen, Zhenghao Lin, Xiao Liu +3
cs.CLarXiv:2602.08321v22026ChatGPT: Jack of all trades, master of none
Jan Kocoń, Igor Cichecki, Oliwier Kaszyca +17
cs.CLcs.AIcs.CYarXiv:2302.10724v42023Mitigating Reasoning-Induced Misalignment via Safety-Direction Penalty
Yipeng Zhao, Qishun Yang, Shenzhe Zhu +2
cs.AIcs.CLarXiv:2608.23497v12026Multimodal Fact-Level Attribution for Verifiable Reasoning
David Wan, Han Wang, Ziyang Wang +3
cs.CLcs.AIcs.CVarXiv:2602.11509v22026Exploring Models and Data for Image Question Answering
Mengye Ren, Ryan Kiros, Richard Zemel
cs.LGcs.AIcs.CLarXiv:1505.02074v42015LLM-Planner: Few-Shot Grounded Planning for Embodied Agents with Large Language Models
Chan Hee Song, Jiaman Wu, Clayton Washington +3
cs.AIcs.CLcs.CVarXiv:2212.04088v32022