Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,641 to 8,700 of 11,247

  1. CodeT5+: Open Code Large Language Models for Code Understanding and Generation

    Yue Wang, Hung Le, Akhilesh Deepak Gotmare +3

    cs.CLcs.LGcs.PLarXiv:2305.07922v22023
  2. An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning

    Yun Luo, Zhen Yang, Fandong Meng +3

    cs.CLarXiv:2308.08747v52023
  3. Visualizing and Understanding Neural Models in NLP

    Jiwei Li, Xinlei Chen, Eduard Hovy +1

    cs.CLarXiv:1506.01066v22015
  4. Recursive Experiential-Working Memory Evolution for Long-Horizon Agent Harnesses

    Zhaochen Yu, Yingcheng Wu, Zhenfei Yin +5

    cs.AIcs.CLarXiv:2608.24876v12026
  5. Privacy Collapse: Benign Fine-Tuning Can Break Contextual Privacy in Language Models

    Anmol Goel, Cornelius Emde, Sangdoo Yun +2

    cs.CLarXiv:2601.15220v22026
  6. Memorization Dynamics in Knowledge Distillation for Language Models

    Jaydeep Borkar, Karan Chadha, Niloofar Mireshghallah +6

    cs.CLarXiv:2601.15394v22026
  7. All four leading LLMs talk more than they listen to personality-verified synthetic help-seekers

    Pablo A. Fonseca, Raquel Rodríguez-Carvajal, Rafael A. Calvo

    cs.HCcs.CLcs.CYarXiv:2608.22425v12026
  8. JudgeRLVR: Judge First, Generate Second for Efficient Reasoning

    Jiangshan Duo, Hanyu Li, Hailin Zhang +3

    cs.CLcs.AIcs.LGarXiv:2601.08468v12026
  9. Whitewashing Hate, Smearing Harmless Content: Annotator-Style Rebuttal Attacks on LLM-Based Moderation

    Junyu Lu, Kaiyuan Liu, Jingyi Kang +7

    cs.CLarXiv:2608.22230v12026
  10. LM-Nav: Robotic Navigation with Large Pre-Trained Models of Language, Vision, and Action

    Dhruv Shah, Blazej Osinski, Brian Ichter +1

    cs.ROcs.AIcs.CLarXiv:2207.04429v22022
  11. Meta$^n$: Recursive Self-Improvement through Emergent Depth

    Zae Myung Kim, Young-Jun Lee, Seungyeon Jwa +1

    cs.AIcs.CLeess.SYarXiv:2608.24735v12026
  12. Lost in the Prompt Order: Revealing the Limitations of Causal Attention in Language Models

    Hyunjong Ok, Jaeho Lee

    cs.CLcs.AIcs.LGarXiv:2601.14152v22026
  13. Training Reasoning Models on Saturated Problems via Failure-Prefix Conditioning

    Minwu Kim, Safal Shrestha, Anubhav Shrestha +1

    cs.LGcs.AIcs.CLarXiv:2601.20829v22026
  14. GDCNet: Generative Discrepancy Comparison Network for Multimodal Sarcasm Detection

    Shuguang Zhang, Junhong Lian, Guoxin Yu +2

    cs.CVcs.AIcs.CLarXiv:2601.20618v12026
  15. ECO: Quantized Training without Full-Precision Master Weights

    Mahdi Nikdan, Amir Zandieh, Dan Alistarh +1

    cs.CLcs.AIcs.LGarXiv:2601.22101v12026
  16. DIFFA-2: A Practical Diffusion Large Language Model for General Audio Understanding

    Jiaming Zhou, Xuxin Cheng, Shiwan Zhao +5

    cs.SDcs.CLarXiv:2601.23161v12026
  17. One Adapts to Any: Meta Reward Modeling for Personalized LLM Alignment

    Hongru Cai, Yongqi Li, Tiezheng Yu +4

    cs.CLcs.AIarXiv:2601.18731v22026
  18. BMAM: Brain-inspired Multi-Agent Memory Framework

    Yang Li, Jiaxiang Liu, Yusong Wang +2

    cs.CLarXiv:2601.20465v12026
  19. Grad-TTS: A Diffusion Probabilistic Model for Text-to-Speech

    Vadim Popov, Ivan Vovk, Vladimir Gogoryan +2

    cs.LGcs.CLstat.MLarXiv:2105.06337v22021
  20. Iteration Without Elaboration: A Simple ReAct Architecture Suffices for Text-to-SQL Generation

    Jian Lu, Haiwei Yu, Raymond M Xiong +2

    cs.CLarXiv:2608.22651v12026
  21. Why Attention Patterns Exist: A Unifying Temporal Perspective Analysis

    Qingyue Yang, Jie Wang, Xing Li +6

    cs.CLarXiv:2601.21709v12026
  22. RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment

    Hanze Dong, Wei Xiong, Deepanshu Goyal +7

    cs.LGcs.AIcs.CLarXiv:2304.06767v42023
  23. Automatic Prompt Optimization with "Gradient Descent" and Beam Search

    Reid Pryzant, Dan Iter, Jerry Li +3

    cs.CLcs.AIcs.LGarXiv:2305.03495v22023
  24. RocketQA: An Optimized Training Approach to Dense Passage Retrieval for Open-Domain Question Answering

    Yingqi Qu, Yuchen Ding, Jing Liu +6

    cs.CLcs.IRarXiv:2010.08191v22020
  25. Where Cognition Lives: Dissecting Emergent from Computed Function in a Minimal Complete Cognitive Architecture

    Francisco M. Arrabal-Campos, Francisco G. Montoya, Alfredo Alcayde +1

    cs.AIcs.CLcs.LGarXiv:2608.22347v12026
  26. Automatic detection of Gen-AI texts: A comparative framework of neural models

    Cristian Buttaro, Irene Amerini

    cs.CLarXiv:2603.18750v12026
  27. TyDi QA: A Benchmark for Information-Seeking Question Answering in Typologically Diverse Languages

    Jonathan H. Clark, Eunsol Choi, Michael Collins +4

    cs.CLcs.LGarXiv:2003.05002v12020
    Summaries:한국어
  28. Decomposed Prompting: A Modular Approach for Solving Complex Tasks

    Tushar Khot, Harsh Trivedi, Matthew Finlayson +4

    cs.CLarXiv:2210.02406v22022
    Summaries:한국어
  29. Scaling Small Agents Through Strategy Auctions

    Lisa Alazraki, William F. Shen, Yoram Bachrach +1

    cs.MAcs.AIcs.CLarXiv:2602.02751v32026
  30. Language Models are Super Mario: Absorbing Abilities from Homologous Models as a Free Lunch

    Le Yu, Bowen Yu, Haiyang Yu +2

    cs.CLcs.LGarXiv:2311.03099v32023
  31. SimpleGPT: Improving GPT via A Simple Normalization Strategy

    Marco Chen, Xianbiao Qi, Yelin He +2

    cs.LGcs.CLcs.CVarXiv:2602.01212v12026
  32. Dynamic Memory Networks for Visual and Textual Question Answering

    Caiming Xiong, Stephen Merity, Richard Socher

    cs.NEcs.CLcs.CVarXiv:1603.01417v12016
  33. WideSeek: Advancing Wide Research via Multi-Agent Scaling

    Ziyang Huang, Haolin Ren, Xiaowei Yuan +6

    cs.CLcs.AIcs.IRarXiv:2602.02636v12026
  34. What learning algorithm is in-context learning? Investigations with linear models

    Ekin Akyürek, Dale Schuurmans, Jacob Andreas +2

    cs.LGcs.CLarXiv:2211.15661v32022
  35. Adaptive Ability Decomposing for Unlocking Large Reasoning Model Effective Reinforcement Learning

    Zhipeng Chen, Xiaobo Qin, Wayne Xin Zhao +2

    cs.CLcs.AIarXiv:2602.00759v12026
  36. Rethinking Selective Knowledge Distillation

    Almog Tavor, Itay Ebenspanger, Neil Cnaan +1

    cs.CLarXiv:2602.01395v12026
  37. Making Avatars Interact: Towards Text-Driven Human-Object Interaction for Controllable Talking Avatars

    Youliang Zhang, Zhengguang Zhou, Zhentao Yu +11

    cs.CVcs.AIcs.CLarXiv:2602.01538v12026
  38. SEA-Guard: Culturally Grounded Multilingual Safeguard for Southeast Asia

    Panuthep Tasawong, Jian Gang Ngui, Alham Fikri Aji +2

    cs.CLarXiv:2602.01618v12026
  39. Hunt Instead of Wait: Evaluating Deep Data Research on Large Language Models

    Wei Liu, Peijie Yu, Michele Orini +2

    cs.AIcs.CLcs.DBarXiv:2602.02039v22026
  40. WildGraphBench: Benchmarking GraphRAG with Wild-Source Corpora

    Pengyu Wang, Benfeng Xu, Licheng Zhang +4

    cs.CLarXiv:2602.02053v22026
  41. LiT: Zero-Shot Transfer with Locked-image text Tuning

    Xiaohua Zhai, Xiao Wang, Basil Mustafa +4

    cs.CVcs.CLcs.LGarXiv:2111.07991v32021
  42. Echoes as Anchors: Probabilistic Costs and Attention Refocusing in LLM Reasoning

    Zhuoyuan Hao, Zhuo Li, Wu Li +3

    cs.CLarXiv:2602.06600v12026
  43. ReMiT: RL-Guided Mid-Training for Iterative LLM Evolution

    Junjie Huang, Jiarui Qin, Di Yin +4

    cs.CLarXiv:2602.03075v12026
  44. EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding

    Karttikeya Mangalam, Raiymbek Akshulakov, Jitendra Malik

    cs.CVcs.AIcs.CLarXiv:2308.09126v12023
  45. Fix the Structural Bottleneck: Context Compression via Explicit Information Transmission

    Jiangnan Ye, Hanqi Yan, Zhenyi Shen +3

    cs.CLarXiv:2602.03784v42026
  46. Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference

    Benjamin Warner, Antoine Chaffin, Benjamin Clavié +11

    cs.CLcs.AIarXiv:2412.13663v22024
  47. Language Chain in Alignment: Cross-Lingual Ranking Preference Optimization

    Seungyoon Lee, Minhyuk Kim, Jungseob Lee +1

    cs.CLcs.AIarXiv:2608.23149v12026
  48. FNet: Mixing Tokens with Fourier Transforms

    James Lee-Thorp, Joshua Ainslie, Ilya Eckstein +1

    cs.CLcs.LGarXiv:2105.03824v42021
  49. Fundamental Reasoning Paradigms Induce Out-of-Domain Generalization in Language Models

    Mingzi Cao, Xingwei Tan, Mahmud Elahi Akhter +4

    cs.CLarXiv:2602.08658v22026
  50. SocialVeil: Probing Social Intelligence of Language Agents under Communication Barriers

    Keyang Xuan, Pengda Wang, Chongrui Ye +3

    cs.AIcs.CLarXiv:2602.05115v12026
  51. Uncovering Cross-Objective Interference in Multi-Objective Alignment

    Yining Lu, Meng Jiang

    cs.CLcs.LGarXiv:2602.06869v22026
  52. Secure Code Generation via Online Reinforcement Learning with Vulnerability Reward Model

    Tianyi Wu, Mingzhe Du, Yue Liu +4

    cs.CRcs.AIcs.CLarXiv:2602.07422v12026
  53. A Survey on Automated Fact-Checking

    Zhijiang Guo, Michael Schlichtkrull, Andreas Vlachos

    cs.CLarXiv:2108.11896v32021
  54. compar:IA: The French Government's LLM arena to collect French-language human prompts and preference data

    Lucie Termignon, Simonas Zilinskas, Hadrien Pélissier +3

    cs.CLcs.AIarXiv:2602.06669v12026
  55. Improving Data and Reward Design for Scientific Reasoning in Large Language Models

    Zijie Chen, Zhenghao Lin, Xiao Liu +3

    cs.CLarXiv:2602.08321v22026
  56. ChatGPT: Jack of all trades, master of none

    Jan Kocoń, Igor Cichecki, Oliwier Kaszyca +17

    cs.CLcs.AIcs.CYarXiv:2302.10724v42023
  57. Mitigating Reasoning-Induced Misalignment via Safety-Direction Penalty

    Yipeng Zhao, Qishun Yang, Shenzhe Zhu +2

    cs.AIcs.CLarXiv:2608.23497v12026
  58. Multimodal Fact-Level Attribution for Verifiable Reasoning

    David Wan, Han Wang, Ziyang Wang +3

    cs.CLcs.AIcs.CVarXiv:2602.11509v22026
  59. Exploring Models and Data for Image Question Answering

    Mengye Ren, Ryan Kiros, Richard Zemel

    cs.LGcs.AIcs.CLarXiv:1505.02074v42015
  60. LLM-Planner: Few-Shot Grounded Planning for Embodied Agents with Large Language Models

    Chan Hee Song, Jiaman Wu, Clayton Washington +3

    cs.AIcs.CLcs.CVarXiv:2212.04088v32022