Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1 to 60 of 11,225

  1. Benign Samples Matter! Fine-tuning On Outlier Benign Samples Severely Breaks Safety

    Zihan Guan, Mengxuan Hu, Ronghang Zhu +2

    cs.LGcs.CLarXiv:2505.06843v22025
  2. SpecExec: Massively Parallel Speculative Decoding for Interactive LLM Inference on Consumer Devices

    Ruslan Svirschevski, Avner May, Zhuoming Chen +3

    cs.CLarXiv:2406.02532v32024
  3. TriForce: Lossless Acceleration of Long Sequence Generation with Hierarchical Speculative Decoding

    Hanshi Sun, Zhuoming Chen, Xinyu Yang +2

    cs.CLcs.LGarXiv:2404.11912v32024
  4. OPT-Tree: Speculative Decoding with Adaptive Draft Tree Structure

    Jikai Wang, Yi Su, Juntao Li +5

    cs.CLarXiv:2406.17276v42024
  5. Cascade Speculative Drafting for Even Faster LLM Inference

    Ziyi Chen, Xiaocong Yang, Jiacheng Lin +3

    cs.LGcs.CLarXiv:2312.11462v52023
  6. Fine-tune Bert for DocRED with Two-step Process

    Hong Wang, Christfried Focke, Rob Sylvester +2

    cs.CLarXiv:1909.11898v12019
  7. HABERTOR: An Efficient and Effective Deep Hatespeech Detector

    Thanh Tran, Yifan Hu, Changwei Hu +4

    cs.CLcs.AIcs.IRarXiv:2010.08865v12020
  8. Detecting Data Contamination from Reinforcement Learning Post-training for Large Language Models

    Yongding Tao, Tian Wang, Yihong Dong +4

    cs.CLcs.AIcs.LGarXiv:2510.09259v22025
  9. LLMs as Scalable, General-Purpose Simulators For Evolving Digital Agent Training

    Yiming Wang, Da Yin, Yuedong Cui +8

    cs.CLcs.AIcs.LGarXiv:2510.14969v12025
  10. Counterfactual reasoning: Testing language models' understanding of hypothetical scenarios

    Jiaxuan Li, Lang Yu, Allyson Ettinger

    cs.CLarXiv:2305.16572v12023
  11. Agentic Knowledgeable Self-awareness

    Shuofei Qiao, Zhisong Qiu, Baochang Ren +8

    cs.CLcs.AIcs.CVarXiv:2504.03553v22025
  12. Training Neural Speech Recognition Systems with Synthetic Speech Augmentation

    Jason Li, Ravi Gadde, Boris Ginsburg +1

    cs.CLcs.LGcs.SDarXiv:1811.00707v12018
  13. End-To-End Memory Networks

    Sainbayar Sukhbaatar, Arthur Szlam, Jason Weston +1

    cs.NEcs.CLarXiv:1503.08895v52015
  14. Finding Generalizable Evidence by Learning to Convince Q&A Models

    Ethan Perez, Siddharth Karamcheti, Rob Fergus +3

    cs.CLcs.AIcs.IRarXiv:1909.05863v12019
  15. Exploring Content Selection in Summarization of Novel Chapters

    Faisal Ladhak, Bryan Li, Yaser Al-Onaizan +1

    cs.CLarXiv:2005.01840v32020
  16. The Evolution of Thought: Tracking LLM Overthinking via Reasoning Dynamics Analysis

    Zihao Wei, Liang Pang, Jiahao Liu +7

    cs.CLcs.AIarXiv:2508.17627v22025
  17. Diving Deep into Modes of Fact Hallucinations in Dialogue Systems

    Souvik Das, Sougata Saha, Rohini K. Srihari

    cs.CLcs.AIarXiv:2301.04449v12023
  18. Intern-S1: A Scientific Multimodal Foundation Model

    Lei Bai, Zhongrui Cai, Yuhang Cao +174

    cs.LGcs.CLcs.CVarXiv:2508.15763v22025
  19. Masked Language Modeling for Proteins via Linearly Scalable Long-Context Transformers

    Krzysztof Choromanski, Valerii Likhosherstov, David Dohan +8

    cs.LGcs.CLstat.MLarXiv:2006.03555v32020
  20. Seeing Beyond Words: Self-Supervised Visual Learning for Multimodal Large Language Models

    Davide Caffagni, Sara Sarto, Marcella Cornia +5

    cs.CVcs.AIcs.CLarXiv:2512.15885v12025
  21. Less is More: Mitigating Multimodal Hallucination from an EOS Decision Perspective

    Zihao Yue, Liang Zhang, Qin Jin

    cs.CLcs.CVarXiv:2402.14545v22024
  22. Predictive Concept Decoders: Training Scalable End-to-End Interpretability Assistants

    Vincent Huang, Dami Choi, Daniel D. Johnson +2

    cs.AIcs.CLcs.LGarXiv:2512.15712v12025
  23. Activation Oracles: Training and Evaluating LLMs as General-Purpose Activation Explainers

    Adam Karvonen, James Chua, Clément Dumas +8

    cs.CLcs.AIcs.LGarXiv:2512.15674v22025
  24. Universal Activation Verbalizer: A Unified Framework for Cross-Model Activation Explanation

    Haiyan Zhao, Zirui He, Guanchu Wang +3

    cs.CLcs.LGarXiv:2605.25903v22026
  25. WenetSpeech-Chuan: A Large-Scale Sichuanese Corpus with Rich Annotation for Dialectal Speech Processing

    Yuhang Dai, Ziyu Zhang, Shuai Wang +13

    cs.CLcs.SDarXiv:2509.18004v12025
  26. Fun-ASR Technical Report

    Keyu An, Yanni Chen, Zhigao Chen +35

    cs.CLcs.AIcs.SDarXiv:2509.12508v42025
  27. SubtleMemory: A Benchmark for Fine-Grained Relational Memory Discrimination in Long-Horizon AI Agents

    Wenxuan Wang, Haoyu Sun, Fukuan Hou +4

    cs.AIcs.CLarXiv:2606.05761v22026
  28. The Amazing Agent Race: Strong Tool Users, Weak Navigators

    Zae Myung Kim, Dongseok Lee, Jaehyung Kim +2

    cs.AIcs.CLcs.LGarXiv:2604.10261v22026
  29. Ask or Assume? Uncertainty-Aware Clarification-Seeking in Coding Agents

    Nicholas Edwards, Sebastian Schuster

    cs.CLarXiv:2603.26233v22026
  30. Strategic Navigation or Stochastic Search? How Agents and Humans Reason Over Document Collections

    Łukasz Borchmann, Jordy Van Landeghem, Michał Turski +12

    cs.CLcs.AIarXiv:2603.12180v22026
  31. Generative AI Act II: Test Time Scaling Drives Cognition Engineering

    Shijie Xia, Yiwei Qin, Xuefeng Li +11

    cs.CLcs.AIarXiv:2504.13828v32025
  32. MobileGUI-RL: Advancing Mobile GUI Agent through Reinforcement Learning in Online Environment

    Yucheng Shi, Wenhao Yu, Zaitang Li +5

    cs.LGcs.CLarXiv:2507.05720v12025
  33. Contrastive Instruction Tuning

    Tianyi Lorena Yan, Fei Wang, James Y. Huang +5

    cs.CLcs.AIcs.LGarXiv:2402.11138v22024
  34. A Fast Post-Training Pruning Framework for Transformers

    Woosuk Kwon, Sehoon Kim, Michael W. Mahoney +3

    cs.CLcs.LGarXiv:2204.09656v22022
  35. Hybrid Transformer with Multi-level Fusion for Multimodal Knowledge Graph Completion

    Xiang Chen, Ningyu Zhang, Lei Li +6

    cs.CLcs.AIcs.CVarXiv:2205.02357v52022
  36. The Unreliability of Explanations in Few-shot Prompting for Textual Reasoning

    Xi Ye, Greg Durrett

    cs.CLarXiv:2205.03401v22022
  37. Maieutic Prompting: Logically Consistent Reasoning with Recursive Explanations

    Jaehun Jung, Lianhui Qin, Sean Welleck +4

    cs.CLarXiv:2205.11822v22022
  38. Adding Interpretable Attention to Neural Translation Models Improves Word Alignment

    Thomas Zenkel, Joern Wuebker, John DeNero

    cs.CLarXiv:1901.11359v12019
  39. Comprehending and Ordering Semantics for Image Captioning

    Yehao Li, Yingwei Pan, Ting Yao +1

    cs.CVcs.AIcs.CLarXiv:2206.06930v12022
  40. Can large language models reason about medical questions?

    Valentin Liévin, Christoffer Egeberg Hother, Andreas Geert Motzfeldt +1

    cs.CLcs.AIcs.LGarXiv:2207.08143v42022
  41. iTool: Reinforced Fine-Tuning with Dynamic Deficiency Calibration for Advanced Tool Use

    Yirong Zeng, Xiao Ding, Yuxian Wang +8

    cs.CLcs.AIcs.LGarXiv:2501.09766v52025
  42. Language Models Are Greedy Reasoners: A Systematic Formal Analysis of Chain-of-Thought

    Abulhair Saparov, He He

    cs.CLarXiv:2210.01240v42022
  43. Language Models as Agent Models

    Jacob Andreas

    cs.CLcs.MAarXiv:2212.01681v12022
    Summaries:한국어
  44. Multi-modal Molecule Structure-text Model for Text-based Retrieval and Editing

    Shengchao Liu, Weili Nie, Chengpeng Wang +6

    cs.LGcs.CLq-bio.QMarXiv:2212.10789v32022
  45. Reinforcement Fine-Tuning Naturally Mitigates Forgetting in Continual Post-Training

    Song Lai, Haohan Zhao, Rong Feng +9

    cs.LGcs.AIcs.CLarXiv:2507.05386v62025
  46. Can Pre-trained Vision and Language Models Answer Visual Information-Seeking Questions?

    Yang Chen, Hexiang Hu, Yi Luan +4

    cs.CVcs.AIcs.CLarXiv:2302.11713v52023
  47. Learning by Distilling Context

    Charlie Snell, Dan Klein, Ruiqi Zhong

    cs.CLcs.AIarXiv:2209.15189v12022
  48. GPT-4 Technical Report

    OpenAI, Josh Achiam, Steven Adler +278

    cs.CLcs.AIarXiv:2303.08774v62023
  49. LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention

    Renrui Zhang, Jiaming Han, Chris Liu +7

    cs.CVcs.AIcs.CLarXiv:2303.16199v32023
  50. Can ChatGPT Forecast Stock Price Movements? Return Predictability and Large Language Models

    Alejandro Lopez-Lira, Yuehua Tang

    q-fin.STcs.CLarXiv:2304.07619v62023
  51. LLaMA-Adapter V2: Parameter-Efficient Visual Instruction Model

    Peng Gao, Jiaming Han, Renrui Zhang +9

    cs.CVcs.AIcs.CLarXiv:2304.15010v12023
  52. Enabling Large Language Models to Generate Text with Citations

    Tianyu Gao, Howard Yen, Jiatong Yu +1

    cs.CLcs.IRcs.LGarXiv:2305.14627v22023
  53. RS5M and GeoRSCLIP: A Large Scale Vision-Language Dataset and A Large Vision-Language Model for Remote Sensing

    Zilun Zhang, Tiancheng Zhao, Yulong Guo +1

    cs.CVcs.AIcs.CLarXiv:2306.11300v52023
  54. Matching Patients to Clinical Trials with Large Language Models

    Qiao Jin, Zifeng Wang, Charalampos S. Floudas +7

    cs.CLcs.AIarXiv:2307.15051v52023
  55. Studying Large Language Model Generalization with Influence Functions

    Roger Grosse, Juhan Bae, Cem Anil +14

    cs.LGcs.CLstat.MLarXiv:2308.03296v12023
  56. Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation

    Yangsibo Huang, Samyak Gupta, Mengzhou Xia +2

    cs.CLcs.AIcs.CRarXiv:2310.06987v12023
  57. Linear Representations of Sentiment in Large Language Models

    Curt Tigges, Oskar John Hollinsworth, Atticus Geiger +1

    cs.LGcs.AIcs.CLarXiv:2310.15154v12023
  58. Large Legal Fictions: Profiling Legal Hallucinations in Large Language Models

    Matthew Dahl, Varun Magesh, Mirac Suzgun +1

    cs.CLcs.AIcs.CYarXiv:2401.01301v22024
  59. One Polluted Page Is Enough: Evaluating Web Content Pollution in LLM Recommenders

    Minghao Luo, Liang Chen

    cs.CLcs.AIarXiv:2606.13610v22026
  60. SkillAdaptor: Self-Adapting Skills for LLM Agents from Trajectories

    Zhuoyun Yu, Xin Xie, Wuguannan Yao +4

    cs.CLcs.AIcs.LGarXiv:2606.01311v12026