Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

10,441 to 10,500 of 11,299

  1. RewardHarness: Self-Evolving Agentic Post-Training

    Yuxuan Zhang, Penghui Du, Bo Li +11

    cs.AIcs.CLcs.CVarXiv:2605.08703v12026
  2. A Single Layer to Explain Them All:Understanding Massive Activations in Large Language Models

    Zeru Shi, Zhenting Wang, Fan Yang +2

    cs.CLarXiv:2605.08504v22026
  3. InterLV-Search: Benchmarking Interleaved Multimodal Agentic Search

    Bohan Hou, Jiuning Gu, Jiayan Guo +5

    cs.CVcs.CLcs.IRarXiv:2605.07510v12026
    Summaries:한국어
  4. Teaching Language Models to Think in Code

    Hyeon Hwang, Jiwoo Lee, Jaewoo Kang

    cs.CLarXiv:2605.07237v22026
    Summaries:한국어
  5. RaguTeam at SemEval-2026 Task 8: Meno and Friends in a Judge-Orchestrated LLM Ensemble for Faithful Multi-Turn Response Generation

    Ivan Bondarenko, Roman Derunets, Oleg Sedukhin +3

    cs.CLcs.AIcs.LGarXiv:2605.04523v12026
  6. Reinforcement Learning for LLM-based Multi-Agent Systems through Orchestration Traces

    Chenchen Zhang

    cs.CLarXiv:2605.02801v12026
  7. InfoLaw: Information Scaling Laws for Large Language Models with Quality-Weighted Mixture Data and Repetition

    Fengze Liu, Weidong Zhou, Binbin Liu +7

    cs.CLarXiv:2605.02364v12026
    Summaries:한국어
  8. FAMA: Failure-Aware Meta-Agentic Framework for Open-Source LLMs in Interactive Tool Use Environments

    Amir Saeidi, Venkatesh Mishra, Souradeep Mukhopadhyay +4

    cs.CLarXiv:2604.25135v12026
  9. The Power of Scale for Parameter-Efficient Prompt Tuning

    Brian Lester, Rami Al-Rfou, Noah Constant

    cs.CLarXiv:2104.08691v22021
  10. RoBERTa: A Robustly Optimized BERT Pretraining Approach

    Yinhan Liu, Myle Ott, Naman Goyal +7

    cs.CLarXiv:1907.11692v12019
  11. Parameter-Efficient Transfer Learning for NLP

    Neil Houlsby, Andrei Giurgiu, Stanislaw Jastrzebski +5

    cs.LGcs.CLstat.MLarXiv:1902.00751v22019
  12. Prompt-Level Distillation: A Non-Parametric Alternative to Model Fine-Tuning for Efficient Reasoning

    Sanket Badhe, Deep Shah

    cs.CLcs.IRarXiv:2602.21103v22026
  13. Diffusion-Pretrained Dense and Contextual Embeddings

    Sedigheh Eslami, Maksim Gaiduk, Markus Krimmel +3

    cs.LGcs.CLcs.IRarXiv:2602.11151v22026
    Summaries:한국어
  14. FollowUpBot: An LLM-Based Conversational Robot for Automatic Postoperative Follow-up

    Chen Chen, Jianing Yin, Jiannong Cao +5

    cs.HCcs.CLcs.ROarXiv:2507.15502v12025
  15. DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

    Zhihong Shao, Peiyi Wang, Qihao Zhu +8

    cs.CLcs.AIcs.LGarXiv:2402.03300v32024
  16. Qwen Technical Report

    Jinze Bai, Shuai Bai, Yunfei Chu +45

    cs.CLarXiv:2309.16609v12023
  17. RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

    Anthony Brohan, Noah Brown, Justice Carbajal +51

    cs.ROcs.CLcs.CVarXiv:2307.15818v12023
  18. Large Language Models Encode Clinical Knowledge

    Karan Singhal, Shekoofeh Azizi, Tao Tu +27

    cs.CLarXiv:2212.13138v12022
  19. Robust Speech Recognition via Large-Scale Weak Supervision

    Alec Radford, Jong Wook Kim, Tao Xu +3

    eess.AScs.CLcs.LGarXiv:2212.04356v12022
  20. Large Language Models are Zero-Shot Reasoners

    Takeshi Kojima, Shixiang Shane Gu, Machel Reid +2

    cs.CLcs.AIcs.LGarXiv:2205.11916v42022
  21. Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

    Yuntao Bai, Andy Jones, Kamal Ndousse +28

    cs.CLcs.LGarXiv:2204.05862v12022
  22. TruthfulQA: Measuring How Models Mimic Human Falsehoods

    Stephanie Lin, Jacob Hilton, Owain Evans

    cs.CLcs.AIcs.CYarXiv:2109.07958v22021
  23. Don't Stop Pretraining: Adapt Language Models to Domains and Tasks

    Suchin Gururangan, Ana Marasović, Swabha Swayamdipta +4

    cs.CLcs.LGarXiv:2004.10964v32020
  24. Dense Passage Retrieval for Open-Domain Question Answering

    Vladimir Karpukhin, Barlas Oğuz, Sewon Min +5

    cs.CLarXiv:2004.04906v32020
  25. BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

    Jacob Devlin, Ming-Wei Chang, Kenton Lee +1

    cs.CLarXiv:1810.04805v22018
  26. An Empirical Evaluation of Generic Convolutional and Recurrent Networks for Sequence Modeling

    Shaojie Bai, J. Zico Kolter, Vladlen Koltun

    cs.LGcs.AIcs.CLarXiv:1803.01271v22018
  27. DeepTCM1.0: A Multi-Expert AI Agent for Deciphering Mechanisms of Chinese Herbal Formulae Based on General Large Language Models

    Wenxin Duan, Hanwei Wang, Zhongying Peng +4

    cs.CLcs.AIarXiv:2608.18103v12026
  28. LLMs for Medical Consultation Are Evaluated Too Late: The Preformulation Gap

    Yining Hua, Cyrus Ayubcha, Hongbin Na +4

    cs.AIcs.CLcs.CYarXiv:2608.17330v12026
  29. Perturbation-based Regional Interpretability through Subtraction Mapping (PRISM): naming-error dissociations in language models and post-stroke aphasia

    Xiang Guan, Roger D. Newman-Norlund, Yong Yang +8

    cs.LGcs.CLarXiv:2608.12717v12026
  30. Intent Speaks Louder: Controllable User Simulation Beyond Response Imitation

    Bo Wang, Ruixing Zhang, Yunqi Liu +4

    cs.CLarXiv:2608.09420v12026
  31. Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application

    Jiachun Li, Zhuoran Jin, Tianyi Men +12

    cs.CLcs.AIarXiv:2606.12191v12026
  32. Manifold Bandits: Bayesian Curriculum Learning over the Latent Geometry of Large Language Models

    Darrien McKenzie, Nicklas Hansen, Xiaolong Wang

    cs.LGcs.AIcs.CLarXiv:2606.19750v12026
  33. The Limits of Binding in Dual Encoders

    Kin Ian Lo

    cs.LGcs.CLcs.CVarXiv:2608.15971v12026
  34. OScaR: The Occam's Razor for Extreme KV Cache Quantization in LLMs and Beyond

    Zunhai Su, Rui Yang, Chao Zhang +11

    cs.LGcs.CLarXiv:2605.19660v12026
  35. PAAC: Privacy-Aware Agentic Device-Cloud Collaboration

    Liangqi Yuan, Wenzhi Fang, Shiqiang Wang +1

    cs.LGcs.CLcs.DCarXiv:2605.08646v12026
  36. Continuous Latent Diffusion Language Model

    Hongcan Guo, Qinyu Zhao, Yian Zhao +8

    cs.CLcs.AIcs.CVarXiv:2605.06548v12026
    Summaries:한국어
  37. EVA-Bench: A New End-to-end Framework for Evaluating Voice Agents

    Tara Bogavelli, Gabrielle Gauthier Melançon, Katrina Stankiewicz +10

    cs.SDcs.AIcs.CLarXiv:2605.13841v22026
  38. Bag of Tricks for Efficient Text Classification

    Armand Joulin, Edouard Grave, Piotr Bojanowski +1

    cs.CLarXiv:1607.01759v32016
  39. MineExplorer: Evaluating Open-World Exploration of MLLM Agents in Minecraft

    Tianjie Ju, Yueqing Sun, Zheng Wu +7

    cs.CLarXiv:2605.30931v22026
  40. ChLogic: Evaluating Robustness of Logical Reasoning in Chinese Expressions

    Peixian Zhou, Yuxu Chen, Chaorui Zhang +3

    cs.CLarXiv:2606.17905v12026
  41. Budget-First Tariff Recommendation (BFTR): A Complete Algorithmic Framework for Telecom Plan Recommendation without Overcharging

    Ghislain Dorian Tchuente Mondjo

    cs.CLcs.AIarXiv:2608.18723v12026
  42. Matryoshka Language Model Suites

    Nathan Godey, Yoav Artzi

    cs.AIcs.CLarXiv:2608.09703v12026
    Summaries:한국어
  43. GPQA: A Graduate-Level Google-Proof Q&A Benchmark

    David Rein, Betty Li Hou, Asa Cooper Stickland +5

    cs.AIcs.CLarXiv:2311.12022v12023
  44. Deep Speech 2: End-to-End Speech Recognition in English and Mandarin

    Dario Amodei, Rishita Anubhai, Eric Battenberg +31

    cs.CLarXiv:1512.02595v12015
  45. CLIPScore: A Reference-free Evaluation Metric for Image Captioning

    Jack Hessel, Ari Holtzman, Maxwell Forbes +2

    cs.CVcs.CLarXiv:2104.08718v32021
  46. DeepFM: A Factorization-Machine based Neural Network for CTR Prediction

    Huifeng Guo, Ruiming Tang, Yunming Ye +2

    cs.IRcs.CLarXiv:1703.04247v12017
  47. Semantics derived automatically from language corpora contain human-like biases

    Aylin Caliskan, Joanna J. Bryson, Arvind Narayanan

    cs.AIcs.CLcs.CYarXiv:1608.07187v42016
  48. Natural TTS Synthesis by Conditioning WaveNet on Mel Spectrogram Predictions

    Jonathan Shen, Ruoming Pang, Ron J. Weiss +10

    cs.CLarXiv:1712.05884v22017
  49. Extracting Training Data from Large Language Models

    Nicholas Carlini, Florian Tramer, Eric Wallace +9

    cs.CRcs.CLcs.LGarXiv:2012.07805v22020
  50. mT5: A massively multilingual pre-trained text-to-text transformer

    Linting Xue, Noah Constant, Adam Roberts +5

    cs.CLarXiv:2010.11934v32020
  51. REALM: Retrieval-Augmented Language Model Pre-Training

    Kelvin Guu, Kenton Lee, Zora Tung +2

    cs.CLcs.LGarXiv:2002.08909v12020
  52. PIQA: Reasoning about Physical Commonsense in Natural Language

    Yonatan Bisk, Rowan Zellers, Ronan Le Bras +2

    cs.CLcs.AIcs.LGarXiv:1911.11641v12019
  53. Automated Hate Speech Detection and the Problem of Offensive Language

    Thomas Davidson, Dana Warmsley, Michael Macy +1

    cs.CLarXiv:1703.04009v12017
  54. Improved Semantic Representations From Tree-Structured Long Short-Term Memory Networks

    Kai Sheng Tai, Richard Socher, Christopher D. Manning

    cs.CLcs.AIcs.LGarXiv:1503.00075v32015
  55. InstructPix2Pix: Learning to Follow Image Editing Instructions

    Tim Brooks, Aleksander Holynski, Alexei A. Efros

    cs.CVcs.AIcs.CLarXiv:2211.09800v22022
  56. Language Models as Knowledge Bases?

    Fabio Petroni, Tim Rocktäschel, Patrick Lewis +4

    cs.CLarXiv:1909.01066v22019
  57. ConceptNet 5.5: An Open Multilingual Graph of General Knowledge

    Robyn Speer, Joshua Chin, Catherine Havasi

    cs.CLarXiv:1612.03975v22016
  58. SWE-bench: Can Language Models Resolve Real-World GitHub Issues?

    Carlos E. Jimenez, John Yang, Alexander Wettig +4

    cs.CLcs.AIcs.SEarXiv:2310.06770v32023
  59. Know What You Don't Know: Unanswerable Questions for SQuAD

    Pranav Rajpurkar, Robin Jia, Percy Liang

    cs.CLarXiv:1806.03822v12018
  60. fairseq: A Fast, Extensible Toolkit for Sequence Modeling

    Myle Ott, Sergey Edunov, Alexei Baevski +5

    cs.CLarXiv:1904.01038v12019