Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

121 to 180 of 11,233

  1. Toward Auto-Research: Mining Falsifiable Research Ideas from Paper Knowledge Graphs with Categorical Structure

    Yuchen Wang, Zhongzhi Luan

    cs.CLcs.AIarXiv:2608.20361v12026
  2. How to Train a Real-World Silicon Concierge? Internalizing Complex Business Workflow to Only OneModel

    Chang Liu, Chaoyang Ning, Dayi Jiang +32

    cs.CLcs.AIarXiv:2608.20350v12026
  3. What is Missing from AI Post-Training AI: An Empirical Analysis

    Joy Jia Yin Lim, Xin Huang, Hao Peng +5

    cs.AIcs.CLcs.LGarXiv:2608.19072v12026
  4. The Plot Thins: Uniformity and Linearity in Literary Summaries

    Rebecca M. M. Hicke, Sil Hamilton, David Mimno +1

    cs.CLarXiv:2608.17218v12026
  5. A Survey of Reinforcement Learning for Large Reasoning Models

    Kaiyan Zhang, Yuxin Zuo, Bingxiang He +36

    cs.CLcs.AIcs.LGarXiv:2509.08827v32025
  6. Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

    Gheorghe Comanici, Eric Bieber, Mike Schaekermann +3432

    cs.CLcs.AIarXiv:2507.06261v62025
  7. Pre-Training BERT on Arabic Tweets: Practical Considerations

    Ahmed Abdelali, Sabit Hassan, Hamdy Mubarak +2

    cs.CLcs.AIarXiv:2102.10684v12021
  8. Moral Stories: Situated Reasoning about Norms, Intents, Actions, and their Consequences

    Denis Emelin, Ronan Le Bras, Jena D. Hwang +2

    cs.CLcs.AIarXiv:2012.15738v12020
  9. Mapping the Space of Chemical Reactions Using Attention-Based Neural Networks

    Philippe Schwaller, Daniel Probst, Alain C. Vaucher +4

    physics.chem-phcs.CLcs.LGarXiv:2012.06051v12020
  10. J1: Incentivizing Thinking in LLM-as-a-Judge via Reinforcement Learning

    Chenxi Whitehouse, Tianlu Wang, Ping Yu +4

    cs.CLcs.AIcs.LGarXiv:2505.10320v32025
  11. Absolute Zero: Reinforced Self-play Reasoning with Zero Data

    Andrew Zhao, Yiran Wu, Yang Yue +8

    cs.LGcs.AIcs.CLarXiv:2505.03335v32025
  12. STAIR: Semantic-Temporal Automaton for Interpretable Reasoning in Temporal Question Answering

    Xinlong Dai, Jinchuan Zhang, Lei Gao +3

    cs.CLcs.AIarXiv:2608.16224v12026
  13. Social Chemistry 101: Learning to Reason about Social and Moral Norms

    Maxwell Forbes, Jena D. Hwang, Vered Shwartz +2

    cs.CLcs.AIarXiv:2011.00620v32020
  14. Open Question Answering over Tables and Text

    Wenhu Chen, Ming-Wei Chang, Eva Schlinger +2

    cs.CLcs.AIarXiv:2010.10439v22020
  15. Dynamic Early Exit in Reasoning Models

    Chenxu Yang, Qingyi Si, Yongjie Duan +6

    cs.CLcs.AIarXiv:2504.15895v32025
  16. Interpreting Graph Neural Networks for NLP With Differentiable Edge Masking

    Michael Sejr Schlichtkrull, Nicola De Cao, Ivan Titov

    cs.CLcs.LGstat.MLarXiv:2010.00577v32020
  17. KG-BART: Knowledge Graph-Augmented BART for Generative Commonsense Reasoning

    Ye Liu, Yao Wan, Lifang He +2

    cs.CLcs.SCarXiv:2009.12677v22020
  18. DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning

    Zhiwei He, Tian Liang, Jiahao Xu +12

    cs.CLcs.AIarXiv:2504.11456v22025
  19. GraphCodeBERT: Pre-training Code Representations with Data Flow

    Daya Guo, Shuo Ren, Shuai Lu +15

    cs.SEcs.CLarXiv:2009.08366v42020
  20. Language models suffer from a curse of ambiguity

    Nicolas Zucchet, Hyun Dong Lee, Scott Linderman

    cs.CLcs.LGcs.NEarXiv:2608.15448v12026
  21. SEAL: Steerable Reasoning Calibration of Large Language Models for Free

    Runjin Chen, Zhenyu Zhang, Junyuan Hong +2

    cs.CLcs.AIarXiv:2504.07986v32025
  22. Calibrated Trust, Not Sharper Prediction: An Empirical Test of Uncertainty Fusion

    Surya Saka

    cs.LGcs.AIcs.CLarXiv:2608.14617v12026
  23. Plausible but Not Valid: A Psychometric Audit of LLMs as Synthetic Survey Respondents

    Mantas Lukauskas, Viktorija Šarkauskaitė

    cs.CYcs.AIcs.CLarXiv:2608.14606v12026
  24. Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus

    Zunhai Su, Bohan Sun, Xialie Zhuang +8

    cs.CLarXiv:2608.12149v12026
  25. Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities

    Sreyan Ghosh, Zhifeng Kong, Sonal Kumar +6

    cs.SDcs.CLcs.LGarXiv:2503.03983v12025
  26. Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks

    Patrick Lewis, Ethan Perez, Aleksandra Piktus +9

    cs.CLcs.LGarXiv:2005.11401v42020
    Summaries:한국어
  27. When Agents Learn to Be You: Benchmarking Privacy Leakage, Impersonation Risk, and Defenses in Persona Skills

    Yongli Xiang, Zhifang Zhang, Bojun Yang +4

    cs.CRcs.CLcs.CYarXiv:2608.03700v12026
  28. Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos

    Kairui Hu, Penghao Wu, Fanyi Pu +5

    cs.CVcs.CLarXiv:2501.13826v12025
  29. Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought

    Violet Xiang, Charlie Snell, Kanishk Gandhi +11

    cs.AIcs.CLarXiv:2501.04682v12025
  30. SWE-Pruner Pro: The Coder LLM Already Knows What to Prune

    Yuhang Wang, Yuling Shi, Shaoqiu Zhang +6

    cs.CLcs.SEarXiv:2607.18213v12026
  31. Improving Medical Large Vision-Language Models with Abnormal-Aware Feedback

    Yucheng Zhou, Lingran Song, Jianbing Shen

    cs.CLcs.AIcs.CVarXiv:2501.01377v22025
  32. Is One Layer Enough? Training A Single Transformer Layer Can Match Full-Parameter RL Training

    Zijian Zhang, Rizhen Hu, Athanasios Glentis +4

    cs.LGcs.CLarXiv:2607.01232v22026
  33. Large Language Model-Brained GUI Agents: A Survey

    Chaoyun Zhang, Shilin He, Jiaxu Qian +10

    cs.AIcs.CLcs.HCarXiv:2411.18279v122024
  34. LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts

    Yijia Xiao, Edward Sun, Tianyu Liu +1

    cs.AIcs.CLcs.CVarXiv:2407.04973v12024
  35. A Survey on Knowledge Graphs: Representation, Acquisition and Applications

    Shaoxiong Ji, Shirui Pan, Erik Cambria +2

    cs.CLcs.AIarXiv:2002.00388v42020
  36. Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs

    Xin Lai, Zhuotao Tian, Yukang Chen +3

    cs.LGcs.AIcs.CLarXiv:2406.18629v12024
  37. Glottal Closure and Opening Instant Detection from Speech Signals

    Thomas Drugman, Thierry Dutoit

    cs.SDcs.CLeess.ASarXiv:2001.00841v12019
  38. MAIRA-2: Grounded Radiology Report Generation

    Shruthi Bannur, Kenza Bouzid, Daniel C. Castro +18

    cs.CLcs.CVarXiv:2406.04449v22024
  39. Detection of Glottal Closure Instants from Speech Signals: a Quantitative Review

    Thomas Drugman, Mark Thomas, Jon Gudnason +2

    cs.SDcs.CLeess.ASarXiv:2001.00473v12019
  40. AI translation of literary texts is "fine", but readers still prefer human translations

    Yves Ferstler, Adam Podoxin, Ty Brassington +3

    cs.CLarXiv:2606.26040v12026
  41. OR-Bench: An Over-Refusal Benchmark for Large Language Models

    Justin Cui, Wei-Lin Chiang, Ion Stoica +1

    cs.CLcs.AIarXiv:2405.20947v52024
  42. SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering

    John Yang, Carlos E. Jimenez, Alexander Wettig +4

    cs.SEcs.AIcs.CLarXiv:2405.15793v32024
  43. CAVEWOMAN: How Large Language Models Behave Under Linguistic Input and Output Compression

    Morayo Danielle Adeyemi, Ryan A. Rossi, Franck Dernoncourt

    cs.CLcs.AIcs.LGarXiv:2606.24083v12026
  44. LayerSkip: Enabling Early Exit Inference and Self-Speculative Decoding

    Mostafa Elhoushi, Akshat Shrivastava, Diana Liskovich +10

    cs.CLcs.AIcs.LGarXiv:2404.16710v42024
  45. Graph-Based Reasoning over Heterogeneous External Knowledge for Commonsense Question Answering

    Shangwen Lv, Daya Guo, Jingjing Xu +7

    cs.CLarXiv:1909.05311v22019
  46. CODA-BENCH: Can Code Agents Handle Data-Intensive Tasks?

    Yuxin Zhang, Ju Fan, Meihao Fan +2

    cs.AIcs.CLarXiv:2606.15300v12026
  47. Do Large Language Models Latently Perform Multi-Hop Reasoning?

    Sohee Yang, Elena Gribovskaya, Nora Kassner +2

    cs.CLarXiv:2402.16837v22024
  48. Attacks, Defenses and Evaluations for LLM Conversation Safety: A Survey

    Zhichen Dong, Zhanhui Zhou, Chao Yang +2

    cs.CLcs.AIcs.CYarXiv:2402.09283v32024
  49. SPHINX-X: Scaling Data and Parameters for a Family of Multi-modal Large Language Models

    Dongyang Liu, Renrui Zhang, Longtian Qiu +16

    cs.CVcs.AIcs.CLarXiv:2402.05935v32024
  50. Reliable Fine-Grained Evaluation of Natural Language Math Proofs

    Wenjie Ma, Andrei Cojocaru, Neel Kolhe +6

    cs.CLcs.AIarXiv:2510.13888v22025
  51. GeoBrowse: A Geolocation Benchmark for Agentic Tool Use with Expert-Annotated Reasoning Traces

    Xinyu Geng, Yanjing Xiao, Yuyang Zhang +5

    cs.CLarXiv:2604.04017v12026
  52. A Nested Attention Neural Hybrid Model for Grammatical Error Correction

    Jianshu Ji, Qinlong Wang, Kristina Toutanova +3

    cs.CLarXiv:1707.02026v22017
  53. Improving Factual Consistency of Abstractive Summarization via Question Answering

    Feng Nan, Cicero Nogueira dos Santos, Henghui Zhu +7

    cs.CLcs.AIarXiv:2105.04623v12021
  54. Emoti-Attack: Zero-Perturbation Adversarial Attacks on NLP Systems via Emoji Sequences

    Yangshijie Zhang

    cs.AIcs.CLcs.CRarXiv:2502.17392v12025
  55. Fair Transfer of Multiple Style Attributes in Text

    Karan Dabas, Nishtha Madan, Vijay Arya +3

    cs.CLcs.AIcs.LGarXiv:2001.06693v12020
  56. Understanding Back-Translation at Scale

    Sergey Edunov, Myle Ott, Michael Auli +1

    cs.CLarXiv:1808.09381v22018
  57. PEER: A Collaborative Language Model

    Timo Schick, Jane Dwivedi-Yu, Zhengbao Jiang +7

    cs.CLarXiv:2208.11663v12022
  58. Training Language Models with Language Feedback

    Jérémy Scheurer, Jon Ander Campos, Jun Shern Chan +3

    cs.CLcs.AIcs.LGarXiv:2204.14146v42022
  59. Rainier: Reinforced Knowledge Introspector for Commonsense Question Answering

    Jiacheng Liu, Skyler Hallinan, Ximing Lu +4

    cs.CLcs.AIarXiv:2210.03078v22022
  60. Towards Empathetic Open-domain Conversation Models: a New Benchmark and Dataset

    Hannah Rashkin, Eric Michael Smith, Margaret Li +1

    cs.CLarXiv:1811.00207v52018