Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

9,841 to 9,900 of 11,411

  1. Describing Videos by Exploiting Temporal Structure

    Li Yao, Atousa Torabi, Kyunghyun Cho +4

    stat.MLcs.AIcs.CLarXiv:1502.08029v52015
  2. ProofJudge: Tool-Grounded LLM Evaluation of Formal Proof Quality in Mathlib

    Shane Caldwell

    cs.LOcs.AIcs.CLarXiv:2608.20432v12026
  3. LingShu: A Large-Scale Symptom-Centric Contextualized Knowledge Graph Bridging Traditional Chinese Medicine and Modern Biomedicine

    Rui Hua, Zixin Shu, Kai Chang +18

    cs.CLcs.AIarXiv:2608.20402v12026
  4. Knowledge-Graph-Gated Defactualization for Style-Controllable and Fact-Preserving Generation in Agentic Conversational AI

    Tanmay Kumar Shrivastava, Darsh Rohit Nandu, Rajesh Kumar Mundotiya

    cs.CLcs.AIarXiv:2608.20393v12026
  5. Evaluation-as-Search: Adaptive Discovery of Grounding Failures in Meeting Assistants

    Sami Khairy, Yasaman Hosseinkashi, Vishak Gopal +1

    cs.CLcs.AIarXiv:2608.20392v12026
  6. EditPPT: Faithful Long-Deck Slide Editing via Structured Tool-Using Multi-Agent with Dual-Modal Validators

    Jiheon Kim, Kyudan Jung, Jaegul Choo

    cs.CLcs.AIcs.HCarXiv:2608.20381v12026
  7. Ansari: A Retrieval-Grounded Islamic AI Assistant -- Architecture, Deployment, and Lessons from 140,000 Conversations

    M Waleed Kadous, Amr Elsayed, Abdullah Al Nahas +1

    cs.CLcs.AIcs.CYarXiv:2608.20390v12026
  8. ASTAR: Automated induction of STAndardized radiology Reporting templates from large-scale clinical free-text corpora

    Xinfeng Zhang, Mingxuan Liu, Yifei Chen +9

    cs.CLcs.AIarXiv:2608.20369v12026
  9. VA-DPO: Valence-Arousal Direct Preference Optimization for Controllable Emotion Generation in Language Models

    Hyunwoo Kim

    cs.CLcs.AIcs.LGarXiv:2608.20374v12026
  10. Let's Scale Step by Step: Compute-Efficient Hyperparameter Transfer for Large-Scale Mixture-of-Experts

    Nayeon Kim, Hojin Lee, Yunju Bak +2

    cs.LGcs.AIcs.CLarXiv:2608.20061v12026
  11. BARTScore: Evaluating Generated Text as Text Generation

    Weizhe Yuan, Graham Neubig, Pengfei Liu

    cs.CLarXiv:2106.11520v22021
  12. ExpertIVS: Sociological Expert Driven Individual Value Simulation in Large Language Models

    Zhen Wang, Yuqi Ren, Yuehan Cui +7

    cs.CLcs.AIarXiv:2608.20355v12026
  13. The Divergence Hypothesis: Unmasking Lexical Interference and Label Bias in Mental Health NLP

    Moustafa Yehia Hassan

    cs.CLcs.AIarXiv:2608.20353v12026
  14. Beyond Prompt Engineering: A Systematic Analysis of Prompt Lexical Sensitivity and Its Impacts on Quality

    Qipeng Xie, Zi Liang, Jiafei Wu +6

    cs.CLcs.AIarXiv:2608.20349v12026
  15. Inhibitory Attention for Clinical Long-Context Reasoning: Characterizing and Mitigating Lost-in-the-Middle Effects in EHR Processing

    Sanjay Basu

    cs.CLcs.AIarXiv:2608.20348v12026
  16. When Vocabulary Comprehension Fails Clinical Reasoning: Evaluating Therapy Bots' Safety Risks for Generation Alpha

    Manisha Mehta, Virendra Mehta

    cs.CLcs.AIcs.CYarXiv:2608.20345v12026
  17. Who Do Language Models Think Is Competent? A Mechanistic Analysis of Occupational Bias

    Keren Fuentes, Aaron Mueller

    cs.CLcs.AIcs.CYarXiv:2608.20347v12026
  18. Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP Tasks

    Yizhong Wang, Swaroop Mishra, Pegah Alipoormolabashi +37

    cs.CLcs.AIarXiv:2204.07705v32022
  19. Personalized Privacy Control in LLMs via Attention Head Intervention

    Junseok Kim, Nakyeong Yang, Kyomin Jung

    cs.AIcs.CLcs.LGarXiv:2608.21209v12026
  20. Enhancing LLMs in Predictive Political QA with Semi-Structured Data

    Yinan Liu, Zihan Zhou, Zichun Jin +3

    cs.AIcs.CLcs.IRarXiv:2608.21218v12026
  21. Meshed-Memory Transformer for Image Captioning

    Marcella Cornia, Matteo Stefanini, Lorenzo Baraldi +1

    cs.CVcs.CLarXiv:1912.08226v22019
  22. Generalization through Memorization: Nearest Neighbor Language Models

    Urvashi Khandelwal, Omer Levy, Dan Jurafsky +2

    cs.CLarXiv:1911.00172v22019
  23. TreeWY: Speculative Verification for Gated DeltaNet Hybrids

    Sneha Murthy Ghantasala

    cs.AIcs.CLcs.DCarXiv:2608.20961v12026
  24. Calibrating Criterion Revision in LLM Agents: Failure Modes and a Trace-Anchored Protocol

    Guodong Xu

    cs.AIcs.CLarXiv:2608.20729v12026
  25. Why2Speak: Faithful Reasoning for Abstaining Action Policies

    Shreya Mendi, Brinnae Bent

    cs.AIcs.CLarXiv:2608.20670v12026
  26. CamemBERT: a Tasty French Language Model

    Louis Martin, Benjamin Muller, Pedro Javier Ortiz Suárez +5

    cs.CLarXiv:1911.03894v32019
  27. Can LLM Already Serve as A Database Interface? A BIg Bench for Large-Scale Database Grounded Text-to-SQLs

    Jinyang Li, Binyuan Hui, Ge Qu +15

    cs.CLarXiv:2305.03111v32023
  28. Universal Adversarial Triggers for Attacking and Analyzing NLP

    Eric Wallace, Shi Feng, Nikhil Kandpal +2

    cs.CLcs.LGarXiv:1908.07125v32019
  29. Reasoning with Language Model is Planning with World Model

    Shibo Hao, Yi Gu, Haodi Ma +4

    cs.CLcs.AIcs.LGarXiv:2305.14992v22023
  30. Long Short-Term Memory Based Recurrent Neural Network Architectures for Large Vocabulary Speech Recognition

    Haşim Sak, Andrew Senior, Françoise Beaufays

    cs.NEcs.CLcs.LGarXiv:1402.1128v12014
  31. SimPO: Simple Preference Optimization with a Reference-Free Reward

    Yu Meng, Mengzhou Xia, Danqi Chen

    cs.CLcs.LGarXiv:2405.14734v32024
  32. Siren's Song in the AI Ocean: A Survey on Hallucination in Large Language Models

    Yue Zhang, Yafu Li, Leyang Cui +13

    cs.CLcs.AIcs.CYarXiv:2309.01219v32023
  33. Auditable by Construction: An Ontology-Driven Framework for Trustworthy LLM Analytics in Enterprise Finance

    Sergiy Lunyakin

    cs.AIcs.CEcs.CLarXiv:2608.20661v12026
  34. Open-Weight Masked Introspection: Measuring What Language Models Can Report About Their Own Computation

    Emilio Ferrara

    cs.AIcs.CLarXiv:2608.20569v12026
  35. It's Not Just Size That Matters: Small Language Models Are Also Few-Shot Learners

    Timo Schick, Hinrich Schütze

    cs.CLcs.AIcs.LGarXiv:2009.07118v22020
  36. P-Tuning v2: Prompt Tuning Can Be Comparable to Fine-tuning Universally Across Scales and Tasks

    Xiao Liu, Kaixuan Ji, Yicheng Fu +4

    cs.CLarXiv:2110.07602v32021
  37. Contrastive Learning of Medical Visual Representations from Paired Images and Text

    Yuhao Zhang, Hang Jiang, Yasuhide Miura +2

    cs.CVcs.CLcs.LGarXiv:2010.00747v22020
  38. QANet: Combining Local Convolution with Global Self-Attention for Reading Comprehension

    Adams Wei Yu, David Dohan, Minh-Thang Luong +4

    cs.CLcs.AIcs.LGarXiv:1804.09541v12018
  39. BERTweet: A pre-trained language model for English Tweets

    Dat Quoc Nguyen, Thanh Vu, Anh Tuan Nguyen

    cs.CLcs.LGarXiv:2005.10200v22020
  40. Large Language Models Cannot Self-Correct Reasoning Yet

    Jie Huang, Xinyun Chen, Swaroop Mishra +4

    cs.CLcs.AIarXiv:2310.01798v22023
  41. A Generalist Agent

    Scott Reed, Konrad Zolna, Emilio Parisotto +17

    cs.AIcs.CLcs.LGarXiv:2205.06175v32022
  42. When Retrieval Fails Before It Begins: Structurally Indirect Prerequisite Eviction as a Retention Failure in Agentic Memory

    Minkyu Song

    cs.AIcs.CLcs.LGarXiv:2608.20400v12026
  43. A Survey on In-context Learning

    Qingxiu Dong, Lei Li, Damai Dai +11

    cs.CLcs.AIarXiv:2301.00234v62022
  44. XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

    Arun Babu, Changhan Wang, Andros Tjandra +10

    cs.CLcs.SDeess.ASarXiv:2111.09296v32021
  45. A Hierarchical Latent Variable Encoder-Decoder Model for Generating Dialogues

    Iulian Vlad Serban, Alessandro Sordoni, Ryan Lowe +4

    cs.CLcs.AIcs.LGarXiv:1605.06069v32016
  46. Large Language Models Can Be Easily Distracted by Irrelevant Context

    Freda Shi, Xinyun Chen, Kanishka Misra +5

    cs.CLcs.AIarXiv:2302.00093v32023
  47. SelfCheckGPT: Zero-Resource Black-Box Hallucination Detection for Generative Large Language Models

    Potsawee Manakul, Adian Liusie, Mark J. F. Gales

    cs.CLarXiv:2303.08896v32023
  48. Did Aristotle Use a Laptop? A Question Answering Benchmark with Implicit Reasoning Strategies

    Mor Geva, Daniel Khashabi, Elad Segal +3

    cs.CLarXiv:2101.02235v12021
  49. An Explanation of In-context Learning as Implicit Bayesian Inference

    Sang Michael Xie, Aditi Raghunathan, Percy Liang +1

    cs.CLcs.LGarXiv:2111.02080v62021
  50. Visualizing and Understanding Recurrent Networks

    Andrej Karpathy, Justin Johnson, Li Fei-Fei

    cs.LGcs.CLcs.NEarXiv:1506.02078v22015
  51. Towards Understanding the Robustness of Sparse Autoencoders

    Ahson Saiyed, Sabrina Sadiekh, Chirag Agarwal

    cs.LGcs.AIcs.CLarXiv:2604.18756v12026
  52. Hadith computational science in the age of large language models: a critical narrative review

    Md. Ashraful Haque, Riasat Islam

    cs.CLcs.AIarXiv:2608.20364v12026
  53. DetectGPT: Zero-Shot Machine-Generated Text Detection using Probability Curvature

    Eric Mitchell, Yoonho Lee, Alexander Khazatsky +2

    cs.CLcs.AIarXiv:2301.11305v22023
  54. Neural Responding Machine for Short-Text Conversation

    Lifeng Shang, Zhengdong Lu, Hang Li

    cs.CLcs.AIcs.NEarXiv:1503.02364v22015
  55. AgentMercury: Your Agent Can Synthesize Verifiable Environments for Business Scenarios at scale

    Minbyul Jeong, Chanwoong Yoon

    cs.CLcs.AIarXiv:2608.20634v12026
  56. MelGAN: Generative Adversarial Networks for Conditional Waveform Synthesis

    Kundan Kumar, Rithesh Kumar, Thibault de Boissiere +6

    eess.AScs.CLcs.LGarXiv:1910.06711v32019
  57. Recipes for building an open-domain chatbot

    Stephen Roller, Emily Dinan, Naman Goyal +9

    cs.CLcs.AIarXiv:2004.13637v22020
  58. Unsupervised Machine Translation Using Monolingual Corpora Only

    Guillaume Lample, Alexis Conneau, Ludovic Denoyer +1

    cs.CLcs.AIarXiv:1711.00043v22017
  59. A Network-based End-to-End Trainable Task-oriented Dialogue System

    Tsung-Hsien Wen, David Vandyke, Nikola Mrksic +5

    cs.CLcs.AIcs.NEarXiv:1604.04562v32016
  60. CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

    Wenyi Hong, Ming Ding, Wendi Zheng +2

    cs.CVcs.CLcs.LGarXiv:2205.15868v12022