Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,341 to 8,400 of 11,248

  1. Contextual Augmentation: Data Augmentation by Words with Paradigmatic Relations

    Sosuke Kobayashi

    cs.CLcs.LGarXiv:1805.06201v12018
  2. Yor-Sarc: A gold-standard dataset for sarcasm detection in a low-resource African language

    Toheeb Aduramomi Jimoh, Tabea De Wille, Nikola S. Nikolov

    cs.CLarXiv:2602.18964v12026
  3. Small Language Models for Privacy-Preserving Clinical Information Extraction in Low-Resource Languages

    Mohammadreza Ghaffarzadeh-Esfahani, Nahid Yousefian, Ebrahim Heidari-Farsani +4

    cs.CLcs.AIcs.LGarXiv:2602.21374v12026
  4. Humans and LLMs Diverge on Probabilistic Inferences

    Gaurav Kamath, Sreenath Madathil, Sebastian Schuster +2

    cs.CLcs.AIarXiv:2602.23546v12026
  5. Q-BERT: Hessian Based Ultra Low Precision Quantization of BERT

    Sheng Shen, Zhen Dong, Jiayu Ye +5

    cs.CLcs.LGarXiv:1909.05840v22019
  6. This Just In: Fake News Packs a Lot in Title, Uses Simpler, Repetitive Content in Text Body, More Similar to Satire than Real News

    Benjamin D. Horne, Sibel Adali

    cs.SIcs.CLarXiv:1703.09398v12017
  7. MOSI: Multimodal Corpus of Sentiment Intensity and Subjectivity Analysis in Online Opinion Videos

    Amir Zadeh, Rowan Zellers, Eli Pincus +1

    cs.CLcs.MMarXiv:1606.06259v22016
  8. The RAT: A Unified Bayesian Model for RAG Evaluation

    Pius von Däniken, Felix Matthias Saaro, Mark Cieliebak +1

    cs.CLcs.AIarXiv:2608.24753v12026
  9. Legal RAG Bench: an end-to-end benchmark for legal RAG

    Abdur-Rahman Butler, Umar Butler

    cs.CLcs.IRcs.LGarXiv:2603.01710v12026
  10. Who is the Agent to Blame? Localizing Faithfulness and Citation Mistakes in Agentic Deep Research

    Eran Hirsch, David Wan, Han Wang +3

    cs.CLarXiv:2608.24306v12026
  11. Surgical Post-Training: Proximal On-Policy Distillation for Reasoning with Knowledge Retention

    Wenye Lin, Kai Han

    cs.CLcs.AIarXiv:2603.01683v22026
  12. OpenAutoNLU: Open Source AutoML Library for NLU

    Grigory Arshinov, Aleksandr Boriskin, Sergey Senichev +4

    cs.CLcs.LGarXiv:2603.01824v12026
  13. CharacterFlywheel: Scaling Iterative Improvement of Engaging and Steerable LLMs in Production

    Yixin Nie, Lin Guan, Zhongyao Ma +19

    cs.CLcs.AIcs.SIarXiv:2603.01973v12026
  14. Linear Probing Provides Robust and Efficient Detection of Machine-Generated Text

    Gerrit Quaremba, Hanqi Yan, Elizabeth Black +2

    cs.CLarXiv:2608.24780v12026
  15. MUSE: A Run-Centric Platform for Multimodal Unified Safety Evaluation of Large Language Models

    Zhongxi Wang, Yueqian Lin, Jingyang Zhang +2

    cs.LGcs.CLcs.CVarXiv:2603.02482v12026
  16. Expectation, Backlash, Recovery, and Excitement: How Model Releases Shape Reddit Perceptions of Conversational AI Systems

    Vahid Rahimzadeh, Yury Zhauniarovich, Savvas Zannettou

    cs.CLcs.CYarXiv:2608.24654v12026
  17. HateMirage: An Explainable Multi-Dimensional Dataset for Decoding Faux Hate and Subtle Online Abuse

    Sai Kartheek Reddy Kasu, Shankar Biradar, Sunil Saumya +1

    cs.CLcs.SIarXiv:2603.02684v12026
  18. APRES: An Agentic Paper Revision and Evaluation System

    Bingchen Zhao, Jenny Zhang, Chenxi Whitehouse +8

    cs.CLcs.AIarXiv:2603.03142v12026
  19. Simple and Accurate Dependency Parsing Using Bidirectional LSTM Feature Representations

    Eliyahu Kiperwasser, Yoav Goldberg

    cs.CLarXiv:1603.04351v32016
  20. Entity, Relation, and Event Extraction with Contextualized Span Representations

    David Wadden, Ulme Wennberg, Yi Luan +1

    cs.CLarXiv:1909.03546v22019
  21. UTS at CheckThat! 2026: Cite-Frame Engineering for Generated Fact-Checking Articles

    Dima Galat, Marian-Andrei Rizoiu

    cs.DLcs.CLarXiv:2608.24466v12026
  22. Self-Rewarding Language Models

    Weizhe Yuan, Richard Yuanzhe Pang, Kyunghyun Cho +4

    cs.CLcs.AIarXiv:2401.10020v32024
  23. Bolbosh: Script-Aware Flow Matching for Kashmiri Text-to-Speech

    Tajamul Ashraf, Burhaan Rasheed Zargar, Saeed Abdul Muizz +5

    cs.CLarXiv:2603.07513v12026
  24. Training Large Language Models to Reason in a Continuous Latent Space

    Shibo Hao, Sainbayar Sukhbaatar, DiJia Su +4

    cs.CLarXiv:2412.06769v42024
  25. On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes

    Rishabh Agarwal, Nino Vieillard, Yongchao Zhou +4

    cs.LGcs.AIcs.CLarXiv:2306.13649v32023
  26. Constraint-Guided Enterprise Data Mapping with Large Language Models

    Sebastian Monka, Pramod Anantharam, Thien Vo Minh +1

    cs.AIcs.CLarXiv:2608.24218v12026
  27. Breaking Training Bottlenecks: Effective and Stable Reinforcement Learning for Coding Models

    Zongqian Li, Shaohan Huang, Zewen Chi +5

    cs.LGcs.CLcs.GLarXiv:2603.07777v12026
  28. Scaling Data Difficulty: Improving Coding Models via Reinforcement Learning on Fresh and Challenging Problems

    Zongqian Li, Tengchao Lv, Shaohan Huang +8

    cs.CLcs.GLcs.LGarXiv:2603.07779v12026
  29. Classifying Relations via Long Short Term Memory Networks along Shortest Dependency Path

    Xu Yan, Lili Mou, Ge Li +3

    cs.CLcs.LGarXiv:1508.03720v12015
  30. ConFu: Contemplate the Future for Better Speculative Sampling

    Zongyue Qin, Raghavv Goel, Mukul Gagrani +3

    cs.CLcs.LGarXiv:2603.08899v32026
  31. Qwen2-Audio Technical Report

    Yunfei Chu, Jin Xu, Qian Yang +9

    eess.AScs.CLcs.LGarXiv:2407.10759v12024
  32. Truth as a Compression Artifact in Language Model Training

    Konstantin Krestnikov

    cs.CLcs.AIarXiv:2603.11749v32026
  33. Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge

    Junjie Wu, Xuan Kan, Zihao He +3

    cs.CLarXiv:2603.11665v22026
  34. MHPO: Modulated Hazard-aware Policy Optimization for Stable Reinforcement Learning

    Hongjun Wang, Wei Liu, Weibo Gu +2

    cs.LGcs.AIcs.CLarXiv:2603.16929v22026
  35. FewRel: A Large-Scale Supervised Few-Shot Relation Classification Dataset with State-of-the-Art Evaluation

    Xu Han, Hao Zhu, Pengfei Yu +4

    cs.LGcs.AIcs.CLarXiv:1810.10147v22018
  36. COMIC: Agentic Sketch Comedy Generation

    Susung Hong, Brian Curless, Ira Kemelmacher-Shlizerman +1

    cs.CVcs.AIcs.CLarXiv:2603.11048v12026
  37. SimulU: Training-free Policy for Long-form Simultaneous Speech-to-Speech Translation

    Amirbek Djanibekov, Luisa Bentivogli, Matteo Negri +1

    eess.AScs.AIcs.CLarXiv:2603.16924v12026
  38. AI Chains: Transparent and Controllable Human-AI Interaction by Chaining Large Language Model Prompts

    Tongshuang Wu, Michael Terry, Carrie J. Cai

    cs.HCcs.CLarXiv:2110.01691v32021
  39. On the Automatic Generation of Medical Imaging Reports

    Baoyu Jing, Pengtao Xie, Eric Xing

    cs.CLcs.CVarXiv:1711.08195v32017
  40. DialogueGCN: A Graph Convolutional Neural Network for Emotion Recognition in Conversation

    Deepanway Ghosal, Navonil Majumder, Soujanya Poria +2

    cs.CLcs.LGarXiv:1908.11540v12019
  41. How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs

    Yi Zeng, Hongpeng Lin, Jingwen Zhang +3

    cs.CLcs.AIarXiv:2401.06373v22024
  42. Decoupled Physical Modeling and Execution for Physics Reasoning

    Ye Zhang, Xuehang Guo, Rui Pan +4

    cs.LGcs.CLarXiv:2608.22126v12026
  43. Probing Cultural Signals in Large Language Models through Author Profiling

    Valentin Lafargue, Ariel Guerra-Adames, Emmanuelle Claeys +2

    cs.CLcs.LGarXiv:2603.16749v22026
  44. A Stylometric Inquiry into Hyperpartisan and Fake News

    Martin Potthast, Johannes Kiesel, Kevin Reinartz +2

    cs.CLarXiv:1702.05638v12017
  45. SentEval: An Evaluation Toolkit for Universal Sentence Representations

    Alexis Conneau, Douwe Kiela

    cs.CLarXiv:1803.05449v12018
  46. sebis at ArchEHR-QA 2026: How Much Can You Do Locally? Evaluating Grounded EHR QA on a Single Notebook

    Ibrahim Ebrar Yurt, Fabian Karl, Tejaswi Choppa +1

    cs.CLarXiv:2603.13962v22026
  47. Mind the Shift: Decoding Monetary Policy Stance from FOMC Statements with Large Language Models

    Yixuan Tang, Yi Yang

    cs.CLarXiv:2603.14313v12026
  48. Polyglot-Lion: Efficient Multilingual ASR for Singapore via Balanced Fine-Tuning of Qwen3-ASR

    Quy-Anh Dang, Chris Ngo

    cs.CLarXiv:2603.16184v12026
  49. When AI Navigates the Fog of War

    Ming Li, Xirui Li, Tianyi Zhou

    cs.AIcs.CLcs.CYarXiv:2603.16642v12026
  50. DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models

    Boxin Wang, Weixin Chen, Hengzhi Pei +16

    cs.CLcs.AIcs.CRarXiv:2306.11698v52023
  51. Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset

    Ke Wang, Junting Pan, Weikang Shi +3

    cs.CVcs.AIcs.CLarXiv:2402.14804v12024
  52. Dynamic Coattention Networks For Question Answering

    Caiming Xiong, Victor Zhong, Richard Socher

    cs.CLcs.AIarXiv:1611.01604v42016
  53. Discrete Diffusion Modeling by Estimating the Ratios of the Data Distribution

    Aaron Lou, Chenlin Meng, Stefano Ermon

    stat.MLcs.CLcs.LGarXiv:2310.16834v32023
  54. Hungry Hungry Hippos: Towards Language Modeling with State Space Models

    Daniel Y. Fu, Tri Dao, Khaled K. Saab +3

    cs.LGcs.CLarXiv:2212.14052v32022
  55. BLADE: Bilevel Low-rank Augmented-Lagrangian Erasure for LLM Unlearning

    Md Toufikuzzaman, Ahmad Mousavi, Dongwon Lee

    cs.LGcs.AIcs.CLarXiv:2608.22557v12026
  56. Benchmarking Zero-shot Text Classification: Datasets, Evaluation and Entailment Approach

    Wenpeng Yin, Jamaal Hay, Dan Roth

    cs.CLarXiv:1909.00161v12019
  57. Progressive Training for Explainable Citation-Grounded Dialogue: Reducing Hallucination to Zero in English-Hindi LLMs

    Vedant Pandya

    cs.CLcs.AIarXiv:2603.18911v12026
  58. BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

    Sheng Zhang, Yanbo Xu, Naoto Usuyama +21

    cs.CVcs.CLarXiv:2303.00915v32023
  59. What Really Controls Temporal Reasoning in Large Language Models: Tokenisation or Representation of Time?

    Gagan Bhatia, Ahmad Muhammad Isa, Maxime Peyrard +1

    cs.CLcs.AIarXiv:2603.19017v12026
  60. RAG Collapse: LLM Responses Collapse When Retrieved Documents Are Self-Authored

    Gregory Druck, Ethan Smith

    cs.CLcs.IRarXiv:2608.22118v12026