Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,341 to 8,400 of 11,248
Contextual Augmentation: Data Augmentation by Words with Paradigmatic Relations
Sosuke Kobayashi
cs.CLcs.LGarXiv:1805.06201v12018Yor-Sarc: A gold-standard dataset for sarcasm detection in a low-resource African language
Toheeb Aduramomi Jimoh, Tabea De Wille, Nikola S. Nikolov
cs.CLarXiv:2602.18964v12026Small Language Models for Privacy-Preserving Clinical Information Extraction in Low-Resource Languages
Mohammadreza Ghaffarzadeh-Esfahani, Nahid Yousefian, Ebrahim Heidari-Farsani +4
cs.CLcs.AIcs.LGarXiv:2602.21374v12026Humans and LLMs Diverge on Probabilistic Inferences
Gaurav Kamath, Sreenath Madathil, Sebastian Schuster +2
cs.CLcs.AIarXiv:2602.23546v12026Q-BERT: Hessian Based Ultra Low Precision Quantization of BERT
Sheng Shen, Zhen Dong, Jiayu Ye +5
cs.CLcs.LGarXiv:1909.05840v22019This Just In: Fake News Packs a Lot in Title, Uses Simpler, Repetitive Content in Text Body, More Similar to Satire than Real News
Benjamin D. Horne, Sibel Adali
cs.SIcs.CLarXiv:1703.09398v12017MOSI: Multimodal Corpus of Sentiment Intensity and Subjectivity Analysis in Online Opinion Videos
Amir Zadeh, Rowan Zellers, Eli Pincus +1
cs.CLcs.MMarXiv:1606.06259v22016The RAT: A Unified Bayesian Model for RAG Evaluation
Pius von Däniken, Felix Matthias Saaro, Mark Cieliebak +1
cs.CLcs.AIarXiv:2608.24753v12026Legal RAG Bench: an end-to-end benchmark for legal RAG
Abdur-Rahman Butler, Umar Butler
cs.CLcs.IRcs.LGarXiv:2603.01710v12026Who is the Agent to Blame? Localizing Faithfulness and Citation Mistakes in Agentic Deep Research
Eran Hirsch, David Wan, Han Wang +3
cs.CLarXiv:2608.24306v12026Surgical Post-Training: Proximal On-Policy Distillation for Reasoning with Knowledge Retention
Wenye Lin, Kai Han
cs.CLcs.AIarXiv:2603.01683v22026OpenAutoNLU: Open Source AutoML Library for NLU
Grigory Arshinov, Aleksandr Boriskin, Sergey Senichev +4
cs.CLcs.LGarXiv:2603.01824v12026CharacterFlywheel: Scaling Iterative Improvement of Engaging and Steerable LLMs in Production
Yixin Nie, Lin Guan, Zhongyao Ma +19
cs.CLcs.AIcs.SIarXiv:2603.01973v12026Linear Probing Provides Robust and Efficient Detection of Machine-Generated Text
Gerrit Quaremba, Hanqi Yan, Elizabeth Black +2
cs.CLarXiv:2608.24780v12026MUSE: A Run-Centric Platform for Multimodal Unified Safety Evaluation of Large Language Models
Zhongxi Wang, Yueqian Lin, Jingyang Zhang +2
cs.LGcs.CLcs.CVarXiv:2603.02482v12026Expectation, Backlash, Recovery, and Excitement: How Model Releases Shape Reddit Perceptions of Conversational AI Systems
Vahid Rahimzadeh, Yury Zhauniarovich, Savvas Zannettou
cs.CLcs.CYarXiv:2608.24654v12026HateMirage: An Explainable Multi-Dimensional Dataset for Decoding Faux Hate and Subtle Online Abuse
Sai Kartheek Reddy Kasu, Shankar Biradar, Sunil Saumya +1
cs.CLcs.SIarXiv:2603.02684v12026APRES: An Agentic Paper Revision and Evaluation System
Bingchen Zhao, Jenny Zhang, Chenxi Whitehouse +8
cs.CLcs.AIarXiv:2603.03142v12026Simple and Accurate Dependency Parsing Using Bidirectional LSTM Feature Representations
Eliyahu Kiperwasser, Yoav Goldberg
cs.CLarXiv:1603.04351v32016Entity, Relation, and Event Extraction with Contextualized Span Representations
David Wadden, Ulme Wennberg, Yi Luan +1
cs.CLarXiv:1909.03546v22019UTS at CheckThat! 2026: Cite-Frame Engineering for Generated Fact-Checking Articles
Dima Galat, Marian-Andrei Rizoiu
cs.DLcs.CLarXiv:2608.24466v12026Self-Rewarding Language Models
Weizhe Yuan, Richard Yuanzhe Pang, Kyunghyun Cho +4
cs.CLcs.AIarXiv:2401.10020v32024Bolbosh: Script-Aware Flow Matching for Kashmiri Text-to-Speech
Tajamul Ashraf, Burhaan Rasheed Zargar, Saeed Abdul Muizz +5
cs.CLarXiv:2603.07513v12026Training Large Language Models to Reason in a Continuous Latent Space
Shibo Hao, Sainbayar Sukhbaatar, DiJia Su +4
cs.CLarXiv:2412.06769v42024On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes
Rishabh Agarwal, Nino Vieillard, Yongchao Zhou +4
cs.LGcs.AIcs.CLarXiv:2306.13649v32023Constraint-Guided Enterprise Data Mapping with Large Language Models
Sebastian Monka, Pramod Anantharam, Thien Vo Minh +1
cs.AIcs.CLarXiv:2608.24218v12026Breaking Training Bottlenecks: Effective and Stable Reinforcement Learning for Coding Models
Zongqian Li, Shaohan Huang, Zewen Chi +5
cs.LGcs.CLcs.GLarXiv:2603.07777v12026Scaling Data Difficulty: Improving Coding Models via Reinforcement Learning on Fresh and Challenging Problems
Zongqian Li, Tengchao Lv, Shaohan Huang +8
cs.CLcs.GLcs.LGarXiv:2603.07779v12026Classifying Relations via Long Short Term Memory Networks along Shortest Dependency Path
Xu Yan, Lili Mou, Ge Li +3
cs.CLcs.LGarXiv:1508.03720v12015ConFu: Contemplate the Future for Better Speculative Sampling
Zongyue Qin, Raghavv Goel, Mukul Gagrani +3
cs.CLcs.LGarXiv:2603.08899v32026Qwen2-Audio Technical Report
Yunfei Chu, Jin Xu, Qian Yang +9
eess.AScs.CLcs.LGarXiv:2407.10759v12024Truth as a Compression Artifact in Language Model Training
Konstantin Krestnikov
cs.CLcs.AIarXiv:2603.11749v32026Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge
Junjie Wu, Xuan Kan, Zihao He +3
cs.CLarXiv:2603.11665v22026MHPO: Modulated Hazard-aware Policy Optimization for Stable Reinforcement Learning
Hongjun Wang, Wei Liu, Weibo Gu +2
cs.LGcs.AIcs.CLarXiv:2603.16929v22026FewRel: A Large-Scale Supervised Few-Shot Relation Classification Dataset with State-of-the-Art Evaluation
Xu Han, Hao Zhu, Pengfei Yu +4
cs.LGcs.AIcs.CLarXiv:1810.10147v22018COMIC: Agentic Sketch Comedy Generation
Susung Hong, Brian Curless, Ira Kemelmacher-Shlizerman +1
cs.CVcs.AIcs.CLarXiv:2603.11048v12026SimulU: Training-free Policy for Long-form Simultaneous Speech-to-Speech Translation
Amirbek Djanibekov, Luisa Bentivogli, Matteo Negri +1
eess.AScs.AIcs.CLarXiv:2603.16924v12026AI Chains: Transparent and Controllable Human-AI Interaction by Chaining Large Language Model Prompts
Tongshuang Wu, Michael Terry, Carrie J. Cai
cs.HCcs.CLarXiv:2110.01691v32021On the Automatic Generation of Medical Imaging Reports
Baoyu Jing, Pengtao Xie, Eric Xing
cs.CLcs.CVarXiv:1711.08195v32017DialogueGCN: A Graph Convolutional Neural Network for Emotion Recognition in Conversation
Deepanway Ghosal, Navonil Majumder, Soujanya Poria +2
cs.CLcs.LGarXiv:1908.11540v12019How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs
Yi Zeng, Hongpeng Lin, Jingwen Zhang +3
cs.CLcs.AIarXiv:2401.06373v22024Decoupled Physical Modeling and Execution for Physics Reasoning
Ye Zhang, Xuehang Guo, Rui Pan +4
cs.LGcs.CLarXiv:2608.22126v12026Probing Cultural Signals in Large Language Models through Author Profiling
Valentin Lafargue, Ariel Guerra-Adames, Emmanuelle Claeys +2
cs.CLcs.LGarXiv:2603.16749v22026A Stylometric Inquiry into Hyperpartisan and Fake News
Martin Potthast, Johannes Kiesel, Kevin Reinartz +2
cs.CLarXiv:1702.05638v12017SentEval: An Evaluation Toolkit for Universal Sentence Representations
Alexis Conneau, Douwe Kiela
cs.CLarXiv:1803.05449v12018sebis at ArchEHR-QA 2026: How Much Can You Do Locally? Evaluating Grounded EHR QA on a Single Notebook
Ibrahim Ebrar Yurt, Fabian Karl, Tejaswi Choppa +1
cs.CLarXiv:2603.13962v22026Mind the Shift: Decoding Monetary Policy Stance from FOMC Statements with Large Language Models
Yixuan Tang, Yi Yang
cs.CLarXiv:2603.14313v12026Polyglot-Lion: Efficient Multilingual ASR for Singapore via Balanced Fine-Tuning of Qwen3-ASR
Quy-Anh Dang, Chris Ngo
cs.CLarXiv:2603.16184v12026When AI Navigates the Fog of War
Ming Li, Xirui Li, Tianyi Zhou
cs.AIcs.CLcs.CYarXiv:2603.16642v12026DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models
Boxin Wang, Weixin Chen, Hengzhi Pei +16
cs.CLcs.AIcs.CRarXiv:2306.11698v52023Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset
Ke Wang, Junting Pan, Weikang Shi +3
cs.CVcs.AIcs.CLarXiv:2402.14804v12024Dynamic Coattention Networks For Question Answering
Caiming Xiong, Victor Zhong, Richard Socher
cs.CLcs.AIarXiv:1611.01604v42016Discrete Diffusion Modeling by Estimating the Ratios of the Data Distribution
Aaron Lou, Chenlin Meng, Stefano Ermon
stat.MLcs.CLcs.LGarXiv:2310.16834v32023Hungry Hungry Hippos: Towards Language Modeling with State Space Models
Daniel Y. Fu, Tri Dao, Khaled K. Saab +3
cs.LGcs.CLarXiv:2212.14052v32022BLADE: Bilevel Low-rank Augmented-Lagrangian Erasure for LLM Unlearning
Md Toufikuzzaman, Ahmad Mousavi, Dongwon Lee
cs.LGcs.AIcs.CLarXiv:2608.22557v12026Benchmarking Zero-shot Text Classification: Datasets, Evaluation and Entailment Approach
Wenpeng Yin, Jamaal Hay, Dan Roth
cs.CLarXiv:1909.00161v12019Progressive Training for Explainable Citation-Grounded Dialogue: Reducing Hallucination to Zero in English-Hindi LLMs
Vedant Pandya
cs.CLcs.AIarXiv:2603.18911v12026BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs
Sheng Zhang, Yanbo Xu, Naoto Usuyama +21
cs.CVcs.CLarXiv:2303.00915v32023What Really Controls Temporal Reasoning in Large Language Models: Tokenisation or Representation of Time?
Gagan Bhatia, Ahmad Muhammad Isa, Maxime Peyrard +1
cs.CLcs.AIarXiv:2603.19017v12026RAG Collapse: LLM Responses Collapse When Retrieved Documents Are Self-Authored
Gregory Druck, Ethan Smith
cs.CLcs.IRarXiv:2608.22118v12026