Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
9,661 to 9,720 of 11,270
SAC-Copula: Quality-Preserving Watermarking for Diffusion Language Models via Smooth Correlated Gumbel Fields
Baixin Li, Haiyun He
cs.CLcs.CRcs.LGarXiv:2608.20839v12026Ontology-Driven Structural Regularization for Document-Level Relation Extraction
Laura Menotti, Stefano Marchesin, Gianmaria Silvello
cs.CLarXiv:2608.20856v12026AsmEvo: Agentic Assembly-Level Optimization of AMD GPU Kernels with Functional Equivalence Verification
Ji Liu, Puyuan Yang, Rongzhang Zheng +18
cs.CLarXiv:2608.20711v12026MIL-BERT: Classification of Arbitrarily Large Text with Performance and Explanatory Guarantees
John Cadigan, Dayne Freitag, Eric Yeh
cs.CLcs.LGarXiv:2608.20636v12026Sparse Token Routing in Efficient Transformers
Sai Krishna Arthanari, JaeHyeong Chang, Chengzhe Sun +1
cs.CLarXiv:2608.20632v12026LiLiCorr: Lightweight Likelihood Correlation of Parallel Drafts for Speculative Decoding
Matan Rusanovsky, Yoav Miron, Roy Uziel +4
cs.CLarXiv:2608.20530v12026ImmigrationReason: A Structured Dataset of U.S. Immigration Appeals for Legal Reasoning Research
Amirhossein Afsharrad, Seyed Shahabeddin Mousavi
cs.CLarXiv:2608.20391v12026Self-Supervised Speech Representations Track Spoken Language Convergence to Adult Models in Infants and Children Who Are Deaf/Hard-of-Hearing
L. Choy, A. S. Khan, S. Patrizi +3
cs.CLcs.SDarXiv:2608.20396v12026ARGUS: Theory-of-Mind Guided Argument Generation with Strategy-Aware Planning and Knowledge Grounding
Zhe Hu
cs.CLarXiv:2608.20405v12026Using Human-LLM Disagreement to Improve Checklist-Based Quality Appraisal
Timo van der Kuil, Bruno Messina Coimbra, Mirjam van Zuiden +7
cs.CLarXiv:2608.20385v12026Research Paper Quality Recognition Through Textual Feature Analysis
Saikiran Korla, Sadwik Gummadavelli, Trung-Nghia Le +2
cs.CLarXiv:2608.20368v12026Multilingual Verifier Bias in RLVR: Benchmark, Rollout Diagnosis, and the Cross-Lingual Selection Bottleneck
Chenyu Zhou, Qiliang Jiang, Xu Zhou
cs.CLcs.LGarXiv:2608.20362v12026Self-Speculation for Faster Reasoning Models
Ravisri Valluri, Tung Nguyen, Aditya Grover
cs.CLarXiv:2608.20359v12026TriPLU: Bypassing the Gate with Direct Trilinear Product FFNs in Tiny Language Models
He Zhang
cs.CLcs.LGarXiv:2608.20360v12026TurboBias 2.0: Streaming Context-Biasing for Production-Efficient ASR Systems
Vladimir Bataev, Lilit Grigoryan, Andrei Andrusenko +3
eess.AScs.AIcs.CLarXiv:2608.21343v12026Target-Aware Calibration Data Selection for Preserving Uncertainty in Quantized Language Models
Zhen Yang, Sizai Hou, Kaiwen Zheng +4
cs.CLcs.AIarXiv:2608.21019v12026EnSI-RAG: Entity-Structure-Indexed Retrieval-Augmented Generation for Long-Document Question Answering
Xuanyu Meng, Jiashuo Sun, Jash Rajesh Parekh +1
cs.CLcs.AIcs.DBarXiv:2608.21252v12026PromptResponse: Optimizing Prompts for LLM Coding Tasks
Erik Thureck, Robert Kühnen, Tim Jacobowitz
cs.CLcs.AIcs.HCarXiv:2608.21074v12026Free-Text Evaluation of LLMs for 5G Domain Knowledge and Fault Analysis using LLM-as-Judge
Rishiraj Sengupta, Sotiris Chatzimiltis, Mohammad Shojafar +1
cs.CLcs.AIcs.NIarXiv:2608.21021v12026Quantization-Aware Healing: A Practical Recipe for Recovering Compressed, 4-Bit LLMs
Bakbergen Ryskulov, Iker García-Ferrero, David Montero +5
cs.CLcs.AIcs.LGarXiv:2608.20953v12026Extractive Summarization for Arabic Documents Using SAraBERT with a Semantic Siamese Similarity Evaluation Metric
Sami Shames El Deen, Mariette Awad
cs.CLcs.AIarXiv:2608.20964v12026PSK at WMT 2026 MIST: Task-Specialized QLoRA Adapters for Multilingual Summarization and Question Answering
Srikar Kashyap Pulipaka
cs.CLcs.AIcs.LGarXiv:2608.20757v12026Temporal Validity on Real Software Histories: Eliminating Stale-Fact Errors in Code-Assistant Memory over GitHub Fixes
Neeraj Yadav
cs.SEcs.AIcs.CLarXiv:2608.20685v12026No PUN Intended: Plausible Unknown Names for Person-Centred LLM Evaluation
Dimitri Staufer, David Hartmann, Ibrahim Baroud
cs.CLcs.AIcs.LGarXiv:2608.21206v12026Trustworthy RAG: An Evaluation Agent for Detecting Misinformation and Knowledge Poisoning in Generative AI Systems
Balkrishna Giri, Md Toufique Hasan, Jussi Rasku +2
cs.SEcs.AIcs.CLarXiv:2608.21095v12026GAIA: a benchmark for General AI Assistants
Grégoire Mialon, Clémentine Fourrier, Craig Swift +3
cs.CLcs.AIarXiv:2311.12983v12023Source-Free MT Evaluation Is Not MT Evaluation
Baban Gain, Ramakrishna Appicharla, Asif Ekbal
cs.CLcs.AIarXiv:2608.20925v12026KREL: Automatic Medical Coding via Knowledge-Guided Reasoning over Clinical Evidence with LLMs
Xubin Chen, Yipeng Zhou, Wen Sun +3
cs.CLcs.AIarXiv:2608.20887v12026STAR-OPD: Structured Aspect-Cascade-Aware On-Policy Reward Distillation for ABSA Quadruple Extraction
Tong Sun, Mingyang Ma, Jiayang Yu
cs.CLcs.AIarXiv:2608.20831v12026Denoising the Future: Context-Aware Spectral Diffusion for Temporal Knowledge Graph Extrapolation
Yanglei Gan, Peng He, Run Lin +3
cs.CLcs.AIarXiv:2608.20804v12026Profiling What Matters: Context-Aware Item Profiles from Large-Scale Metadata for LLM Recommenders
Dojun Hwang, Seunghan Lee, Cheonyoung Park +2
cs.IRcs.AIcs.CLarXiv:2608.20801v12026Beyond English-Centric Multilingual Machine Translation
Angela Fan, Shruti Bhosale, Holger Schwenk +14
cs.CLcs.LGarXiv:2010.11125v12020JuryProbe: An Empirical Consensus-Risk Diagnostic for Routing Reference-Free Factuality Judge Panels to Grounded Verification
Tianxin Zhou, Ruixi Lin
cs.CLcs.AIcs.LGarXiv:2608.20607v12026Word Embeddings Quantify 100 Years of Gender and Ethnic Stereotypes
Nikhil Garg, Londa Schiebinger, Dan Jurafsky +1
cs.CLcs.CYarXiv:1711.08412v12017Poly-InstructTTS: Learning In-the-Wild Expressive Speech Synthesis from Open-Ended Instructions
Junhui Zhang, Qianhui Xu, Qingxiang Guo +4
cs.CLcs.AIarXiv:2608.20387v12026Trilingual Topic Modeling of Sri Lankan Parliamentary Debates
Himath Dhanapala, Haren Daishika, Himandhi Kuruppu +6
cs.CLcs.AIarXiv:2608.20365v12026When Do LLMs Replace Fine-Tuned NLU? A Decision Framework for Intent Detection in Production Conversational Systems
Carson Rodrigues, Oysturn Vas
cs.CLcs.AIarXiv:2608.20371v12026When Failures Propagate: Causal Failure Attribution in Agentic Retrieval-Augmented Generation
Lauren Pothuru
cs.CLcs.AIarXiv:2608.20627v12026A Persona-Based Neural Conversation Model
Jiwei Li, Michel Galley, Chris Brockett +3
cs.CLarXiv:1603.06155v22016Describing Videos by Exploiting Temporal Structure
Li Yao, Atousa Torabi, Kyunghyun Cho +4
stat.MLcs.AIcs.CLarXiv:1502.08029v52015ProofJudge: Tool-Grounded LLM Evaluation of Formal Proof Quality in Mathlib
Shane Caldwell
cs.LOcs.AIcs.CLarXiv:2608.20432v12026LingShu: A Large-Scale Symptom-Centric Contextualized Knowledge Graph Bridging Traditional Chinese Medicine and Modern Biomedicine
Rui Hua, Zixin Shu, Kai Chang +18
cs.CLcs.AIarXiv:2608.20402v12026Knowledge-Graph-Gated Defactualization for Style-Controllable and Fact-Preserving Generation in Agentic Conversational AI
Tanmay Kumar Shrivastava, Darsh Rohit Nandu, Rajesh Kumar Mundotiya
cs.CLcs.AIarXiv:2608.20393v12026Evaluation-as-Search: Adaptive Discovery of Grounding Failures in Meeting Assistants
Sami Khairy, Yasaman Hosseinkashi, Vishak Gopal +1
cs.CLcs.AIarXiv:2608.20392v12026EditPPT: Faithful Long-Deck Slide Editing via Structured Tool-Using Multi-Agent with Dual-Modal Validators
Jiheon Kim, Kyudan Jung, Jaegul Choo
cs.CLcs.AIcs.HCarXiv:2608.20381v12026Ansari: A Retrieval-Grounded Islamic AI Assistant -- Architecture, Deployment, and Lessons from 140,000 Conversations
M Waleed Kadous, Amr Elsayed, Abdullah Al Nahas +1
cs.CLcs.AIcs.CYarXiv:2608.20390v12026ASTAR: Automated induction of STAndardized radiology Reporting templates from large-scale clinical free-text corpora
Xinfeng Zhang, Mingxuan Liu, Yifei Chen +9
cs.CLcs.AIarXiv:2608.20369v12026VA-DPO: Valence-Arousal Direct Preference Optimization for Controllable Emotion Generation in Language Models
Hyunwoo Kim
cs.CLcs.AIcs.LGarXiv:2608.20374v12026Let's Scale Step by Step: Compute-Efficient Hyperparameter Transfer for Large-Scale Mixture-of-Experts
Nayeon Kim, Hojin Lee, Yunju Bak +2
cs.LGcs.AIcs.CLarXiv:2608.20061v12026BARTScore: Evaluating Generated Text as Text Generation
Weizhe Yuan, Graham Neubig, Pengfei Liu
cs.CLarXiv:2106.11520v22021ExpertIVS: Sociological Expert Driven Individual Value Simulation in Large Language Models
Zhen Wang, Yuqi Ren, Yuehan Cui +7
cs.CLcs.AIarXiv:2608.20355v12026The Divergence Hypothesis: Unmasking Lexical Interference and Label Bias in Mental Health NLP
Moustafa Yehia Hassan
cs.CLcs.AIarXiv:2608.20353v12026Beyond Prompt Engineering: A Systematic Analysis of Prompt Lexical Sensitivity and Its Impacts on Quality
Qipeng Xie, Zi Liang, Jiafei Wu +6
cs.CLcs.AIarXiv:2608.20349v12026Inhibitory Attention for Clinical Long-Context Reasoning: Characterizing and Mitigating Lost-in-the-Middle Effects in EHR Processing
Sanjay Basu
cs.CLcs.AIarXiv:2608.20348v12026When Vocabulary Comprehension Fails Clinical Reasoning: Evaluating Therapy Bots' Safety Risks for Generation Alpha
Manisha Mehta, Virendra Mehta
cs.CLcs.AIcs.CYarXiv:2608.20345v12026Who Do Language Models Think Is Competent? A Mechanistic Analysis of Occupational Bias
Keren Fuentes, Aaron Mueller
cs.CLcs.AIcs.CYarXiv:2608.20347v12026Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP Tasks
Yizhong Wang, Swaroop Mishra, Pegah Alipoormolabashi +37
cs.CLcs.AIarXiv:2204.07705v32022Personalized Privacy Control in LLMs via Attention Head Intervention
Junseok Kim, Nakyeong Yang, Kyomin Jung
cs.AIcs.CLcs.LGarXiv:2608.21209v12026Enhancing LLMs in Predictive Political QA with Semi-Structured Data
Yinan Liu, Zihan Zhou, Zichun Jin +3
cs.AIcs.CLcs.IRarXiv:2608.21218v12026Meshed-Memory Transformer for Image Captioning
Marcella Cornia, Matteo Stefanini, Lorenzo Baraldi +1
cs.CVcs.CLarXiv:1912.08226v22019