Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
9,601 to 9,660 of 11,241
How Do Large Language Models Learn Concepts During Continual Pre-Training?
Barry Menglong Yao, Sha Li, Yunzhi Yao +4
cs.CLarXiv:2601.03570v22026Hybrid Policy Distillation for LLMs
Wenhong Zhu, Ruobing Xie, Rui Wang +1
cs.CLcs.AIarXiv:2604.20244v22026Building a Precise Video Language with Human-AI Oversight
Zhiqiu Lin, Chancharik Mitra, Siyuan Cen +13
cs.CVcs.AIcs.CLarXiv:2604.21718v22026Explainable Disentangled Representation Learning for Generalizable Authorship Attribution in the Era of Generative AI
Hieu Man, Van-Cuong Pham, Nghia Trung Ngo +2
cs.CLcs.IRcs.LGarXiv:2604.21300v12026SLIDERS: Systematic Reviews via Automated Evidence Synthesis and Reconciliation
Harshit Joshi, Priyank Shethia, Jadelynn Dao +1
cs.CLcs.AIarXiv:2604.22294v22026The Geometric Canary: Predicting Steerability and Detecting Drift via Representational Stability
Prashant C. Raju
cs.LGcs.CLstat.MLarXiv:2604.17698v22026Sparse Reward Subsystem in Large Language Models
Guowei Xu, Mert Yuksekgonul, James Zou
cs.CLarXiv:2602.00986v22026Future-KL Regularized GRPO: Process-Level Credit Assignment from $f$-Divergence Regularization
Jiarui Yao, Ruida Wang, Hao Bai +1
cs.LGcs.AIcs.CLarXiv:2601.10201v22026PRISM-$Δ$: Differential Subspace Steering for Prompt Highlighting in Large Language Models
Yuyao Ge, Shenghua Liu, Yiwei Wang +5
cs.CLarXiv:2603.10705v22026XSkill: Continual Learning from Experience and Skills in Multimodal Agents
Guanyu Jiang, Zhaochen Su, Xiaoye Qu +1
cs.AIcs.CLarXiv:2603.12056v32026Geometric Stability: The Missing Axis of Representations
Prashant C. Raju
cs.LGcs.CLq-bio.QMarXiv:2601.09173v52026Decoupled Vision-Language System for Multimodal Understanding and Generation
Yifan Xu, Baochen Xiong, Xiaoshan Yang +3
cs.CLcs.CVarXiv:2608.20382v12026COMET: Contrastive Motion-Enhanced Temporal Reasoning for Video Multimodal Large Language Models
Chenghua Zhu, Zhaolu Kang, Qifan Shi +8
cs.CVcs.CLcs.LGarXiv:2608.21030v12026MigrationNarrate: A Dataset for Detection of Migration Narratives in YouTube Videos
Fatima Haouari, Carolina Scarton, Kalina Bontcheva
cs.CVcs.CLcs.CYarXiv:2608.20984v12026Prompt-Model Interaction Reaches the Fixed Points: A deterministic, task-free structural readout -- and the factorizations of it that failed
Nicolás Vera Zúñiga
cs.CLarXiv:2608.21315v12026Tree-of-Concerns: Hierarchical Multi-Agent Debate for Unstated-Limitation Extraction in Scientific Critique
Sahil Mishra, Niranjan Rajeev, Tanmoy Chakraborty
cs.CLarXiv:2608.20777v12026When the Feature Pool Goes Algorithmic: Extending Mufwene's Ecology of Language Evolution to LLM-Mediated Exposure
Kunmei Han
cs.CLarXiv:2608.21088v12026GRAFT: Adaptive DLM-Based Draft Tree Construction with Target-Distilled Edge Scoring
Xuming Ye, Zeming Ma, Runjie Yu +5
cs.CLarXiv:2608.20375v12026An ambiguity taxonomy for evaluating large language model performance on clinical registry abstraction: a multi-site prospective study
James Matheson, Betsy Castillo, Andrew Y. Shin +1
cs.CLarXiv:2608.20373v12026Beyond Raw Transcripts: Structured Persona Extraction for LLM-Based Digital Twins
Iris Ye, Tianze Deng, Ozan Candogan
cs.CLcs.CYarXiv:2608.20344v12026Building and Evaluating a Synthetic Bengali Speech Resource for Telecom Customer Care
Kawshik Kumar Paul, Md. Nafiul Alam Fuji
cs.CLcs.SDeess.ASarXiv:2608.20346v12026Exploratory As-Analyzed No-Detection of Culturally-Marked Predicate-Triggered PII Amplification in a Synthetic-English RAG Probe: A Predicate-Resource-Confounded Audit
Yanhang Li, Zhichao Fan, Zexin Zhuang
cs.CLcs.CYcs.LGarXiv:2608.20351v12026Identify, Locate, Link: End-to-End Key-Value Extraction from Document Images
A. Said Gurbuz, Ahmed Nassar, Christoph Auer +8
cs.CVcs.CLarXiv:2608.20868v12026A Factorial Ablation of a Speech-to-SFT Pipeline: Differential Effects on Data Quality and Downstream Transfer
Wonsup Shin, Jingu Kim
cs.SDcs.CLarXiv:2608.20394v12026Move by Move: Measuring and Steering How LLMs Conduct Psychotherapy
Afonso Baldo, Hugo Pitorro, Areti Vassilopoulos +5
cs.CLarXiv:2608.21325v12026Memory Augmentation Unlocks Efficient Chain-of-Thought Reasoning
Simeng Zhang, Yilong Chen, Wenyuan Zhang +4
cs.CLarXiv:2608.21265v12026Benchmarking Patent Drafting from Inventor-Style Disclosures
Lekang Jiang, Wenjun Sun, Stephan Goetz
cs.CLarXiv:2608.21249v12026Affective Context Amplifies Sycophancy in LLM Responses
Jiayi Li, Sanjana Menon, Brett Frischmann +2
cs.CLarXiv:2608.21242v12026RARE: Decoupling Representation Steering from Expert Routing in Mixture-of-Experts Language Models
Zhibo Zhang, Zhen Ouyang, Ling Shi +1
cs.CLarXiv:2608.21236v12026Evidence-Consistent Generative Detection under Scenario-Level Distribution Shift
San Kim, JinYeong Bak
cs.CLarXiv:2608.21043v12026ForeDreamer: A Self-Evolving Dual-Agent Memory Architecture for Future Event Prediction
Linhao Zhong, Zongze Du, Linyu Wu +6
cs.CLarXiv:2608.20920v12026SAC-Copula: Quality-Preserving Watermarking for Diffusion Language Models via Smooth Correlated Gumbel Fields
Baixin Li, Haiyun He
cs.CLcs.CRcs.LGarXiv:2608.20839v12026Ontology-Driven Structural Regularization for Document-Level Relation Extraction
Laura Menotti, Stefano Marchesin, Gianmaria Silvello
cs.CLarXiv:2608.20856v12026AsmEvo: Agentic Assembly-Level Optimization of AMD GPU Kernels with Functional Equivalence Verification
Ji Liu, Puyuan Yang, Rongzhang Zheng +18
cs.CLarXiv:2608.20711v12026MIL-BERT: Classification of Arbitrarily Large Text with Performance and Explanatory Guarantees
John Cadigan, Dayne Freitag, Eric Yeh
cs.CLcs.LGarXiv:2608.20636v12026Sparse Token Routing in Efficient Transformers
Sai Krishna Arthanari, JaeHyeong Chang, Chengzhe Sun +1
cs.CLarXiv:2608.20632v12026LiLiCorr: Lightweight Likelihood Correlation of Parallel Drafts for Speculative Decoding
Matan Rusanovsky, Yoav Miron, Roy Uziel +4
cs.CLarXiv:2608.20530v12026ImmigrationReason: A Structured Dataset of U.S. Immigration Appeals for Legal Reasoning Research
Amirhossein Afsharrad, Seyed Shahabeddin Mousavi
cs.CLarXiv:2608.20391v12026Self-Supervised Speech Representations Track Spoken Language Convergence to Adult Models in Infants and Children Who Are Deaf/Hard-of-Hearing
L. Choy, A. S. Khan, S. Patrizi +3
cs.CLcs.SDarXiv:2608.20396v12026ARGUS: Theory-of-Mind Guided Argument Generation with Strategy-Aware Planning and Knowledge Grounding
Zhe Hu
cs.CLarXiv:2608.20405v12026Using Human-LLM Disagreement to Improve Checklist-Based Quality Appraisal
Timo van der Kuil, Bruno Messina Coimbra, Mirjam van Zuiden +7
cs.CLarXiv:2608.20385v12026Research Paper Quality Recognition Through Textual Feature Analysis
Saikiran Korla, Sadwik Gummadavelli, Trung-Nghia Le +2
cs.CLarXiv:2608.20368v12026Multilingual Verifier Bias in RLVR: Benchmark, Rollout Diagnosis, and the Cross-Lingual Selection Bottleneck
Chenyu Zhou, Qiliang Jiang, Xu Zhou
cs.CLcs.LGarXiv:2608.20362v12026Self-Speculation for Faster Reasoning Models
Ravisri Valluri, Tung Nguyen, Aditya Grover
cs.CLarXiv:2608.20359v12026TriPLU: Bypassing the Gate with Direct Trilinear Product FFNs in Tiny Language Models
He Zhang
cs.CLcs.LGarXiv:2608.20360v12026TurboBias 2.0: Streaming Context-Biasing for Production-Efficient ASR Systems
Vladimir Bataev, Lilit Grigoryan, Andrei Andrusenko +3
eess.AScs.AIcs.CLarXiv:2608.21343v12026Target-Aware Calibration Data Selection for Preserving Uncertainty in Quantized Language Models
Zhen Yang, Sizai Hou, Kaiwen Zheng +4
cs.CLcs.AIarXiv:2608.21019v12026EnSI-RAG: Entity-Structure-Indexed Retrieval-Augmented Generation for Long-Document Question Answering
Xuanyu Meng, Jiashuo Sun, Jash Rajesh Parekh +1
cs.CLcs.AIcs.DBarXiv:2608.21252v12026PromptResponse: Optimizing Prompts for LLM Coding Tasks
Erik Thureck, Robert Kühnen, Tim Jacobowitz
cs.CLcs.AIcs.HCarXiv:2608.21074v12026Free-Text Evaluation of LLMs for 5G Domain Knowledge and Fault Analysis using LLM-as-Judge
Rishiraj Sengupta, Sotiris Chatzimiltis, Mohammad Shojafar +1
cs.CLcs.AIcs.NIarXiv:2608.21021v12026Quantization-Aware Healing: A Practical Recipe for Recovering Compressed, 4-Bit LLMs
Bakbergen Ryskulov, Iker García-Ferrero, David Montero +5
cs.CLcs.AIcs.LGarXiv:2608.20953v12026Extractive Summarization for Arabic Documents Using SAraBERT with a Semantic Siamese Similarity Evaluation Metric
Sami Shames El Deen, Mariette Awad
cs.CLcs.AIarXiv:2608.20964v12026PSK at WMT 2026 MIST: Task-Specialized QLoRA Adapters for Multilingual Summarization and Question Answering
Srikar Kashyap Pulipaka
cs.CLcs.AIcs.LGarXiv:2608.20757v12026Temporal Validity on Real Software Histories: Eliminating Stale-Fact Errors in Code-Assistant Memory over GitHub Fixes
Neeraj Yadav
cs.SEcs.AIcs.CLarXiv:2608.20685v12026No PUN Intended: Plausible Unknown Names for Person-Centred LLM Evaluation
Dimitri Staufer, David Hartmann, Ibrahim Baroud
cs.CLcs.AIcs.LGarXiv:2608.21206v12026Trustworthy RAG: An Evaluation Agent for Detecting Misinformation and Knowledge Poisoning in Generative AI Systems
Balkrishna Giri, Md Toufique Hasan, Jussi Rasku +2
cs.SEcs.AIcs.CLarXiv:2608.21095v12026GAIA: a benchmark for General AI Assistants
Grégoire Mialon, Clémentine Fourrier, Craig Swift +3
cs.CLcs.AIarXiv:2311.12983v12023Source-Free MT Evaluation Is Not MT Evaluation
Baban Gain, Ramakrishna Appicharla, Asif Ekbal
cs.CLcs.AIarXiv:2608.20925v12026KREL: Automatic Medical Coding via Knowledge-Guided Reasoning over Clinical Evidence with LLMs
Xubin Chen, Yipeng Zhou, Wen Sun +3
cs.CLcs.AIarXiv:2608.20887v12026STAR-OPD: Structured Aspect-Cascade-Aware On-Policy Reward Distillation for ABSA Quadruple Extraction
Tong Sun, Mingyang Ma, Jiayang Yu
cs.CLcs.AIarXiv:2608.20831v12026