Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

9,601 to 9,660 of 11,241

  1. How Do Large Language Models Learn Concepts During Continual Pre-Training?

    Barry Menglong Yao, Sha Li, Yunzhi Yao +4

    cs.CLarXiv:2601.03570v22026
  2. Hybrid Policy Distillation for LLMs

    Wenhong Zhu, Ruobing Xie, Rui Wang +1

    cs.CLcs.AIarXiv:2604.20244v22026
  3. Building a Precise Video Language with Human-AI Oversight

    Zhiqiu Lin, Chancharik Mitra, Siyuan Cen +13

    cs.CVcs.AIcs.CLarXiv:2604.21718v22026
  4. Explainable Disentangled Representation Learning for Generalizable Authorship Attribution in the Era of Generative AI

    Hieu Man, Van-Cuong Pham, Nghia Trung Ngo +2

    cs.CLcs.IRcs.LGarXiv:2604.21300v12026
  5. SLIDERS: Systematic Reviews via Automated Evidence Synthesis and Reconciliation

    Harshit Joshi, Priyank Shethia, Jadelynn Dao +1

    cs.CLcs.AIarXiv:2604.22294v22026
  6. The Geometric Canary: Predicting Steerability and Detecting Drift via Representational Stability

    Prashant C. Raju

    cs.LGcs.CLstat.MLarXiv:2604.17698v22026
  7. Sparse Reward Subsystem in Large Language Models

    Guowei Xu, Mert Yuksekgonul, James Zou

    cs.CLarXiv:2602.00986v22026
  8. Future-KL Regularized GRPO: Process-Level Credit Assignment from $f$-Divergence Regularization

    Jiarui Yao, Ruida Wang, Hao Bai +1

    cs.LGcs.AIcs.CLarXiv:2601.10201v22026
  9. PRISM-$Δ$: Differential Subspace Steering for Prompt Highlighting in Large Language Models

    Yuyao Ge, Shenghua Liu, Yiwei Wang +5

    cs.CLarXiv:2603.10705v22026
  10. XSkill: Continual Learning from Experience and Skills in Multimodal Agents

    Guanyu Jiang, Zhaochen Su, Xiaoye Qu +1

    cs.AIcs.CLarXiv:2603.12056v32026
  11. Geometric Stability: The Missing Axis of Representations

    Prashant C. Raju

    cs.LGcs.CLq-bio.QMarXiv:2601.09173v52026
  12. Decoupled Vision-Language System for Multimodal Understanding and Generation

    Yifan Xu, Baochen Xiong, Xiaoshan Yang +3

    cs.CLcs.CVarXiv:2608.20382v12026
  13. COMET: Contrastive Motion-Enhanced Temporal Reasoning for Video Multimodal Large Language Models

    Chenghua Zhu, Zhaolu Kang, Qifan Shi +8

    cs.CVcs.CLcs.LGarXiv:2608.21030v12026
  14. MigrationNarrate: A Dataset for Detection of Migration Narratives in YouTube Videos

    Fatima Haouari, Carolina Scarton, Kalina Bontcheva

    cs.CVcs.CLcs.CYarXiv:2608.20984v12026
  15. Prompt-Model Interaction Reaches the Fixed Points: A deterministic, task-free structural readout -- and the factorizations of it that failed

    Nicolás Vera Zúñiga

    cs.CLarXiv:2608.21315v12026
  16. Tree-of-Concerns: Hierarchical Multi-Agent Debate for Unstated-Limitation Extraction in Scientific Critique

    Sahil Mishra, Niranjan Rajeev, Tanmoy Chakraborty

    cs.CLarXiv:2608.20777v12026
  17. When the Feature Pool Goes Algorithmic: Extending Mufwene's Ecology of Language Evolution to LLM-Mediated Exposure

    Kunmei Han

    cs.CLarXiv:2608.21088v12026
  18. GRAFT: Adaptive DLM-Based Draft Tree Construction with Target-Distilled Edge Scoring

    Xuming Ye, Zeming Ma, Runjie Yu +5

    cs.CLarXiv:2608.20375v12026
  19. An ambiguity taxonomy for evaluating large language model performance on clinical registry abstraction: a multi-site prospective study

    James Matheson, Betsy Castillo, Andrew Y. Shin +1

    cs.CLarXiv:2608.20373v12026
  20. Beyond Raw Transcripts: Structured Persona Extraction for LLM-Based Digital Twins

    Iris Ye, Tianze Deng, Ozan Candogan

    cs.CLcs.CYarXiv:2608.20344v12026
  21. Building and Evaluating a Synthetic Bengali Speech Resource for Telecom Customer Care

    Kawshik Kumar Paul, Md. Nafiul Alam Fuji

    cs.CLcs.SDeess.ASarXiv:2608.20346v12026
  22. Exploratory As-Analyzed No-Detection of Culturally-Marked Predicate-Triggered PII Amplification in a Synthetic-English RAG Probe: A Predicate-Resource-Confounded Audit

    Yanhang Li, Zhichao Fan, Zexin Zhuang

    cs.CLcs.CYcs.LGarXiv:2608.20351v12026
  23. Identify, Locate, Link: End-to-End Key-Value Extraction from Document Images

    A. Said Gurbuz, Ahmed Nassar, Christoph Auer +8

    cs.CVcs.CLarXiv:2608.20868v12026
  24. A Factorial Ablation of a Speech-to-SFT Pipeline: Differential Effects on Data Quality and Downstream Transfer

    Wonsup Shin, Jingu Kim

    cs.SDcs.CLarXiv:2608.20394v12026
  25. Move by Move: Measuring and Steering How LLMs Conduct Psychotherapy

    Afonso Baldo, Hugo Pitorro, Areti Vassilopoulos +5

    cs.CLarXiv:2608.21325v12026
  26. Memory Augmentation Unlocks Efficient Chain-of-Thought Reasoning

    Simeng Zhang, Yilong Chen, Wenyuan Zhang +4

    cs.CLarXiv:2608.21265v12026
  27. Benchmarking Patent Drafting from Inventor-Style Disclosures

    Lekang Jiang, Wenjun Sun, Stephan Goetz

    cs.CLarXiv:2608.21249v12026
  28. Affective Context Amplifies Sycophancy in LLM Responses

    Jiayi Li, Sanjana Menon, Brett Frischmann +2

    cs.CLarXiv:2608.21242v12026
  29. RARE: Decoupling Representation Steering from Expert Routing in Mixture-of-Experts Language Models

    Zhibo Zhang, Zhen Ouyang, Ling Shi +1

    cs.CLarXiv:2608.21236v12026
  30. Evidence-Consistent Generative Detection under Scenario-Level Distribution Shift

    San Kim, JinYeong Bak

    cs.CLarXiv:2608.21043v12026
  31. ForeDreamer: A Self-Evolving Dual-Agent Memory Architecture for Future Event Prediction

    Linhao Zhong, Zongze Du, Linyu Wu +6

    cs.CLarXiv:2608.20920v12026
  32. SAC-Copula: Quality-Preserving Watermarking for Diffusion Language Models via Smooth Correlated Gumbel Fields

    Baixin Li, Haiyun He

    cs.CLcs.CRcs.LGarXiv:2608.20839v12026
  33. Ontology-Driven Structural Regularization for Document-Level Relation Extraction

    Laura Menotti, Stefano Marchesin, Gianmaria Silvello

    cs.CLarXiv:2608.20856v12026
  34. AsmEvo: Agentic Assembly-Level Optimization of AMD GPU Kernels with Functional Equivalence Verification

    Ji Liu, Puyuan Yang, Rongzhang Zheng +18

    cs.CLarXiv:2608.20711v12026
  35. MIL-BERT: Classification of Arbitrarily Large Text with Performance and Explanatory Guarantees

    John Cadigan, Dayne Freitag, Eric Yeh

    cs.CLcs.LGarXiv:2608.20636v12026
  36. Sparse Token Routing in Efficient Transformers

    Sai Krishna Arthanari, JaeHyeong Chang, Chengzhe Sun +1

    cs.CLarXiv:2608.20632v12026
  37. LiLiCorr: Lightweight Likelihood Correlation of Parallel Drafts for Speculative Decoding

    Matan Rusanovsky, Yoav Miron, Roy Uziel +4

    cs.CLarXiv:2608.20530v12026
  38. ImmigrationReason: A Structured Dataset of U.S. Immigration Appeals for Legal Reasoning Research

    Amirhossein Afsharrad, Seyed Shahabeddin Mousavi

    cs.CLarXiv:2608.20391v12026
  39. Self-Supervised Speech Representations Track Spoken Language Convergence to Adult Models in Infants and Children Who Are Deaf/Hard-of-Hearing

    L. Choy, A. S. Khan, S. Patrizi +3

    cs.CLcs.SDarXiv:2608.20396v12026
  40. ARGUS: Theory-of-Mind Guided Argument Generation with Strategy-Aware Planning and Knowledge Grounding

    Zhe Hu

    cs.CLarXiv:2608.20405v12026
  41. Using Human-LLM Disagreement to Improve Checklist-Based Quality Appraisal

    Timo van der Kuil, Bruno Messina Coimbra, Mirjam van Zuiden +7

    cs.CLarXiv:2608.20385v12026
  42. Research Paper Quality Recognition Through Textual Feature Analysis

    Saikiran Korla, Sadwik Gummadavelli, Trung-Nghia Le +2

    cs.CLarXiv:2608.20368v12026
  43. Multilingual Verifier Bias in RLVR: Benchmark, Rollout Diagnosis, and the Cross-Lingual Selection Bottleneck

    Chenyu Zhou, Qiliang Jiang, Xu Zhou

    cs.CLcs.LGarXiv:2608.20362v12026
  44. Self-Speculation for Faster Reasoning Models

    Ravisri Valluri, Tung Nguyen, Aditya Grover

    cs.CLarXiv:2608.20359v12026
  45. TriPLU: Bypassing the Gate with Direct Trilinear Product FFNs in Tiny Language Models

    He Zhang

    cs.CLcs.LGarXiv:2608.20360v12026
  46. TurboBias 2.0: Streaming Context-Biasing for Production-Efficient ASR Systems

    Vladimir Bataev, Lilit Grigoryan, Andrei Andrusenko +3

    eess.AScs.AIcs.CLarXiv:2608.21343v12026
  47. Target-Aware Calibration Data Selection for Preserving Uncertainty in Quantized Language Models

    Zhen Yang, Sizai Hou, Kaiwen Zheng +4

    cs.CLcs.AIarXiv:2608.21019v12026
  48. EnSI-RAG: Entity-Structure-Indexed Retrieval-Augmented Generation for Long-Document Question Answering

    Xuanyu Meng, Jiashuo Sun, Jash Rajesh Parekh +1

    cs.CLcs.AIcs.DBarXiv:2608.21252v12026
  49. PromptResponse: Optimizing Prompts for LLM Coding Tasks

    Erik Thureck, Robert Kühnen, Tim Jacobowitz

    cs.CLcs.AIcs.HCarXiv:2608.21074v12026
  50. Free-Text Evaluation of LLMs for 5G Domain Knowledge and Fault Analysis using LLM-as-Judge

    Rishiraj Sengupta, Sotiris Chatzimiltis, Mohammad Shojafar +1

    cs.CLcs.AIcs.NIarXiv:2608.21021v12026
  51. Quantization-Aware Healing: A Practical Recipe for Recovering Compressed, 4-Bit LLMs

    Bakbergen Ryskulov, Iker García-Ferrero, David Montero +5

    cs.CLcs.AIcs.LGarXiv:2608.20953v12026
  52. Extractive Summarization for Arabic Documents Using SAraBERT with a Semantic Siamese Similarity Evaluation Metric

    Sami Shames El Deen, Mariette Awad

    cs.CLcs.AIarXiv:2608.20964v12026
  53. PSK at WMT 2026 MIST: Task-Specialized QLoRA Adapters for Multilingual Summarization and Question Answering

    Srikar Kashyap Pulipaka

    cs.CLcs.AIcs.LGarXiv:2608.20757v12026
  54. Temporal Validity on Real Software Histories: Eliminating Stale-Fact Errors in Code-Assistant Memory over GitHub Fixes

    Neeraj Yadav

    cs.SEcs.AIcs.CLarXiv:2608.20685v12026
  55. No PUN Intended: Plausible Unknown Names for Person-Centred LLM Evaluation

    Dimitri Staufer, David Hartmann, Ibrahim Baroud

    cs.CLcs.AIcs.LGarXiv:2608.21206v12026
  56. Trustworthy RAG: An Evaluation Agent for Detecting Misinformation and Knowledge Poisoning in Generative AI Systems

    Balkrishna Giri, Md Toufique Hasan, Jussi Rasku +2

    cs.SEcs.AIcs.CLarXiv:2608.21095v12026
  57. GAIA: a benchmark for General AI Assistants

    Grégoire Mialon, Clémentine Fourrier, Craig Swift +3

    cs.CLcs.AIarXiv:2311.12983v12023
  58. Source-Free MT Evaluation Is Not MT Evaluation

    Baban Gain, Ramakrishna Appicharla, Asif Ekbal

    cs.CLcs.AIarXiv:2608.20925v12026
  59. KREL: Automatic Medical Coding via Knowledge-Guided Reasoning over Clinical Evidence with LLMs

    Xubin Chen, Yipeng Zhou, Wen Sun +3

    cs.CLcs.AIarXiv:2608.20887v12026
  60. STAR-OPD: Structured Aspect-Cascade-Aware On-Policy Reward Distillation for ABSA Quadruple Extraction

    Tong Sun, Mingyang Ma, Jiayang Yu

    cs.CLcs.AIarXiv:2608.20831v12026