Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

781 to 840 of 11,199

  1. Data Scarcity and Model Sparsity: Mixtures-of-Experts Overfit More to Repeated Data

    Atindra Jha, Margaret Li, Jure Leskovec +2

    cs.LGcs.CLarXiv:2609.11917v12026
  2. Biology-in-the-loop: Amortized Adaptive Hit Discovery in CRISPR Screens

    Carl Edwards, Edward De Brouwer, Xiner Li +5

    q-bio.QMcs.AIcs.CLarXiv:2609.11877v12026
  3. Abstraction Agent

    Boning Li, Longbo Huang

    cs.MAcs.AIcs.CLarXiv:2609.04303v12026
  4. BERT for Evidence Retrieval and Claim Verification

    Amir Soleimani, Christof Monz, Marcel Worring

    cs.CLarXiv:1910.02655v12019
  5. RetroThinker: Enabling Retrospective Thinking in Speech LLMs

    Yi-Jen Shih, Puyuan Peng, Abdelrahman Mohamed +1

    eess.AScs.AIcs.CLarXiv:2609.11864v12026
  6. A Unified Per-Token Gating Family for On-Policy Distillation: FKL/RKL Mixing with Multi-Channel and Bias Coefficients

    Suwan Wu, Yumeng Lin, Pengcheng Yuan +1

    cs.AIcs.CLcs.LGarXiv:2609.11768v12026
  7. SIRF: A Spec-Internalized Risk Foundation Model for Industrial Content Risk Control

    Suwan Wu, Yumeng Lin, Pengcheng Yuan +1

    cs.AIcs.CLcs.LGarXiv:2609.11752v12026
  8. A Systematic Survey and Critical Review on Evaluating Large Language Models: Challenges, Limitations, and Recommendations

    Md Tahmid Rahman Laskar, Sawsan Alqahtani, M Saiful Bari +10

    cs.CLcs.AIcs.LGarXiv:2407.04069v22024
  9. The Semantic Elevation Operator and the Closure of the Undecidable Class under Preservation

    Jose Pascual Gumbau Mezquita

    cs.LOcs.AIcs.CLarXiv:2609.11326v12026
  10. The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement

    Yi Duan, Ying Liu, Zirui Tang +30

    cs.LGcs.AIcs.CLarXiv:2609.11873v12026
  11. SpecGuard: Inference-Time Backdoor Detection For Free

    Rui Wen, Ahmed Salem, Andrew Paverd +2

    cs.CRcs.CLarXiv:2609.11799v12026
  12. Nemotron-CC: Transforming Common Crawl into a Refined Long-Horizon Pretraining Dataset

    Dan Su, Kezhi Kong, Ying Lin +6

    cs.CLarXiv:2412.02595v22024
  13. Neural Segmental Hypergraphs for Overlapping Mention Recognition

    Bailin Wang, Wei Lu

    cs.CLarXiv:1810.01817v12018
  14. Whisper-Based Speech Transcription from Videos Across Multiple Languages for Cross-Cultural Understanding

    Michael Picheny

    eess.AScs.CLarXiv:2609.11772v12026
  15. InFoBench: Evaluating Instruction Following Ability in Large Language Models

    Yiwei Qin, Kaiqiang Song, Yebowen Hu +7

    cs.CLcs.AIarXiv:2401.03601v12024
  16. A Large-Scale Chinese Short-Text Conversation Dataset

    Yida Wang, Pei Ke, Yinhe Zheng +4

    cs.CLarXiv:2008.03946v22020
  17. Why Does Post-Training Quantization Work?

    Yuxiang Chen, Michael Beyer, Jun Zhu +1

    cs.LGcs.CLarXiv:2609.11716v12026
  18. Survey on reinforcement learning for language processing

    Victor Uc-Cetina, Nicolas Navarro-Guerrero, Anabel Martin-Gonzalez +2

    cs.CLcs.AIcs.LGarXiv:2104.05565v32021
  19. VikingRAG: Accurate and Token-efficient Retrieval-augmented Generation over Structured Documents

    Peiyuan Gao, Gaoyuan Zhang, Haojie Qin +5

    cs.IRcs.AIcs.CLarXiv:2609.11390v12026
  20. Xiaomi-CocktailASR-1 Technical Report

    Yiru Zhang, Hang Su, Lichun Fan +10

    cs.SDcs.CLeess.ASarXiv:2609.11274v12026
  21. (Whose defaults?) Is artificial intelligence reorienting archaeological methods?

    Lorenzo Cardarelli, Roberto Ragno

    cs.CYcs.AIcs.CLarXiv:2609.11198v12026
  22. LILA: Calibration-Free Structured Pruning of Large Language Models via Latent Spectral Geometry

    Sankar Behera, Dhruv Singh, Anshika Agnihotri +3

    cs.LGcs.CLarXiv:2609.11163v12026
  23. Learning a bidirectional mapping between human whole-body motion and natural language using deep recurrent neural networks

    Matthias Plappert, Christian Mandery, Tamim Asfour

    cs.LGcs.CLcs.ROarXiv:1705.06400v22017
  24. Learning to Parse and Translate Improves Neural Machine Translation

    Akiko Eriguchi, Yoshimasa Tsuruoka, Kyunghyun Cho

    cs.CLarXiv:1702.03525v22017
  25. INDRA: A New AI Tool for Exploring Tobacco, Fossil Fuel, and Chemical Industry Archives

    Daniel Akselrad, Robert N. Proctor

    cs.DLcs.CLcs.CYarXiv:2609.11261v12026
  26. MUtE: A Dual Framework for Concept Erasure and Counterfactual Interventions

    Antoine Saillenfest

    cs.LGcs.CLarXiv:2609.11253v12026
  27. A Voice-Interactive Multi-Agent System for Smart Operating Rooms: Architecture Design and Key Technologies

    Tianxiang Zhou

    cs.AIcs.CLcs.HCarXiv:2609.11231v12026
  28. REVA: Reusable Evidence View Aggregation for Context-Efficient RAG Serving

    Tuan Nguyen, Qiran Hu, Banruo Liu +3

    cs.LGcs.CLcs.IRarXiv:2609.11209v12026
  29. ONCE: Boosting Content-based Recommendation with Both Open- and Closed-source Large Language Models

    Qijiong Liu, Nuo Chen, Tetsuya Sakai +1

    cs.IRcs.CLarXiv:2305.06566v42023
  30. FedBiOT: LLM Local Fine-tuning in Federated Learning without Full Model

    Feijie Wu, Zitao Li, Yaliang Li +2

    cs.LGcs.CLcs.DCarXiv:2406.17706v12024
  31. The Oligarch Barely Steers Model Collapse in Multi-Model Ecosystems

    Yangze Liu, Zhongyi Han

    cs.AIcs.CLcs.LGarXiv:2609.11146v12026
  32. Same Day, Same Story; One Day Ahead, a Different Signal: The Dual Validity of Financial Sentiment

    AS Aravinthkakshan, Laven Srivastava, Harsh Nandwani

    cs.AIcs.CLcs.SIarXiv:2609.11144v12026
  33. The information geometry of large language models is shared, learned, and controllable

    Dario Picozzi

    cs.LGcs.CLarXiv:2609.11063v12026
  34. ChatGPT Needs SPADE (Sustainability, PrivAcy, Digital divide, and Ethics) Evaluation: A Review

    Sunder Ali Khowaja, Parus Khuwaja, Kapal Dev +2

    cs.CYcs.AIcs.CLarXiv:2305.03123v42023
  35. Story Imprinting: AI Assistants Absorb Traits from Human Characters They Resemble

    Jorio Cocola, Lev McKinney, Harry Mayne +2

    cs.LGcs.AIcs.CLarXiv:2609.10883v12026
  36. Optimization Methods for Personalizing Large Language Models through Retrieval Augmentation

    Alireza Salemi, Surya Kallumadi, Hamed Zamani

    cs.CLcs.IRarXiv:2404.05970v12024
  37. BodyCam-VQA: Enhanced Body-Worn Camera Video Captioning via Multimodal Reasoning and Probe Question Generation

    Karish Gupta, Matthew Alex, Alex Li +6

    cs.CVcs.CLarXiv:2609.10815v12026
  38. More than half of recent astronomy papers are written with language-model assistance

    Serat M. Saad, Yuan-Sen Ting

    astro-ph.IMcs.CLcs.DLarXiv:2609.10664v12026
  39. KuaiRP Series Role-playing Models Technical Report

    Yipeng Wang, Ziwei Zhang, Jiahui Zhang +2

    cs.AIcs.CLarXiv:2609.11127v12026
  40. Beyond Solver Verdicts: Generative Reward Models for Autoformalization

    Vikash Singh, Debargha Ganguly, Aman Goel +5

    cs.LGcs.CLarXiv:2609.11085v12026
  41. Relying on the Unreliable: The Impact of Language Models' Reluctance to Express Uncertainty

    Kaitlyn Zhou, Jena D. Hwang, Xiang Ren +1

    cs.CLcs.AIcs.HCarXiv:2401.06730v22024
  42. New Evidence, Same Choice: Testing Physical Experiment Selection in Vision Language Models

    Sourajit Saha, Shubhashis Roy Dipta, Nobin Sarwar +4

    cs.CVcs.AIcs.CLarXiv:2609.11022v12026
  43. Empirical Evaluation of Membership Inference Attacks on NLP Text Classifiers: A Baseline Study on SST-2

    William Novak, Muhammad Abusaqer

    cs.CRcs.CLcs.LGarXiv:2609.10935v12026
  44. Detection of Hate Speech using BERT and Hate Speech Word Embedding with Deep Model

    Hind Saleh, Areej Alhothali, Kawthar Moria

    cs.CLarXiv:2111.01515v12021
  45. Studying Without a Syllabus: Task-Agnostic Environment Preprocessing

    Vinay Samuel, Varun Ursekar, Vijay S. Kalmath +3

    cs.AIcs.CLcs.LGarXiv:2609.10824v12026
  46. The Truth Was Never Gone: Perfect Aliasing in Compliant-Context Truth Probes

    Dylan Jayabahu

    cs.LGcs.AIcs.CLarXiv:2609.10739v12026
  47. The Lazy Neuron Phenomenon: On Emergence of Activation Sparsity in Transformers

    Zonglin Li, Chong You, Srinadh Bhojanapalli +8

    cs.LGcs.CLcs.CVarXiv:2210.06313v22022
  48. Bi-Directional Block Self-Attention for Fast and Memory-Efficient Sequence Modeling

    Tao Shen, Tianyi Zhou, Guodong Long +2

    cs.CLcs.AIarXiv:1804.00857v12018
  49. Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA

    Minzheng Wang, Longze Chen, Cheng Fu +11

    cs.CLcs.AIarXiv:2406.17419v22024
  50. IndicTriMix: Developing Language Identification Datasets and Models for Tri-Language Code-Mixing

    Pruthwik Mishra, Rudra Trivedi, Avi Patel +2

    cs.CLarXiv:2609.11851v12026
  51. Target leakage, not model class, explains reported accuracy in survey-based cardiovascular screening: a leakage-tiered audit of glass-box and tabular foundation models

    Raad Bin Tareaf, Murad Al-Rajab, Samia Loucif +2

    cs.CLarXiv:2609.11838v12026
  52. Probing Natural Language Inference Models through Semantic Fragments

    Kyle Richardson, Hai Hu, Lawrence S. Moss +1

    cs.CLarXiv:1909.07521v22019
  53. Beyond Word Error Rate: A Switch Aware Evaluation of ASR and Audio Language Models on English Yoruba Code-Switched Speech

    Chibuzor Okocha, Christan Earl Grant

    cs.CLcs.AIarXiv:2609.11786v12026
  54. Distance generalization in transformers: why bother with positional encoding?

    Daniel Henrik Nevermann, Claudius Gros

    cs.CLarXiv:2609.11913v12026
  55. Nuha-Speech: Building General-Purpose Arabic Speech-LLMs

    Yingzhi Wang, Reem Alhazzani, Muhammad Alqurishi

    cs.CLarXiv:2609.11892v12026
  56. Domain-Specific Hallucination Detection in Large Language Models

    Varun Teja Chundru, Debasmita Biswas

    cs.CLcs.AIcs.LGarXiv:2609.11878v12026
  57. Augustinian BabyLM: What Ostensive Definition Can and Cannot Teach a Small Language Model

    Lisa Bylinina

    cs.CLarXiv:2609.11870v12026
  58. Epistemic orientation predicts legislative effectiveness among members of the US Congress

    Segun Aroyehun, Stephan Lewandowsky, David Garcia

    cs.CLarXiv:2609.11865v12026
  59. AnglE-optimized Text Embeddings

    Xianming Li, Jing Li

    cs.CLcs.AIcs.LGarXiv:2309.12871v92023
  60. The widening evaluation gap in medical large language model research 2023 to 2026

    Raad Bin Tareaf, Murad Al-Rajab, Samia Loucif

    cs.CLarXiv:2609.11770v12026