Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

841 to 900 of 11,277

  1. wav2vec: Unsupervised Pre-training for Speech Recognition

    Steffen Schneider, Alexei Baevski, Ronan Collobert +1

    cs.CLarXiv:1904.05862v42019
  2. Rethinking Interpretability in the Era of Large Language Models

    Chandan Singh, Jeevana Priya Inala, Michel Galley +2

    cs.CLcs.AIcs.LGarXiv:2402.01761v12024
  3. Compression of Deep Learning Models for Text: A Survey

    Manish Gupta, Puneet Agrawal

    cs.CLcs.AIcs.CVarXiv:2008.05221v42020
  4. A Panoramic Survey of Natural Language Processing in the Arab World

    Kareem Darwish, Nizar Habash, Mourad Abbas +9

    cs.CLarXiv:2011.12631v32020
  5. SpikeGPT: Generative Pre-trained Language Model with Spiking Neural Networks

    Rui-Jie Zhu, Qihang Zhao, Guoqi Li +1

    cs.CLcs.LGcs.NEarXiv:2302.13939v52023
  6. ClinicalGPT: Large Language Models Finetuned with Diverse Medical Data and Comprehensive Evaluation

    Guangyu Wang, Guoxing Yang, Zongxin Du +2

    cs.CLarXiv:2306.09968v12023
  7. The Ultimate Guide to Fine-Tuning LLMs from Basics to Breakthroughs: An Exhaustive Review of Technologies, Research, Best Practices, Applied Research Challenges and Opportunities

    Venkatesh Balavadhani Parthasarathy, Ahtsham Zafar, Aafaq Khan +1

    cs.LGcs.CLarXiv:2408.13296v32024
  8. LongRAG: Enhancing Retrieval-Augmented Generation with Long-context LLMs

    Ziyan Jiang, Xueguang Ma, Wenhu Chen

    cs.CLcs.AIarXiv:2406.15319v32024
  9. Exploring LLM-based Agents for Root Cause Analysis

    Devjeet Roy, Xuchao Zhang, Rashi Bhave +4

    cs.SEcs.CLcs.LGarXiv:2403.04123v12024
  10. Lifted Rule Injection for Relation Embeddings

    Thomas Demeester, Tim Rocktäschel, Sebastian Riedel

    cs.LGcs.AIcs.CLarXiv:1606.08359v22016
  11. Speech Recognition with Augmented Synthesized Speech

    Andrew Rosenberg, Yu Zhang, Bhuvana Ramabhadran +4

    cs.CLcs.SDeess.ASarXiv:1909.11699v12019
  12. An analysis of Twitter messages in the 2011 Tohoku Earthquake

    Son Doan, Bao-Khanh Ho Vo, Nigel Collier

    cs.SIcs.CLphysics.soc-pharXiv:1109.1618v12011
  13. Drug-Drug Interaction Extraction from Biomedical Text Using Long Short Term Memory Network

    Sunil Kumar Sahu, Ashish Anand

    cs.CLarXiv:1701.08303v22017
  14. "HOT" ChatGPT: The promise of ChatGPT in detecting and discriminating hateful, offensive, and toxic comments on social media

    Lingyao Li, Lizhou Fan, Shubham Atreja +1

    cs.CLcs.AIcs.HCarXiv:2304.10619v12023
  15. Backpropagation through Signal Temporal Logic Specifications: Infusing Logical Structure into Gradient-Based Methods

    Karen Leung, Nikos Aréchiga, Marco Pavone

    eess.SYcs.CLcs.LOarXiv:2008.00097v32020
  16. Imagination improves Multimodal Translation

    Desmond Elliott, Ákos Kádár

    cs.CLcs.CVarXiv:1705.04350v22017
  17. M6: A Chinese Multimodal Pretrainer

    Junyang Lin, Rui Men, An Yang +22

    cs.CLarXiv:2103.00823v42021
  18. Deep Multitask Learning for Semantic Dependency Parsing

    Hao Peng, Sam Thomson, Noah A. Smith

    cs.CLarXiv:1704.06855v22017
  19. Contextual Parameter Generation for Universal Neural Machine Translation

    Emmanouil Antonios Platanios, Mrinmaya Sachan, Graham Neubig +1

    cs.CLcs.LGstat.MLarXiv:1808.08493v12018
  20. Language Models that Seek for Knowledge: Modular Search & Generation for Dialogue and Prompt Completion

    Kurt Shuster, Mojtaba Komeili, Leonard Adolphs +3

    cs.CLcs.AIarXiv:2203.13224v22022
  21. ChatTime: A Unified Multimodal Time Series Foundation Model Bridging Numerical and Textual Data

    Chengsen Wang, Qi Qi, Jingyu Wang +5

    cs.CLcs.LGarXiv:2412.11376v12024
  22. Neural data-to-text generation: A comparison between pipeline and end-to-end architectures

    Thiago Castro Ferreira, Chris van der Lee, Emiel van Miltenburg +1

    cs.CLarXiv:1908.09022v22019
  23. Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs

    Rui Yang, Ruomeng Ding, Yong Lin +2

    cs.CLcs.AIarXiv:2406.10216v22024
  24. MindTopo: Can Foundation Models Reason in Topological Space?

    Yunfei Ge, Anbang Liu, Qineng Wang +9

    cs.AIcs.CLcs.CVarXiv:2609.11900v12026
  25. Improving speaker discrimination of target speech extraction with time-domain SpeakerBeam

    Marc Delcroix, Tsubasa Ochiai, Katerina Zmolikova +4

    eess.AScs.CLcs.SDarXiv:2001.08378v12020
  26. How would Stance Detection Techniques Evolve after the Launch of ChatGPT?

    Bowen Zhang, Daijun Ding, Liwen Jing +2

    cs.CLarXiv:2212.14548v42022
  27. UmlsBERT: Clinical Domain Knowledge Augmentation of Contextual Embeddings Using the Unified Medical Language System Metathesaurus

    George Michalopoulos, Yuanxin Wang, Hussam Kaka +2

    cs.CLcs.AIcs.LGarXiv:2010.10391v52020
  28. EmoBERTa: Speaker-Aware Emotion Recognition in Conversation with RoBERTa

    Taewoon Kim, Piek Vossen

    cs.CLarXiv:2108.12009v12021
  29. Universal Jailbreak Backdoors from Poisoned Human Feedback

    Javier Rando, Florian Tramèr

    cs.AIcs.CLcs.CRarXiv:2311.14455v42023
  30. A Question Answering Approach to Emotion Cause Extraction

    Lin Gui, Jiannan Hu, Yulan He +3

    cs.CLarXiv:1708.05482v22017
  31. CALF: Aligning LLMs for Time Series Forecasting via Cross-modal Fine-Tuning

    Peiyuan Liu, Hang Guo, Tao Dai +5

    cs.LGcs.CLarXiv:2403.07300v32024
  32. Neural Language Correction with Character-Based Attention

    Ziang Xie, Anand Avati, Naveen Arivazhagan +2

    cs.CLcs.AIarXiv:1603.09727v12016
  33. Data Scarcity and Model Sparsity: Mixtures-of-Experts Overfit More to Repeated Data

    Atindra Jha, Margaret Li, Jure Leskovec +2

    cs.LGcs.CLarXiv:2609.11917v12026
  34. Biology-in-the-loop: Amortized Adaptive Hit Discovery in CRISPR Screens

    Carl Edwards, Edward De Brouwer, Xiner Li +5

    q-bio.QMcs.AIcs.CLarXiv:2609.11877v12026
  35. Abstraction Agent

    Boning Li, Longbo Huang

    cs.MAcs.AIcs.CLarXiv:2609.04303v12026
  36. BERT for Evidence Retrieval and Claim Verification

    Amir Soleimani, Christof Monz, Marcel Worring

    cs.CLarXiv:1910.02655v12019
  37. RetroThinker: Enabling Retrospective Thinking in Speech LLMs

    Yi-Jen Shih, Puyuan Peng, Abdelrahman Mohamed +1

    eess.AScs.AIcs.CLarXiv:2609.11864v12026
  38. A Unified Per-Token Gating Family for On-Policy Distillation: FKL/RKL Mixing with Multi-Channel and Bias Coefficients

    Suwan Wu, Yumeng Lin, Pengcheng Yuan +1

    cs.AIcs.CLcs.LGarXiv:2609.11768v12026
  39. SIRF: A Spec-Internalized Risk Foundation Model for Industrial Content Risk Control

    Suwan Wu, Yumeng Lin, Pengcheng Yuan +1

    cs.AIcs.CLcs.LGarXiv:2609.11752v12026
  40. A Systematic Survey and Critical Review on Evaluating Large Language Models: Challenges, Limitations, and Recommendations

    Md Tahmid Rahman Laskar, Sawsan Alqahtani, M Saiful Bari +10

    cs.CLcs.AIcs.LGarXiv:2407.04069v22024
  41. The Semantic Elevation Operator and the Closure of the Undecidable Class under Preservation

    Jose Pascual Gumbau Mezquita

    cs.LOcs.AIcs.CLarXiv:2609.11326v12026
  42. The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement

    Yi Duan, Ying Liu, Zirui Tang +30

    cs.LGcs.AIcs.CLarXiv:2609.11873v12026
  43. SpecGuard: Inference-Time Backdoor Detection For Free

    Rui Wen, Ahmed Salem, Andrew Paverd +2

    cs.CRcs.CLarXiv:2609.11799v12026
  44. Nemotron-CC: Transforming Common Crawl into a Refined Long-Horizon Pretraining Dataset

    Dan Su, Kezhi Kong, Ying Lin +6

    cs.CLarXiv:2412.02595v22024
  45. Neural Segmental Hypergraphs for Overlapping Mention Recognition

    Bailin Wang, Wei Lu

    cs.CLarXiv:1810.01817v12018
  46. Whisper-Based Speech Transcription from Videos Across Multiple Languages for Cross-Cultural Understanding

    Michael Picheny

    eess.AScs.CLarXiv:2609.11772v12026
  47. InFoBench: Evaluating Instruction Following Ability in Large Language Models

    Yiwei Qin, Kaiqiang Song, Yebowen Hu +7

    cs.CLcs.AIarXiv:2401.03601v12024
  48. A Large-Scale Chinese Short-Text Conversation Dataset

    Yida Wang, Pei Ke, Yinhe Zheng +4

    cs.CLarXiv:2008.03946v22020
  49. Why Does Post-Training Quantization Work?

    Yuxiang Chen, Michael Beyer, Jun Zhu +1

    cs.LGcs.CLarXiv:2609.11716v12026
  50. Survey on reinforcement learning for language processing

    Victor Uc-Cetina, Nicolas Navarro-Guerrero, Anabel Martin-Gonzalez +2

    cs.CLcs.AIcs.LGarXiv:2104.05565v32021
  51. VikingRAG: Accurate and Token-efficient Retrieval-augmented Generation over Structured Documents

    Peiyuan Gao, Gaoyuan Zhang, Haojie Qin +5

    cs.IRcs.AIcs.CLarXiv:2609.11390v12026
  52. Xiaomi-CocktailASR-1 Technical Report

    Yiru Zhang, Hang Su, Lichun Fan +10

    cs.SDcs.CLeess.ASarXiv:2609.11274v12026
  53. (Whose defaults?) Is artificial intelligence reorienting archaeological methods?

    Lorenzo Cardarelli, Roberto Ragno

    cs.CYcs.AIcs.CLarXiv:2609.11198v12026
  54. LILA: Calibration-Free Structured Pruning of Large Language Models via Latent Spectral Geometry

    Sankar Behera, Dhruv Singh, Anshika Agnihotri +3

    cs.LGcs.CLarXiv:2609.11163v12026
  55. Learning a bidirectional mapping between human whole-body motion and natural language using deep recurrent neural networks

    Matthias Plappert, Christian Mandery, Tamim Asfour

    cs.LGcs.CLcs.ROarXiv:1705.06400v22017
  56. Learning to Parse and Translate Improves Neural Machine Translation

    Akiko Eriguchi, Yoshimasa Tsuruoka, Kyunghyun Cho

    cs.CLarXiv:1702.03525v22017
  57. INDRA: A New AI Tool for Exploring Tobacco, Fossil Fuel, and Chemical Industry Archives

    Daniel Akselrad, Robert N. Proctor

    cs.DLcs.CLcs.CYarXiv:2609.11261v12026
  58. MUtE: A Dual Framework for Concept Erasure and Counterfactual Interventions

    Antoine Saillenfest

    cs.LGcs.CLarXiv:2609.11253v12026
  59. A Voice-Interactive Multi-Agent System for Smart Operating Rooms: Architecture Design and Key Technologies

    Tianxiang Zhou

    cs.AIcs.CLcs.HCarXiv:2609.11231v12026
  60. REVA: Reusable Evidence View Aggregation for Context-Efficient RAG Serving

    Tuan Nguyen, Qiran Hu, Banruo Liu +3

    cs.LGcs.CLcs.IRarXiv:2609.11209v12026