Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
841 to 900 of 11,277
wav2vec: Unsupervised Pre-training for Speech Recognition
Steffen Schneider, Alexei Baevski, Ronan Collobert +1
cs.CLarXiv:1904.05862v42019Rethinking Interpretability in the Era of Large Language Models
Chandan Singh, Jeevana Priya Inala, Michel Galley +2
cs.CLcs.AIcs.LGarXiv:2402.01761v12024Compression of Deep Learning Models for Text: A Survey
Manish Gupta, Puneet Agrawal
cs.CLcs.AIcs.CVarXiv:2008.05221v42020A Panoramic Survey of Natural Language Processing in the Arab World
Kareem Darwish, Nizar Habash, Mourad Abbas +9
cs.CLarXiv:2011.12631v32020SpikeGPT: Generative Pre-trained Language Model with Spiking Neural Networks
Rui-Jie Zhu, Qihang Zhao, Guoqi Li +1
cs.CLcs.LGcs.NEarXiv:2302.13939v52023ClinicalGPT: Large Language Models Finetuned with Diverse Medical Data and Comprehensive Evaluation
Guangyu Wang, Guoxing Yang, Zongxin Du +2
cs.CLarXiv:2306.09968v12023The Ultimate Guide to Fine-Tuning LLMs from Basics to Breakthroughs: An Exhaustive Review of Technologies, Research, Best Practices, Applied Research Challenges and Opportunities
Venkatesh Balavadhani Parthasarathy, Ahtsham Zafar, Aafaq Khan +1
cs.LGcs.CLarXiv:2408.13296v32024LongRAG: Enhancing Retrieval-Augmented Generation with Long-context LLMs
Ziyan Jiang, Xueguang Ma, Wenhu Chen
cs.CLcs.AIarXiv:2406.15319v32024Exploring LLM-based Agents for Root Cause Analysis
Devjeet Roy, Xuchao Zhang, Rashi Bhave +4
cs.SEcs.CLcs.LGarXiv:2403.04123v12024Lifted Rule Injection for Relation Embeddings
Thomas Demeester, Tim Rocktäschel, Sebastian Riedel
cs.LGcs.AIcs.CLarXiv:1606.08359v22016Speech Recognition with Augmented Synthesized Speech
Andrew Rosenberg, Yu Zhang, Bhuvana Ramabhadran +4
cs.CLcs.SDeess.ASarXiv:1909.11699v12019An analysis of Twitter messages in the 2011 Tohoku Earthquake
Son Doan, Bao-Khanh Ho Vo, Nigel Collier
cs.SIcs.CLphysics.soc-pharXiv:1109.1618v12011Drug-Drug Interaction Extraction from Biomedical Text Using Long Short Term Memory Network
Sunil Kumar Sahu, Ashish Anand
cs.CLarXiv:1701.08303v22017"HOT" ChatGPT: The promise of ChatGPT in detecting and discriminating hateful, offensive, and toxic comments on social media
Lingyao Li, Lizhou Fan, Shubham Atreja +1
cs.CLcs.AIcs.HCarXiv:2304.10619v12023Backpropagation through Signal Temporal Logic Specifications: Infusing Logical Structure into Gradient-Based Methods
Karen Leung, Nikos Aréchiga, Marco Pavone
eess.SYcs.CLcs.LOarXiv:2008.00097v32020Imagination improves Multimodal Translation
Desmond Elliott, Ákos Kádár
cs.CLcs.CVarXiv:1705.04350v22017M6: A Chinese Multimodal Pretrainer
Junyang Lin, Rui Men, An Yang +22
cs.CLarXiv:2103.00823v42021Deep Multitask Learning for Semantic Dependency Parsing
Hao Peng, Sam Thomson, Noah A. Smith
cs.CLarXiv:1704.06855v22017Contextual Parameter Generation for Universal Neural Machine Translation
Emmanouil Antonios Platanios, Mrinmaya Sachan, Graham Neubig +1
cs.CLcs.LGstat.MLarXiv:1808.08493v12018Language Models that Seek for Knowledge: Modular Search & Generation for Dialogue and Prompt Completion
Kurt Shuster, Mojtaba Komeili, Leonard Adolphs +3
cs.CLcs.AIarXiv:2203.13224v22022ChatTime: A Unified Multimodal Time Series Foundation Model Bridging Numerical and Textual Data
Chengsen Wang, Qi Qi, Jingyu Wang +5
cs.CLcs.LGarXiv:2412.11376v12024Neural data-to-text generation: A comparison between pipeline and end-to-end architectures
Thiago Castro Ferreira, Chris van der Lee, Emiel van Miltenburg +1
cs.CLarXiv:1908.09022v22019Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs
Rui Yang, Ruomeng Ding, Yong Lin +2
cs.CLcs.AIarXiv:2406.10216v22024MindTopo: Can Foundation Models Reason in Topological Space?
Yunfei Ge, Anbang Liu, Qineng Wang +9
cs.AIcs.CLcs.CVarXiv:2609.11900v12026Improving speaker discrimination of target speech extraction with time-domain SpeakerBeam
Marc Delcroix, Tsubasa Ochiai, Katerina Zmolikova +4
eess.AScs.CLcs.SDarXiv:2001.08378v12020How would Stance Detection Techniques Evolve after the Launch of ChatGPT?
Bowen Zhang, Daijun Ding, Liwen Jing +2
cs.CLarXiv:2212.14548v42022UmlsBERT: Clinical Domain Knowledge Augmentation of Contextual Embeddings Using the Unified Medical Language System Metathesaurus
George Michalopoulos, Yuanxin Wang, Hussam Kaka +2
cs.CLcs.AIcs.LGarXiv:2010.10391v52020EmoBERTa: Speaker-Aware Emotion Recognition in Conversation with RoBERTa
Taewoon Kim, Piek Vossen
cs.CLarXiv:2108.12009v12021Universal Jailbreak Backdoors from Poisoned Human Feedback
Javier Rando, Florian Tramèr
cs.AIcs.CLcs.CRarXiv:2311.14455v42023A Question Answering Approach to Emotion Cause Extraction
Lin Gui, Jiannan Hu, Yulan He +3
cs.CLarXiv:1708.05482v22017CALF: Aligning LLMs for Time Series Forecasting via Cross-modal Fine-Tuning
Peiyuan Liu, Hang Guo, Tao Dai +5
cs.LGcs.CLarXiv:2403.07300v32024Neural Language Correction with Character-Based Attention
Ziang Xie, Anand Avati, Naveen Arivazhagan +2
cs.CLcs.AIarXiv:1603.09727v12016Data Scarcity and Model Sparsity: Mixtures-of-Experts Overfit More to Repeated Data
Atindra Jha, Margaret Li, Jure Leskovec +2
cs.LGcs.CLarXiv:2609.11917v12026Biology-in-the-loop: Amortized Adaptive Hit Discovery in CRISPR Screens
Carl Edwards, Edward De Brouwer, Xiner Li +5
q-bio.QMcs.AIcs.CLarXiv:2609.11877v12026Abstraction Agent
Boning Li, Longbo Huang
cs.MAcs.AIcs.CLarXiv:2609.04303v12026BERT for Evidence Retrieval and Claim Verification
Amir Soleimani, Christof Monz, Marcel Worring
cs.CLarXiv:1910.02655v12019RetroThinker: Enabling Retrospective Thinking in Speech LLMs
Yi-Jen Shih, Puyuan Peng, Abdelrahman Mohamed +1
eess.AScs.AIcs.CLarXiv:2609.11864v12026A Unified Per-Token Gating Family for On-Policy Distillation: FKL/RKL Mixing with Multi-Channel and Bias Coefficients
Suwan Wu, Yumeng Lin, Pengcheng Yuan +1
cs.AIcs.CLcs.LGarXiv:2609.11768v12026SIRF: A Spec-Internalized Risk Foundation Model for Industrial Content Risk Control
Suwan Wu, Yumeng Lin, Pengcheng Yuan +1
cs.AIcs.CLcs.LGarXiv:2609.11752v12026A Systematic Survey and Critical Review on Evaluating Large Language Models: Challenges, Limitations, and Recommendations
Md Tahmid Rahman Laskar, Sawsan Alqahtani, M Saiful Bari +10
cs.CLcs.AIcs.LGarXiv:2407.04069v22024The Semantic Elevation Operator and the Closure of the Undecidable Class under Preservation
Jose Pascual Gumbau Mezquita
cs.LOcs.AIcs.CLarXiv:2609.11326v12026The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement
Yi Duan, Ying Liu, Zirui Tang +30
cs.LGcs.AIcs.CLarXiv:2609.11873v12026SpecGuard: Inference-Time Backdoor Detection For Free
Rui Wen, Ahmed Salem, Andrew Paverd +2
cs.CRcs.CLarXiv:2609.11799v12026Nemotron-CC: Transforming Common Crawl into a Refined Long-Horizon Pretraining Dataset
Dan Su, Kezhi Kong, Ying Lin +6
cs.CLarXiv:2412.02595v22024Neural Segmental Hypergraphs for Overlapping Mention Recognition
Bailin Wang, Wei Lu
cs.CLarXiv:1810.01817v12018Whisper-Based Speech Transcription from Videos Across Multiple Languages for Cross-Cultural Understanding
Michael Picheny
eess.AScs.CLarXiv:2609.11772v12026InFoBench: Evaluating Instruction Following Ability in Large Language Models
Yiwei Qin, Kaiqiang Song, Yebowen Hu +7
cs.CLcs.AIarXiv:2401.03601v12024A Large-Scale Chinese Short-Text Conversation Dataset
Yida Wang, Pei Ke, Yinhe Zheng +4
cs.CLarXiv:2008.03946v22020Why Does Post-Training Quantization Work?
Yuxiang Chen, Michael Beyer, Jun Zhu +1
cs.LGcs.CLarXiv:2609.11716v12026Survey on reinforcement learning for language processing
Victor Uc-Cetina, Nicolas Navarro-Guerrero, Anabel Martin-Gonzalez +2
cs.CLcs.AIcs.LGarXiv:2104.05565v32021VikingRAG: Accurate and Token-efficient Retrieval-augmented Generation over Structured Documents
Peiyuan Gao, Gaoyuan Zhang, Haojie Qin +5
cs.IRcs.AIcs.CLarXiv:2609.11390v12026Xiaomi-CocktailASR-1 Technical Report
Yiru Zhang, Hang Su, Lichun Fan +10
cs.SDcs.CLeess.ASarXiv:2609.11274v12026(Whose defaults?) Is artificial intelligence reorienting archaeological methods?
Lorenzo Cardarelli, Roberto Ragno
cs.CYcs.AIcs.CLarXiv:2609.11198v12026LILA: Calibration-Free Structured Pruning of Large Language Models via Latent Spectral Geometry
Sankar Behera, Dhruv Singh, Anshika Agnihotri +3
cs.LGcs.CLarXiv:2609.11163v12026Learning a bidirectional mapping between human whole-body motion and natural language using deep recurrent neural networks
Matthias Plappert, Christian Mandery, Tamim Asfour
cs.LGcs.CLcs.ROarXiv:1705.06400v22017Learning to Parse and Translate Improves Neural Machine Translation
Akiko Eriguchi, Yoshimasa Tsuruoka, Kyunghyun Cho
cs.CLarXiv:1702.03525v22017INDRA: A New AI Tool for Exploring Tobacco, Fossil Fuel, and Chemical Industry Archives
Daniel Akselrad, Robert N. Proctor
cs.DLcs.CLcs.CYarXiv:2609.11261v12026MUtE: A Dual Framework for Concept Erasure and Counterfactual Interventions
Antoine Saillenfest
cs.LGcs.CLarXiv:2609.11253v12026A Voice-Interactive Multi-Agent System for Smart Operating Rooms: Architecture Design and Key Technologies
Tianxiang Zhou
cs.AIcs.CLcs.HCarXiv:2609.11231v12026REVA: Reusable Evidence View Aggregation for Context-Efficient RAG Serving
Tuan Nguyen, Qiran Hu, Banruo Liu +3
cs.LGcs.CLcs.IRarXiv:2609.11209v12026