Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
5,101 to 5,160 of 11,259
MTOP: A Comprehensive Multilingual Task-Oriented Semantic Parsing Benchmark
Haoran Li, Abhinav Arora, Shuohui Chen +3
cs.CLcs.LGarXiv:2008.09335v22020From Production Traffic to Post-Training: Building a Self-Hosted LLM That Covers the Corporate Request Mix
Olga Tsymboi, Dmitrii Stoianov, Ramil Latypov +11
cs.CLarXiv:2609.01572v12026Handling Divergent Reference Texts when Evaluating Table-to-Text Generation
Bhuwan Dhingra, Manaal Faruqui, Ankur Parikh +3
cs.CLarXiv:1906.01081v12019A Streaming On-Device End-to-End Model Surpassing Server-Side Conventional Model Quality and Latency
Tara N. Sainath, Yanzhang He, Bo Li +26
cs.CLcs.LGcs.SDarXiv:2003.12710v22020Quantifying the Persona Effect in LLM Simulations
Tiancheng Hu, Nigel Collier
cs.CLcs.CYarXiv:2402.10811v22024Primer: Searching for Efficient Transformers for Language Modeling
David R. So, Wojciech Mańke, Hanxiao Liu +3
cs.LGcs.AIcs.CLarXiv:2109.08668v22021Language as a Latent Variable: Discrete Generative Models for Sentence Compression
Yishu Miao, Phil Blunsom
cs.CLcs.AIarXiv:1609.07317v22016Unsupervised pretraining transfers well across languages
Morgane Rivière, Armand Joulin, Pierre-Emmanuel Mazaré +1
eess.AScs.CLcs.LGarXiv:2002.02848v12020ChatGPT-4 Outperforms Experts and Crowd Workers in Annotating Political Twitter Messages with Zero-Shot Learning
Petter Törnberg
cs.CLcs.AIcs.SIarXiv:2304.06588v12023Do the Rewards Justify the Means? Measuring Trade-Offs Between Rewards and Ethical Behavior in the MACHIAVELLI Benchmark
Alexander Pan, Jun Shern Chan, Andy Zou +7
cs.LGcs.AIcs.CLarXiv:2304.03279v42023Pretraining-Based Natural Language Generation for Text Summarization
Haoyu Zhang, Jianjun Xu, Ji Wang
cs.CLcs.AIarXiv:1902.09243v22019Raise a Child in Large Language Model: Towards Effective and Generalizable Fine-tuning
Runxin Xu, Fuli Luo, Zhiyuan Zhang +4
cs.CLcs.AIarXiv:2109.05687v12021Neural Natural Language Inference Models Enhanced with External Knowledge
Qian Chen, Xiaodan Zhu, Zhen-Hua Ling +2
cs.CLarXiv:1711.04289v32017On the Predictive Power of Neural Language Models for Human Real-Time Comprehension Behavior
Ethan Gotlieb Wilcox, Jon Gauthier, Jennifer Hu +2
cs.CLarXiv:2006.01912v12020Recurrent Dropout without Memory Loss
Stanislau Semeniuta, Aliaksei Severyn, Erhardt Barth
cs.CLarXiv:1603.05118v22016Privacy- and Utility-Preserving Textual Analysis via Calibrated Multivariate Perturbations
Oluwaseyi Feyisetan, Borja Balle, Thomas Drake +1
cs.LGcs.CLcs.CRarXiv:1910.08902v12019Improving Entity Linking by Modeling Latent Relations between Mentions
Phong Le, Ivan Titov
cs.CLarXiv:1804.10637v12018MiNER: Fine-Tuned Biomedical Natural Language Processing for Malaria Disease Entity Recognition in Clinical Texts
V. S. Anoop, Devika N
cs.AIcs.CLarXiv:2609.00073v12026An Analysis of Hierarchical Text Classification Using Word Embeddings
Roger A. Stein, Patricia A. Jaques, Joao F. Valiati
cs.CLcs.AIcs.LGarXiv:1809.01771v12018Augmenting Language Models with Long-Term Memory
Weizhi Wang, Li Dong, Hao Cheng +4
cs.CLarXiv:2306.07174v12023Vision-Language Pre-training: Basics, Recent Advances, and Future Trends
Zhe Gan, Linjie Li, Chunyuan Li +3
cs.CVcs.CLarXiv:2210.09263v12022COCO-LM: Correcting and Contrasting Text Sequences for Language Model Pretraining
Yu Meng, Chenyan Xiong, Payal Bajaj +4
cs.CLcs.LGarXiv:2102.08473v22021BERT-of-Theseus: Compressing BERT by Progressive Module Replacing
Canwen Xu, Wangchunshu Zhou, Tao Ge +2
cs.CLcs.LGarXiv:2002.02925v42020Between words and characters: A Brief History of Open-Vocabulary Modeling and Tokenization in NLP
Sabrina J. Mielke, Zaid Alyafeai, Elizabeth Salesky +8
cs.CLcs.LGarXiv:2112.10508v12021Colossal-AI: A Unified Deep Learning System For Large-Scale Parallel Training
Shenggui Li, Hongxin Liu, Zhengda Bian +5
cs.LGcs.AIcs.CLarXiv:2110.14883v32021InCharacter: Evaluating Personality Fidelity in Role-Playing Agents through Psychological Interviews
Xintao Wang, Yunze Xiao, Jen-tse Huang +10
cs.CLarXiv:2310.17976v42023The Language Interpretability Tool: Extensible, Interactive Visualizations and Analysis for NLP Models
Ian Tenney, James Wexler, Jasmijn Bastings +8
cs.CLarXiv:2008.05122v12020Automatically Identifying Words That Can Serve as Labels for Few-Shot Text Classification
Timo Schick, Helmut Schmid, Hinrich Schütze
cs.CLcs.AIcs.LGarXiv:2010.13641v12020Detecting Hidden Behaviors in LLMs via Activation-matched Finetuning
Robin Haselhorst, Lucie Flek, Florian Mai
cs.CLcs.AIarXiv:2609.00351v12026From Distillation to Hard Negative Sampling: Making Sparse Neural IR Models More Effective
Thibault Formal, Carlos Lassance, Benjamin Piwowarski +1
cs.IRcs.CLarXiv:2205.04733v22022Mobile-Agent-v2: Mobile Device Operation Assistant with Effective Navigation via Multi-Agent Collaboration
Junyang Wang, Haiyang Xu, Haitao Jia +6
cs.CLcs.CVarXiv:2406.01014v12024Strategies for Structuring Story Generation
Angela Fan, Mike Lewis, Yann Dauphin
cs.CLarXiv:1902.01109v22019FlashRAG: A Modular Toolkit for Efficient Retrieval-Augmented Generation Research
Jiajie Jin, Yutao Zhu, Guanting Dong +7
cs.CLcs.IRarXiv:2405.13576v22024Multilingual Language Processing From Bytes
Dan Gillick, Cliff Brunk, Oriol Vinyals +1
cs.CLarXiv:1512.00103v22015Open-Domain Targeted Sentiment Analysis via Span-Based Extraction and Classification
Minghao Hu, Yuxing Peng, Zhen Huang +2
cs.CLarXiv:1906.03820v12019Identifying beneficial task relations for multi-task learning in deep neural networks
Joachim Bingel, Anders Søgaard
cs.CLarXiv:1702.08303v12017DEGREE: A Data-Efficient Generation-Based Event Extraction Model
I-Hung Hsu, Kuan-Hao Huang, Elizabeth Boschee +4
cs.CLcs.AIarXiv:2108.12724v32021Connecting the Dots: Document-level Neural Relation Extraction with Edge-oriented Graphs
Fenia Christopoulou, Makoto Miwa, Sophia Ananiadou
cs.CLarXiv:1909.00228v12019Fake News Early Detection: An Interdisciplinary Study
Xinyi Zhou, Atishay Jain, Vir V. Phoha +1
cs.CLcs.SIarXiv:1904.11679v22019Normalized and Geometry-Aware Self-Attention Network for Image Captioning
Longteng Guo, Jing Liu, Xinxin Zhu +3
cs.CVcs.CLcs.MMarXiv:2003.08897v12020CCAligned: A Massive Collection of Cross-Lingual Web-Document Pairs
Ahmed El-Kishky, Vishrav Chaudhary, Francisco Guzman +1
cs.CLcs.LGstat.MLarXiv:1911.06154v22019RENSA: Rich Environment Metadata to Navigate Shared and Distributed Endpoints for Automated Federated SPARQL Query Generation
Victor Eiti Yamamoto, Takeda Hideaki, Yamamoto Yasunori
cs.DBcs.CLarXiv:2608.28963v12026Towards Crafting Text Adversarial Samples
Suranjana Samanta, Sameep Mehta
cs.LGcs.AIcs.CLarXiv:1707.02812v12017Deductive Verification of Chain-of-Thought Reasoning
Zhan Ling, Yunhao Fang, Xuanlin Li +4
cs.CLcs.AIcs.LGarXiv:2306.03872v32023Attention Correctness in Neural Image Captioning
Chenxi Liu, Junhua Mao, Fei Sha +1
cs.CVcs.CLcs.LGarXiv:1605.09553v22016On the Impact of Various Types of Noise on Neural Machine Translation
Huda Khayrallah, Philipp Koehn
cs.CLarXiv:1805.12282v12018Knowledge Matters: Radiology Report Generation with General and Specific Knowledge
Shuxin Yang, Xian Wu, Shen Ge +2
eess.IVcs.CLcs.CVarXiv:2112.15009v22021When Are Tree Structures Necessary for Deep Learning of Representations?
Jiwei Li, Minh-Thang Luong, Dan Jurafsky +1
cs.AIcs.CLarXiv:1503.00185v52015The Parallelism Tradeoff: Limitations of Log-Precision Transformers
William Merrill, Ashish Sabharwal
cs.CCcs.CLarXiv:2207.00729v42022RankT5: Fine-Tuning T5 for Text Ranking with Ranking Losses
Honglei Zhuang, Zhen Qin, Rolf Jagerman +6
cs.IRcs.CLarXiv:2210.10634v12022Rate-Coding Bundle Memory: A Unified Model of Memory and Control for Symbolic Computation in the Brain
Teun van Gils, Rowan P. Sommers, Markus Ostarek +1
q-bio.NCcs.AIcs.CLarXiv:2608.29189v12026BIRD-History: A Benchmark for History-Driven Text-to-SQL with Fine-Grained Knowledge Annotations
Yunfan Zhou, Qiming Shi, Yizhou Yang +2
cs.AIcs.CLarXiv:2608.29345v12026Meta-StyleSpeech : Multi-Speaker Adaptive Text-to-Speech Generation
Dongchan Min, Dong Bok Lee, Eunho Yang +1
eess.AScs.CLcs.LGarXiv:2106.03153v32021Neurosymbolics for Data Engineering: Achieving Long Context Token Reduction Without Finetuning
Vishvesh Bhat
cs.CLcs.AIcs.LGarXiv:2609.00367v12026Concepts and Their Dynamics: A Quantum-Theoretic Modeling of Human Thought
Diederik Aerts, Liane Gabora, Sandro Sozzo
cs.AIcs.CLquant-pharXiv:1206.1069v22012Ultra-Fine Entity Typing
Eunsol Choi, Omer Levy, Yejin Choi +1
cs.CLcs.AIcs.LGarXiv:1807.04905v12018Removable and Irreducible: A Token-Cost Ledger for the Multilingual Tokenization Tax
Madhulatha Mandarapu, Sandeep Kunkunuru
cs.CLarXiv:2609.00378v12026Evaluating the Underlying Gender Bias in Contextualized Word Embeddings
Christine Basta, Marta R. Costa-jussà, Noe Casas
cs.CLcs.LGarXiv:1904.08783v12019Contextual LSTM (CLSTM) models for Large scale NLP tasks
Shalini Ghosh, Oriol Vinyals, Brian Strope +3
cs.CLarXiv:1602.06291v22016A Careful Examination of Large Language Model Performance on Grade School Arithmetic
Hugh Zhang, Jeff Da, Dean Lee +12
cs.CLcs.AIcs.LGarXiv:2405.00332v42024