Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,781 to 6,840 of 11,218
NaturalSpeech: End-to-End Text to Speech Synthesis with Human-Level Quality
Xu Tan, Jiawei Chen, Haohe Liu +11
eess.AScs.AIcs.CLarXiv:2205.04421v22022CLEAR: Contrastive Learning for Sentence Representation
Zhuofeng Wu, Sinong Wang, Jiatao Gu +3
cs.CLarXiv:2012.15466v12020German's Next Language Model
Branden Chan, Stefan Schweter, Timo Möller
cs.CLcs.LGarXiv:2010.10906v42020From Words to Watts: Benchmarking the Energy Costs of Large Language Model Inference
Siddharth Samsi, Dan Zhao, Joseph McDonald +7
cs.CLcs.DCarXiv:2310.03003v12023Empower Sequence Labeling with Task-Aware Neural Language Model
Liyuan Liu, Jingbo Shang, Frank F. Xu +4
cs.CLcs.LGarXiv:1709.04109v42017Combining Fact Extraction and Verification with Neural Semantic Matching Networks
Yixin Nie, Haonan Chen, Mohit Bansal
cs.CLcs.AIarXiv:1811.07039v12018CLIPstyler: Image Style Transfer with a Single Text Condition
Gihyun Kwon, Jong Chul Ye
cs.CVcs.CLeess.IVarXiv:2112.00374v32021To CoT or not to CoT? Chain-of-thought helps mainly on math and symbolic reasoning
Zayne Sprague, Fangcong Yin, Juan Diego Rodriguez +7
cs.CLcs.AIcs.LGarXiv:2409.12183v32024mGTE: Generalized Long-Context Text Representation and Reranking Models for Multilingual Text Retrieval
Xin Zhang, Yanzhao Zhang, Dingkun Long +10
cs.CLcs.IRarXiv:2407.19669v22024Baseline Needs More Love: On Simple Word-Embedding-Based Models and Associated Pooling Mechanisms
Dinghan Shen, Guoyin Wang, Wenlin Wang +6
cs.CLcs.AIcs.LGarXiv:1805.09843v12018Break the Sequential Dependency of LLM Inference Using Lookahead Decoding
Yichao Fu, Peter Bailis, Ion Stoica +1
cs.LGcs.CLarXiv:2402.02057v12024WinoGrande: An Adversarial Winograd Schema Challenge at Scale
Keisuke Sakaguchi, Ronan Le Bras, Chandra Bhagavatula +1
cs.CLarXiv:1907.10641v22019Recent Advances in Deep Learning Based Dialogue Systems: A Systematic Survey
Jinjie Ni, Tom Young, Vlad Pandelea +2
cs.CLcs.AIcs.IRarXiv:2105.04387v52021ReCoRD: Bridging the Gap between Human and Machine Commonsense Reading Comprehension
Sheng Zhang, Xiaodong Liu, Jingjing Liu +3
cs.CLarXiv:1810.12885v12018Multimodal Speech Emotion Recognition Using Audio and Text
Seunghyun Yoon, Seokhyun Byun, Kyomin Jung
cs.CLarXiv:1810.04635v12018Just Ask: Learning to Answer Questions from Millions of Narrated Videos
Antoine Yang, Antoine Miech, Josef Sivic +2
cs.CVcs.CLcs.LGarXiv:2012.00451v32020The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants
Lucas Bandarkar, Davis Liang, Benjamin Muller +7
cs.CLcs.AIcs.LGarXiv:2308.16884v22023TOD-BERT: Pre-trained Natural Language Understanding for Task-Oriented Dialogue
Chien-Sheng Wu, Steven Hoi, Richard Socher +1
cs.CLarXiv:2004.06871v32020Leak, Cheat, Repeat: Data Contamination and Evaluation Malpractices in Closed-Source LLMs
Simone Balloccu, Patrícia Schmidtová, Mateusz Lango +1
cs.CLcs.AIarXiv:2402.03927v22024Open Question Answering with Weakly Supervised Embedding Models
Antoine Bordes, Jason Weston, Nicolas Usunier
cs.CLcs.LGarXiv:1404.4326v12014Structured Pruning of Large Language Models
Ziheng Wang, Jeremy Wohlwend, Tao Lei
cs.CLcs.LGstat.MLarXiv:1910.04732v22019Pre-training of Graph Augmented Transformers for Medication Recommendation
Junyuan Shang, Tengfei Ma, Cao Xiao +1
cs.AIcs.CLcs.LGarXiv:1906.00346v22019Adversarial Removal of Demographic Attributes from Text Data
Yanai Elazar, Yoav Goldberg
cs.CLcs.LGstat.MLarXiv:1808.06640v22018Unified Named Entity Recognition as Word-Word Relation Classification
Jingye Li, Hao Fei, Jiang Liu +5
cs.CLarXiv:2112.10070v12021A Unified Model for Opinion Target Extraction and Target Sentiment Prediction
Xin Li, Lidong Bing, Piji Li +1
cs.CLarXiv:1811.05082v22018INSIDE: LLMs' Internal States Retain the Power of Hallucination Detection
Chao Chen, Kai Liu, Ze Chen +5
cs.CLarXiv:2402.03744v22024RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models
Cheng Niu, Yuanhao Wu, Juno Zhu +5
cs.CLarXiv:2401.00396v22023Effective Long-Context Scaling of Foundation Models
Wenhan Xiong, Jingyu Liu, Igor Molybog +18
cs.CLarXiv:2309.16039v32023Transfer Learning for Sequence Tagging with Hierarchical Recurrent Networks
Zhilin Yang, Ruslan Salakhutdinov, William W. Cohen
cs.CLcs.LGarXiv:1703.06345v12017COGS: A Compositional Generalization Challenge Based on Semantic Interpretation
Najoung Kim, Tal Linzen
cs.CLarXiv:2010.05465v12020Understanding Dataset Difficulty with $\mathcal{V}$-Usable Information
Kawin Ethayarajh, Yejin Choi, Swabha Swayamdipta
cs.CLcs.AIcs.LGarXiv:2110.08420v32021Ghost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory
Xizhou Zhu, Yuntao Chen, Hao Tian +10
cs.AIcs.CLcs.CVarXiv:2305.17144v22023xCOMET: Transparent Machine Translation Evaluation through Fine-grained Error Detection
Nuno M. Guerreiro, Ricardo Rei, Daan van Stigt +3
cs.CLarXiv:2310.10482v12023Double Embeddings and CNN-based Sequence Labeling for Aspect Extraction
Hu Xu, Bing Liu, Lei Shu +1
cs.CLarXiv:1805.04601v12018Are We Modeling the Task or the Annotator? An Investigation of Annotator Bias in Natural Language Understanding Datasets
Mor Geva, Yoav Goldberg, Jonathan Berant
cs.CLarXiv:1908.07898v22019The political ideology of conversational AI: Converging evidence on ChatGPT's pro-environmental, left-libertarian orientation
Jochen Hartmann, Jasper Schwenzow, Maximilian Witte
cs.CLcs.CYarXiv:2301.01768v12023Commonsense Knowledge Mining from Pretrained Models
Joshua Feldman, Joe Davison, Alexander M. Rush
cs.CLcs.AIcs.LGarXiv:1909.00505v12019The Effect of Sampling Temperature on Problem Solving in Large Language Models
Matthew Renze, Erhan Guven
cs.CLcs.AIarXiv:2402.05201v32024Improving Topic Models with Latent Feature Word Representations
Dat Quoc Nguyen, Richard Billingsley, Lan Du +1
cs.CLcs.IRcs.LGarXiv:1810.06306v12018Deep Joint Entity Disambiguation with Local Neural Attention
Octavian-Eugen Ganea, Thomas Hofmann
cs.CLarXiv:1704.04920v32017Towards Making the Most of ChatGPT for Machine Translation
Keqin Peng, Liang Ding, Qihuang Zhong +5
cs.CLarXiv:2303.13780v42023Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Xidong Feng, Ziyu Wan, Muning Wen +4
cs.LGcs.AIcs.CLarXiv:2309.17179v22023A Deep Architecture for Semantic Matching with Multiple Positional Sentence Representations
Shengxian Wan, Yanyan Lan, Jiafeng Guo +3
cs.AIcs.CLcs.NEarXiv:1511.08277v12015Relation Classification via Recurrent Neural Network
Dongxu Zhang, Dong Wang
cs.CLcs.LGcs.NEarXiv:1508.01006v22015Syntax-Directed Variational Autoencoder for Structured Data
Hanjun Dai, Yingtao Tian, Bo Dai +2
cs.LGcs.CLarXiv:1802.08786v12018Multilingual and Multi-Aspect Hate Speech Analysis
Nedjma Ousidhoum, Zizheng Lin, Hongming Zhang +2
cs.CLarXiv:1908.11049v12019Twitter as a Lifeline: Human-annotated Twitter Corpora for NLP of Crisis-related Messages
Muhammad Imran, Prasenjit Mitra, Carlos Castillo
cs.CLcs.CYcs.SIarXiv:1605.05894v22016Aspect Level Sentiment Classification with Attention-over-Attention Neural Networks
Binxuan Huang, Yanglan Ou, Kathleen M. Carley
cs.CLarXiv:1804.06536v12018Large Language Models for Mathematical Reasoning: Progresses and Challenges
Janice Ahn, Rishu Verma, Renze Lou +3
cs.CLarXiv:2402.00157v42024A Survey of Available Corpora for Building Data-Driven Dialogue Systems
Iulian Vlad Serban, Ryan Lowe, Peter Henderson +2
cs.CLcs.AIcs.HCarXiv:1512.05742v32015Clinically Accurate Chest X-Ray Report Generation
Guanxiong Liu, Tzu-Ming Harry Hsu, Matthew McDermott +4
cs.CVcs.CLarXiv:1904.02633v22019Data Selection for Language Models via Importance Resampling
Sang Michael Xie, Shibani Santurkar, Tengyu Ma +1
cs.CLcs.LGarXiv:2302.03169v32023Generating Sequences by Learning to Self-Correct
Sean Welleck, Ximing Lu, Peter West +4
cs.CLarXiv:2211.00053v12022OpenAGI: When LLM Meets Domain Experts
Yingqiang Ge, Wenyue Hua, Kai Mei +5
cs.AIcs.CLcs.LGarXiv:2304.04370v62023Analyzing and Mitigating Object Hallucination in Large Vision-Language Models
Yiyang Zhou, Chenhang Cui, Jaehong Yoon +5
cs.LGcs.CLcs.CVarXiv:2310.00754v22023Lawformer: A Pre-trained Language Model for Chinese Legal Long Documents
Chaojun Xiao, Xueyu Hu, Zhiyuan Liu +2
cs.CLarXiv:2105.03887v12021MQuAKE: Assessing Knowledge Editing in Language Models via Multi-Hop Questions
Zexuan Zhong, Zhengxuan Wu, Christopher D. Manning +2
cs.CLarXiv:2305.14795v32023SqueezeLLM: Dense-and-Sparse Quantization
Sehoon Kim, Coleman Hooper, Amir Gholami +5
cs.CLcs.LGarXiv:2306.07629v42023TrustLLM: Trustworthiness in Large Language Models
Yue Huang, Lichao Sun, Haoran Wang +67
cs.CLarXiv:2401.05561v62024RARR: Researching and Revising What Language Models Say, Using Language Models
Luyu Gao, Zhuyun Dai, Panupong Pasupat +8
cs.CLcs.AIcs.IRarXiv:2210.08726v32022