Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,841 to 6,900 of 11,260
Leak, Cheat, Repeat: Data Contamination and Evaluation Malpractices in Closed-Source LLMs
Simone Balloccu, Patrícia Schmidtová, Mateusz Lango +1
cs.CLcs.AIarXiv:2402.03927v22024Open Question Answering with Weakly Supervised Embedding Models
Antoine Bordes, Jason Weston, Nicolas Usunier
cs.CLcs.LGarXiv:1404.4326v12014Structured Pruning of Large Language Models
Ziheng Wang, Jeremy Wohlwend, Tao Lei
cs.CLcs.LGstat.MLarXiv:1910.04732v22019Pre-training of Graph Augmented Transformers for Medication Recommendation
Junyuan Shang, Tengfei Ma, Cao Xiao +1
cs.AIcs.CLcs.LGarXiv:1906.00346v22019Adversarial Removal of Demographic Attributes from Text Data
Yanai Elazar, Yoav Goldberg
cs.CLcs.LGstat.MLarXiv:1808.06640v22018Unified Named Entity Recognition as Word-Word Relation Classification
Jingye Li, Hao Fei, Jiang Liu +5
cs.CLarXiv:2112.10070v12021A Unified Model for Opinion Target Extraction and Target Sentiment Prediction
Xin Li, Lidong Bing, Piji Li +1
cs.CLarXiv:1811.05082v22018INSIDE: LLMs' Internal States Retain the Power of Hallucination Detection
Chao Chen, Kai Liu, Ze Chen +5
cs.CLarXiv:2402.03744v22024RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models
Cheng Niu, Yuanhao Wu, Juno Zhu +5
cs.CLarXiv:2401.00396v22023Effective Long-Context Scaling of Foundation Models
Wenhan Xiong, Jingyu Liu, Igor Molybog +18
cs.CLarXiv:2309.16039v32023Transfer Learning for Sequence Tagging with Hierarchical Recurrent Networks
Zhilin Yang, Ruslan Salakhutdinov, William W. Cohen
cs.CLcs.LGarXiv:1703.06345v12017COGS: A Compositional Generalization Challenge Based on Semantic Interpretation
Najoung Kim, Tal Linzen
cs.CLarXiv:2010.05465v12020Understanding Dataset Difficulty with $\mathcal{V}$-Usable Information
Kawin Ethayarajh, Yejin Choi, Swabha Swayamdipta
cs.CLcs.AIcs.LGarXiv:2110.08420v32021Ghost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory
Xizhou Zhu, Yuntao Chen, Hao Tian +10
cs.AIcs.CLcs.CVarXiv:2305.17144v22023xCOMET: Transparent Machine Translation Evaluation through Fine-grained Error Detection
Nuno M. Guerreiro, Ricardo Rei, Daan van Stigt +3
cs.CLarXiv:2310.10482v12023Double Embeddings and CNN-based Sequence Labeling for Aspect Extraction
Hu Xu, Bing Liu, Lei Shu +1
cs.CLarXiv:1805.04601v12018Are We Modeling the Task or the Annotator? An Investigation of Annotator Bias in Natural Language Understanding Datasets
Mor Geva, Yoav Goldberg, Jonathan Berant
cs.CLarXiv:1908.07898v22019The political ideology of conversational AI: Converging evidence on ChatGPT's pro-environmental, left-libertarian orientation
Jochen Hartmann, Jasper Schwenzow, Maximilian Witte
cs.CLcs.CYarXiv:2301.01768v12023Commonsense Knowledge Mining from Pretrained Models
Joshua Feldman, Joe Davison, Alexander M. Rush
cs.CLcs.AIcs.LGarXiv:1909.00505v12019The Effect of Sampling Temperature on Problem Solving in Large Language Models
Matthew Renze, Erhan Guven
cs.CLcs.AIarXiv:2402.05201v32024Improving Topic Models with Latent Feature Word Representations
Dat Quoc Nguyen, Richard Billingsley, Lan Du +1
cs.CLcs.IRcs.LGarXiv:1810.06306v12018Deep Joint Entity Disambiguation with Local Neural Attention
Octavian-Eugen Ganea, Thomas Hofmann
cs.CLarXiv:1704.04920v32017Towards Making the Most of ChatGPT for Machine Translation
Keqin Peng, Liang Ding, Qihuang Zhong +5
cs.CLarXiv:2303.13780v42023Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Xidong Feng, Ziyu Wan, Muning Wen +4
cs.LGcs.AIcs.CLarXiv:2309.17179v22023A Deep Architecture for Semantic Matching with Multiple Positional Sentence Representations
Shengxian Wan, Yanyan Lan, Jiafeng Guo +3
cs.AIcs.CLcs.NEarXiv:1511.08277v12015Relation Classification via Recurrent Neural Network
Dongxu Zhang, Dong Wang
cs.CLcs.LGcs.NEarXiv:1508.01006v22015Syntax-Directed Variational Autoencoder for Structured Data
Hanjun Dai, Yingtao Tian, Bo Dai +2
cs.LGcs.CLarXiv:1802.08786v12018Multilingual and Multi-Aspect Hate Speech Analysis
Nedjma Ousidhoum, Zizheng Lin, Hongming Zhang +2
cs.CLarXiv:1908.11049v12019Twitter as a Lifeline: Human-annotated Twitter Corpora for NLP of Crisis-related Messages
Muhammad Imran, Prasenjit Mitra, Carlos Castillo
cs.CLcs.CYcs.SIarXiv:1605.05894v22016Aspect Level Sentiment Classification with Attention-over-Attention Neural Networks
Binxuan Huang, Yanglan Ou, Kathleen M. Carley
cs.CLarXiv:1804.06536v12018Large Language Models for Mathematical Reasoning: Progresses and Challenges
Janice Ahn, Rishu Verma, Renze Lou +3
cs.CLarXiv:2402.00157v42024A Survey of Available Corpora for Building Data-Driven Dialogue Systems
Iulian Vlad Serban, Ryan Lowe, Peter Henderson +2
cs.CLcs.AIcs.HCarXiv:1512.05742v32015Clinically Accurate Chest X-Ray Report Generation
Guanxiong Liu, Tzu-Ming Harry Hsu, Matthew McDermott +4
cs.CVcs.CLarXiv:1904.02633v22019Data Selection for Language Models via Importance Resampling
Sang Michael Xie, Shibani Santurkar, Tengyu Ma +1
cs.CLcs.LGarXiv:2302.03169v32023Generating Sequences by Learning to Self-Correct
Sean Welleck, Ximing Lu, Peter West +4
cs.CLarXiv:2211.00053v12022OpenAGI: When LLM Meets Domain Experts
Yingqiang Ge, Wenyue Hua, Kai Mei +5
cs.AIcs.CLcs.LGarXiv:2304.04370v62023Analyzing and Mitigating Object Hallucination in Large Vision-Language Models
Yiyang Zhou, Chenhang Cui, Jaehong Yoon +5
cs.LGcs.CLcs.CVarXiv:2310.00754v22023Lawformer: A Pre-trained Language Model for Chinese Legal Long Documents
Chaojun Xiao, Xueyu Hu, Zhiyuan Liu +2
cs.CLarXiv:2105.03887v12021MQuAKE: Assessing Knowledge Editing in Language Models via Multi-Hop Questions
Zexuan Zhong, Zhengxuan Wu, Christopher D. Manning +2
cs.CLarXiv:2305.14795v32023SqueezeLLM: Dense-and-Sparse Quantization
Sehoon Kim, Coleman Hooper, Amir Gholami +5
cs.CLcs.LGarXiv:2306.07629v42023TrustLLM: Trustworthiness in Large Language Models
Yue Huang, Lichao Sun, Haoran Wang +67
cs.CLarXiv:2401.05561v62024RARR: Researching and Revising What Language Models Say, Using Language Models
Luyu Gao, Zhuyun Dai, Panupong Pasupat +8
cs.CLcs.AIcs.IRarXiv:2210.08726v32022Nomic Embed: Training a Reproducible Long Context Text Embedder
Zach Nussbaum, John X. Morris, Brandon Duderstadt +1
cs.CLcs.AIarXiv:2402.01613v22024Black-Box Tuning for Language-Model-as-a-Service
Tianxiang Sun, Yunfan Shao, Hong Qian +2
cs.CLcs.AIarXiv:2201.03514v42022MLE-bench: Evaluating Machine Learning Agents on Machine Learning Engineering
Jun Shern Chan, Neil Chowdhury, Oliver Jaffe +9
cs.CLarXiv:2410.07095v62024AgentTuning: Enabling Generalized Agent Abilities for LLMs
Aohan Zeng, Mingdao Liu, Rui Lu +4
cs.CLcs.AIcs.LGarXiv:2310.12823v22023Is GPT-3 a Good Data Annotator?
Bosheng Ding, Chengwei Qin, Linlin Liu +4
cs.CLarXiv:2212.10450v22022Gromov-Wasserstein Alignment of Word Embedding Spaces
David Alvarez-Melis, Tommi S. Jaakkola
cs.CLarXiv:1809.00013v12018Quality at a Glance: An Audit of Web-Crawled Multilingual Datasets
Julia Kreutzer, Isaac Caswell, Lisa Wang +49
cs.CLcs.AIarXiv:2103.12028v42021Attentive Pooling Networks
Cicero dos Santos, Ming Tan, Bing Xiang +1
cs.CLcs.LGarXiv:1602.03609v12016Scaling Relationship on Learning Mathematical Reasoning with Large Language Models
Zheng Yuan, Hongyi Yuan, Chengpeng Li +5
cs.CLarXiv:2308.01825v22023Controlling Linguistic Style Aspects in Neural Language Generation
Jessica Ficler, Yoav Goldberg
cs.CLarXiv:1707.02633v12017Text Generation from Knowledge Graphs with Graph Transformers
Rik Koncel-Kedziorski, Dhanush Bekal, Yi Luan +2
cs.CLarXiv:1904.02342v32019A Survey of Controllable Text Generation using Transformer-based Pre-trained Language Models
Hanqing Zhang, Haolin Song, Shaoyu Li +2
cs.CLarXiv:2201.05337v52022Exploring Architectures, Data and Units For Streaming End-to-End Speech Recognition with RNN-Transducer
Kanishka Rao, Haşim Sak, Rohit Prabhavalkar
cs.CLcs.SDeess.ASarXiv:1801.00841v12018Adversarial Feature Matching for Text Generation
Yizhe Zhang, Zhe Gan, Kai Fan +4
stat.MLcs.CLcs.LGarXiv:1706.03850v32017VeRA: Vector-based Random Matrix Adaptation
Dawid J. Kopiczko, Tijmen Blankevoort, Yuki M. Asano
cs.CLarXiv:2310.11454v22023Temporal Analysis of Language through Neural Language Models
Yoon Kim, Yi-I Chiu, Kentaro Hanaki +2
cs.CLarXiv:1405.3515v12014Robust Distortion-free Watermarks for Language Models
Rohith Kuditipudi, John Thickstun, Tatsunori Hashimoto +1
cs.LGcs.CLcs.CRarXiv:2307.15593v32023Unleashing the potential of prompt engineering for large language models
Banghao Chen, Zhaofeng Zhang, Nicolas Langrené +1
cs.CLcs.AIarXiv:2310.14735v62023