Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,781 to 6,840 of 11,218

  1. NaturalSpeech: End-to-End Text to Speech Synthesis with Human-Level Quality

    Xu Tan, Jiawei Chen, Haohe Liu +11

    eess.AScs.AIcs.CLarXiv:2205.04421v22022
  2. CLEAR: Contrastive Learning for Sentence Representation

    Zhuofeng Wu, Sinong Wang, Jiatao Gu +3

    cs.CLarXiv:2012.15466v12020
  3. German's Next Language Model

    Branden Chan, Stefan Schweter, Timo Möller

    cs.CLcs.LGarXiv:2010.10906v42020
  4. From Words to Watts: Benchmarking the Energy Costs of Large Language Model Inference

    Siddharth Samsi, Dan Zhao, Joseph McDonald +7

    cs.CLcs.DCarXiv:2310.03003v12023
  5. Empower Sequence Labeling with Task-Aware Neural Language Model

    Liyuan Liu, Jingbo Shang, Frank F. Xu +4

    cs.CLcs.LGarXiv:1709.04109v42017
  6. Combining Fact Extraction and Verification with Neural Semantic Matching Networks

    Yixin Nie, Haonan Chen, Mohit Bansal

    cs.CLcs.AIarXiv:1811.07039v12018
  7. CLIPstyler: Image Style Transfer with a Single Text Condition

    Gihyun Kwon, Jong Chul Ye

    cs.CVcs.CLeess.IVarXiv:2112.00374v32021
  8. To CoT or not to CoT? Chain-of-thought helps mainly on math and symbolic reasoning

    Zayne Sprague, Fangcong Yin, Juan Diego Rodriguez +7

    cs.CLcs.AIcs.LGarXiv:2409.12183v32024
  9. mGTE: Generalized Long-Context Text Representation and Reranking Models for Multilingual Text Retrieval

    Xin Zhang, Yanzhao Zhang, Dingkun Long +10

    cs.CLcs.IRarXiv:2407.19669v22024
  10. Baseline Needs More Love: On Simple Word-Embedding-Based Models and Associated Pooling Mechanisms

    Dinghan Shen, Guoyin Wang, Wenlin Wang +6

    cs.CLcs.AIcs.LGarXiv:1805.09843v12018
  11. Break the Sequential Dependency of LLM Inference Using Lookahead Decoding

    Yichao Fu, Peter Bailis, Ion Stoica +1

    cs.LGcs.CLarXiv:2402.02057v12024
  12. WinoGrande: An Adversarial Winograd Schema Challenge at Scale

    Keisuke Sakaguchi, Ronan Le Bras, Chandra Bhagavatula +1

    cs.CLarXiv:1907.10641v22019
  13. Recent Advances in Deep Learning Based Dialogue Systems: A Systematic Survey

    Jinjie Ni, Tom Young, Vlad Pandelea +2

    cs.CLcs.AIcs.IRarXiv:2105.04387v52021
  14. ReCoRD: Bridging the Gap between Human and Machine Commonsense Reading Comprehension

    Sheng Zhang, Xiaodong Liu, Jingjing Liu +3

    cs.CLarXiv:1810.12885v12018
  15. Multimodal Speech Emotion Recognition Using Audio and Text

    Seunghyun Yoon, Seokhyun Byun, Kyomin Jung

    cs.CLarXiv:1810.04635v12018
  16. Just Ask: Learning to Answer Questions from Millions of Narrated Videos

    Antoine Yang, Antoine Miech, Josef Sivic +2

    cs.CVcs.CLcs.LGarXiv:2012.00451v32020
  17. The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

    Lucas Bandarkar, Davis Liang, Benjamin Muller +7

    cs.CLcs.AIcs.LGarXiv:2308.16884v22023
  18. TOD-BERT: Pre-trained Natural Language Understanding for Task-Oriented Dialogue

    Chien-Sheng Wu, Steven Hoi, Richard Socher +1

    cs.CLarXiv:2004.06871v32020
  19. Leak, Cheat, Repeat: Data Contamination and Evaluation Malpractices in Closed-Source LLMs

    Simone Balloccu, Patrícia Schmidtová, Mateusz Lango +1

    cs.CLcs.AIarXiv:2402.03927v22024
  20. Open Question Answering with Weakly Supervised Embedding Models

    Antoine Bordes, Jason Weston, Nicolas Usunier

    cs.CLcs.LGarXiv:1404.4326v12014
  21. Structured Pruning of Large Language Models

    Ziheng Wang, Jeremy Wohlwend, Tao Lei

    cs.CLcs.LGstat.MLarXiv:1910.04732v22019
  22. Pre-training of Graph Augmented Transformers for Medication Recommendation

    Junyuan Shang, Tengfei Ma, Cao Xiao +1

    cs.AIcs.CLcs.LGarXiv:1906.00346v22019
  23. Adversarial Removal of Demographic Attributes from Text Data

    Yanai Elazar, Yoav Goldberg

    cs.CLcs.LGstat.MLarXiv:1808.06640v22018
  24. Unified Named Entity Recognition as Word-Word Relation Classification

    Jingye Li, Hao Fei, Jiang Liu +5

    cs.CLarXiv:2112.10070v12021
  25. A Unified Model for Opinion Target Extraction and Target Sentiment Prediction

    Xin Li, Lidong Bing, Piji Li +1

    cs.CLarXiv:1811.05082v22018
  26. INSIDE: LLMs' Internal States Retain the Power of Hallucination Detection

    Chao Chen, Kai Liu, Ze Chen +5

    cs.CLarXiv:2402.03744v22024
  27. RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models

    Cheng Niu, Yuanhao Wu, Juno Zhu +5

    cs.CLarXiv:2401.00396v22023
  28. Effective Long-Context Scaling of Foundation Models

    Wenhan Xiong, Jingyu Liu, Igor Molybog +18

    cs.CLarXiv:2309.16039v32023
  29. Transfer Learning for Sequence Tagging with Hierarchical Recurrent Networks

    Zhilin Yang, Ruslan Salakhutdinov, William W. Cohen

    cs.CLcs.LGarXiv:1703.06345v12017
  30. COGS: A Compositional Generalization Challenge Based on Semantic Interpretation

    Najoung Kim, Tal Linzen

    cs.CLarXiv:2010.05465v12020
  31. Understanding Dataset Difficulty with $\mathcal{V}$-Usable Information

    Kawin Ethayarajh, Yejin Choi, Swabha Swayamdipta

    cs.CLcs.AIcs.LGarXiv:2110.08420v32021
  32. Ghost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory

    Xizhou Zhu, Yuntao Chen, Hao Tian +10

    cs.AIcs.CLcs.CVarXiv:2305.17144v22023
  33. xCOMET: Transparent Machine Translation Evaluation through Fine-grained Error Detection

    Nuno M. Guerreiro, Ricardo Rei, Daan van Stigt +3

    cs.CLarXiv:2310.10482v12023
  34. Double Embeddings and CNN-based Sequence Labeling for Aspect Extraction

    Hu Xu, Bing Liu, Lei Shu +1

    cs.CLarXiv:1805.04601v12018
  35. Are We Modeling the Task or the Annotator? An Investigation of Annotator Bias in Natural Language Understanding Datasets

    Mor Geva, Yoav Goldberg, Jonathan Berant

    cs.CLarXiv:1908.07898v22019
  36. The political ideology of conversational AI: Converging evidence on ChatGPT's pro-environmental, left-libertarian orientation

    Jochen Hartmann, Jasper Schwenzow, Maximilian Witte

    cs.CLcs.CYarXiv:2301.01768v12023
  37. Commonsense Knowledge Mining from Pretrained Models

    Joshua Feldman, Joe Davison, Alexander M. Rush

    cs.CLcs.AIcs.LGarXiv:1909.00505v12019
  38. The Effect of Sampling Temperature on Problem Solving in Large Language Models

    Matthew Renze, Erhan Guven

    cs.CLcs.AIarXiv:2402.05201v32024
  39. Improving Topic Models with Latent Feature Word Representations

    Dat Quoc Nguyen, Richard Billingsley, Lan Du +1

    cs.CLcs.IRcs.LGarXiv:1810.06306v12018
  40. Deep Joint Entity Disambiguation with Local Neural Attention

    Octavian-Eugen Ganea, Thomas Hofmann

    cs.CLarXiv:1704.04920v32017
  41. Towards Making the Most of ChatGPT for Machine Translation

    Keqin Peng, Liang Ding, Qihuang Zhong +5

    cs.CLarXiv:2303.13780v42023
  42. Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training

    Xidong Feng, Ziyu Wan, Muning Wen +4

    cs.LGcs.AIcs.CLarXiv:2309.17179v22023
  43. A Deep Architecture for Semantic Matching with Multiple Positional Sentence Representations

    Shengxian Wan, Yanyan Lan, Jiafeng Guo +3

    cs.AIcs.CLcs.NEarXiv:1511.08277v12015
  44. Relation Classification via Recurrent Neural Network

    Dongxu Zhang, Dong Wang

    cs.CLcs.LGcs.NEarXiv:1508.01006v22015
  45. Syntax-Directed Variational Autoencoder for Structured Data

    Hanjun Dai, Yingtao Tian, Bo Dai +2

    cs.LGcs.CLarXiv:1802.08786v12018
  46. Multilingual and Multi-Aspect Hate Speech Analysis

    Nedjma Ousidhoum, Zizheng Lin, Hongming Zhang +2

    cs.CLarXiv:1908.11049v12019
  47. Twitter as a Lifeline: Human-annotated Twitter Corpora for NLP of Crisis-related Messages

    Muhammad Imran, Prasenjit Mitra, Carlos Castillo

    cs.CLcs.CYcs.SIarXiv:1605.05894v22016
  48. Aspect Level Sentiment Classification with Attention-over-Attention Neural Networks

    Binxuan Huang, Yanglan Ou, Kathleen M. Carley

    cs.CLarXiv:1804.06536v12018
  49. Large Language Models for Mathematical Reasoning: Progresses and Challenges

    Janice Ahn, Rishu Verma, Renze Lou +3

    cs.CLarXiv:2402.00157v42024
  50. A Survey of Available Corpora for Building Data-Driven Dialogue Systems

    Iulian Vlad Serban, Ryan Lowe, Peter Henderson +2

    cs.CLcs.AIcs.HCarXiv:1512.05742v32015
  51. Clinically Accurate Chest X-Ray Report Generation

    Guanxiong Liu, Tzu-Ming Harry Hsu, Matthew McDermott +4

    cs.CVcs.CLarXiv:1904.02633v22019
  52. Data Selection for Language Models via Importance Resampling

    Sang Michael Xie, Shibani Santurkar, Tengyu Ma +1

    cs.CLcs.LGarXiv:2302.03169v32023
  53. Generating Sequences by Learning to Self-Correct

    Sean Welleck, Ximing Lu, Peter West +4

    cs.CLarXiv:2211.00053v12022
  54. OpenAGI: When LLM Meets Domain Experts

    Yingqiang Ge, Wenyue Hua, Kai Mei +5

    cs.AIcs.CLcs.LGarXiv:2304.04370v62023
  55. Analyzing and Mitigating Object Hallucination in Large Vision-Language Models

    Yiyang Zhou, Chenhang Cui, Jaehong Yoon +5

    cs.LGcs.CLcs.CVarXiv:2310.00754v22023
  56. Lawformer: A Pre-trained Language Model for Chinese Legal Long Documents

    Chaojun Xiao, Xueyu Hu, Zhiyuan Liu +2

    cs.CLarXiv:2105.03887v12021
  57. MQuAKE: Assessing Knowledge Editing in Language Models via Multi-Hop Questions

    Zexuan Zhong, Zhengxuan Wu, Christopher D. Manning +2

    cs.CLarXiv:2305.14795v32023
  58. SqueezeLLM: Dense-and-Sparse Quantization

    Sehoon Kim, Coleman Hooper, Amir Gholami +5

    cs.CLcs.LGarXiv:2306.07629v42023
  59. TrustLLM: Trustworthiness in Large Language Models

    Yue Huang, Lichao Sun, Haoran Wang +67

    cs.CLarXiv:2401.05561v62024
  60. RARR: Researching and Revising What Language Models Say, Using Language Models

    Luyu Gao, Zhuyun Dai, Panupong Pasupat +8

    cs.CLcs.AIcs.IRarXiv:2210.08726v32022