Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,841 to 6,900 of 11,260

  1. Leak, Cheat, Repeat: Data Contamination and Evaluation Malpractices in Closed-Source LLMs

    Simone Balloccu, Patrícia Schmidtová, Mateusz Lango +1

    cs.CLcs.AIarXiv:2402.03927v22024
  2. Open Question Answering with Weakly Supervised Embedding Models

    Antoine Bordes, Jason Weston, Nicolas Usunier

    cs.CLcs.LGarXiv:1404.4326v12014
  3. Structured Pruning of Large Language Models

    Ziheng Wang, Jeremy Wohlwend, Tao Lei

    cs.CLcs.LGstat.MLarXiv:1910.04732v22019
  4. Pre-training of Graph Augmented Transformers for Medication Recommendation

    Junyuan Shang, Tengfei Ma, Cao Xiao +1

    cs.AIcs.CLcs.LGarXiv:1906.00346v22019
  5. Adversarial Removal of Demographic Attributes from Text Data

    Yanai Elazar, Yoav Goldberg

    cs.CLcs.LGstat.MLarXiv:1808.06640v22018
  6. Unified Named Entity Recognition as Word-Word Relation Classification

    Jingye Li, Hao Fei, Jiang Liu +5

    cs.CLarXiv:2112.10070v12021
  7. A Unified Model for Opinion Target Extraction and Target Sentiment Prediction

    Xin Li, Lidong Bing, Piji Li +1

    cs.CLarXiv:1811.05082v22018
  8. INSIDE: LLMs' Internal States Retain the Power of Hallucination Detection

    Chao Chen, Kai Liu, Ze Chen +5

    cs.CLarXiv:2402.03744v22024
  9. RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models

    Cheng Niu, Yuanhao Wu, Juno Zhu +5

    cs.CLarXiv:2401.00396v22023
  10. Effective Long-Context Scaling of Foundation Models

    Wenhan Xiong, Jingyu Liu, Igor Molybog +18

    cs.CLarXiv:2309.16039v32023
  11. Transfer Learning for Sequence Tagging with Hierarchical Recurrent Networks

    Zhilin Yang, Ruslan Salakhutdinov, William W. Cohen

    cs.CLcs.LGarXiv:1703.06345v12017
  12. COGS: A Compositional Generalization Challenge Based on Semantic Interpretation

    Najoung Kim, Tal Linzen

    cs.CLarXiv:2010.05465v12020
  13. Understanding Dataset Difficulty with $\mathcal{V}$-Usable Information

    Kawin Ethayarajh, Yejin Choi, Swabha Swayamdipta

    cs.CLcs.AIcs.LGarXiv:2110.08420v32021
  14. Ghost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory

    Xizhou Zhu, Yuntao Chen, Hao Tian +10

    cs.AIcs.CLcs.CVarXiv:2305.17144v22023
  15. xCOMET: Transparent Machine Translation Evaluation through Fine-grained Error Detection

    Nuno M. Guerreiro, Ricardo Rei, Daan van Stigt +3

    cs.CLarXiv:2310.10482v12023
  16. Double Embeddings and CNN-based Sequence Labeling for Aspect Extraction

    Hu Xu, Bing Liu, Lei Shu +1

    cs.CLarXiv:1805.04601v12018
  17. Are We Modeling the Task or the Annotator? An Investigation of Annotator Bias in Natural Language Understanding Datasets

    Mor Geva, Yoav Goldberg, Jonathan Berant

    cs.CLarXiv:1908.07898v22019
  18. The political ideology of conversational AI: Converging evidence on ChatGPT's pro-environmental, left-libertarian orientation

    Jochen Hartmann, Jasper Schwenzow, Maximilian Witte

    cs.CLcs.CYarXiv:2301.01768v12023
  19. Commonsense Knowledge Mining from Pretrained Models

    Joshua Feldman, Joe Davison, Alexander M. Rush

    cs.CLcs.AIcs.LGarXiv:1909.00505v12019
  20. The Effect of Sampling Temperature on Problem Solving in Large Language Models

    Matthew Renze, Erhan Guven

    cs.CLcs.AIarXiv:2402.05201v32024
  21. Improving Topic Models with Latent Feature Word Representations

    Dat Quoc Nguyen, Richard Billingsley, Lan Du +1

    cs.CLcs.IRcs.LGarXiv:1810.06306v12018
  22. Deep Joint Entity Disambiguation with Local Neural Attention

    Octavian-Eugen Ganea, Thomas Hofmann

    cs.CLarXiv:1704.04920v32017
  23. Towards Making the Most of ChatGPT for Machine Translation

    Keqin Peng, Liang Ding, Qihuang Zhong +5

    cs.CLarXiv:2303.13780v42023
  24. Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training

    Xidong Feng, Ziyu Wan, Muning Wen +4

    cs.LGcs.AIcs.CLarXiv:2309.17179v22023
  25. A Deep Architecture for Semantic Matching with Multiple Positional Sentence Representations

    Shengxian Wan, Yanyan Lan, Jiafeng Guo +3

    cs.AIcs.CLcs.NEarXiv:1511.08277v12015
  26. Relation Classification via Recurrent Neural Network

    Dongxu Zhang, Dong Wang

    cs.CLcs.LGcs.NEarXiv:1508.01006v22015
  27. Syntax-Directed Variational Autoencoder for Structured Data

    Hanjun Dai, Yingtao Tian, Bo Dai +2

    cs.LGcs.CLarXiv:1802.08786v12018
  28. Multilingual and Multi-Aspect Hate Speech Analysis

    Nedjma Ousidhoum, Zizheng Lin, Hongming Zhang +2

    cs.CLarXiv:1908.11049v12019
  29. Twitter as a Lifeline: Human-annotated Twitter Corpora for NLP of Crisis-related Messages

    Muhammad Imran, Prasenjit Mitra, Carlos Castillo

    cs.CLcs.CYcs.SIarXiv:1605.05894v22016
  30. Aspect Level Sentiment Classification with Attention-over-Attention Neural Networks

    Binxuan Huang, Yanglan Ou, Kathleen M. Carley

    cs.CLarXiv:1804.06536v12018
  31. Large Language Models for Mathematical Reasoning: Progresses and Challenges

    Janice Ahn, Rishu Verma, Renze Lou +3

    cs.CLarXiv:2402.00157v42024
  32. A Survey of Available Corpora for Building Data-Driven Dialogue Systems

    Iulian Vlad Serban, Ryan Lowe, Peter Henderson +2

    cs.CLcs.AIcs.HCarXiv:1512.05742v32015
  33. Clinically Accurate Chest X-Ray Report Generation

    Guanxiong Liu, Tzu-Ming Harry Hsu, Matthew McDermott +4

    cs.CVcs.CLarXiv:1904.02633v22019
  34. Data Selection for Language Models via Importance Resampling

    Sang Michael Xie, Shibani Santurkar, Tengyu Ma +1

    cs.CLcs.LGarXiv:2302.03169v32023
  35. Generating Sequences by Learning to Self-Correct

    Sean Welleck, Ximing Lu, Peter West +4

    cs.CLarXiv:2211.00053v12022
  36. OpenAGI: When LLM Meets Domain Experts

    Yingqiang Ge, Wenyue Hua, Kai Mei +5

    cs.AIcs.CLcs.LGarXiv:2304.04370v62023
  37. Analyzing and Mitigating Object Hallucination in Large Vision-Language Models

    Yiyang Zhou, Chenhang Cui, Jaehong Yoon +5

    cs.LGcs.CLcs.CVarXiv:2310.00754v22023
  38. Lawformer: A Pre-trained Language Model for Chinese Legal Long Documents

    Chaojun Xiao, Xueyu Hu, Zhiyuan Liu +2

    cs.CLarXiv:2105.03887v12021
  39. MQuAKE: Assessing Knowledge Editing in Language Models via Multi-Hop Questions

    Zexuan Zhong, Zhengxuan Wu, Christopher D. Manning +2

    cs.CLarXiv:2305.14795v32023
  40. SqueezeLLM: Dense-and-Sparse Quantization

    Sehoon Kim, Coleman Hooper, Amir Gholami +5

    cs.CLcs.LGarXiv:2306.07629v42023
  41. TrustLLM: Trustworthiness in Large Language Models

    Yue Huang, Lichao Sun, Haoran Wang +67

    cs.CLarXiv:2401.05561v62024
  42. RARR: Researching and Revising What Language Models Say, Using Language Models

    Luyu Gao, Zhuyun Dai, Panupong Pasupat +8

    cs.CLcs.AIcs.IRarXiv:2210.08726v32022
  43. Nomic Embed: Training a Reproducible Long Context Text Embedder

    Zach Nussbaum, John X. Morris, Brandon Duderstadt +1

    cs.CLcs.AIarXiv:2402.01613v22024
  44. Black-Box Tuning for Language-Model-as-a-Service

    Tianxiang Sun, Yunfan Shao, Hong Qian +2

    cs.CLcs.AIarXiv:2201.03514v42022
  45. MLE-bench: Evaluating Machine Learning Agents on Machine Learning Engineering

    Jun Shern Chan, Neil Chowdhury, Oliver Jaffe +9

    cs.CLarXiv:2410.07095v62024
  46. AgentTuning: Enabling Generalized Agent Abilities for LLMs

    Aohan Zeng, Mingdao Liu, Rui Lu +4

    cs.CLcs.AIcs.LGarXiv:2310.12823v22023
  47. Is GPT-3 a Good Data Annotator?

    Bosheng Ding, Chengwei Qin, Linlin Liu +4

    cs.CLarXiv:2212.10450v22022
  48. Gromov-Wasserstein Alignment of Word Embedding Spaces

    David Alvarez-Melis, Tommi S. Jaakkola

    cs.CLarXiv:1809.00013v12018
  49. Quality at a Glance: An Audit of Web-Crawled Multilingual Datasets

    Julia Kreutzer, Isaac Caswell, Lisa Wang +49

    cs.CLcs.AIarXiv:2103.12028v42021
  50. Attentive Pooling Networks

    Cicero dos Santos, Ming Tan, Bing Xiang +1

    cs.CLcs.LGarXiv:1602.03609v12016
  51. Scaling Relationship on Learning Mathematical Reasoning with Large Language Models

    Zheng Yuan, Hongyi Yuan, Chengpeng Li +5

    cs.CLarXiv:2308.01825v22023
  52. Controlling Linguistic Style Aspects in Neural Language Generation

    Jessica Ficler, Yoav Goldberg

    cs.CLarXiv:1707.02633v12017
  53. Text Generation from Knowledge Graphs with Graph Transformers

    Rik Koncel-Kedziorski, Dhanush Bekal, Yi Luan +2

    cs.CLarXiv:1904.02342v32019
  54. A Survey of Controllable Text Generation using Transformer-based Pre-trained Language Models

    Hanqing Zhang, Haolin Song, Shaoyu Li +2

    cs.CLarXiv:2201.05337v52022
  55. Exploring Architectures, Data and Units For Streaming End-to-End Speech Recognition with RNN-Transducer

    Kanishka Rao, Haşim Sak, Rohit Prabhavalkar

    cs.CLcs.SDeess.ASarXiv:1801.00841v12018
  56. Adversarial Feature Matching for Text Generation

    Yizhe Zhang, Zhe Gan, Kai Fan +4

    stat.MLcs.CLcs.LGarXiv:1706.03850v32017
  57. VeRA: Vector-based Random Matrix Adaptation

    Dawid J. Kopiczko, Tijmen Blankevoort, Yuki M. Asano

    cs.CLarXiv:2310.11454v22023
  58. Temporal Analysis of Language through Neural Language Models

    Yoon Kim, Yi-I Chiu, Kentaro Hanaki +2

    cs.CLarXiv:1405.3515v12014
  59. Robust Distortion-free Watermarks for Language Models

    Rohith Kuditipudi, John Thickstun, Tatsunori Hashimoto +1

    cs.LGcs.CLcs.CRarXiv:2307.15593v32023
  60. Unleashing the potential of prompt engineering for large language models

    Banghao Chen, Zhaofeng Zhang, Nicolas Langrené +1

    cs.CLcs.AIarXiv:2310.14735v62023