Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,301 to 6,360 of 11,299

  1. L-Eval: Instituting Standardized Evaluation for Long Context Language Models

    Chenxin An, Shansan Gong, Ming Zhong +5

    cs.CLarXiv:2307.11088v32023
  2. Query-Key Normalization for Transformers

    Alex Henry, Prudhvi Raj Dachapally, Shubham Pawar +1

    cs.CLcs.AIcs.LGarXiv:2010.04245v12020
  3. Data Contamination: From Memorization to Exploitation

    Inbal Magar, Roy Schwartz

    cs.CLcs.LGarXiv:2203.08242v12022
  4. AnnoLLM: Making Large Language Models to Be Better Crowdsourced Annotators

    Xingwei He, Zhenghao Lin, Yeyun Gong +7

    cs.CLarXiv:2303.16854v22023
  5. The Microsoft 2016 Conversational Speech Recognition System

    W. Xiong, J. Droppo, X. Huang +5

    cs.CLeess.ASarXiv:1609.03528v22016
  6. Art or Artifice? Large Language Models and the False Promise of Creativity

    Tuhin Chakrabarty, Philippe Laban, Divyansh Agarwal +2

    cs.CLcs.AIcs.HCarXiv:2309.14556v32023
  7. Generative Representational Instruction Tuning

    Niklas Muennighoff, Hongjin Su, Liang Wang +5

    cs.CLcs.AIcs.LGarXiv:2402.09906v32024
  8. Fine-tuning Language Models for Factuality

    Katherine Tian, Eric Mitchell, Huaxiu Yao +2

    cs.CLcs.AIcs.LGarXiv:2311.08401v12023
  9. OpenCodeInterpreter: Integrating Code Generation with Execution and Refinement

    Tianyu Zheng, Ge Zhang, Tianhao Shen +5

    cs.SEcs.AIcs.CLarXiv:2402.14658v32024
  10. A Hierarchical Model of Reviews for Aspect-based Sentiment Analysis

    Sebastian Ruder, Parsa Ghaffari, John G. Breslin

    cs.CLcs.LGarXiv:1609.02745v12016
  11. GPT3Mix: Leveraging Large-scale Language Models for Text Augmentation

    Kang Min Yoo, Dongju Park, Jaewook Kang +2

    cs.CLcs.AIarXiv:2104.08826v22021
  12. Emilia: An Extensive, Multilingual, and Diverse Speech Dataset for Large-Scale Speech Generation

    Haorui He, Zengqiang Shang, Chaoren Wang +11

    eess.AScs.CLarXiv:2407.05361v32024
  13. Deliberative Alignment: Reasoning Enables Safer Language Models

    Melody Y. Guan, Manas Joglekar, Eric Wallace +12

    cs.CLcs.AIcs.CYarXiv:2412.16339v22024
  14. SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

    Zhangchen Xu, Fengqing Jiang, Luyao Niu +3

    cs.CRcs.AIcs.CLarXiv:2402.08983v42024
  15. FinanceBench: A New Benchmark for Financial Question Answering

    Pranab Islam, Anand Kannappan, Douwe Kiela +3

    cs.CLcs.AIcs.CEarXiv:2311.11944v12023
  16. Star-Transformer

    Qipeng Guo, Xipeng Qiu, Pengfei Liu +3

    cs.CLarXiv:1902.09113v32019
  17. Probing Pretrained Language Models for Lexical Semantics

    Ivan Vulić, Edoardo Maria Ponti, Robert Litschko +2

    cs.CLarXiv:2010.05731v12020
  18. Deep Compositional Captioning: Describing Novel Object Categories without Paired Training Data

    Lisa Anne Hendricks, Subhashini Venugopalan, Marcus Rohrbach +3

    cs.CVcs.CLarXiv:1511.05284v22015
  19. Societal Biases in Language Generation: Progress and Challenges

    Emily Sheng, Kai-Wei Chang, Premkumar Natarajan +1

    cs.CLarXiv:2105.04054v32021
  20. Unsupervised Pretraining for Sequence to Sequence Learning

    Prajit Ramachandran, Peter J. Liu, Quoc V. Le

    cs.CLcs.LGcs.NEarXiv:1611.02683v22016
  21. Multilingual Speech Recognition With A Single End-To-End Model

    Shubham Toshniwal, Tara N. Sainath, Ron J. Weiss +4

    eess.AScs.AIcs.CLarXiv:1711.01694v22017
  22. R-Judge: Benchmarking Safety Risk Awareness for LLM Agents

    Tongxin Yuan, Zhiwei He, Lingzhong Dong +9

    cs.CLcs.AIarXiv:2401.10019v32024
  23. Tensor Graph Convolutional Networks for Text Classification

    Xien Liu, Xinxin You, Xiao Zhang +2

    cs.CLcs.IRcs.LGarXiv:2001.05313v12020
  24. TransNets: Learning to Transform for Recommendation

    Rose Catherine, William Cohen

    cs.IRcs.CLcs.LGarXiv:1704.02298v22017
  25. What to talk about and how? Selective Generation using LSTMs with Coarse-to-Fine Alignment

    Hongyuan Mei, Mohit Bansal, Matthew R. Walter

    cs.CLcs.AIcs.LGarXiv:1509.00838v22015
  26. AugGPT: Leveraging ChatGPT for Text Data Augmentation

    Haixing Dai, Zhengliang Liu, Wenxiong Liao +15

    cs.CLcs.AIcs.LGarXiv:2302.13007v32023
  27. Language Models for Image Captioning: The Quirks and What Works

    Jacob Devlin, Hao Cheng, Hao Fang +5

    cs.CLcs.AIcs.CVarXiv:1505.01809v32015
  28. Analyzing Uncertainty in Neural Machine Translation

    Myle Ott, Michael Auli, David Grangier +1

    cs.CLarXiv:1803.00047v42018
  29. Information-Theoretic Probing for Linguistic Structure

    Tiago Pimentel, Josef Valvoda, Rowan Hall Maudslay +3

    cs.CLcs.LGarXiv:2004.03061v22020
  30. SemEval-2020 Task 1: Unsupervised Lexical Semantic Change Detection

    Dominik Schlechtweg, Barbara McGillivray, Simon Hengchen +2

    cs.CLarXiv:2007.11464v22020
  31. Video Understanding with Large Language Models: A Survey

    Yolo Y. Tang, Jing Bi, Siting Xu +17

    cs.CVcs.CLarXiv:2312.17432v82023
  32. Scaling Down to Scale Up: A Guide to Parameter-Efficient Fine-Tuning

    Vladislav Lialin, Vijeta Deshpande, Xiaowei Yao +1

    cs.CLarXiv:2303.15647v22023
  33. What can Large Language Models do in chemistry? A comprehensive benchmark on eight tasks

    Taicheng Guo, Kehan Guo, Bozhao Nan +5

    cs.CLcs.AIarXiv:2305.18365v32023
  34. Label Words are Anchors: An Information Flow Perspective for Understanding In-Context Learning

    Lean Wang, Lei Li, Damai Dai +5

    cs.CLcs.LGarXiv:2305.14160v42023
  35. Supporting Qualitative Analysis with Large Language Models: Combining Codebook with GPT-3 for Deductive Coding

    Ziang Xiao, Xingdi Yuan, Q. Vera Liao +2

    cs.CLcs.AIcs.HCarXiv:2304.10548v12023
  36. Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging

    Joel Jang, Seungone Kim, Bill Yuchen Lin +6

    cs.CLarXiv:2310.11564v12023
  37. Detecting Emergent Intersectional Biases: Contextualized Word Embeddings Contain a Distribution of Human-like Biases

    Wei Guo, Aylin Caliskan

    cs.CYcs.AIcs.CLarXiv:2006.03955v52020
  38. Get Your Vitamin C! Robust Fact Verification with Contrastive Evidence

    Tal Schuster, Adam Fisch, Regina Barzilay

    cs.CLcs.IRcs.LGarXiv:2103.08541v12021
  39. Function Vectors in Large Language Models

    Eric Todd, Millicent L. Li, Arnab Sen Sharma +3

    cs.CLcs.LGarXiv:2310.15213v22023
  40. Towards Debiasing Sentence Representations

    Paul Pu Liang, Irene Mengze Li, Emily Zheng +3

    cs.CLcs.LGarXiv:2007.08100v12020
  41. A Joint Speaker-Listener-Reinforcer Model for Referring Expressions

    Licheng Yu, Hao Tan, Mohit Bansal +1

    cs.CVcs.AIcs.CLarXiv:1612.09542v22016
  42. Sentiment Analysis of Review Datasets Using Naive Bayes and K-NN Classifier

    Lopamudra Dey, Sanjay Chakraborty, Anuraag Biswas +2

    cs.IRcs.CLarXiv:1610.09982v12016
  43. Active Example Selection for In-Context Learning

    Yiming Zhang, Shi Feng, Chenhao Tan

    cs.CLcs.AIarXiv:2211.04486v12022
  44. Document-Level Neural Machine Translation with Hierarchical Attention Networks

    Lesly Miculicich, Dhananjay Ram, Nikolaos Pappas +1

    cs.CLarXiv:1809.01576v22018
  45. Drive Like a Human: Rethinking Autonomous Driving with Large Language Models

    Daocheng Fu, Xin Li, Licheng Wen +4

    cs.ROcs.CLarXiv:2307.07162v12023
  46. Automatically Correcting Large Language Models: Surveying the landscape of diverse self-correction strategies

    Liangming Pan, Michael Saxon, Wenda Xu +3

    cs.CLcs.AIcs.LGarXiv:2308.03188v22023
  47. Predictive Biases in Natural Language Processing Models: A Conceptual Framework and Overview

    Deven Shah, H. Andrew Schwartz, Dirk Hovy

    cs.CLcs.AIcs.LGarXiv:1912.11078v22019
  48. Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities

    Enneng Yang, Li Shen, Guibing Guo +4

    cs.LGcs.AIcs.CLarXiv:2408.07666v52024
  49. Chain-of-Thought Reasoning Without Prompting

    Xuezhi Wang, Denny Zhou

    cs.CLarXiv:2402.10200v22024
  50. Explicit Knowledge-based Reasoning for Visual Question Answering

    Peng Wang, Qi Wu, Chunhua Shen +2

    cs.CVcs.CLarXiv:1511.02570v22015
  51. Do Large Language Models Know What They Don't Know?

    Zhangyue Yin, Qiushi Sun, Qipeng Guo +3

    cs.CLarXiv:2305.18153v22023
  52. SKEP: Sentiment Knowledge Enhanced Pre-training for Sentiment Analysis

    Hao Tian, Can Gao, Xinyan Xiao +5

    cs.CLarXiv:2005.05635v22020
  53. GSum: A General Framework for Guided Neural Abstractive Summarization

    Zi-Yi Dou, Pengfei Liu, Hiroaki Hayashi +2

    cs.CLarXiv:2010.08014v32020
  54. Scene Text Recognition with Permuted Autoregressive Sequence Models

    Darwin Bautista, Rowel Atienza

    cs.CVcs.CLarXiv:2207.06966v12022
  55. Large Language Models Sensitivity to The Order of Options in Multiple-Choice Questions

    Pouya Pezeshkpour, Estevam Hruschka

    cs.CLcs.AIcs.LGarXiv:2308.11483v12023
  56. A Survey of Domain Adaptation for Neural Machine Translation

    Chenhui Chu, Rui Wang

    cs.CLcs.AIcs.LGarXiv:1806.00258v12018
  57. A Long Way to Go: Investigating Length Correlations in RLHF

    Prasann Singhal, Tanya Goyal, Jiacheng Xu +1

    cs.CLcs.LGarXiv:2310.03716v22023
  58. SimKGC: Simple Contrastive Knowledge Graph Completion with Pre-trained Language Models

    Liang Wang, Wei Zhao, Zhuoyu Wei +1

    cs.CLarXiv:2203.02167v12022
  59. RADAR: Robust AI-Text Detection via Adversarial Learning

    Xiaomeng Hu, Pin-Yu Chen, Tsung-Yi Ho

    cs.CLcs.AIcs.LGarXiv:2307.03838v22023
  60. HAT: Hardware-Aware Transformers for Efficient Natural Language Processing

    Hanrui Wang, Zhanghao Wu, Zhijian Liu +4

    cs.CLcs.LGcs.NEarXiv:2005.14187v12020