Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

7,561 to 7,620 of 11,208

  1. AudioPaLM: A Large Language Model That Can Speak and Listen

    Paul K. Rubenstein, Chulayuth Asawaroengchai, Duc Dung Nguyen +27

    cs.CLcs.AIcs.SDarXiv:2306.12925v12023
  2. BanglaVeilGuard: Cross-Script Safety Benchmarking and Lightweight Guardrails for Bangla Large Language Models

    Md. Rakibul Hassan, Muhammad Iqbal Hossain

    cs.CLcs.CRarXiv:2608.21880v12026
  3. Cognitive Architectures for Language Agents

    Theodore R. Sumers, Shunyu Yao, Karthik Narasimhan +1

    cs.AIcs.CLcs.LGarXiv:2309.02427v32023
  4. The Communication Map of a Transformer

    Richard Zhe Wang

    cs.LGcs.CLarXiv:2608.22007v12026
  5. Layer-wise Analysis of a Self-supervised Speech Representation Model

    Ankita Pasad, Ju-Chieh Chou, Karen Livescu

    cs.CLcs.LGeess.ASarXiv:2107.04734v32021
  6. XSTest: A Test Suite for Identifying Exaggerated Safety Behaviours in Large Language Models

    Paul Röttger, Hannah Rose Kirk, Bertie Vidgen +3

    cs.CLcs.AIarXiv:2308.01263v32023
  7. TextWorld: A Learning Environment for Text-based Games

    Marc-Alexandre Côté, Ákos Kádár, Xingdi Yuan +10

    cs.LGcs.CLstat.MLarXiv:1806.11532v22018
  8. Analog Bits: Generating Discrete Data using Diffusion Models with Self-Conditioning

    Ting Chen, Ruixiang Zhang, Geoffrey Hinton

    cs.CVcs.AIcs.CLarXiv:2208.04202v22022
  9. Polyglot: Distributed Word Representations for Multilingual NLP

    Rami Al-Rfou, Bryan Perozzi, Steven Skiena

    cs.CLcs.LGarXiv:1307.1662v22013
  10. A Diverse Corpus for Evaluating and Developing English Math Word Problem Solvers

    Shen-Yun Miao, Chao-Chun Liang, Keh-Yih Su

    cs.AIcs.CLarXiv:2106.15772v12021
  11. Convergence in Science, Divergence in Religion: Calibrated Framing Differences Across Wikipedia's Language Editions

    Hung-Hsuan Chen

    cs.CLcs.CYcs.SIarXiv:2608.21821v12026
  12. GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher

    Youliang Yuan, Wenxiang Jiao, Wenxuan Wang +4

    cs.CLarXiv:2308.06463v22023
  13. GreenLeaf Law Embed Tiny: A Compact Embedding Model for Legal Domain Retrieval

    Surya Saka

    cs.LGcs.AIcs.CLarXiv:2608.24936v12026
  14. Auditing the Synthetic Memoir: Measuring Scene-Level Confabulation in LLM-Generated Autobiography Against the Documented Record of the Life It Describes

    Heather Renze

    cs.AIcs.CLcs.CYarXiv:2608.23640v12026
  15. Unnatural Instructions: Tuning Language Models with (Almost) No Human Labor

    Or Honovich, Thomas Scialom, Omer Levy +1

    cs.CLcs.AIcs.LGarXiv:2212.09689v12022
  16. Contrastive Preference Optimization: Pushing the Boundaries of LLM Performance in Machine Translation

    Haoran Xu, Amr Sharaf, Yunmo Chen +5

    cs.CLarXiv:2401.08417v42024
  17. From Exposure to Expectation: Frequency, Surprisal, and Language Across Development in Spanish

    Francisco Portillo López

    cs.CLarXiv:2608.22452v12026
  18. Structured Attention Networks

    Yoon Kim, Carl Denton, Luong Hoang +1

    cs.CLcs.LGcs.NEarXiv:1702.00887v32017
  19. Improved Image Captioning via Policy Gradient optimization of SPIDEr

    Siqi Liu, Zhenhai Zhu, Ning Ye +2

    cs.CVcs.CLarXiv:1612.00370v42016
  20. Multi-Agent Cooperation and the Emergence of (Natural) Language

    Angeliki Lazaridou, Alexander Peysakhovich, Marco Baroni

    cs.CLcs.CVcs.GTarXiv:1612.07182v22016
  21. Character-LLM: A Trainable Agent for Role-Playing

    Yunfan Shao, Linyang Li, Junqi Dai +1

    cs.CLcs.AIarXiv:2310.10158v22023
  22. One Embedder, Any Task: Instruction-Finetuned Text Embeddings

    Hongjin Su, Weijia Shi, Jungo Kasai +7

    cs.CLarXiv:2212.09741v32022
  23. Fine-Grained Human Feedback Gives Better Rewards for Language Model Training

    Zeqiu Wu, Yushi Hu, Weijia Shi +6

    cs.CLarXiv:2306.01693v22023
  24. MedAgents: Large Language Models as Collaborators for Zero-shot Medical Reasoning

    Xiangru Tang, Anni Zou, Zhuosheng Zhang +5

    cs.CLcs.AIarXiv:2311.10537v42023
  25. Large Language Models in Finance: A Survey

    Yinheng Li, Shaofei Wang, Han Ding +1

    q-fin.GNcs.AIcs.CLarXiv:2311.10723v22023
  26. The Stack: 3 TB of permissively licensed source code

    Denis Kocetkov, Raymond Li, Loubna Ben Allal +10

    cs.CLcs.AIarXiv:2211.15533v12022
  27. Multi-step Jailbreaking Privacy Attacks on ChatGPT

    Haoran Li, Dadi Guo, Wei Fan +4

    cs.CLcs.CRarXiv:2304.05197v32023
  28. A Comprehensive Capability Analysis of GPT-3 and GPT-3.5 Series Models

    Junjie Ye, Xuanting Chen, Nuo Xu +12

    cs.CLarXiv:2303.10420v22023
  29. Measuring and Improving Consistency in Pretrained Language Models

    Yanai Elazar, Nora Kassner, Shauli Ravfogel +4

    cs.CLarXiv:2102.01017v22021
  30. Event Extraction by Answering (Almost) Natural Questions

    Xinya Du, Claire Cardie

    cs.CLarXiv:2004.13625v22020
  31. Higher-order Coreference Resolution with Coarse-to-fine Inference

    Kenton Lee, Luheng He, Luke Zettlemoyer

    cs.CLarXiv:1804.05392v12018
  32. LegalBench: A Collaboratively Built Benchmark for Measuring Legal Reasoning in Large Language Models

    Neel Guha, Julian Nyarko, Daniel E. Ho +37

    cs.CLcs.AIcs.CYarXiv:2308.11462v12023
  33. A Categorical Archive of ChatGPT Failures

    Ali Borji

    cs.CLcs.AIcs.LGarXiv:2302.03494v82023
  34. Pretrained Transformers Improve Out-of-Distribution Robustness

    Dan Hendrycks, Xiaoyuan Liu, Eric Wallace +3

    cs.CLcs.LGarXiv:2004.06100v22020
  35. Dissecting Contextual Word Embeddings: Architecture and Representation

    Matthew E. Peters, Mark Neumann, Luke Zettlemoyer +1

    cs.CLarXiv:1808.08949v22018
  36. A Survey on Model Compression for Large Language Models

    Xunyu Zhu, Jian Li, Yong Liu +2

    cs.CLcs.AIarXiv:2308.07633v42023
  37. Jamba: A Hybrid Transformer-Mamba Language Model

    Opher Lieber, Barak Lenz, Hofit Bata +19

    cs.CLcs.LGarXiv:2403.19887v22024
  38. PPT: Pre-trained Prompt Tuning for Few-shot Learning

    Yuxian Gu, Xu Han, Zhiyuan Liu +1

    cs.CLarXiv:2109.04332v32021
  39. Probing Neural Network Comprehension of Natural Language Arguments

    Timothy Niven, Hung-Yu Kao

    cs.CLarXiv:1907.07355v22019
  40. Self-Diagnosis and Self-Debiasing: A Proposal for Reducing Corpus-Based Bias in NLP

    Timo Schick, Sahana Udupa, Hinrich Schütze

    cs.CLarXiv:2103.00453v22021
  41. Plan-And-Write: Towards Better Automatic Storytelling

    Lili Yao, Nanyun Peng, Ralph Weischedel +3

    cs.CLarXiv:1811.05701v32018
  42. The Evolved Transformer

    David R. So, Chen Liang, Quoc V. Le

    cs.LGcs.CLcs.NEarXiv:1901.11117v42019
  43. Talking About Large Language Models

    Murray Shanahan

    cs.CLcs.LGarXiv:2212.03551v52022
  44. A Computational Approach to Politeness with Application to Social Factors

    Cristian Danescu-Niculescu-Mizil, Moritz Sudhof, Dan Jurafsky +2

    cs.CLcs.SIphysics.soc-pharXiv:1306.6078v12013
  45. BigVGAN: A Universal Neural Vocoder with Large-Scale Training

    Sang-gil Lee, Wei Ping, Boris Ginsburg +2

    cs.SDcs.CLcs.LGarXiv:2206.04658v22022
  46. Large language models in healthcare and medical domain: A review

    Zabir Al Nazi, Wei Peng

    cs.CLcs.AIarXiv:2401.06775v22023
  47. Hate Speech Dataset from a White Supremacy Forum

    Ona de Gibert, Naiara Perez, Aitor García-Pablos +1

    cs.CLarXiv:1809.04444v12018
  48. Understanding the planning of LLM agents: A survey

    Xu Huang, Weiwen Liu, Xiaolong Chen +6

    cs.AIcs.CLcs.LGarXiv:2402.02716v12024
  49. Self-Alignment Pretraining for Biomedical Entity Representations

    Fangyu Liu, Ehsan Shareghi, Zaiqiao Meng +2

    cs.CLcs.AIcs.LGarXiv:2010.11784v22020
  50. Statistically Significant Detection of Linguistic Change

    Vivek Kulkarni, Rami Al-Rfou, Bryan Perozzi +1

    cs.CLcs.IRcs.LGarXiv:1411.3315v12014
  51. Deal or No Deal? End-to-End Learning for Negotiation Dialogues

    Mike Lewis, Denis Yarats, Yann N. Dauphin +2

    cs.AIcs.CLarXiv:1706.05125v12017
  52. TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding

    Shuhuai Ren, Linli Yao, Shicheng Li +2

    cs.CVcs.AIcs.CLarXiv:2312.02051v22023
  53. Scissorhands: Exploiting the Persistence of Importance Hypothesis for LLM KV Cache Compression at Test Time

    Zichang Liu, Aditya Desai, Fangshuo Liao +5

    cs.LGcs.CLarXiv:2305.17118v22023
  54. InjecAgent: Benchmarking Indirect Prompt Injections in Tool-Integrated Large Language Model Agents

    Qiusi Zhan, Zhixiang Liang, Zifan Ying +1

    cs.CLcs.CRarXiv:2403.02691v32024
  55. Named Entity Recognition as Dependency Parsing

    Juntao Yu, Bernd Bohnet, Massimo Poesio

    cs.CLarXiv:2005.07150v32020
  56. LLM-Adapters: An Adapter Family for Parameter-Efficient Fine-Tuning of Large Language Models

    Zhiqiang Hu, Lei Wang, Yihuai Lan +6

    cs.CLarXiv:2304.01933v32023
  57. A Comprehensive Survey on Applications of Transformers for Deep Learning Tasks

    Saidul Islam, Hanae Elmekki, Ahmed Elsebai +4

    cs.LGcs.CLarXiv:2306.07303v12023
  58. LinkBERT: Pretraining Language Models with Document Links

    Michihiro Yasunaga, Jure Leskovec, Percy Liang

    cs.CLcs.LGarXiv:2203.15827v12022
  59. From Crowdsourced Data to High-Quality Benchmarks: Arena-Hard and BenchBuilder Pipeline

    Tianle Li, Wei-Lin Chiang, Evan Frick +5

    cs.LGcs.AIcs.CLarXiv:2406.11939v22024
  60. Revisiting Few-sample BERT Fine-tuning

    Tianyi Zhang, Felix Wu, Arzoo Katiyar +2

    cs.CLcs.LGarXiv:2006.05987v32020