Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
7,561 to 7,620 of 11,208
AudioPaLM: A Large Language Model That Can Speak and Listen
Paul K. Rubenstein, Chulayuth Asawaroengchai, Duc Dung Nguyen +27
cs.CLcs.AIcs.SDarXiv:2306.12925v12023BanglaVeilGuard: Cross-Script Safety Benchmarking and Lightweight Guardrails for Bangla Large Language Models
Md. Rakibul Hassan, Muhammad Iqbal Hossain
cs.CLcs.CRarXiv:2608.21880v12026Cognitive Architectures for Language Agents
Theodore R. Sumers, Shunyu Yao, Karthik Narasimhan +1
cs.AIcs.CLcs.LGarXiv:2309.02427v32023The Communication Map of a Transformer
Richard Zhe Wang
cs.LGcs.CLarXiv:2608.22007v12026Layer-wise Analysis of a Self-supervised Speech Representation Model
Ankita Pasad, Ju-Chieh Chou, Karen Livescu
cs.CLcs.LGeess.ASarXiv:2107.04734v32021XSTest: A Test Suite for Identifying Exaggerated Safety Behaviours in Large Language Models
Paul Röttger, Hannah Rose Kirk, Bertie Vidgen +3
cs.CLcs.AIarXiv:2308.01263v32023TextWorld: A Learning Environment for Text-based Games
Marc-Alexandre Côté, Ákos Kádár, Xingdi Yuan +10
cs.LGcs.CLstat.MLarXiv:1806.11532v22018Analog Bits: Generating Discrete Data using Diffusion Models with Self-Conditioning
Ting Chen, Ruixiang Zhang, Geoffrey Hinton
cs.CVcs.AIcs.CLarXiv:2208.04202v22022Polyglot: Distributed Word Representations for Multilingual NLP
Rami Al-Rfou, Bryan Perozzi, Steven Skiena
cs.CLcs.LGarXiv:1307.1662v22013A Diverse Corpus for Evaluating and Developing English Math Word Problem Solvers
Shen-Yun Miao, Chao-Chun Liang, Keh-Yih Su
cs.AIcs.CLarXiv:2106.15772v12021Convergence in Science, Divergence in Religion: Calibrated Framing Differences Across Wikipedia's Language Editions
Hung-Hsuan Chen
cs.CLcs.CYcs.SIarXiv:2608.21821v12026GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher
Youliang Yuan, Wenxiang Jiao, Wenxuan Wang +4
cs.CLarXiv:2308.06463v22023GreenLeaf Law Embed Tiny: A Compact Embedding Model for Legal Domain Retrieval
Surya Saka
cs.LGcs.AIcs.CLarXiv:2608.24936v12026Auditing the Synthetic Memoir: Measuring Scene-Level Confabulation in LLM-Generated Autobiography Against the Documented Record of the Life It Describes
Heather Renze
cs.AIcs.CLcs.CYarXiv:2608.23640v12026Unnatural Instructions: Tuning Language Models with (Almost) No Human Labor
Or Honovich, Thomas Scialom, Omer Levy +1
cs.CLcs.AIcs.LGarXiv:2212.09689v12022Contrastive Preference Optimization: Pushing the Boundaries of LLM Performance in Machine Translation
Haoran Xu, Amr Sharaf, Yunmo Chen +5
cs.CLarXiv:2401.08417v42024From Exposure to Expectation: Frequency, Surprisal, and Language Across Development in Spanish
Francisco Portillo López
cs.CLarXiv:2608.22452v12026Structured Attention Networks
Yoon Kim, Carl Denton, Luong Hoang +1
cs.CLcs.LGcs.NEarXiv:1702.00887v32017Improved Image Captioning via Policy Gradient optimization of SPIDEr
Siqi Liu, Zhenhai Zhu, Ning Ye +2
cs.CVcs.CLarXiv:1612.00370v42016Multi-Agent Cooperation and the Emergence of (Natural) Language
Angeliki Lazaridou, Alexander Peysakhovich, Marco Baroni
cs.CLcs.CVcs.GTarXiv:1612.07182v22016Character-LLM: A Trainable Agent for Role-Playing
Yunfan Shao, Linyang Li, Junqi Dai +1
cs.CLcs.AIarXiv:2310.10158v22023One Embedder, Any Task: Instruction-Finetuned Text Embeddings
Hongjin Su, Weijia Shi, Jungo Kasai +7
cs.CLarXiv:2212.09741v32022Fine-Grained Human Feedback Gives Better Rewards for Language Model Training
Zeqiu Wu, Yushi Hu, Weijia Shi +6
cs.CLarXiv:2306.01693v22023MedAgents: Large Language Models as Collaborators for Zero-shot Medical Reasoning
Xiangru Tang, Anni Zou, Zhuosheng Zhang +5
cs.CLcs.AIarXiv:2311.10537v42023Large Language Models in Finance: A Survey
Yinheng Li, Shaofei Wang, Han Ding +1
q-fin.GNcs.AIcs.CLarXiv:2311.10723v22023The Stack: 3 TB of permissively licensed source code
Denis Kocetkov, Raymond Li, Loubna Ben Allal +10
cs.CLcs.AIarXiv:2211.15533v12022Multi-step Jailbreaking Privacy Attacks on ChatGPT
Haoran Li, Dadi Guo, Wei Fan +4
cs.CLcs.CRarXiv:2304.05197v32023A Comprehensive Capability Analysis of GPT-3 and GPT-3.5 Series Models
Junjie Ye, Xuanting Chen, Nuo Xu +12
cs.CLarXiv:2303.10420v22023Measuring and Improving Consistency in Pretrained Language Models
Yanai Elazar, Nora Kassner, Shauli Ravfogel +4
cs.CLarXiv:2102.01017v22021Event Extraction by Answering (Almost) Natural Questions
Xinya Du, Claire Cardie
cs.CLarXiv:2004.13625v22020Higher-order Coreference Resolution with Coarse-to-fine Inference
Kenton Lee, Luheng He, Luke Zettlemoyer
cs.CLarXiv:1804.05392v12018LegalBench: A Collaboratively Built Benchmark for Measuring Legal Reasoning in Large Language Models
Neel Guha, Julian Nyarko, Daniel E. Ho +37
cs.CLcs.AIcs.CYarXiv:2308.11462v12023A Categorical Archive of ChatGPT Failures
Ali Borji
cs.CLcs.AIcs.LGarXiv:2302.03494v82023Pretrained Transformers Improve Out-of-Distribution Robustness
Dan Hendrycks, Xiaoyuan Liu, Eric Wallace +3
cs.CLcs.LGarXiv:2004.06100v22020Dissecting Contextual Word Embeddings: Architecture and Representation
Matthew E. Peters, Mark Neumann, Luke Zettlemoyer +1
cs.CLarXiv:1808.08949v22018A Survey on Model Compression for Large Language Models
Xunyu Zhu, Jian Li, Yong Liu +2
cs.CLcs.AIarXiv:2308.07633v42023Jamba: A Hybrid Transformer-Mamba Language Model
Opher Lieber, Barak Lenz, Hofit Bata +19
cs.CLcs.LGarXiv:2403.19887v22024PPT: Pre-trained Prompt Tuning for Few-shot Learning
Yuxian Gu, Xu Han, Zhiyuan Liu +1
cs.CLarXiv:2109.04332v32021Probing Neural Network Comprehension of Natural Language Arguments
Timothy Niven, Hung-Yu Kao
cs.CLarXiv:1907.07355v22019Self-Diagnosis and Self-Debiasing: A Proposal for Reducing Corpus-Based Bias in NLP
Timo Schick, Sahana Udupa, Hinrich Schütze
cs.CLarXiv:2103.00453v22021Plan-And-Write: Towards Better Automatic Storytelling
Lili Yao, Nanyun Peng, Ralph Weischedel +3
cs.CLarXiv:1811.05701v32018The Evolved Transformer
David R. So, Chen Liang, Quoc V. Le
cs.LGcs.CLcs.NEarXiv:1901.11117v42019Talking About Large Language Models
Murray Shanahan
cs.CLcs.LGarXiv:2212.03551v52022A Computational Approach to Politeness with Application to Social Factors
Cristian Danescu-Niculescu-Mizil, Moritz Sudhof, Dan Jurafsky +2
cs.CLcs.SIphysics.soc-pharXiv:1306.6078v12013BigVGAN: A Universal Neural Vocoder with Large-Scale Training
Sang-gil Lee, Wei Ping, Boris Ginsburg +2
cs.SDcs.CLcs.LGarXiv:2206.04658v22022Large language models in healthcare and medical domain: A review
Zabir Al Nazi, Wei Peng
cs.CLcs.AIarXiv:2401.06775v22023Hate Speech Dataset from a White Supremacy Forum
Ona de Gibert, Naiara Perez, Aitor García-Pablos +1
cs.CLarXiv:1809.04444v12018Understanding the planning of LLM agents: A survey
Xu Huang, Weiwen Liu, Xiaolong Chen +6
cs.AIcs.CLcs.LGarXiv:2402.02716v12024Self-Alignment Pretraining for Biomedical Entity Representations
Fangyu Liu, Ehsan Shareghi, Zaiqiao Meng +2
cs.CLcs.AIcs.LGarXiv:2010.11784v22020Statistically Significant Detection of Linguistic Change
Vivek Kulkarni, Rami Al-Rfou, Bryan Perozzi +1
cs.CLcs.IRcs.LGarXiv:1411.3315v12014Deal or No Deal? End-to-End Learning for Negotiation Dialogues
Mike Lewis, Denis Yarats, Yann N. Dauphin +2
cs.AIcs.CLarXiv:1706.05125v12017TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Shuhuai Ren, Linli Yao, Shicheng Li +2
cs.CVcs.AIcs.CLarXiv:2312.02051v22023Scissorhands: Exploiting the Persistence of Importance Hypothesis for LLM KV Cache Compression at Test Time
Zichang Liu, Aditya Desai, Fangshuo Liao +5
cs.LGcs.CLarXiv:2305.17118v22023InjecAgent: Benchmarking Indirect Prompt Injections in Tool-Integrated Large Language Model Agents
Qiusi Zhan, Zhixiang Liang, Zifan Ying +1
cs.CLcs.CRarXiv:2403.02691v32024Named Entity Recognition as Dependency Parsing
Juntao Yu, Bernd Bohnet, Massimo Poesio
cs.CLarXiv:2005.07150v32020LLM-Adapters: An Adapter Family for Parameter-Efficient Fine-Tuning of Large Language Models
Zhiqiang Hu, Lei Wang, Yihuai Lan +6
cs.CLarXiv:2304.01933v32023A Comprehensive Survey on Applications of Transformers for Deep Learning Tasks
Saidul Islam, Hanae Elmekki, Ahmed Elsebai +4
cs.LGcs.CLarXiv:2306.07303v12023LinkBERT: Pretraining Language Models with Document Links
Michihiro Yasunaga, Jure Leskovec, Percy Liang
cs.CLcs.LGarXiv:2203.15827v12022From Crowdsourced Data to High-Quality Benchmarks: Arena-Hard and BenchBuilder Pipeline
Tianle Li, Wei-Lin Chiang, Evan Frick +5
cs.LGcs.AIcs.CLarXiv:2406.11939v22024Revisiting Few-sample BERT Fine-tuning
Tianyi Zhang, Felix Wu, Arzoo Katiyar +2
cs.CLcs.LGarXiv:2006.05987v32020