Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
5,041 to 5,100 of 11,259
MetaTool Benchmark for Large Language Models: Deciding Whether to Use Tools and Which to Use
Yue Huang, Jiawen Shi, Yuan Li +8
cs.SEcs.CLarXiv:2310.03128v62023CoLAKE: Contextualized Language and Knowledge Embedding
Tianxiang Sun, Yunfan Shao, Xipeng Qiu +4
cs.CLcs.AIarXiv:2010.00309v12020Out-of-Distribution Detection and Selective Generation for Conditional Language Models
Jie Ren, Jiaming Luo, Yao Zhao +4
cs.CLarXiv:2209.15558v22022Margin-based Parallel Corpus Mining with Multilingual Sentence Embeddings
Mikel Artetxe, Holger Schwenk
cs.CLcs.AIcs.LGarXiv:1811.01136v22018Artificial Intelligence, speech and language processing approaches to monitoring Alzheimer's Disease: a systematic review
Sofia de la Fuente Garcia, Craig Ritchie, Saturnino Luz
cs.AIcs.CLeess.ASarXiv:2010.06047v12020Scaling Up Models and Data with $\texttt{t5x}$ and $\texttt{seqio}$
Adam Roberts, Hyung Won Chung, Anselm Levskaya +40
cs.LGcs.CLarXiv:2203.17189v12022Don't Hallucinate, Abstain: Identifying LLM Knowledge Gaps via Multi-LLM Collaboration
Shangbin Feng, Weijia Shi, Yike Wang +3
cs.CLarXiv:2402.00367v22024Locate and Label: A Two-stage Identifier for Nested Named Entity Recognition
Yongliang Shen, Xinyin Ma, Zeqi Tan +3
cs.CLarXiv:2105.06804v22021Adversarial Watermarking Transformer: Towards Tracing Text Provenance with Data Hiding
Sahar Abdelnabi, Mario Fritz
cs.CRcs.CLcs.CYarXiv:2009.03015v22020ToolSandbox: A Stateful, Conversational, Interactive Evaluation Benchmark for LLM Tool Use Capabilities
Jiarui Lu, Thomas Holleis, Yizhe Zhang +9
cs.CLcs.AIcs.LGarXiv:2408.04682v22024Verify-and-Edit: A Knowledge-Enhanced Chain-of-Thought Framework
Ruochen Zhao, Xingxuan Li, Shafiq Joty +2
cs.CLarXiv:2305.03268v12023On decoder-only architecture for speech-to-text and large language model integration
Jian Wu, Yashesh Gaur, Zhuo Chen +8
eess.AScs.CLcs.SDarXiv:2307.03917v32023Causal Parrots: Large Language Models May Talk Causality But Are Not Causal
Matej Zečević, Moritz Willig, Devendra Singh Dhami +1
cs.AIcs.CLarXiv:2308.13067v12023Defending Large Language Models Against Jailbreaking Attacks Through Goal Prioritization
Zhexin Zhang, Junxiao Yang, Pei Ke +3
cs.CLarXiv:2311.09096v22023Rapid Adaptation of Neural Machine Translation to New Languages
Graham Neubig, Junjie Hu
cs.CLarXiv:1808.04189v12018Fooling End-to-end Speaker Verification by Adversarial Examples
Felix Kreuk, Yossi Adi, Moustapha Cisse +1
cs.LGcs.CLarXiv:1801.03339v22018Diverse Few-Shot Text Classification with Multiple Metrics
Mo Yu, Xiaoxiao Guo, Jinfeng Yi +6
cs.CLcs.LGarXiv:1805.07513v12018Aya Dataset: An Open-Access Collection for Multilingual Instruction Tuning
Shivalika Singh, Freddie Vargus, Daniel Dsouza +30
cs.CLcs.AIarXiv:2402.06619v12024What Is One Grain of Sand in the Desert? Analyzing Individual Neurons in Deep NLP Models
Fahim Dalvi, Nadir Durrani, Hassan Sajjad +3
cs.CLarXiv:1812.09355v12018An Emotional Analysis of False Information in Social Media and News Articles
Bilal Ghanem, Paolo Rosso, Francisco Rangel
cs.CLcs.IRcs.SIarXiv:1908.09951v12019Sentence-State LSTM for Text Representation
Yue Zhang, Qi Liu, Linfeng Song
cs.CLcs.LGstat.MLarXiv:1805.02474v12018Sparse Upcycling: Training Mixture-of-Experts from Dense Checkpoints
Aran Komatsuzaki, Joan Puigcerver, James Lee-Thorp +6
cs.LGcs.CLcs.CVarXiv:2212.05055v22022Learning to Count Objects in Natural Images for Visual Question Answering
Yan Zhang, Jonathon Hare, Adam Prügel-Bennett
cs.CVcs.CLarXiv:1802.05766v12018Paraphrase Generation with Deep Reinforcement Learning
Zichao Li, Xin Jiang, Lifeng Shang +1
cs.CLarXiv:1711.00279v32017Exploring the Limits of ChatGPT for Query or Aspect-based Text Summarization
Xianjun Yang, Yan Li, Xinlu Zhang +2
cs.CLcs.AIarXiv:2302.08081v12023A Roadmap to Pluralistic Alignment
Taylor Sorensen, Jared Moore, Jillian Fisher +9
cs.AIcs.CLcs.IRarXiv:2402.05070v32024Does Writing with Language Models Reduce Content Diversity?
Vishakh Padmakumar, He He
cs.CLcs.CYcs.HCarXiv:2309.05196v32023Learning to Paraphrase for Question Answering
Li Dong, Jonathan Mallinson, Siva Reddy +1
cs.CLarXiv:1708.06022v12017Spherical Latent Spaces for Stable Variational Autoencoders
Jiacheng Xu, Greg Durrett
cs.CLarXiv:1808.10805v22018Compression of Neural Machine Translation Models via Pruning
Abigail See, Minh-Thang Luong, Christopher D. Manning
cs.AIcs.CLcs.NEarXiv:1606.09274v12016Compressing Context to Enhance Inference Efficiency of Large Language Models
Yucheng Li, Bo Dong, Chenghua Lin +1
cs.CLarXiv:2310.06201v12023FLASK: Fine-grained Language Model Evaluation based on Alignment Skill Sets
Seonghyeon Ye, Doyoung Kim, Sungdong Kim +6
cs.CLcs.AIarXiv:2307.10928v42023Search, Inspect, Fetch: Exploiting Structure-Aware Boolean Retrieval for Deep-Search Agents
Shuai Wang, Haodong Chen, Yu Yin +3
cs.IRcs.AIcs.CLarXiv:2608.02751v32026JEC-QA: A Legal-Domain Question Answering Dataset
Haoxi Zhong, Chaojun Xiao, Cunchao Tu +3
cs.CLarXiv:1911.12011v12019Bidirectional Machine Reading Comprehension for Aspect Sentiment Triplet Extraction
Shaowei Chen, Yu Wang, Jie Liu +1
cs.CLarXiv:2103.07665v12021Prompting Large Language Models with Speech Recognition Abilities
Yassir Fathullah, Chunyang Wu, Egor Lakomkin +9
eess.AScs.AIcs.CLarXiv:2307.11795v12023DePlot: One-shot visual language reasoning by plot-to-table translation
Fangyu Liu, Julian Martin Eisenschlos, Francesco Piccinno +7
cs.CLcs.AIcs.CVarXiv:2212.10505v22022SciReC: Diagnostic Evaluation of Multimodal, Multi-Turn Relational Reasoning with Adaptive Interaction
Nilay Yilmaz, Naga Sai Abhiram Kusumba, Stella Wenxing Liu +1
cs.CLcs.AIcs.LGarXiv:2608.27461v12026UBAR: Towards Fully End-to-End Task-Oriented Dialog Systems with GPT-2
Yunyi Yang, Yunhao Li, Xiaojun Quan
cs.CLarXiv:2012.03539v22020KAT: A Knowledge Augmented Transformer for Vision-and-Language
Liangke Gui, Borui Wang, Qiuyuan Huang +3
cs.CLarXiv:2112.08614v22021Semantic Sentence Matching with Densely-connected Recurrent and Co-attentive Information
Seonhoon Kim, Inho Kang, Nojun Kwak
cs.CLarXiv:1805.11360v22018Compositional Explanations of Neurons
Jesse Mu, Jacob Andreas
cs.LGcs.AIcs.CLarXiv:2006.14032v22020CulturaX: A Cleaned, Enormous, and Multilingual Dataset for Large Language Models in 167 Languages
Thuat Nguyen, Chien Van Nguyen, Viet Dac Lai +5
cs.CLcs.AIarXiv:2309.09400v12023DistilHuBERT: Speech Representation Learning by Layer-wise Distillation of Hidden-unit BERT
Heng-Jui Chang, Shu-wen Yang, Hung-yi Lee
cs.CLeess.ASarXiv:2110.01900v42021MAVEN: A Massive General Domain Event Detection Dataset
Xiaozhi Wang, Ziqi Wang, Xu Han +7
cs.CLarXiv:2004.13590v22020Very Deep Multilingual Convolutional Neural Networks for LVCSR
Tom Sercu, Christian Puhrsch, Brian Kingsbury +1
cs.CLcs.NEarXiv:1509.08967v22015Look Before You Leap: Bridging Model-Free and Model-Based Reinforcement Learning for Planned-Ahead Vision-and-Language Navigation
Xin Wang, Wenhan Xiong, Hongmin Wang +1
cs.CVcs.AIcs.CLarXiv:1803.07729v22018MASSIVE: A 1M-Example Multilingual Natural Language Understanding Dataset with 51 Typologically-Diverse Languages
Jack FitzGerald, Christopher Hench, Charith Peris +13
cs.CLcs.AIcs.LGarXiv:2204.08582v22022BLEnD: A Benchmark for LLMs on Everyday Knowledge in Diverse Cultures and Languages
Junho Myung, Nayeon Lee, Yi Zhou +19
cs.CLarXiv:2406.09948v22024RedditBias: A Real-World Resource for Bias Evaluation and Debiasing of Conversational Language Models
Soumya Barikeri, Anne Lauscher, Ivan Vulić +1
cs.CLarXiv:2106.03521v12021OpenTag: Open Attribute Value Extraction from Product Profiles [Deep Learning, Active Learning, Named Entity Recognition]
Guineng Zheng, Subhabrata Mukherjee, Xin Luna Dong +1
cs.CLcs.AIcs.IRarXiv:1806.01264v22018TableLlama: Towards Open Large Generalist Models for Tables
Tianshu Zhang, Xiang Yue, Yifei Li +1
cs.CLcs.AIcs.DBarXiv:2311.09206v32023Conversations Gone Awry: Detecting Early Signs of Conversational Failure
Justine Zhang, Jonathan P. Chang, Cristian Danescu-Niculescu-Mizil +4
cs.CLcs.AIcs.CYarXiv:1805.05345v12018BLEU is Not Suitable for the Evaluation of Text Simplification
Elior Sulem, Omri Abend, Ari Rappoport
cs.CLarXiv:1810.05995v12018SecureBERT: A Domain-Specific Language Model for Cybersecurity
Ehsan Aghaei, Xi Niu, Waseem Shadid +1
cs.CLcs.AIcs.CRarXiv:2204.02685v32022Learning from Context or Names? An Empirical Study on Neural Relation Extraction
Hao Peng, Tianyu Gao, Xu Han +5
cs.CLarXiv:2010.01923v22020Multi-Task Cross-Lingual Sequence Tagging from Scratch
Zhilin Yang, Ruslan Salakhutdinov, William Cohen
cs.CLcs.LGarXiv:1603.06270v22016Deep Recurrent Models with Fast-Forward Connections for Neural Machine Translation
Jie Zhou, Ying Cao, Xuguang Wang +2
cs.CLcs.LGarXiv:1606.04199v32016RUBER: An Unsupervised Method for Automatic Evaluation of Open-Domain Dialog Systems
Chongyang Tao, Lili Mou, Dongyan Zhao +1
cs.CLcs.HCcs.IRarXiv:1701.03079v22017Learning Named Entity Tagger using Domain-Specific Dictionary
Jingbo Shang, Liyuan Liu, Xiang Ren +3
cs.CLarXiv:1809.03599v12018