Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,601 to 6,660 of 11,332
Federated Learning for Keyword Spotting
David Leroy, Alice Coucke, Thibaut Lavril +2
eess.AScs.CLcs.LGarXiv:1810.05512v42018Comparing BERT against traditional machine learning text classification
Santiago González-Carvajal, Eduardo C. Garrido-Merchán
cs.CLcs.LGstat.MLarXiv:2005.13012v22020Neural Speech Recognizer: Acoustic-to-Word LSTM Model for Large Vocabulary Speech Recognition
Hagen Soltau, Hank Liao, Hasim Sak
cs.CLcs.LGcs.NEarXiv:1610.09975v12016Jailbreak in pieces: Compositional Adversarial Attacks on Multi-Modal Language Models
Erfan Shayegani, Yue Dong, Nael Abu-Ghazaleh
cs.CRcs.CLarXiv:2307.14539v22023Generating Natural Questions About an Image
Nasrin Mostafazadeh, Ishan Misra, Jacob Devlin +3
cs.CLcs.AIcs.CVarXiv:1603.06059v32016RocketQAv2: A Joint Training Method for Dense Passage Retrieval and Passage Re-ranking
Ruiyang Ren, Yingqi Qu, Jing Liu +5
cs.CLarXiv:2110.07367v22021Depthwise Separable Convolutions for Neural Machine Translation
Lukasz Kaiser, Aidan N. Gomez, Francois Chollet
cs.CLcs.LGarXiv:1706.03059v22017On the Robustness of ChatGPT: An Adversarial and Out-of-distribution Perspective
Jindong Wang, Xixu Hu, Wenxin Hou +10
cs.AIcs.CLcs.LGarXiv:2302.12095v52023IconQA: A New Benchmark for Abstract Diagram Understanding and Visual Language Reasoning
Pan Lu, Liang Qiu, Jiaqi Chen +6
cs.CVcs.AIcs.CLarXiv:2110.13214v42021Multimodal Sentiment Analysis with Word-Level Fusion and Reinforcement Learning
Minghai Chen, Sen Wang, Paul Pu Liang +3
cs.LGcs.AIcs.CLarXiv:1802.00924v12018Unleashing the Emergent Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration
Zhenhailong Wang, Shaoguang Mao, Wenshan Wu +3
cs.AIcs.CLarXiv:2307.05300v42023Toward Abstractive Summarization Using Semantic Representations
Fei Liu, Jeffrey Flanigan, Sam Thomson +2
cs.CLarXiv:1805.10399v12018oLMpics -- On what Language Model Pre-training Captures
Alon Talmor, Yanai Elazar, Yoav Goldberg +1
cs.CLcs.AIcs.LGarXiv:1912.13283v22019Grounded Language Learning in a Simulated 3D World
Karl Moritz Hermann, Felix Hill, Simon Green +11
cs.CLcs.LGstat.MLarXiv:1706.06551v22017Short Text Topic Modeling Techniques, Applications, and Performance: A Survey
Qiang Jipeng, Qian Zhenyu, Li Yun +2
cs.IRcs.CLarXiv:1904.07695v12019LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models
Huiqiang Jiang, Qianhui Wu, Chin-Yew Lin +2
cs.CLcs.LGarXiv:2310.05736v22023Augmented SBERT: Data Augmentation Method for Improving Bi-Encoders for Pairwise Sentence Scoring Tasks
Nandan Thakur, Nils Reimers, Johannes Daxenberger +1
cs.CLarXiv:2010.08240v22020Sparse Sequence-to-Sequence Models
Ben Peters, Vlad Niculae, André F. T. Martins
cs.CLcs.LGarXiv:1905.05702v22019Text Level Graph Neural Network for Text Classification
Lianzhe Huang, Dehong Ma, Sujian Li +2
cs.CLarXiv:1910.02356v22019Automated Fact-Checking for Assisting Human Fact-Checkers
Preslav Nakov, David Corney, Maram Hasanain +6
cs.AIcs.CLcs.CRarXiv:2103.07769v22021Scientific Paper Summarization Using Citation Summary Networks
Vahed Qazvinian, Dragomir R. Radev
cs.IRcs.CLarXiv:0807.1560v12008LLaVAR: Enhanced Visual Instruction Tuning for Text-Rich Image Understanding
Yanzhe Zhang, Ruiyi Zhang, Jiuxiang Gu +4
cs.CVcs.CLarXiv:2306.17107v22023Deep Communicating Agents for Abstractive Summarization
Asli Celikyilmaz, Antoine Bosselut, Xiaodong He +1
cs.CLarXiv:1803.10357v32018End-to-End Neural Speaker Diarization with Permutation-Free Objectives
Yusuke Fujita, Naoyuki Kanda, Shota Horiguchi +2
eess.AScs.CLcs.SDarXiv:1909.05952v12019Generating Informative and Diverse Conversational Responses via Adversarial Information Maximization
Yizhe Zhang, Michel Galley, Jianfeng Gao +4
cs.CLcs.AIarXiv:1809.05972v52018Measuring and Reducing Gendered Correlations in Pre-trained Models
Kellie Webster, Xuezhi Wang, Ian Tenney +6
cs.CLarXiv:2010.06032v22020Unsupervised Quality Estimation for Neural Machine Translation
Marina Fomicheva, Shuo Sun, Lisa Yankovskaya +6
cs.CLarXiv:2005.10608v22020Text Understanding with the Attention Sum Reader Network
Rudolf Kadlec, Martin Schmid, Ondrej Bajgar +1
cs.CLarXiv:1603.01547v22016Holodeck: Language Guided Generation of 3D Embodied AI Environments
Yue Yang, Fan-Yun Sun, Luca Weihs +11
cs.CVcs.AIcs.CLarXiv:2312.09067v22023Attentional Encoder Network for Targeted Sentiment Classification
Youwei Song, Jiahai Wang, Tao Jiang +2
cs.CLarXiv:1902.09314v22019Neural AMR: Sequence-to-Sequence Models for Parsing and Generation
Ioannis Konstas, Srinivasan Iyer, Mark Yatskar +2
cs.CLarXiv:1704.08381v32017Structural Scaffolds for Citation Intent Classification in Scientific Publications
Arman Cohan, Waleed Ammar, Madeleine van Zuylen +1
cs.CLarXiv:1904.01608v22019DyLoRA: Parameter Efficient Tuning of Pre-trained Models using Dynamic Search-Free Low-Rank Adaptation
Mojtaba Valipour, Mehdi Rezagholizadeh, Ivan Kobyzev +1
cs.CLcs.LGarXiv:2210.07558v22022mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models
Jiabo Ye, Haiyang Xu, Haowei Liu +6
cs.CVcs.AIcs.CLarXiv:2408.04840v22024Interpreting and improving natural-language processing (in machines) with natural language-processing (in the brain)
Mariya Toneva, Leila Wehbe
cs.CLcs.AIcs.LGarXiv:1905.11833v42019Towards Best Practices of Activation Patching in Language Models: Metrics and Methods
Fred Zhang, Neel Nanda
cs.LGcs.AIcs.CLarXiv:2309.16042v22023JudgeBench: A Benchmark for Evaluating LLM-based Judges
Sijun Tan, Siyuan Zhuang, Kyle Montgomery +5
cs.AIcs.CLcs.LGarXiv:2410.12784v22024DREAM: A Challenge Dataset and Models for Dialogue-Based Reading Comprehension
Kai Sun, Dian Yu, Jianshu Chen +3
cs.CLarXiv:1902.00164v12019StyleGAN-NADA: CLIP-Guided Domain Adaptation of Image Generators
Rinon Gal, Or Patashnik, Haggai Maron +2
cs.CVcs.CLcs.GRarXiv:2108.00946v22021Deep Semantic Role Labeling with Self-Attention
Zhixing Tan, Mingxuan Wang, Jun Xie +2
cs.CLarXiv:1712.01586v12017How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition
Guanting Dong, Hongyi Yuan, Keming Lu +7
cs.CLcs.AIcs.LGarXiv:2310.05492v42023Unlocking Efficiency in Large Language Model Inference: A Comprehensive Survey of Speculative Decoding
Heming Xia, Zhe Yang, Qingxiu Dong +6
cs.CLarXiv:2401.07851v32024Object Relational Graph with Teacher-Recommended Learning for Video Captioning
Ziqi Zhang, Yaya Shi, Chunfeng Yuan +4
cs.CVcs.CLarXiv:2002.11566v12020CoType: Joint Extraction of Typed Entities and Relations with Knowledge Bases
Xiang Ren, Zeqiu Wu, Wenqi He +5
cs.CLcs.LGarXiv:1610.08763v22016Multilingual Models for Compositional Distributed Semantics
Karl Moritz Hermann, Phil Blunsom
cs.CLarXiv:1404.4641v12014Building Cooperative Embodied Agents Modularly with Large Language Models
Hongxin Zhang, Weihua Du, Jiaming Shan +5
cs.AIcs.CLcs.CVarXiv:2307.02485v22023Hierarchical Transformers for Multi-Document Summarization
Yang Liu, Mirella Lapata
cs.CLcs.AIarXiv:1905.13164v12019Roget's Thesaurus and Semantic Similarity
Mario Jarmasz, Stan Szpakowicz
cs.CLarXiv:1204.0245v12012ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving
Zhibin Gou, Zhihong Shao, Yeyun Gong +5
cs.CLcs.AIarXiv:2309.17452v42023Think-on-Graph: Deep and Responsible Reasoning of Large Language Model on Knowledge Graph
Jiashuo Sun, Chengjin Xu, Lumingyuan Tang +6
cs.CLarXiv:2307.07697v62023Thinking Fast and Slow in Large Language Models
Thilo Hagendorff, Sarah Fabi, Michal Kosinski
cs.CLcs.AIcs.LGarXiv:2212.05206v22022Can ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERT
Qihuang Zhong, Liang Ding, Juhua Liu +2
cs.CLarXiv:2302.10198v22023Cross-Lingual Transfer Learning for Multilingual Task Oriented Dialog
Sebastian Schuster, Sonal Gupta, Rushin Shah +1
cs.CLarXiv:1810.13327v22018A New Generation of Perspective API: Efficient Multilingual Character-level Transformers
Alyssa Lees, Vinh Q. Tran, Yi Tay +4
cs.CLcs.AIcs.CYarXiv:2202.11176v12022Demystifying Prompts in Language Models via Perplexity Estimation
Hila Gonen, Srini Iyer, Terra Blevins +2
cs.CLarXiv:2212.04037v22022Understanding the Difficulty of Training Transformers
Liyuan Liu, Xiaodong Liu, Jianfeng Gao +2
cs.LGcs.CLstat.MLarXiv:2004.08249v32020Mem2Seq: Effectively Incorporating Knowledge Bases into End-to-End Task-Oriented Dialog Systems
Andrea Madotto, Chien-Sheng Wu, Pascale Fung
cs.CLarXiv:1804.08217v32018Show Your Work: Improved Reporting of Experimental Results
Jesse Dodge, Suchin Gururangan, Dallas Card +2
cs.LGcs.CLstat.MEarXiv:1909.03004v12019MultiHop-RAG: Benchmarking Retrieval-Augmented Generation for Multi-Hop Queries
Yixuan Tang, Yi Yang
cs.CLarXiv:2401.15391v12024Deep Voice 3: Scaling Text-to-Speech with Convolutional Sequence Learning
Wei Ping, Kainan Peng, Andrew Gibiansky +5
cs.SDcs.AIcs.CLarXiv:1710.07654v32017