Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
7,681 to 7,740 of 11,272
A Comprehensive Survey on Applications of Transformers for Deep Learning Tasks
Saidul Islam, Hanae Elmekki, Ahmed Elsebai +4
cs.LGcs.CLarXiv:2306.07303v12023LinkBERT: Pretraining Language Models with Document Links
Michihiro Yasunaga, Jure Leskovec, Percy Liang
cs.CLcs.LGarXiv:2203.15827v12022From Crowdsourced Data to High-Quality Benchmarks: Arena-Hard and BenchBuilder Pipeline
Tianle Li, Wei-Lin Chiang, Evan Frick +5
cs.LGcs.AIcs.CLarXiv:2406.11939v22024Revisiting Few-sample BERT Fine-tuning
Tianyi Zhang, Felix Wu, Arzoo Katiyar +2
cs.CLcs.LGarXiv:2006.05987v32020Detecting Language Model Attacks with Perplexity
Gabriel Alon, Michael Kamfonas
cs.CLcs.AIcs.CRarXiv:2308.14132v32023ProphetNet: Predicting Future N-gram for Sequence-to-Sequence Pre-training
Weizhen Qi, Yu Yan, Yeyun Gong +5
cs.CLarXiv:2001.04063v32020A Comprehensive Survey of Hallucination Mitigation Techniques in Large Language Models
S. M Towhidul Islam Tonmoy, S M Mehedi Zaman, Vinija Jain +4
cs.CLarXiv:2401.01313v32024VL-Adapter: Parameter-Efficient Transfer Learning for Vision-and-Language Tasks
Yi-Lin Sung, Jaemin Cho, Mohit Bansal
cs.CVcs.AIcs.CLarXiv:2112.06825v22021QMSum: A New Benchmark for Query-based Multi-domain Meeting Summarization
Ming Zhong, Da Yin, Tao Yu +8
cs.CLarXiv:2104.05938v12021SearchQA: A New Q&A Dataset Augmented with Context from a Search Engine
Matthew Dunn, Levent Sagun, Mike Higgins +3
cs.CLarXiv:1704.05179v32017Sequential Matching Network: A New Architecture for Multi-turn Response Selection in Retrieval-based Chatbots
Yu Wu, Wei Wu, Chen Xing +2
cs.CLarXiv:1612.01627v22016KnowPrompt: Knowledge-aware Prompt-tuning with Synergistic Optimization for Relation Extraction
Xiang Chen, Ningyu Zhang, Xin Xie +6
cs.CLcs.AIcs.IRarXiv:2104.07650v72021Dealing with Disagreements: Looking Beyond the Majority Vote in Subjective Annotations
Aida Mostafazadeh Davani, Mark Díaz, Vinodkumar Prabhakaran
cs.CLcs.CYarXiv:2110.05719v12021OCRBench: On the Hidden Mystery of OCR in Large Multimodal Models
Yuliang Liu, Zhang Li, Mingxin Huang +7
cs.CVcs.CLarXiv:2305.07895v72023ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search
Dan Zhang, Sining Zhoubian, Ziniu Hu +3
cs.CLarXiv:2406.03816v32024Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Zeming Wei, Yifei Wang, Ang Li +2
cs.LGcs.AIcs.CLarXiv:2310.06387v32023Learning Hierarchy-Aware Knowledge Graph Embeddings for Link Prediction
Zhanqiu Zhang, Jianyu Cai, Yongdong Zhang +1
cs.LGcs.CLstat.MLarXiv:1911.09419v32019LMSYS-Chat-1M: A Large-Scale Real-World LLM Conversation Dataset
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng +10
cs.CLcs.AIarXiv:2309.11998v42023Challenges and Applications of Large Language Models
Jean Kaddour, Joshua Harris, Maximilian Mozes +3
cs.CLcs.AIcs.LGarXiv:2307.10169v12023The Role of AI in Drug Discovery: Challenges, Opportunities, and Strategies
Alexandre Blanco-Gonzalez, Alfonso Cabezon, Alejandro Seco-Gonzalez +4
cs.CLcs.AIcs.CYarXiv:2212.08104v12022News Summarization and Evaluation in the Era of GPT-3
Tanya Goyal, Junyi Jessy Li, Greg Durrett
cs.CLarXiv:2209.12356v2202212-in-1: Multi-Task Vision and Language Representation Learning
Jiasen Lu, Vedanuj Goswami, Marcus Rohrbach +2
cs.CVcs.CLcs.LGarXiv:1912.02315v22019Deeper Text Understanding for IR with Contextual Neural Language Modeling
Zhuyun Dai, Jamie Callan
cs.IRcs.CLarXiv:1905.09217v12019Prometheus 2: An Open Source Language Model Specialized in Evaluating Other Language Models
Seungone Kim, Juyoung Suk, Shayne Longpre +7
cs.CLarXiv:2405.01535v22024Hello Edge: Keyword Spotting on Microcontrollers
Yundong Zhang, Naveen Suda, Liangzhen Lai +1
cs.SDcs.CLcs.LGarXiv:1711.07128v32017Neural Belief Tracker: Data-Driven Dialogue State Tracking
Nikola Mrkšić, Diarmuid Ó Séaghdha, Tsung-Hsien Wen +2
cs.CLcs.AIcs.LGarXiv:1606.03777v22016Towards Understanding and Mitigating Social Biases in Language Models
Paul Pu Liang, Chiyu Wu, Louis-Philippe Morency +1
cs.CLcs.AIcs.CYarXiv:2106.13219v12021The KIT Motion-Language Dataset
Matthias Plappert, Christian Mandery, Tamim Asfour
cs.ROcs.CLcs.CVarXiv:1607.03827v22016LEGAL-BERT: The Muppets straight out of Law School
Ilias Chalkidis, Manos Fergadiotis, Prodromos Malakasiotis +2
cs.CLarXiv:2010.02559v12020Knowledge Unlearning for Mitigating Privacy Risks in Language Models
Joel Jang, Dongkeun Yoon, Sohee Yang +4
cs.CLarXiv:2210.01504v22022HateBERT: Retraining BERT for Abusive Language Detection in English
Tommaso Caselli, Valerio Basile, Jelena Mitrović +1
cs.CLarXiv:2010.12472v22020AmbigQA: Answering Ambiguous Open-domain Questions
Sewon Min, Julian Michael, Hannaneh Hajishirzi +1
cs.CLcs.AIarXiv:2004.10645v22020WHAM!: Extending Speech Separation to Noisy Environments
Gordon Wichern, Joe Antognini, Michael Flynn +5
cs.SDcs.CLcs.LGarXiv:1907.01160v12019Gated Linear Attention Transformers with Hardware-Efficient Training
Songlin Yang, Bailin Wang, Yikang Shen +2
cs.LGcs.CLarXiv:2312.06635v62023Large Language Models Empowered Agent-based Modeling and Simulation: A Survey and Perspectives
Chen Gao, Xiaochong Lan, Nian Li +5
cs.AIcs.CLcs.CYarXiv:2312.11970v12023Reducing Activation Recomputation in Large Transformer Models
Vijay Korthikanti, Jared Casper, Sangkug Lym +4
cs.LGcs.CLarXiv:2205.05198v12022TinyStories: How Small Can Language Models Be and Still Speak Coherent English?
Ronen Eldan, Yuanzhi Li
cs.CLcs.AIcs.LGarXiv:2305.07759v22023The WMDP Benchmark: Measuring and Reducing Malicious Use With Unlearning
Nathaniel Li, Alexander Pan, Anjali Gopal +54
cs.LGcs.AIcs.CLarXiv:2403.03218v72024Automatic Generation of Programming Exercises and Code Explanations using Large Language Models
Sami Sarsa, Paul Denny, Arto Hellas +1
cs.SEcs.AIcs.CLarXiv:2206.11861v22022Chat-REC: Towards Interactive and Explainable LLMs-Augmented Recommender System
Yunfan Gao, Tao Sheng, Youlin Xiang +3
cs.IRcs.CLcs.LGarXiv:2303.14524v22023Making the Most of Text Semantics to Improve Biomedical Vision--Language Processing
Benedikt Boecking, Naoto Usuyama, Shruthi Bannur +9
cs.CVcs.CLarXiv:2204.09817v42022Cosmos QA: Machine Reading Comprehension with Contextual Commonsense Reasoning
Lifu Huang, Ronan Le Bras, Chandra Bhagavatula +1
cs.CLcs.AIarXiv:1909.00277v22019Graph Convolutional Encoders for Syntax-aware Neural Machine Translation
Jasmijn Bastings, Ivan Titov, Wilker Aziz +2
cs.CLarXiv:1704.04675v42017Efficient Natural Language Response Suggestion for Smart Reply
Matthew Henderson, Rami Al-Rfou, Brian Strope +6
cs.CLarXiv:1705.00652v12017A Dataset of Information-Seeking Questions and Answers Anchored in Research Papers
Pradeep Dasigi, Kyle Lo, Iz Beltagy +3
cs.CLarXiv:2105.03011v12021Negative Preference Optimization: From Catastrophic Collapse to Effective Unlearning
Ruiqi Zhang, Licong Lin, Yu Bai +1
cs.LGcs.AIcs.CLarXiv:2404.05868v22024IndoNLU: Benchmark and Resources for Evaluating Indonesian Natural Language Understanding
Bryan Wilie, Karissa Vincentio, Genta Indra Winata +8
cs.CLarXiv:2009.05387v32020Data Augmentation for Low-Resource Neural Machine Translation
Marzieh Fadaee, Arianna Bisazza, Christof Monz
cs.CLarXiv:1705.00440v12017Beam Search Strategies for Neural Machine Translation
Markus Freitag, Yaser Al-Onaizan
cs.CLarXiv:1702.01806v22017Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs
Suyu Ge, Yunan Zhang, Liyuan Liu +3
cs.CLarXiv:2310.01801v42023Sparse, Dense, and Attentional Representations for Text Retrieval
Yi Luan, Jacob Eisenstein, Kristina Toutanova +1
cs.CLarXiv:2005.00181v32020Measuring Bias in Contextualized Word Representations
Keita Kurita, Nidhi Vyas, Ayush Pareek +2
cs.CLarXiv:1906.07337v12019Counter-fitting Word Vectors to Linguistic Constraints
Nikola Mrkšić, Diarmuid Ó Séaghdha, Blaise Thomson +6
cs.CLcs.LGarXiv:1603.00892v12016CodeRL: Mastering Code Generation through Pretrained Models and Deep Reinforcement Learning
Hung Le, Yue Wang, Akhilesh Deepak Gotmare +2
cs.LGcs.CLcs.PLarXiv:2207.01780v32022Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak Supervision
Collin Burns, Pavel Izmailov, Jan Hendrik Kirchner +9
cs.CLarXiv:2312.09390v12023REVERIE: Remote Embodied Visual Referring Expression in Real Indoor Environments
Yuankai Qi, Qi Wu, Peter Anderson +4
cs.CVcs.CLarXiv:1904.10151v22019Extractive Summarization as Text Matching
Ming Zhong, Pengfei Liu, Yiran Chen +3
cs.CLarXiv:2004.08795v12020Top2Vec: Distributed Representations of Topics
Dimo Angelov
cs.CLcs.LGstat.MLarXiv:2008.09470v12020SQLNet: Generating Structured Queries From Natural Language Without Reinforcement Learning
Xiaojun Xu, Chang Liu, Dawn Song
cs.CLcs.AIcs.DBarXiv:1711.04436v12017Self-Supervised Speech Representation Learning: A Review
Abdelrahman Mohamed, Hung-yi Lee, Lasse Borgholt +9
cs.CLcs.SDeess.ASarXiv:2205.10643v32022