Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

7,681 to 7,740 of 11,272

  1. A Comprehensive Survey on Applications of Transformers for Deep Learning Tasks

    Saidul Islam, Hanae Elmekki, Ahmed Elsebai +4

    cs.LGcs.CLarXiv:2306.07303v12023
  2. LinkBERT: Pretraining Language Models with Document Links

    Michihiro Yasunaga, Jure Leskovec, Percy Liang

    cs.CLcs.LGarXiv:2203.15827v12022
  3. From Crowdsourced Data to High-Quality Benchmarks: Arena-Hard and BenchBuilder Pipeline

    Tianle Li, Wei-Lin Chiang, Evan Frick +5

    cs.LGcs.AIcs.CLarXiv:2406.11939v22024
  4. Revisiting Few-sample BERT Fine-tuning

    Tianyi Zhang, Felix Wu, Arzoo Katiyar +2

    cs.CLcs.LGarXiv:2006.05987v32020
  5. Detecting Language Model Attacks with Perplexity

    Gabriel Alon, Michael Kamfonas

    cs.CLcs.AIcs.CRarXiv:2308.14132v32023
  6. ProphetNet: Predicting Future N-gram for Sequence-to-Sequence Pre-training

    Weizhen Qi, Yu Yan, Yeyun Gong +5

    cs.CLarXiv:2001.04063v32020
  7. A Comprehensive Survey of Hallucination Mitigation Techniques in Large Language Models

    S. M Towhidul Islam Tonmoy, S M Mehedi Zaman, Vinija Jain +4

    cs.CLarXiv:2401.01313v32024
  8. VL-Adapter: Parameter-Efficient Transfer Learning for Vision-and-Language Tasks

    Yi-Lin Sung, Jaemin Cho, Mohit Bansal

    cs.CVcs.AIcs.CLarXiv:2112.06825v22021
  9. QMSum: A New Benchmark for Query-based Multi-domain Meeting Summarization

    Ming Zhong, Da Yin, Tao Yu +8

    cs.CLarXiv:2104.05938v12021
  10. SearchQA: A New Q&A Dataset Augmented with Context from a Search Engine

    Matthew Dunn, Levent Sagun, Mike Higgins +3

    cs.CLarXiv:1704.05179v32017
  11. Sequential Matching Network: A New Architecture for Multi-turn Response Selection in Retrieval-based Chatbots

    Yu Wu, Wei Wu, Chen Xing +2

    cs.CLarXiv:1612.01627v22016
  12. KnowPrompt: Knowledge-aware Prompt-tuning with Synergistic Optimization for Relation Extraction

    Xiang Chen, Ningyu Zhang, Xin Xie +6

    cs.CLcs.AIcs.IRarXiv:2104.07650v72021
  13. Dealing with Disagreements: Looking Beyond the Majority Vote in Subjective Annotations

    Aida Mostafazadeh Davani, Mark Díaz, Vinodkumar Prabhakaran

    cs.CLcs.CYarXiv:2110.05719v12021
  14. OCRBench: On the Hidden Mystery of OCR in Large Multimodal Models

    Yuliang Liu, Zhang Li, Mingxin Huang +7

    cs.CVcs.CLarXiv:2305.07895v72023
  15. ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search

    Dan Zhang, Sining Zhoubian, Ziniu Hu +3

    cs.CLarXiv:2406.03816v32024
  16. Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations

    Zeming Wei, Yifei Wang, Ang Li +2

    cs.LGcs.AIcs.CLarXiv:2310.06387v32023
  17. Learning Hierarchy-Aware Knowledge Graph Embeddings for Link Prediction

    Zhanqiu Zhang, Jianyu Cai, Yongdong Zhang +1

    cs.LGcs.CLstat.MLarXiv:1911.09419v32019
  18. LMSYS-Chat-1M: A Large-Scale Real-World LLM Conversation Dataset

    Lianmin Zheng, Wei-Lin Chiang, Ying Sheng +10

    cs.CLcs.AIarXiv:2309.11998v42023
  19. Challenges and Applications of Large Language Models

    Jean Kaddour, Joshua Harris, Maximilian Mozes +3

    cs.CLcs.AIcs.LGarXiv:2307.10169v12023
  20. The Role of AI in Drug Discovery: Challenges, Opportunities, and Strategies

    Alexandre Blanco-Gonzalez, Alfonso Cabezon, Alejandro Seco-Gonzalez +4

    cs.CLcs.AIcs.CYarXiv:2212.08104v12022
  21. News Summarization and Evaluation in the Era of GPT-3

    Tanya Goyal, Junyi Jessy Li, Greg Durrett

    cs.CLarXiv:2209.12356v22022
  22. 12-in-1: Multi-Task Vision and Language Representation Learning

    Jiasen Lu, Vedanuj Goswami, Marcus Rohrbach +2

    cs.CVcs.CLcs.LGarXiv:1912.02315v22019
  23. Deeper Text Understanding for IR with Contextual Neural Language Modeling

    Zhuyun Dai, Jamie Callan

    cs.IRcs.CLarXiv:1905.09217v12019
  24. Prometheus 2: An Open Source Language Model Specialized in Evaluating Other Language Models

    Seungone Kim, Juyoung Suk, Shayne Longpre +7

    cs.CLarXiv:2405.01535v22024
  25. Hello Edge: Keyword Spotting on Microcontrollers

    Yundong Zhang, Naveen Suda, Liangzhen Lai +1

    cs.SDcs.CLcs.LGarXiv:1711.07128v32017
  26. Neural Belief Tracker: Data-Driven Dialogue State Tracking

    Nikola Mrkšić, Diarmuid Ó Séaghdha, Tsung-Hsien Wen +2

    cs.CLcs.AIcs.LGarXiv:1606.03777v22016
  27. Towards Understanding and Mitigating Social Biases in Language Models

    Paul Pu Liang, Chiyu Wu, Louis-Philippe Morency +1

    cs.CLcs.AIcs.CYarXiv:2106.13219v12021
  28. The KIT Motion-Language Dataset

    Matthias Plappert, Christian Mandery, Tamim Asfour

    cs.ROcs.CLcs.CVarXiv:1607.03827v22016
  29. LEGAL-BERT: The Muppets straight out of Law School

    Ilias Chalkidis, Manos Fergadiotis, Prodromos Malakasiotis +2

    cs.CLarXiv:2010.02559v12020
  30. Knowledge Unlearning for Mitigating Privacy Risks in Language Models

    Joel Jang, Dongkeun Yoon, Sohee Yang +4

    cs.CLarXiv:2210.01504v22022
  31. HateBERT: Retraining BERT for Abusive Language Detection in English

    Tommaso Caselli, Valerio Basile, Jelena Mitrović +1

    cs.CLarXiv:2010.12472v22020
  32. AmbigQA: Answering Ambiguous Open-domain Questions

    Sewon Min, Julian Michael, Hannaneh Hajishirzi +1

    cs.CLcs.AIarXiv:2004.10645v22020
  33. WHAM!: Extending Speech Separation to Noisy Environments

    Gordon Wichern, Joe Antognini, Michael Flynn +5

    cs.SDcs.CLcs.LGarXiv:1907.01160v12019
  34. Gated Linear Attention Transformers with Hardware-Efficient Training

    Songlin Yang, Bailin Wang, Yikang Shen +2

    cs.LGcs.CLarXiv:2312.06635v62023
  35. Large Language Models Empowered Agent-based Modeling and Simulation: A Survey and Perspectives

    Chen Gao, Xiaochong Lan, Nian Li +5

    cs.AIcs.CLcs.CYarXiv:2312.11970v12023
  36. Reducing Activation Recomputation in Large Transformer Models

    Vijay Korthikanti, Jared Casper, Sangkug Lym +4

    cs.LGcs.CLarXiv:2205.05198v12022
  37. TinyStories: How Small Can Language Models Be and Still Speak Coherent English?

    Ronen Eldan, Yuanzhi Li

    cs.CLcs.AIcs.LGarXiv:2305.07759v22023
  38. The WMDP Benchmark: Measuring and Reducing Malicious Use With Unlearning

    Nathaniel Li, Alexander Pan, Anjali Gopal +54

    cs.LGcs.AIcs.CLarXiv:2403.03218v72024
  39. Automatic Generation of Programming Exercises and Code Explanations using Large Language Models

    Sami Sarsa, Paul Denny, Arto Hellas +1

    cs.SEcs.AIcs.CLarXiv:2206.11861v22022
  40. Chat-REC: Towards Interactive and Explainable LLMs-Augmented Recommender System

    Yunfan Gao, Tao Sheng, Youlin Xiang +3

    cs.IRcs.CLcs.LGarXiv:2303.14524v22023
  41. Making the Most of Text Semantics to Improve Biomedical Vision--Language Processing

    Benedikt Boecking, Naoto Usuyama, Shruthi Bannur +9

    cs.CVcs.CLarXiv:2204.09817v42022
  42. Cosmos QA: Machine Reading Comprehension with Contextual Commonsense Reasoning

    Lifu Huang, Ronan Le Bras, Chandra Bhagavatula +1

    cs.CLcs.AIarXiv:1909.00277v22019
  43. Graph Convolutional Encoders for Syntax-aware Neural Machine Translation

    Jasmijn Bastings, Ivan Titov, Wilker Aziz +2

    cs.CLarXiv:1704.04675v42017
  44. Efficient Natural Language Response Suggestion for Smart Reply

    Matthew Henderson, Rami Al-Rfou, Brian Strope +6

    cs.CLarXiv:1705.00652v12017
  45. A Dataset of Information-Seeking Questions and Answers Anchored in Research Papers

    Pradeep Dasigi, Kyle Lo, Iz Beltagy +3

    cs.CLarXiv:2105.03011v12021
  46. Negative Preference Optimization: From Catastrophic Collapse to Effective Unlearning

    Ruiqi Zhang, Licong Lin, Yu Bai +1

    cs.LGcs.AIcs.CLarXiv:2404.05868v22024
  47. IndoNLU: Benchmark and Resources for Evaluating Indonesian Natural Language Understanding

    Bryan Wilie, Karissa Vincentio, Genta Indra Winata +8

    cs.CLarXiv:2009.05387v32020
  48. Data Augmentation for Low-Resource Neural Machine Translation

    Marzieh Fadaee, Arianna Bisazza, Christof Monz

    cs.CLarXiv:1705.00440v12017
  49. Beam Search Strategies for Neural Machine Translation

    Markus Freitag, Yaser Al-Onaizan

    cs.CLarXiv:1702.01806v22017
  50. Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs

    Suyu Ge, Yunan Zhang, Liyuan Liu +3

    cs.CLarXiv:2310.01801v42023
  51. Sparse, Dense, and Attentional Representations for Text Retrieval

    Yi Luan, Jacob Eisenstein, Kristina Toutanova +1

    cs.CLarXiv:2005.00181v32020
  52. Measuring Bias in Contextualized Word Representations

    Keita Kurita, Nidhi Vyas, Ayush Pareek +2

    cs.CLarXiv:1906.07337v12019
  53. Counter-fitting Word Vectors to Linguistic Constraints

    Nikola Mrkšić, Diarmuid Ó Séaghdha, Blaise Thomson +6

    cs.CLcs.LGarXiv:1603.00892v12016
  54. CodeRL: Mastering Code Generation through Pretrained Models and Deep Reinforcement Learning

    Hung Le, Yue Wang, Akhilesh Deepak Gotmare +2

    cs.LGcs.CLcs.PLarXiv:2207.01780v32022
  55. Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak Supervision

    Collin Burns, Pavel Izmailov, Jan Hendrik Kirchner +9

    cs.CLarXiv:2312.09390v12023
  56. REVERIE: Remote Embodied Visual Referring Expression in Real Indoor Environments

    Yuankai Qi, Qi Wu, Peter Anderson +4

    cs.CVcs.CLarXiv:1904.10151v22019
  57. Extractive Summarization as Text Matching

    Ming Zhong, Pengfei Liu, Yiran Chen +3

    cs.CLarXiv:2004.08795v12020
  58. Top2Vec: Distributed Representations of Topics

    Dimo Angelov

    cs.CLcs.LGstat.MLarXiv:2008.09470v12020
  59. SQLNet: Generating Structured Queries From Natural Language Without Reinforcement Learning

    Xiaojun Xu, Chang Liu, Dawn Song

    cs.CLcs.AIcs.DBarXiv:1711.04436v12017
  60. Self-Supervised Speech Representation Learning: A Review

    Abdelrahman Mohamed, Hung-yi Lee, Lasse Borgholt +9

    cs.CLcs.SDeess.ASarXiv:2205.10643v32022