Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,081 to 1,140 of 11,260

  1. Convolutional RNN: an Enhanced Model for Extracting Features from Sequential Data

    Gil Keren, Björn Schuller

    stat.MLcs.CLarXiv:1602.05875v32016
  2. Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers

    Yi Tay, Mostafa Dehghani, Jinfeng Rao +7

    cs.CLcs.AIcs.CVarXiv:2109.10686v22021
  3. Do Androids Laugh at Electric Sheep? Humor "Understanding" Benchmarks from The New Yorker Caption Contest

    Jack Hessel, Ana Marasović, Jena D. Hwang +5

    cs.CLcs.CVarXiv:2209.06293v22022
  4. Training Language Models with Memory Augmentation

    Zexuan Zhong, Tao Lei, Danqi Chen

    cs.CLcs.LGarXiv:2205.12674v32022
  5. Investigating Backtranslation in Neural Machine Translation

    Alberto Poncelas, Dimitar Shterionov, Andy Way +2

    cs.CLarXiv:1804.06189v12018
  6. State-of-the-art generalisation research in NLP: A taxonomy and review

    Dieuwke Hupkes, Mario Giulianelli, Verna Dankers +17

    cs.CLcs.AIarXiv:2210.03050v42022
  7. Detecting Fake News with Capsule Neural Networks

    Mohammad Hadi Goldani, Saeedeh Momtazi, Reza Safabakhsh

    cs.CLcs.CYarXiv:2002.01030v12020
  8. Continuous Speech Separation with Conformer

    Sanyuan Chen, Yu Wu, Zhuo Chen +6

    eess.AScs.CLarXiv:2008.05773v22020
  9. Towards Automated ICD Coding Using Deep Learning

    Haoran Shi, Pengtao Xie, Zhiting Hu +2

    cs.CLarXiv:1711.04075v32017
  10. Discourse-Aware Rumour Stance Classification in Social Media Using Sequential Classifiers

    Arkaitz Zubiaga, Elena Kochkina, Maria Liakata +5

    cs.CLcs.SIarXiv:1712.02223v12017
  11. A Survey on Employing Large Language Models for Text-to-SQL Tasks

    Liang Shi, Zhengju Tang, Nan Zhang +2

    cs.CLarXiv:2407.15186v52024
  12. Benchmarking Large Language Models on Answering and Explaining Challenging Medical Questions

    Hanjie Chen, Zhouxiang Fang, Yash Singla +1

    cs.CLarXiv:2402.18060v62024
  13. The Truth is in There: Improving Reasoning in Language Models with Layer-Selective Rank Reduction

    Pratyusha Sharma, Jordan T. Ash, Dipendra Misra

    cs.LGcs.AIcs.CLarXiv:2312.13558v12023
  14. Sustainable Modular Debiasing of Language Models

    Anne Lauscher, Tobias Lüken, Goran Glavaš

    cs.CLarXiv:2109.03646v12021
  15. Learning Structured Text Representations

    Yang Liu, Mirella Lapata

    cs.CLcs.AIarXiv:1705.09207v42017
  16. LayoutLLM: Layout Instruction Tuning with Large Language Models for Document Understanding

    Chuwei Luo, Yufan Shen, Zhaoqing Zhu +3

    cs.CVcs.CLarXiv:2404.05225v12024
  17. Embedding Multimodal Relational Data for Knowledge Base Completion

    Pouya Pezeshkpour, Liyan Chen, Sameer Singh

    cs.AIcs.CLstat.MLarXiv:1809.01341v22018
  18. Head-Driven Phrase Structure Grammar Parsing on Penn Treebank

    Junru Zhou, Hai Zhao

    cs.CLarXiv:1907.02684v42019
  19. Fully Non-autoregressive Neural Machine Translation: Tricks of the Trade

    Jiatao Gu, Xiang Kong

    cs.CLarXiv:2012.15833v12020
  20. Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

    Zhenting Qi, Mingyuan Ma, Jiahang Xu +3

    cs.CLarXiv:2408.06195v12024
  21. FastIF: Scalable Influence Functions for Efficient Model Interpretation and Debugging

    Han Guo, Nazneen Fatema Rajani, Peter Hase +2

    cs.LGcs.AIcs.CLarXiv:2012.15781v22020
  22. Contrasting Linguistic Patterns in Human and LLM-Generated News Text

    Alberto Muñoz-Ortiz, Carlos Gómez-Rodríguez, David Vilares

    cs.CLarXiv:2308.09067v32023
  23. Modeling Empathy and Distress in Reaction to News Stories

    Sven Buechel, Anneke Buffone, Barry Slaff +2

    cs.CLarXiv:1808.10399v12018
  24. LLM Dataset Inference: Did you train on my dataset?

    Pratyush Maini, Hengrui Jia, Nicolas Papernot +1

    cs.LGcs.CLcs.CRarXiv:2406.06443v12024
  25. CLEVR-Ref+: Diagnosing Visual Reasoning with Referring Expressions

    Runtao Liu, Chenxi Liu, Yutong Bai +1

    cs.CVcs.CLcs.LGarXiv:1901.00850v22019
  26. RelationPrompt: Leveraging Prompts to Generate Synthetic Data for Zero-Shot Relation Triplet Extraction

    Yew Ken Chia, Lidong Bing, Soujanya Poria +1

    cs.CLarXiv:2203.09101v12022
  27. AEGIS: Online Adaptive AI Content Safety Moderation with Ensemble of LLM Experts

    Shaona Ghosh, Prasoon Varshney, Erick Galinkin +1

    cs.LGcs.CLcs.CYarXiv:2404.05993v22024
  28. Simplicity Prevails: Rethinking Negative Preference Optimization for LLM Unlearning

    Chongyu Fan, Jiancheng Liu, Licong Lin +4

    cs.CLcs.AIcs.LGarXiv:2410.07163v42024
  29. A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions

    Ola Shorinwa, Zhiting Mei, Justin Lidard +2

    cs.CLcs.AIarXiv:2412.05563v22024
  30. Monotonic Multihead Attention

    Xutai Ma, Juan Pino, James Cross +2

    cs.CLarXiv:1909.12406v12019
  31. Symbol Emergence in Robotics: A Survey

    Tadahiro Taniguchi, Takayuki Nagai, Tomoaki Nakamura +3

    cs.AIcs.CLcs.CVarXiv:1509.08973v12015
  32. Generalization in Generation: A closer look at Exposure Bias

    Florian Schmidt

    cs.LGcs.CLstat.MLarXiv:1910.00292v22019
  33. Mitigating Large Language Model Hallucinations via Autonomous Knowledge Graph-based Retrofitting

    Xinyan Guan, Yanjiang Liu, Hongyu Lin +4

    cs.CLarXiv:2311.13314v12023
  34. Machine Translationese: Effects of Algorithmic Bias on Linguistic Complexity in Machine Translation

    Eva Vanmassenhove, Dimitar Shterionov, Matthew Gwilliam

    cs.CLcs.AIcs.CYarXiv:2102.00287v12021
  35. Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

    Terry Yue Zhuo, Yujin Huang, Chunyang Chen +1

    cs.CLcs.SEarXiv:2301.12867v42023
  36. GPT-3-driven pedagogical agents for training children's curious question-asking skills

    Rania Abdelghani, Yen-Hsiang Wang, Xingdi Yuan +4

    cs.CLcs.HCarXiv:2211.14228v62022
  37. An Information-theoretic Approach to Prompt Engineering Without Ground Truth Labels

    Taylor Sorensen, Joshua Robinson, Christopher Michael Rytting +6

    cs.CLcs.LGarXiv:2203.11364v12022
  38. Including Signed Languages in Natural Language Processing

    Kayo Yin, Amit Moryossef, Julie Hochgesang +2

    cs.CLcs.AIcs.LGarXiv:2105.05222v22021
  39. PARSER: Read in Parallel, Reason in Depth for Long-Context LLM Agents

    Kun Li, Zexuan Qiu, Tianhua Zhang +2

    cs.CLarXiv:2609.06702v12026
  40. Discuss Before Moving: Visual Language Navigation via Multi-expert Discussions

    Yuxing Long, Xiaoqi Li, Wenzhe Cai +1

    cs.ROcs.AIcs.CLarXiv:2309.11382v12023
  41. ZigMa: A DiT-style Zigzag Mamba Diffusion Model

    Vincent Tao Hu, Stefan Andreas Baumann, Ming Gui +4

    cs.CVcs.AIcs.CLarXiv:2403.13802v32024
  42. Abusing Images and Sounds for Indirect Instruction Injection in Multi-Modal LLMs

    Eugene Bagdasaryan, Tsung-Yin Hsieh, Ben Nassi +1

    cs.CRcs.AIcs.CLarXiv:2307.10490v42023
  43. WavLLM: Towards Robust and Adaptive Speech Large Language Model

    Shujie Hu, Long Zhou, Shujie Liu +9

    cs.CLcs.AIcs.SDarXiv:2404.00656v32024
  44. A Survey of Deep Learning Techniques for Neural Machine Translation

    Shuoheng Yang, Yuxin Wang, Xiaowen Chu

    cs.CLarXiv:2002.07526v12020
  45. A Teacher-Student Framework for Zero-Resource Neural Machine Translation

    Yun Chen, Yang Liu, Yong Cheng +1

    cs.CLarXiv:1705.00753v12017
  46. Deep Encoder, Shallow Decoder: Reevaluating Non-autoregressive Machine Translation

    Jungo Kasai, Nikolaos Pappas, Hao Peng +2

    cs.CLarXiv:2006.10369v42020
  47. Weather impacts expressed sentiment

    Patrick Baylis, Nick Obradovich, Yury Kryvasheyeu +5

    stat.APcs.CLarXiv:1709.00071v12017
  48. UNKs Everywhere: Adapting Multilingual Language Models to New Scripts

    Jonas Pfeiffer, Ivan Vulić, Iryna Gurevych +1

    cs.CLarXiv:2012.15562v32020
  49. Analyzing Transformers in Embedding Space

    Guy Dar, Mor Geva, Ankit Gupta +1

    cs.CLcs.LGarXiv:2209.02535v32022
  50. The VoicePrivacy 2020 Challenge: Results and findings

    Natalia Tomashenko, Xin Wang, Emmanuel Vincent +11

    cs.CLcs.SDeess.ASarXiv:2109.00648v42021
  51. Causal Analysis of Syntactic Agreement Mechanisms in Neural Language Models

    Matthew Finlayson, Aaron Mueller, Sebastian Gehrmann +3

    cs.CLarXiv:2106.06087v32021
  52. WARM: On the Benefits of Weight Averaged Reward Models

    Alexandre Ramé, Nino Vieillard, Léonard Hussenot +4

    cs.LGcs.AIcs.CLarXiv:2401.12187v12024
  53. Drug Repurposing for COVID-19 via Knowledge Graph Completion

    Rui Zhang, Dimitar Hristovski, Dalton Schutte +3

    cs.CLcs.IRarXiv:2010.09600v22020
  54. Auxiliary Signal-Guided Knowledge Encoder-Decoder for Medical Report Generation

    Mingjie Li, Fuyu Wang, Xiaojun Chang +1

    cs.CVcs.CLeess.IVarXiv:2006.03744v12020
  55. Revisiting Out-of-distribution Robustness in NLP: Benchmark, Analysis, and LLMs Evaluations

    Lifan Yuan, Yangyi Chen, Ganqu Cui +6

    cs.CLcs.CRcs.LGarXiv:2306.04618v22023
  56. Progressive-Hint Prompting Improves Reasoning in Large Language Models

    Chuanyang Zheng, Zhengying Liu, Enze Xie +2

    cs.CLcs.LGarXiv:2304.09797v62023
  57. Robust and fine-grained prosody control of end-to-end speech synthesis

    Younggun Lee, Taesu Kim

    cs.CLcs.LGcs.SDarXiv:1811.02122v22018
  58. VGCN-BERT: Augmenting BERT with Graph Embedding for Text Classification

    Zhibin Lu, Pan Du, Jian-Yun Nie

    cs.CLcs.LGstat.MLarXiv:2004.05707v12020
  59. In-Context Learning with Long-Context Models: An In-Depth Exploration

    Amanda Bertsch, Maor Ivgi, Emily Xiao +4

    cs.CLarXiv:2405.00200v22024
  60. Integrating Stance Detection and Fact Checking in a Unified Corpus

    Ramy Baly, Mitra Mohtarami, James Glass +3

    cs.CLarXiv:1804.08012v12018