Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

7,141 to 7,200 of 11,245

  1. Co-Writing Screenplays and Theatre Scripts with Language Models: An Evaluation by Industry Professionals

    Piotr Mirowski, Kory W. Mathewson, Jaylen Pittman +1

    cs.HCcs.CLarXiv:2209.14958v12022
  2. Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages

    Yu Zhang, Wei Han, James Qin +24

    cs.CLcs.SDeess.ASarXiv:2303.01037v32023
  3. ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases

    Qiaoyu Tang, Ziliang Deng, Hongyu Lin +4

    cs.CLarXiv:2306.05301v22023
  4. Linguistic Input Features Improve Neural Machine Translation

    Rico Sennrich, Barry Haddow

    cs.CLarXiv:1606.02892v22016
  5. MMGCN: Multimodal Fusion via Deep Graph Convolution Network for Emotion Recognition in Conversation

    Jingwen Hu, Yuchen Liu, Jinming Zhao +1

    cs.CLcs.SDeess.ASarXiv:2107.06779v12021
  6. The Second Conversational Intelligence Challenge (ConvAI2)

    Emily Dinan, Varvara Logacheva, Valentin Malykh +14

    cs.AIcs.CLcs.HCarXiv:1902.00098v12019
  7. Joint entity recognition and relation extraction as a multi-head selection problem

    Giannis Bekoulis, Johannes Deleu, Thomas Demeester +1

    cs.CLarXiv:1804.07847v32018
  8. Explain Images with Multimodal Recurrent Neural Networks

    Junhua Mao, Wei Xu, Yi Yang +2

    cs.CVcs.CLcs.LGarXiv:1410.1090v12014
  9. Community Interaction and Conflict on the Web

    Srijan Kumar, William L. Hamilton, Jure Leskovec +1

    cs.SIcs.CLcs.HCarXiv:1803.03697v12018
  10. Improving zero-shot learning by mitigating the hubness problem

    Georgiana Dinu, Angeliki Lazaridou, Marco Baroni

    cs.CLcs.LGarXiv:1412.6568v32014
  11. Levenshtein Transformer

    Jiatao Gu, Changhan Wang, Jake Zhao

    cs.CLcs.LGarXiv:1905.11006v22019
  12. A Literature Survey of Recent Advances in Chatbots

    Guendalina Caldarini, Sardar Jaf, Kenneth McGarry

    cs.CLarXiv:2201.06657v12022
  13. Sparse Sinkhorn Attention

    Yi Tay, Dara Bahri, Liu Yang +2

    cs.LGcs.CLarXiv:2002.11296v12020
  14. ToolQA: A Dataset for LLM Question Answering with External Tools

    Yuchen Zhuang, Yue Yu, Kuan Wang +2

    cs.CLcs.AIarXiv:2306.13304v12023
  15. Template-Based Named Entity Recognition Using BART

    Leyang Cui, Yu Wu, Jian Liu +2

    cs.CLarXiv:2106.01760v12021
  16. Exploring and Distilling Posterior and Prior Knowledge for Radiology Report Generation

    Fenglin Liu, Xian Wu, Shen Ge +2

    cs.CVcs.CLarXiv:2106.06963v22021
  17. Nematus: a Toolkit for Neural Machine Translation

    Rico Sennrich, Orhan Firat, Kyunghyun Cho +8

    cs.CLarXiv:1703.04357v12017
  18. A Real-World WebAgent with Planning, Long Context Understanding, and Program Synthesis

    Izzeddin Gur, Hiroki Furuta, Austin Huang +4

    cs.LGcs.AIcs.CLarXiv:2307.12856v42023
  19. Towards Understanding Chain-of-Thought Prompting: An Empirical Study of What Matters

    Boshi Wang, Sewon Min, Xiang Deng +4

    cs.CLarXiv:2212.10001v22022
  20. NaturalSpeech 2: Latent Diffusion Models are Natural and Zero-Shot Speech and Singing Synthesizers

    Kai Shen, Zeqian Ju, Xu Tan +6

    eess.AScs.AIcs.CLarXiv:2304.09116v32023
  21. $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

    Xinrong Zhang, Yingfa Chen, Shengding Hu +8

    cs.CLarXiv:2402.13718v32024
  22. Compositional Generalization via Structural Identification in a Category-Theoretic Framework

    Akihiro Maeda, Thomas Seiller, Yohei Oseki

    cs.CLstat.MLarXiv:2608.26465v12026
  23. Tying Word Vectors and Word Classifiers: A Loss Framework for Language Modeling

    Hakan Inan, Khashayar Khosravi, Richard Socher

    cs.LGcs.CLstat.MLarXiv:1611.01462v32016
  24. Blockwise Parallel Decoding for Deep Autoregressive Models

    Mitchell Stern, Noam Shazeer, Jakob Uszkoreit

    cs.LGcs.CLstat.MLarXiv:1811.03115v12018
  25. Analogical Inference for Multi-Relational Embeddings

    Hanxiao Liu, Yuexin Wu, Yiming Yang

    cs.LGcs.AIcs.CLarXiv:1705.02426v22017
  26. Retrieve, Program, Repeat: Complex Knowledge Base Question Answering via Alternate Meta-learning

    Yuncheng Hua, Yuan-Fang Li, Gholamreza Haffari +2

    cs.AIcs.CLarXiv:2010.15875v12020
  27. Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2

    Tom Lieberum, Senthooran Rajamanoharan, Arthur Conmy +7

    cs.LGcs.AIcs.CLarXiv:2408.05147v22024
  28. Fine-Grained Analysis of Propaganda in News Articles

    Giovanni Da San Martino, Seunghak Yu, Alberto Barrón-Cedeño +2

    cs.CLcs.AIcs.IRarXiv:1910.02517v12019
  29. Improving Conversational Recommender Systems via Knowledge Graph based Semantic Fusion

    Kun Zhou, Wayne Xin Zhao, Shuqing Bian +3

    cs.CLcs.AIcs.IRarXiv:2007.04032v12020
  30. A Survey on Recent Approaches for Natural Language Processing in Low-Resource Scenarios

    Michael A. Hedderich, Lukas Lange, Heike Adel +2

    cs.CLcs.LGarXiv:2010.12309v32020
  31. Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks

    Tomer Ullman

    cs.AIcs.CLarXiv:2302.08399v52023
  32. LoraHub: Efficient Cross-Task Generalization via Dynamic LoRA Composition

    Chengsong Huang, Qian Liu, Bill Yuchen Lin +3

    cs.CLcs.AIarXiv:2307.13269v32023
  33. Demographic Dialectal Variation in Social Media: A Case Study of African-American English

    Su Lin Blodgett, Lisa Green, Brendan O'Connor

    cs.CLarXiv:1608.08868v12016
  34. Assessing Gender Bias in Machine Translation -- A Case Study with Google Translate

    Marcelo O. R. Prates, Pedro H. C. Avelar, Luis Lamb

    cs.CYcs.CLarXiv:1809.02208v42018
  35. Neural Text Summarization: A Critical Evaluation

    Wojciech Kryściński, Nitish Shirish Keskar, Bryan McCann +2

    cs.CLarXiv:1908.08960v12019
  36. Trusting Your Evidence: Hallucinate Less with Context-aware Decoding

    Weijia Shi, Xiaochuang Han, Mike Lewis +3

    cs.CLarXiv:2305.14739v12023
  37. Towards Multimodal Sarcasm Detection (An _Obviously_ Perfect Paper)

    Santiago Castro, Devamanyu Hazarika, Verónica Pérez-Rosas +3

    cs.CLcs.CVarXiv:1906.01815v12019
  38. Span-based Localizing Network for Natural Language Video Localization

    Hao Zhang, Aixin Sun, Wei Jing +1

    cs.CLcs.CVarXiv:2004.13931v22020
  39. Speech Model Pre-training for End-to-End Spoken Language Understanding

    Loren Lugosch, Mirco Ravanelli, Patrick Ignoto +2

    eess.AScs.CLcs.LGarXiv:1904.03670v22019
  40. From Pretraining Data to Language Models to Downstream Tasks: Tracking the Trails of Political Biases Leading to Unfair NLP Models

    Shangbin Feng, Chan Young Park, Yuhan Liu +1

    cs.CLarXiv:2305.08283v32023
  41. The Thousand-Graph Hypothesis: A Testable Hypothesis of Task-Conditioned Relation Materialization in Repository-Level Code Reasoning

    Fei Ding

    cs.SEcs.CLarXiv:2608.26602v12026
  42. Research Design Tracking and Assessment for the Social Sciences

    Marco Rovera, Sergiu Burlacu, Dominique Cappelletti +5

    cs.CLcs.CYarXiv:2608.27049v12026
  43. Instruction Quality Matters: Refining Instructions for Effective Preference Learning

    Seohyeong Lee, Hwaran Lee, Buru Chang

    cs.CLarXiv:2608.26779v12026
  44. NLP Evaluation in trouble: On the Need to Measure LLM Data Contamination for each Benchmark

    Oscar Sainz, Jon Ander Campos, Iker García-Ferrero +3

    cs.CLarXiv:2310.18018v12023
  45. STAR : Sentence Translation Alignment Rate for Document-to-Document Machine Translation

    Yichen Dong, Hao Wang, Junhui Li +3

    cs.CLarXiv:2608.27161v12026
  46. Understanding Factuality in Abstractive Summarization with FRANK: A Benchmark for Factuality Metrics

    Artidoro Pagnoni, Vidhisha Balachandran, Yulia Tsvetkov

    cs.CLarXiv:2104.13346v22021
  47. SemEval-2013 Task 2: Sentiment Analysis in Twitter

    Preslav Nakov, Zornitsa Kozareva, Alan Ritter +3

    cs.CLcs.IRcs.LGarXiv:1912.06806v12019
  48. Multi-Expert Conformal Risk Control for Pairwise LLM Judging in Open-Ended Dialogue

    Ming Cheng, Yusheng Dai, Qiuhong Ke +2

    cs.CLarXiv:2608.26529v12026
  49. SWIFT:A Scalable lightWeight Infrastructure for Fine-Tuning

    Yuze Zhao, Jintao Huang, Jinghan Hu +10

    cs.CLarXiv:2408.05517v42024
  50. FOCUS & RePAIR: Mitigating Text Degeneration via Token-Level Guidance for Pruned Large Language Models

    Junyoung Lee, Sehyeon Park, Shinhyoung Jang +5

    cs.CLcs.AIcs.LGarXiv:2608.26676v12026
  51. Synthesizer: Rethinking Self-Attention in Transformer Models

    Yi Tay, Dara Bahri, Donald Metzler +3

    cs.CLcs.IRcs.LGarXiv:2005.00743v32020
  52. MEDITRON-70B: Scaling Medical Pretraining for Large Language Models

    Zeming Chen, Alejandro Hernández Cano, Angelika Romanou +17

    cs.CLcs.AIcs.LGarXiv:2311.16079v12023
  53. Multi-Hop Knowledge Graph Reasoning with Reward Shaping

    Xi Victoria Lin, Richard Socher, Caiming Xiong

    cs.AIcs.CLcs.LGarXiv:1808.10568v22018
  54. UniVL: A Unified Video and Language Pre-Training Model for Multimodal Understanding and Generation

    Huaishao Luo, Lei Ji, Botian Shi +6

    cs.CVcs.CLcs.LGarXiv:2002.06353v32020
  55. Cross-lingual Representation Learning via Centroid Intervention Fusion

    Wei Sun, Marie-Francine Moens

    cs.CLarXiv:2608.26357v12026
  56. Coarse-to-Fine Decoding for Neural Semantic Parsing

    Li Dong, Mirella Lapata

    cs.CLarXiv:1805.04793v12018
  57. Are All Languages Created Equal in Multilingual BERT?

    Shijie Wu, Mark Dredze

    cs.CLarXiv:2005.09093v22020
  58. CHiME-6 Challenge:Tackling Multispeaker Speech Recognition for Unsegmented Recordings

    Shinji Watanabe, Michael Mandel, Jon Barker +18

    cs.SDcs.CLeess.ASarXiv:2004.09249v22020
  59. Not Enough Data? Deep Learning to the Rescue!

    Ateret Anaby-Tavor, Boaz Carmeli, Esther Goldbraich +5

    cs.CLcs.LGarXiv:1911.03118v22019
  60. Good Friends, Bad News - Affect and Virality in Twitter

    Lars Kai Hansen, Adam Arvidsson, Finn Årup Nielsen +2

    cs.SIcs.CLphysics.soc-pharXiv:1101.0510v12011