Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
7,141 to 7,200 of 11,245
Co-Writing Screenplays and Theatre Scripts with Language Models: An Evaluation by Industry Professionals
Piotr Mirowski, Kory W. Mathewson, Jaylen Pittman +1
cs.HCcs.CLarXiv:2209.14958v12022Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
Yu Zhang, Wei Han, James Qin +24
cs.CLcs.SDeess.ASarXiv:2303.01037v32023ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases
Qiaoyu Tang, Ziliang Deng, Hongyu Lin +4
cs.CLarXiv:2306.05301v22023Linguistic Input Features Improve Neural Machine Translation
Rico Sennrich, Barry Haddow
cs.CLarXiv:1606.02892v22016MMGCN: Multimodal Fusion via Deep Graph Convolution Network for Emotion Recognition in Conversation
Jingwen Hu, Yuchen Liu, Jinming Zhao +1
cs.CLcs.SDeess.ASarXiv:2107.06779v12021The Second Conversational Intelligence Challenge (ConvAI2)
Emily Dinan, Varvara Logacheva, Valentin Malykh +14
cs.AIcs.CLcs.HCarXiv:1902.00098v12019Joint entity recognition and relation extraction as a multi-head selection problem
Giannis Bekoulis, Johannes Deleu, Thomas Demeester +1
cs.CLarXiv:1804.07847v32018Explain Images with Multimodal Recurrent Neural Networks
Junhua Mao, Wei Xu, Yi Yang +2
cs.CVcs.CLcs.LGarXiv:1410.1090v12014Community Interaction and Conflict on the Web
Srijan Kumar, William L. Hamilton, Jure Leskovec +1
cs.SIcs.CLcs.HCarXiv:1803.03697v12018Improving zero-shot learning by mitigating the hubness problem
Georgiana Dinu, Angeliki Lazaridou, Marco Baroni
cs.CLcs.LGarXiv:1412.6568v32014Levenshtein Transformer
Jiatao Gu, Changhan Wang, Jake Zhao
cs.CLcs.LGarXiv:1905.11006v22019A Literature Survey of Recent Advances in Chatbots
Guendalina Caldarini, Sardar Jaf, Kenneth McGarry
cs.CLarXiv:2201.06657v12022Sparse Sinkhorn Attention
Yi Tay, Dara Bahri, Liu Yang +2
cs.LGcs.CLarXiv:2002.11296v12020ToolQA: A Dataset for LLM Question Answering with External Tools
Yuchen Zhuang, Yue Yu, Kuan Wang +2
cs.CLcs.AIarXiv:2306.13304v12023Template-Based Named Entity Recognition Using BART
Leyang Cui, Yu Wu, Jian Liu +2
cs.CLarXiv:2106.01760v12021Exploring and Distilling Posterior and Prior Knowledge for Radiology Report Generation
Fenglin Liu, Xian Wu, Shen Ge +2
cs.CVcs.CLarXiv:2106.06963v22021Nematus: a Toolkit for Neural Machine Translation
Rico Sennrich, Orhan Firat, Kyunghyun Cho +8
cs.CLarXiv:1703.04357v12017A Real-World WebAgent with Planning, Long Context Understanding, and Program Synthesis
Izzeddin Gur, Hiroki Furuta, Austin Huang +4
cs.LGcs.AIcs.CLarXiv:2307.12856v42023Towards Understanding Chain-of-Thought Prompting: An Empirical Study of What Matters
Boshi Wang, Sewon Min, Xiang Deng +4
cs.CLarXiv:2212.10001v22022NaturalSpeech 2: Latent Diffusion Models are Natural and Zero-Shot Speech and Singing Synthesizers
Kai Shen, Zeqian Ju, Xu Tan +6
eess.AScs.AIcs.CLarXiv:2304.09116v32023$\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens
Xinrong Zhang, Yingfa Chen, Shengding Hu +8
cs.CLarXiv:2402.13718v32024Compositional Generalization via Structural Identification in a Category-Theoretic Framework
Akihiro Maeda, Thomas Seiller, Yohei Oseki
cs.CLstat.MLarXiv:2608.26465v12026Tying Word Vectors and Word Classifiers: A Loss Framework for Language Modeling
Hakan Inan, Khashayar Khosravi, Richard Socher
cs.LGcs.CLstat.MLarXiv:1611.01462v32016Blockwise Parallel Decoding for Deep Autoregressive Models
Mitchell Stern, Noam Shazeer, Jakob Uszkoreit
cs.LGcs.CLstat.MLarXiv:1811.03115v12018Analogical Inference for Multi-Relational Embeddings
Hanxiao Liu, Yuexin Wu, Yiming Yang
cs.LGcs.AIcs.CLarXiv:1705.02426v22017Retrieve, Program, Repeat: Complex Knowledge Base Question Answering via Alternate Meta-learning
Yuncheng Hua, Yuan-Fang Li, Gholamreza Haffari +2
cs.AIcs.CLarXiv:2010.15875v12020Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2
Tom Lieberum, Senthooran Rajamanoharan, Arthur Conmy +7
cs.LGcs.AIcs.CLarXiv:2408.05147v22024Fine-Grained Analysis of Propaganda in News Articles
Giovanni Da San Martino, Seunghak Yu, Alberto Barrón-Cedeño +2
cs.CLcs.AIcs.IRarXiv:1910.02517v12019Improving Conversational Recommender Systems via Knowledge Graph based Semantic Fusion
Kun Zhou, Wayne Xin Zhao, Shuqing Bian +3
cs.CLcs.AIcs.IRarXiv:2007.04032v12020A Survey on Recent Approaches for Natural Language Processing in Low-Resource Scenarios
Michael A. Hedderich, Lukas Lange, Heike Adel +2
cs.CLcs.LGarXiv:2010.12309v32020Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks
Tomer Ullman
cs.AIcs.CLarXiv:2302.08399v52023LoraHub: Efficient Cross-Task Generalization via Dynamic LoRA Composition
Chengsong Huang, Qian Liu, Bill Yuchen Lin +3
cs.CLcs.AIarXiv:2307.13269v32023Demographic Dialectal Variation in Social Media: A Case Study of African-American English
Su Lin Blodgett, Lisa Green, Brendan O'Connor
cs.CLarXiv:1608.08868v12016Assessing Gender Bias in Machine Translation -- A Case Study with Google Translate
Marcelo O. R. Prates, Pedro H. C. Avelar, Luis Lamb
cs.CYcs.CLarXiv:1809.02208v42018Neural Text Summarization: A Critical Evaluation
Wojciech Kryściński, Nitish Shirish Keskar, Bryan McCann +2
cs.CLarXiv:1908.08960v12019Trusting Your Evidence: Hallucinate Less with Context-aware Decoding
Weijia Shi, Xiaochuang Han, Mike Lewis +3
cs.CLarXiv:2305.14739v12023Towards Multimodal Sarcasm Detection (An _Obviously_ Perfect Paper)
Santiago Castro, Devamanyu Hazarika, Verónica Pérez-Rosas +3
cs.CLcs.CVarXiv:1906.01815v12019Span-based Localizing Network for Natural Language Video Localization
Hao Zhang, Aixin Sun, Wei Jing +1
cs.CLcs.CVarXiv:2004.13931v22020Speech Model Pre-training for End-to-End Spoken Language Understanding
Loren Lugosch, Mirco Ravanelli, Patrick Ignoto +2
eess.AScs.CLcs.LGarXiv:1904.03670v22019From Pretraining Data to Language Models to Downstream Tasks: Tracking the Trails of Political Biases Leading to Unfair NLP Models
Shangbin Feng, Chan Young Park, Yuhan Liu +1
cs.CLarXiv:2305.08283v32023The Thousand-Graph Hypothesis: A Testable Hypothesis of Task-Conditioned Relation Materialization in Repository-Level Code Reasoning
Fei Ding
cs.SEcs.CLarXiv:2608.26602v12026Research Design Tracking and Assessment for the Social Sciences
Marco Rovera, Sergiu Burlacu, Dominique Cappelletti +5
cs.CLcs.CYarXiv:2608.27049v12026Instruction Quality Matters: Refining Instructions for Effective Preference Learning
Seohyeong Lee, Hwaran Lee, Buru Chang
cs.CLarXiv:2608.26779v12026NLP Evaluation in trouble: On the Need to Measure LLM Data Contamination for each Benchmark
Oscar Sainz, Jon Ander Campos, Iker García-Ferrero +3
cs.CLarXiv:2310.18018v12023STAR : Sentence Translation Alignment Rate for Document-to-Document Machine Translation
Yichen Dong, Hao Wang, Junhui Li +3
cs.CLarXiv:2608.27161v12026Understanding Factuality in Abstractive Summarization with FRANK: A Benchmark for Factuality Metrics
Artidoro Pagnoni, Vidhisha Balachandran, Yulia Tsvetkov
cs.CLarXiv:2104.13346v22021SemEval-2013 Task 2: Sentiment Analysis in Twitter
Preslav Nakov, Zornitsa Kozareva, Alan Ritter +3
cs.CLcs.IRcs.LGarXiv:1912.06806v12019Multi-Expert Conformal Risk Control for Pairwise LLM Judging in Open-Ended Dialogue
Ming Cheng, Yusheng Dai, Qiuhong Ke +2
cs.CLarXiv:2608.26529v12026SWIFT:A Scalable lightWeight Infrastructure for Fine-Tuning
Yuze Zhao, Jintao Huang, Jinghan Hu +10
cs.CLarXiv:2408.05517v42024FOCUS & RePAIR: Mitigating Text Degeneration via Token-Level Guidance for Pruned Large Language Models
Junyoung Lee, Sehyeon Park, Shinhyoung Jang +5
cs.CLcs.AIcs.LGarXiv:2608.26676v12026Synthesizer: Rethinking Self-Attention in Transformer Models
Yi Tay, Dara Bahri, Donald Metzler +3
cs.CLcs.IRcs.LGarXiv:2005.00743v32020MEDITRON-70B: Scaling Medical Pretraining for Large Language Models
Zeming Chen, Alejandro Hernández Cano, Angelika Romanou +17
cs.CLcs.AIcs.LGarXiv:2311.16079v12023Multi-Hop Knowledge Graph Reasoning with Reward Shaping
Xi Victoria Lin, Richard Socher, Caiming Xiong
cs.AIcs.CLcs.LGarXiv:1808.10568v22018UniVL: A Unified Video and Language Pre-Training Model for Multimodal Understanding and Generation
Huaishao Luo, Lei Ji, Botian Shi +6
cs.CVcs.CLcs.LGarXiv:2002.06353v32020Cross-lingual Representation Learning via Centroid Intervention Fusion
Wei Sun, Marie-Francine Moens
cs.CLarXiv:2608.26357v12026Coarse-to-Fine Decoding for Neural Semantic Parsing
Li Dong, Mirella Lapata
cs.CLarXiv:1805.04793v12018Are All Languages Created Equal in Multilingual BERT?
Shijie Wu, Mark Dredze
cs.CLarXiv:2005.09093v22020CHiME-6 Challenge:Tackling Multispeaker Speech Recognition for Unsegmented Recordings
Shinji Watanabe, Michael Mandel, Jon Barker +18
cs.SDcs.CLeess.ASarXiv:2004.09249v22020Not Enough Data? Deep Learning to the Rescue!
Ateret Anaby-Tavor, Boaz Carmeli, Esther Goldbraich +5
cs.CLcs.LGarXiv:1911.03118v22019Good Friends, Bad News - Affect and Virality in Twitter
Lars Kai Hansen, Adam Arvidsson, Finn Årup Nielsen +2
cs.SIcs.CLphysics.soc-pharXiv:1101.0510v12011