Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
7,081 to 7,140 of 11,210
Parsing Argumentation Structures in Persuasive Essays
Christian Stab, Iryna Gurevych
cs.CLarXiv:1604.07370v22016Collaborative Deep Learning for Recommender Systems
Hao Wang, Naiyan Wang, Dit-Yan Yeung
cs.LGcs.CLcs.IRarXiv:1409.2944v22014XGLUE: A New Benchmark Dataset for Cross-lingual Pre-training, Understanding and Generation
Yaobo Liang, Nan Duan, Yeyun Gong +21
cs.CLarXiv:2004.01401v32020Representation Degeneration Problem in Training Natural Language Generation Models
Jun Gao, Di He, Xu Tan +3
cs.CLarXiv:1907.12009v12019Tips and Tricks for Visual Question Answering: Learnings from the 2017 Challenge
Damien Teney, Peter Anderson, Xiaodong He +1
cs.CVcs.CLarXiv:1708.02711v12017A Single Suffix to Break Them All: Basin-Aware Jailbreaks for Merged Model Families
Yu Zhe, Yixin Tan, Junhao Wei +1
cs.LGcs.CLarXiv:2608.26506v12026RESDSQL: Decoupling Schema Linking and Skeleton Parsing for Text-to-SQL
Haoyang Li, Jing Zhang, Cuiping Li +1
cs.CLarXiv:2302.05965v32023Navigating the Digital World as Humans Do: Universal Visual Grounding for GUI Agents
Boyu Gou, Ruohan Wang, Boyuan Zheng +5
cs.AIcs.CLcs.CVarXiv:2410.05243v32024Pair-Level Essay-Scale Republication and Reuse from Fragmented Historical Text Reuse: A Workflow Study on Eighteenth-Century Books and Newspapers
Ke Shu, Kira Hinderks, Eetu Mäkelä +1
cs.CLarXiv:2608.27343v12026Summaries:한국어LLaVA-PruMerge: Adaptive Token Reduction for Efficient Large Multimodal Models
Yuzhang Shang, Mu Cai, Bingxin Xu +2
cs.CVcs.AIcs.CLarXiv:2403.15388v62024Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing
Dujian Ding, Ankur Mallick, Chi Wang +5
cs.LGcs.AIcs.CLarXiv:2404.14618v12024JudgeLM: Fine-tuned Large Language Models are Scalable Judges
Lianghui Zhu, Xinggang Wang, Xinlong Wang
cs.CLcs.AIarXiv:2310.17631v22023Time-Aware Language Models as Temporal Knowledge Bases
Bhuwan Dhingra, Jeremy R. Cole, Julian Martin Eisenschlos +3
cs.CLarXiv:2106.15110v22021Diff Mining: Logit Differences Reveal Finetuning Objectives
Greg Kocher, Robert West, Clément Dumas +1
cs.LGcs.AIcs.CLarXiv:2608.26462v12026ParlAI: A Dialog Research Software Platform
Alexander H. Miller, Will Feng, Adam Fisch +5
cs.CLarXiv:1705.06476v42017Linguistically-Informed Self-Attention for Semantic Role Labeling
Emma Strubell, Patrick Verga, Daniel Andor +2
cs.CLarXiv:1804.08199v32018Squeezing More from Limited Data with Recursive Transformers
Serdar Gülbahar, Lukas Edman, Alexander Fraser
cs.CLcs.LGarXiv:2608.26973v12026Planting a Latent Variable in Natural-Looking Text: a More Realistic Test of Belief States in LLMs and Their Link to Concept Geometry
Alexandru-Iulius Jerpelea
cs.CLarXiv:2608.26887v12026Mockingjay: Unsupervised Speech Representation Learning with Deep Bidirectional Transformer Encoders
Andy T. Liu, Shu-wen Yang, Po-Han Chi +2
eess.AScs.CLcs.LGarXiv:1910.12638v22019Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge
Jiayi Ye, Yanbo Wang, Yue Huang +9
cs.CLcs.AIarXiv:2410.02736v22024Cultural Bias and Cultural Alignment of Large Language Models
Yan Tao, Olga Viberg, Ryan S. Baker +1
cs.CLcs.AIarXiv:2311.14096v22023Knowledge-Verified Emergent Deception in LLM Agents Under Conflicting Incentives
Zheyuan Liu, Weiliang Zhao, Xiangchi Yuan +3
cs.CLcs.AIarXiv:2608.26372v12026Improved Variational Autoencoders for Text Modeling using Dilated Convolutions
Zichao Yang, Zhiting Hu, Ruslan Salakhutdinov +1
cs.NEcs.CLcs.LGarXiv:1702.08139v22017Relation-Aware Entity Alignment for Heterogeneous Knowledge Graphs
Yuting Wu, Xiao Liu, Yansong Feng +3
cs.CLarXiv:1908.08210v12019Large Pre-trained Language Models Contain Human-like Biases of What is Right and Wrong to Do
Patrick Schramowski, Cigdem Turan, Nico Andersen +2
cs.CLcs.CYarXiv:2103.11790v32021Co-Writing Screenplays and Theatre Scripts with Language Models: An Evaluation by Industry Professionals
Piotr Mirowski, Kory W. Mathewson, Jaylen Pittman +1
cs.HCcs.CLarXiv:2209.14958v12022Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
Yu Zhang, Wei Han, James Qin +24
cs.CLcs.SDeess.ASarXiv:2303.01037v32023ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases
Qiaoyu Tang, Ziliang Deng, Hongyu Lin +4
cs.CLarXiv:2306.05301v22023Linguistic Input Features Improve Neural Machine Translation
Rico Sennrich, Barry Haddow
cs.CLarXiv:1606.02892v22016MMGCN: Multimodal Fusion via Deep Graph Convolution Network for Emotion Recognition in Conversation
Jingwen Hu, Yuchen Liu, Jinming Zhao +1
cs.CLcs.SDeess.ASarXiv:2107.06779v12021The Second Conversational Intelligence Challenge (ConvAI2)
Emily Dinan, Varvara Logacheva, Valentin Malykh +14
cs.AIcs.CLcs.HCarXiv:1902.00098v12019Joint entity recognition and relation extraction as a multi-head selection problem
Giannis Bekoulis, Johannes Deleu, Thomas Demeester +1
cs.CLarXiv:1804.07847v32018Explain Images with Multimodal Recurrent Neural Networks
Junhua Mao, Wei Xu, Yi Yang +2
cs.CVcs.CLcs.LGarXiv:1410.1090v12014Community Interaction and Conflict on the Web
Srijan Kumar, William L. Hamilton, Jure Leskovec +1
cs.SIcs.CLcs.HCarXiv:1803.03697v12018Improving zero-shot learning by mitigating the hubness problem
Georgiana Dinu, Angeliki Lazaridou, Marco Baroni
cs.CLcs.LGarXiv:1412.6568v32014Levenshtein Transformer
Jiatao Gu, Changhan Wang, Jake Zhao
cs.CLcs.LGarXiv:1905.11006v22019A Literature Survey of Recent Advances in Chatbots
Guendalina Caldarini, Sardar Jaf, Kenneth McGarry
cs.CLarXiv:2201.06657v12022Sparse Sinkhorn Attention
Yi Tay, Dara Bahri, Liu Yang +2
cs.LGcs.CLarXiv:2002.11296v12020ToolQA: A Dataset for LLM Question Answering with External Tools
Yuchen Zhuang, Yue Yu, Kuan Wang +2
cs.CLcs.AIarXiv:2306.13304v12023Template-Based Named Entity Recognition Using BART
Leyang Cui, Yu Wu, Jian Liu +2
cs.CLarXiv:2106.01760v12021Exploring and Distilling Posterior and Prior Knowledge for Radiology Report Generation
Fenglin Liu, Xian Wu, Shen Ge +2
cs.CVcs.CLarXiv:2106.06963v22021Nematus: a Toolkit for Neural Machine Translation
Rico Sennrich, Orhan Firat, Kyunghyun Cho +8
cs.CLarXiv:1703.04357v12017A Real-World WebAgent with Planning, Long Context Understanding, and Program Synthesis
Izzeddin Gur, Hiroki Furuta, Austin Huang +4
cs.LGcs.AIcs.CLarXiv:2307.12856v42023Towards Understanding Chain-of-Thought Prompting: An Empirical Study of What Matters
Boshi Wang, Sewon Min, Xiang Deng +4
cs.CLarXiv:2212.10001v22022NaturalSpeech 2: Latent Diffusion Models are Natural and Zero-Shot Speech and Singing Synthesizers
Kai Shen, Zeqian Ju, Xu Tan +6
eess.AScs.AIcs.CLarXiv:2304.09116v32023$\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens
Xinrong Zhang, Yingfa Chen, Shengding Hu +8
cs.CLarXiv:2402.13718v32024Compositional Generalization via Structural Identification in a Category-Theoretic Framework
Akihiro Maeda, Thomas Seiller, Yohei Oseki
cs.CLstat.MLarXiv:2608.26465v12026Tying Word Vectors and Word Classifiers: A Loss Framework for Language Modeling
Hakan Inan, Khashayar Khosravi, Richard Socher
cs.LGcs.CLstat.MLarXiv:1611.01462v32016Blockwise Parallel Decoding for Deep Autoregressive Models
Mitchell Stern, Noam Shazeer, Jakob Uszkoreit
cs.LGcs.CLstat.MLarXiv:1811.03115v12018Analogical Inference for Multi-Relational Embeddings
Hanxiao Liu, Yuexin Wu, Yiming Yang
cs.LGcs.AIcs.CLarXiv:1705.02426v22017Retrieve, Program, Repeat: Complex Knowledge Base Question Answering via Alternate Meta-learning
Yuncheng Hua, Yuan-Fang Li, Gholamreza Haffari +2
cs.AIcs.CLarXiv:2010.15875v12020Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2
Tom Lieberum, Senthooran Rajamanoharan, Arthur Conmy +7
cs.LGcs.AIcs.CLarXiv:2408.05147v22024Fine-Grained Analysis of Propaganda in News Articles
Giovanni Da San Martino, Seunghak Yu, Alberto Barrón-Cedeño +2
cs.CLcs.AIcs.IRarXiv:1910.02517v12019Improving Conversational Recommender Systems via Knowledge Graph based Semantic Fusion
Kun Zhou, Wayne Xin Zhao, Shuqing Bian +3
cs.CLcs.AIcs.IRarXiv:2007.04032v12020A Survey on Recent Approaches for Natural Language Processing in Low-Resource Scenarios
Michael A. Hedderich, Lukas Lange, Heike Adel +2
cs.CLcs.LGarXiv:2010.12309v32020Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks
Tomer Ullman
cs.AIcs.CLarXiv:2302.08399v52023LoraHub: Efficient Cross-Task Generalization via Dynamic LoRA Composition
Chengsong Huang, Qian Liu, Bill Yuchen Lin +3
cs.CLcs.AIarXiv:2307.13269v32023Demographic Dialectal Variation in Social Media: A Case Study of African-American English
Su Lin Blodgett, Lisa Green, Brendan O'Connor
cs.CLarXiv:1608.08868v12016Assessing Gender Bias in Machine Translation -- A Case Study with Google Translate
Marcelo O. R. Prates, Pedro H. C. Avelar, Luis Lamb
cs.CYcs.CLarXiv:1809.02208v42018Neural Text Summarization: A Critical Evaluation
Wojciech Kryściński, Nitish Shirish Keskar, Bryan McCann +2
cs.CLarXiv:1908.08960v12019