Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

7,081 to 7,140 of 11,210

  1. Parsing Argumentation Structures in Persuasive Essays

    Christian Stab, Iryna Gurevych

    cs.CLarXiv:1604.07370v22016
  2. Collaborative Deep Learning for Recommender Systems

    Hao Wang, Naiyan Wang, Dit-Yan Yeung

    cs.LGcs.CLcs.IRarXiv:1409.2944v22014
  3. XGLUE: A New Benchmark Dataset for Cross-lingual Pre-training, Understanding and Generation

    Yaobo Liang, Nan Duan, Yeyun Gong +21

    cs.CLarXiv:2004.01401v32020
  4. Representation Degeneration Problem in Training Natural Language Generation Models

    Jun Gao, Di He, Xu Tan +3

    cs.CLarXiv:1907.12009v12019
  5. Tips and Tricks for Visual Question Answering: Learnings from the 2017 Challenge

    Damien Teney, Peter Anderson, Xiaodong He +1

    cs.CVcs.CLarXiv:1708.02711v12017
  6. A Single Suffix to Break Them All: Basin-Aware Jailbreaks for Merged Model Families

    Yu Zhe, Yixin Tan, Junhao Wei +1

    cs.LGcs.CLarXiv:2608.26506v12026
  7. RESDSQL: Decoupling Schema Linking and Skeleton Parsing for Text-to-SQL

    Haoyang Li, Jing Zhang, Cuiping Li +1

    cs.CLarXiv:2302.05965v32023
  8. Navigating the Digital World as Humans Do: Universal Visual Grounding for GUI Agents

    Boyu Gou, Ruohan Wang, Boyuan Zheng +5

    cs.AIcs.CLcs.CVarXiv:2410.05243v32024
  9. Pair-Level Essay-Scale Republication and Reuse from Fragmented Historical Text Reuse: A Workflow Study on Eighteenth-Century Books and Newspapers

    Ke Shu, Kira Hinderks, Eetu Mäkelä +1

    cs.CLarXiv:2608.27343v12026
    Summaries:한국어
  10. LLaVA-PruMerge: Adaptive Token Reduction for Efficient Large Multimodal Models

    Yuzhang Shang, Mu Cai, Bingxin Xu +2

    cs.CVcs.AIcs.CLarXiv:2403.15388v62024
  11. Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

    Dujian Ding, Ankur Mallick, Chi Wang +5

    cs.LGcs.AIcs.CLarXiv:2404.14618v12024
  12. JudgeLM: Fine-tuned Large Language Models are Scalable Judges

    Lianghui Zhu, Xinggang Wang, Xinlong Wang

    cs.CLcs.AIarXiv:2310.17631v22023
  13. Time-Aware Language Models as Temporal Knowledge Bases

    Bhuwan Dhingra, Jeremy R. Cole, Julian Martin Eisenschlos +3

    cs.CLarXiv:2106.15110v22021
  14. Diff Mining: Logit Differences Reveal Finetuning Objectives

    Greg Kocher, Robert West, Clément Dumas +1

    cs.LGcs.AIcs.CLarXiv:2608.26462v12026
  15. ParlAI: A Dialog Research Software Platform

    Alexander H. Miller, Will Feng, Adam Fisch +5

    cs.CLarXiv:1705.06476v42017
  16. Linguistically-Informed Self-Attention for Semantic Role Labeling

    Emma Strubell, Patrick Verga, Daniel Andor +2

    cs.CLarXiv:1804.08199v32018
  17. Squeezing More from Limited Data with Recursive Transformers

    Serdar Gülbahar, Lukas Edman, Alexander Fraser

    cs.CLcs.LGarXiv:2608.26973v12026
  18. Planting a Latent Variable in Natural-Looking Text: a More Realistic Test of Belief States in LLMs and Their Link to Concept Geometry

    Alexandru-Iulius Jerpelea

    cs.CLarXiv:2608.26887v12026
  19. Mockingjay: Unsupervised Speech Representation Learning with Deep Bidirectional Transformer Encoders

    Andy T. Liu, Shu-wen Yang, Po-Han Chi +2

    eess.AScs.CLcs.LGarXiv:1910.12638v22019
  20. Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge

    Jiayi Ye, Yanbo Wang, Yue Huang +9

    cs.CLcs.AIarXiv:2410.02736v22024
  21. Cultural Bias and Cultural Alignment of Large Language Models

    Yan Tao, Olga Viberg, Ryan S. Baker +1

    cs.CLcs.AIarXiv:2311.14096v22023
  22. Knowledge-Verified Emergent Deception in LLM Agents Under Conflicting Incentives

    Zheyuan Liu, Weiliang Zhao, Xiangchi Yuan +3

    cs.CLcs.AIarXiv:2608.26372v12026
  23. Improved Variational Autoencoders for Text Modeling using Dilated Convolutions

    Zichao Yang, Zhiting Hu, Ruslan Salakhutdinov +1

    cs.NEcs.CLcs.LGarXiv:1702.08139v22017
  24. Relation-Aware Entity Alignment for Heterogeneous Knowledge Graphs

    Yuting Wu, Xiao Liu, Yansong Feng +3

    cs.CLarXiv:1908.08210v12019
  25. Large Pre-trained Language Models Contain Human-like Biases of What is Right and Wrong to Do

    Patrick Schramowski, Cigdem Turan, Nico Andersen +2

    cs.CLcs.CYarXiv:2103.11790v32021
  26. Co-Writing Screenplays and Theatre Scripts with Language Models: An Evaluation by Industry Professionals

    Piotr Mirowski, Kory W. Mathewson, Jaylen Pittman +1

    cs.HCcs.CLarXiv:2209.14958v12022
  27. Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages

    Yu Zhang, Wei Han, James Qin +24

    cs.CLcs.SDeess.ASarXiv:2303.01037v32023
  28. ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases

    Qiaoyu Tang, Ziliang Deng, Hongyu Lin +4

    cs.CLarXiv:2306.05301v22023
  29. Linguistic Input Features Improve Neural Machine Translation

    Rico Sennrich, Barry Haddow

    cs.CLarXiv:1606.02892v22016
  30. MMGCN: Multimodal Fusion via Deep Graph Convolution Network for Emotion Recognition in Conversation

    Jingwen Hu, Yuchen Liu, Jinming Zhao +1

    cs.CLcs.SDeess.ASarXiv:2107.06779v12021
  31. The Second Conversational Intelligence Challenge (ConvAI2)

    Emily Dinan, Varvara Logacheva, Valentin Malykh +14

    cs.AIcs.CLcs.HCarXiv:1902.00098v12019
  32. Joint entity recognition and relation extraction as a multi-head selection problem

    Giannis Bekoulis, Johannes Deleu, Thomas Demeester +1

    cs.CLarXiv:1804.07847v32018
  33. Explain Images with Multimodal Recurrent Neural Networks

    Junhua Mao, Wei Xu, Yi Yang +2

    cs.CVcs.CLcs.LGarXiv:1410.1090v12014
  34. Community Interaction and Conflict on the Web

    Srijan Kumar, William L. Hamilton, Jure Leskovec +1

    cs.SIcs.CLcs.HCarXiv:1803.03697v12018
  35. Improving zero-shot learning by mitigating the hubness problem

    Georgiana Dinu, Angeliki Lazaridou, Marco Baroni

    cs.CLcs.LGarXiv:1412.6568v32014
  36. Levenshtein Transformer

    Jiatao Gu, Changhan Wang, Jake Zhao

    cs.CLcs.LGarXiv:1905.11006v22019
  37. A Literature Survey of Recent Advances in Chatbots

    Guendalina Caldarini, Sardar Jaf, Kenneth McGarry

    cs.CLarXiv:2201.06657v12022
  38. Sparse Sinkhorn Attention

    Yi Tay, Dara Bahri, Liu Yang +2

    cs.LGcs.CLarXiv:2002.11296v12020
  39. ToolQA: A Dataset for LLM Question Answering with External Tools

    Yuchen Zhuang, Yue Yu, Kuan Wang +2

    cs.CLcs.AIarXiv:2306.13304v12023
  40. Template-Based Named Entity Recognition Using BART

    Leyang Cui, Yu Wu, Jian Liu +2

    cs.CLarXiv:2106.01760v12021
  41. Exploring and Distilling Posterior and Prior Knowledge for Radiology Report Generation

    Fenglin Liu, Xian Wu, Shen Ge +2

    cs.CVcs.CLarXiv:2106.06963v22021
  42. Nematus: a Toolkit for Neural Machine Translation

    Rico Sennrich, Orhan Firat, Kyunghyun Cho +8

    cs.CLarXiv:1703.04357v12017
  43. A Real-World WebAgent with Planning, Long Context Understanding, and Program Synthesis

    Izzeddin Gur, Hiroki Furuta, Austin Huang +4

    cs.LGcs.AIcs.CLarXiv:2307.12856v42023
  44. Towards Understanding Chain-of-Thought Prompting: An Empirical Study of What Matters

    Boshi Wang, Sewon Min, Xiang Deng +4

    cs.CLarXiv:2212.10001v22022
  45. NaturalSpeech 2: Latent Diffusion Models are Natural and Zero-Shot Speech and Singing Synthesizers

    Kai Shen, Zeqian Ju, Xu Tan +6

    eess.AScs.AIcs.CLarXiv:2304.09116v32023
  46. $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

    Xinrong Zhang, Yingfa Chen, Shengding Hu +8

    cs.CLarXiv:2402.13718v32024
  47. Compositional Generalization via Structural Identification in a Category-Theoretic Framework

    Akihiro Maeda, Thomas Seiller, Yohei Oseki

    cs.CLstat.MLarXiv:2608.26465v12026
  48. Tying Word Vectors and Word Classifiers: A Loss Framework for Language Modeling

    Hakan Inan, Khashayar Khosravi, Richard Socher

    cs.LGcs.CLstat.MLarXiv:1611.01462v32016
  49. Blockwise Parallel Decoding for Deep Autoregressive Models

    Mitchell Stern, Noam Shazeer, Jakob Uszkoreit

    cs.LGcs.CLstat.MLarXiv:1811.03115v12018
  50. Analogical Inference for Multi-Relational Embeddings

    Hanxiao Liu, Yuexin Wu, Yiming Yang

    cs.LGcs.AIcs.CLarXiv:1705.02426v22017
  51. Retrieve, Program, Repeat: Complex Knowledge Base Question Answering via Alternate Meta-learning

    Yuncheng Hua, Yuan-Fang Li, Gholamreza Haffari +2

    cs.AIcs.CLarXiv:2010.15875v12020
  52. Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2

    Tom Lieberum, Senthooran Rajamanoharan, Arthur Conmy +7

    cs.LGcs.AIcs.CLarXiv:2408.05147v22024
  53. Fine-Grained Analysis of Propaganda in News Articles

    Giovanni Da San Martino, Seunghak Yu, Alberto Barrón-Cedeño +2

    cs.CLcs.AIcs.IRarXiv:1910.02517v12019
  54. Improving Conversational Recommender Systems via Knowledge Graph based Semantic Fusion

    Kun Zhou, Wayne Xin Zhao, Shuqing Bian +3

    cs.CLcs.AIcs.IRarXiv:2007.04032v12020
  55. A Survey on Recent Approaches for Natural Language Processing in Low-Resource Scenarios

    Michael A. Hedderich, Lukas Lange, Heike Adel +2

    cs.CLcs.LGarXiv:2010.12309v32020
  56. Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks

    Tomer Ullman

    cs.AIcs.CLarXiv:2302.08399v52023
  57. LoraHub: Efficient Cross-Task Generalization via Dynamic LoRA Composition

    Chengsong Huang, Qian Liu, Bill Yuchen Lin +3

    cs.CLcs.AIarXiv:2307.13269v32023
  58. Demographic Dialectal Variation in Social Media: A Case Study of African-American English

    Su Lin Blodgett, Lisa Green, Brendan O'Connor

    cs.CLarXiv:1608.08868v12016
  59. Assessing Gender Bias in Machine Translation -- A Case Study with Google Translate

    Marcelo O. R. Prates, Pedro H. C. Avelar, Luis Lamb

    cs.CYcs.CLarXiv:1809.02208v42018
  60. Neural Text Summarization: A Critical Evaluation

    Wojciech Kryściński, Nitish Shirish Keskar, Bryan McCann +2

    cs.CLarXiv:1908.08960v12019