Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,221 to 5,280 of 11,259

  1. Deep Bayesian Active Learning for Natural Language Processing: Results of a Large-Scale Empirical Study

    Aditya Siddhant, Zachary C. Lipton

    cs.CLcs.LGstat.MLarXiv:1808.05697v32018
  2. Cost-Efficient Large Language Model Serving for Multi-turn Conversations with CachedAttention

    Bin Gao, Zhuomin He, Puru Sharma +6

    cs.CLcs.LGarXiv:2403.19708v32024
  3. GUI-CC: Benchmarking Contextual Consistency of GUI World Models as Agent Environments

    Lin Fu, Zheyuan Yang, Tianhui Zhang +5

    cs.CLcs.AIarXiv:2609.00048v12026
  4. Findings of the Second Shared Task on Multimodal Machine Translation and Multilingual Image Description

    Desmond Elliott, Stella Frank, Loïc Barrault +2

    cs.CLcs.CVarXiv:1710.07177v12017
  5. On Adversarial Examples for Character-Level Neural Machine Translation

    Javid Ebrahimi, Daniel Lowd, Dejing Dou

    cs.CLcs.AIarXiv:1806.09030v12018
  6. Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data

    Chaoyi Wu, Xiaoman Zhang, Ya Zhang +2

    cs.CVcs.CLarXiv:2308.02463v52023
  7. TimeTraveler: Reinforcement Learning for Temporal Knowledge Graph Forecasting

    Haohai Sun, Jialun Zhong, Yunpu Ma +2

    cs.LGcs.AIcs.CLarXiv:2109.04101v12021
  8. JFLEG: A Fluency Corpus and Benchmark for Grammatical Error Correction

    Courtney Napoles, Keisuke Sakaguchi, Joel Tetreault

    cs.CLarXiv:1702.04066v12017
  9. TMR: Text-to-Motion Retrieval Using Contrastive 3D Human Motion Synthesis

    Mathis Petrovich, Michael J. Black, Gül Varol

    cs.CVcs.CLarXiv:2305.00976v22023
  10. Explaining Answers with Entailment Trees

    Bhavana Dalvi, Peter Jansen, Oyvind Tafjord +4

    cs.CLcs.AIarXiv:2104.08661v32021
  11. Relevance of Unsupervised Metrics in Task-Oriented Dialogue for Evaluating Natural Language Generation

    Shikhar Sharma, Layla El Asri, Hannes Schulz +1

    cs.CLarXiv:1706.09799v12017
  12. Compositional Chain-of-Thought Prompting for Large Multimodal Models

    Chancharik Mitra, Brandon Huang, Trevor Darrell +1

    cs.CVcs.AIcs.CLarXiv:2311.17076v32023
  13. Same Task, More Tokens: the Impact of Input Length on the Reasoning Performance of Large Language Models

    Mosh Levy, Alon Jacoby, Yoav Goldberg

    cs.CLcs.AIarXiv:2402.14848v22024
  14. Automatic Language Identification in Texts: A Survey

    Tommi Jauhiainen, Marco Lui, Marcos Zampieri +2

    cs.CLarXiv:1804.08186v22018
  15. Word Emdeddings through Hellinger PCA

    Rémi Lebret, Ronan Collobert

    cs.CLcs.LGarXiv:1312.5542v32013
  16. Clever Hans or Neural Theory of Mind? Stress Testing Social Reasoning in Large Language Models

    Natalie Shapira, Mosh Levy, Seyed Hossein Alavi +5

    cs.CLarXiv:2305.14763v12023
  17. AgentClinic: a multimodal agent benchmark to evaluate AI in simulated clinical environments

    Samuel Schmidgall, Rojin Ziaei, Carl Harris +3

    cs.HCcs.CLarXiv:2405.07960v52024
  18. Revisiting the "Video" in Video-Language Understanding

    Shyamal Buch, Cristóbal Eyzaguirre, Adrien Gaidon +3

    cs.CVcs.AIcs.CLarXiv:2206.01720v12022
  19. A Multilayer Convolutional Encoder-Decoder Neural Network for Grammatical Error Correction

    Shamil Chollampatt, Hwee Tou Ng

    cs.CLarXiv:1801.08831v12018
  20. Structure-Grounded Pretraining for Text-to-SQL

    Xiang Deng, Ahmed Hassan Awadallah, Christopher Meek +3

    cs.CLcs.AIarXiv:2010.12773v32020
  21. Improving Grammatical Error Correction via Pre-Training a Copy-Augmented Architecture with Unlabeled Data

    Wei Zhao, Liang Wang, Kewei Shen +2

    cs.CLarXiv:1903.00138v32019
  22. Do Multi-Sense Embeddings Improve Natural Language Understanding?

    Jiwei Li, Dan Jurafsky

    cs.CLarXiv:1506.01070v32015
  23. Lipreading with Long Short-Term Memory

    Michael Wand, Jan Koutník, Jürgen Schmidhuber

    cs.CVcs.CLarXiv:1601.08188v12016
  24. Emotion Intensities in Tweets

    Saif M. Mohammad, Felipe Bravo-Marquez

    cs.CLarXiv:1708.03696v12017
  25. On the Reliability of Watermarks for Large Language Models

    John Kirchenbauer, Jonas Geiping, Yuxin Wen +7

    cs.LGcs.CLcs.CRarXiv:2306.04634v42023
  26. Do Membership Inference Attacks Work on Large Language Models?

    Michael Duan, Anshuman Suri, Niloofar Mireshghallah +7

    cs.CLarXiv:2402.07841v22024
  27. Dynamic Word Embeddings for Evolving Semantic Discovery

    Zijun Yao, Yifan Sun, Weicong Ding +2

    cs.CLstat.MLarXiv:1703.00607v22017
  28. Neural Natural Language Processing for Unstructured Data in Electronic Health Records: a Review

    Irene Li, Jessica Pan, Jeremy Goldwasser +10

    cs.CLcs.AIarXiv:2107.02975v12021
  29. Transfer Learning for Speech and Language Processing

    Dong Wang, Thomas Fang Zheng

    cs.CLcs.LGarXiv:1511.06066v12015
  30. What Language Model Architecture and Pretraining Objective Work Best for Zero-Shot Generalization?

    Thomas Wang, Adam Roberts, Daniel Hesslow +5

    cs.CLcs.LGstat.MLarXiv:2204.05832v12022
  31. Chain-of-Note: Enhancing Robustness in Retrieval-Augmented Language Models

    Wenhao Yu, Hongming Zhang, Xiaoman Pan +3

    cs.CLcs.AIarXiv:2311.09210v22023
  32. Yelp Dataset Challenge: Review Rating Prediction

    Nabiha Asghar

    cs.CLcs.IRcs.LGarXiv:1605.05362v12016
  33. Large Language Model Is Not a Good Few-shot Information Extractor, but a Good Reranker for Hard Samples!

    Yubo Ma, Yixin Cao, YongChing Hong +1

    cs.CLcs.AIarXiv:2303.08559v22023
  34. Macaw-LLM: Multi-Modal Language Modeling with Image, Audio, Video, and Text Integration

    Chenyang Lyu, Minghao Wu, Longyue Wang +5

    cs.CLcs.AIcs.CVarXiv:2306.09093v12023
  35. TABBIE: Pretrained Representations of Tabular Data

    Hiroshi Iida, Dung Thai, Varun Manjunatha +1

    cs.CLarXiv:2105.02584v12021
  36. Understanding the Behaviors of BERT in Ranking

    Yifan Qiao, Chenyan Xiong, Zhenghao Liu +1

    cs.IRcs.CLarXiv:1904.07531v42019
  37. Measuring Progress on Scalable Oversight for Large Language Models

    Samuel R. Bowman, Jeeyoon Hyun, Ethan Perez +43

    cs.HCcs.AIcs.CLarXiv:2211.03540v22022
  38. EvoBrowseComp: Benchmarking Search Agents on Evolving Knowledge

    Yunhan Wang, Jiaan Wang, Lianzhe Huang +2

    cs.CLarXiv:2606.13120v12026
  39. D-CORE: Incentivizing Task Decomposition in Large Reasoning Models for Complex Tool Use

    Bowen Xu, Shaoyu Wu, Hao Jiang +4

    cs.CLarXiv:2602.02160v12026
  40. UnifiedSKG: Unifying and Multi-Tasking Structured Knowledge Grounding with Text-to-Text Language Models

    Tianbao Xie, Chen Henry Wu, Peng Shi +20

    cs.CLarXiv:2201.05966v32022
  41. Tracking the World State with Recurrent Entity Networks

    Mikael Henaff, Jason Weston, Arthur Szlam +2

    cs.CLarXiv:1612.03969v32016
  42. Structured Training for Neural Network Transition-Based Parsing

    David Weiss, Chris Alberti, Michael Collins +1

    cs.CLarXiv:1506.06158v12015
  43. End-to-End Referring Video Object Segmentation with Multimodal Transformers

    Adam Botach, Evgenii Zheltonozhskii, Chaim Baskin

    cs.CVcs.CLcs.LGarXiv:2111.14821v22021
  44. BB_twtr at SemEval-2017 Task 4: Twitter Sentiment Analysis with CNNs and LSTMs

    Mathieu Cliche

    cs.CLstat.MLarXiv:1704.06125v12017
  45. Attentive Convolutional Neural Network based Speech Emotion Recognition: A Study on the Impact of Input Features, Signal Length, and Acted Speech

    Michael Neumann, Ngoc Thang Vu

    cs.CLarXiv:1706.00612v12017
  46. Landmark Attention: Random-Access Infinite Context Length for Transformers

    Amirkeivan Mohtashami, Martin Jaggi

    cs.CLcs.LGarXiv:2305.16300v22023
  47. HelpSteer2: Open-source dataset for training top-performing reward models

    Zhilin Wang, Yi Dong, Olivier Delalleau +6

    cs.CLcs.AIcs.LGarXiv:2406.08673v12024
  48. Break It Down: A Question Understanding Benchmark

    Tomer Wolfson, Mor Geva, Ankit Gupta +4

    cs.CLarXiv:2001.11770v12020
  49. Subgraph Retrieval Enhanced Model for Multi-hop Knowledge Base Question Answering

    Jing Zhang, Xiaokang Zhang, Jifan Yu +4

    cs.CLarXiv:2202.13296v22022
  50. Open-Domain Question Answering Goes Conversational via Question Rewriting

    Raviteja Anantha, Svitlana Vakulenko, Zhucheng Tu +3

    cs.IRcs.CLarXiv:2010.04898v32020
  51. Detecting Hallucinated Content in Conditional Neural Sequence Generation

    Chunting Zhou, Graham Neubig, Jiatao Gu +4

    cs.CLcs.AIarXiv:2011.02593v32020
  52. Can large language models replace humans in the systematic review process? Evaluating GPT-4's efficacy in screening and extracting data from peer-reviewed and grey literature in multiple languages

    Qusai Khraisha, Sophie Put, Johanna Kappenberg +2

    cs.CLcs.AIcs.LGarXiv:2310.17526v22023
  53. Sparks: Inspiration for Science Writing using Language Models

    Katy Ilonka Gero, Vivian Liu, Lydia B. Chilton

    cs.HCcs.CLarXiv:2110.07640v12021
  54. Queens are Powerful too: Mitigating Gender Bias in Dialogue Generation

    Emily Dinan, Angela Fan, Adina Williams +3

    cs.CLarXiv:1911.03842v22019
  55. Understanding and Improving Transformer From a Multi-Particle Dynamic System Point of View

    Yiping Lu, Zhuohan Li, Di He +5

    cs.LGcs.CLstat.MLarXiv:1906.02762v12019
  56. DoctorGLM: Fine-tuning your Chinese Doctor is not a Herculean Task

    Honglin Xiong, Sheng Wang, Yitao Zhu +5

    cs.CLarXiv:2304.01097v22023
  57. Symbolic Chain-of-Thought Distillation: Small Models Can Also "Think" Step-by-Step

    Liunian Harold Li, Jack Hessel, Youngjae Yu +3

    cs.CLarXiv:2306.14050v22023
  58. R-Zero: Self-Evolving Reasoning LLM from Zero Data

    Chengsong Huang, Wenhao Yu, Xiaoyang Wang +6

    cs.LGcs.AIcs.CLarXiv:2508.05004v42025
  59. A Dependency-Based Neural Network for Relation Classification

    Yang Liu, Furu Wei, Sujian Li +3

    cs.CLcs.LGcs.NEarXiv:1507.04646v12015
  60. Revisiting Low-Resource Neural Machine Translation: A Case Study

    Rico Sennrich, Biao Zhang

    cs.CLarXiv:1905.11901v12019