Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,101 to 5,160 of 11,259

  1. MTOP: A Comprehensive Multilingual Task-Oriented Semantic Parsing Benchmark

    Haoran Li, Abhinav Arora, Shuohui Chen +3

    cs.CLcs.LGarXiv:2008.09335v22020
  2. From Production Traffic to Post-Training: Building a Self-Hosted LLM That Covers the Corporate Request Mix

    Olga Tsymboi, Dmitrii Stoianov, Ramil Latypov +11

    cs.CLarXiv:2609.01572v12026
  3. Handling Divergent Reference Texts when Evaluating Table-to-Text Generation

    Bhuwan Dhingra, Manaal Faruqui, Ankur Parikh +3

    cs.CLarXiv:1906.01081v12019
  4. A Streaming On-Device End-to-End Model Surpassing Server-Side Conventional Model Quality and Latency

    Tara N. Sainath, Yanzhang He, Bo Li +26

    cs.CLcs.LGcs.SDarXiv:2003.12710v22020
  5. Quantifying the Persona Effect in LLM Simulations

    Tiancheng Hu, Nigel Collier

    cs.CLcs.CYarXiv:2402.10811v22024
  6. Primer: Searching for Efficient Transformers for Language Modeling

    David R. So, Wojciech Mańke, Hanxiao Liu +3

    cs.LGcs.AIcs.CLarXiv:2109.08668v22021
  7. Language as a Latent Variable: Discrete Generative Models for Sentence Compression

    Yishu Miao, Phil Blunsom

    cs.CLcs.AIarXiv:1609.07317v22016
  8. Unsupervised pretraining transfers well across languages

    Morgane Rivière, Armand Joulin, Pierre-Emmanuel Mazaré +1

    eess.AScs.CLcs.LGarXiv:2002.02848v12020
  9. ChatGPT-4 Outperforms Experts and Crowd Workers in Annotating Political Twitter Messages with Zero-Shot Learning

    Petter Törnberg

    cs.CLcs.AIcs.SIarXiv:2304.06588v12023
  10. Do the Rewards Justify the Means? Measuring Trade-Offs Between Rewards and Ethical Behavior in the MACHIAVELLI Benchmark

    Alexander Pan, Jun Shern Chan, Andy Zou +7

    cs.LGcs.AIcs.CLarXiv:2304.03279v42023
  11. Pretraining-Based Natural Language Generation for Text Summarization

    Haoyu Zhang, Jianjun Xu, Ji Wang

    cs.CLcs.AIarXiv:1902.09243v22019
  12. Raise a Child in Large Language Model: Towards Effective and Generalizable Fine-tuning

    Runxin Xu, Fuli Luo, Zhiyuan Zhang +4

    cs.CLcs.AIarXiv:2109.05687v12021
  13. Neural Natural Language Inference Models Enhanced with External Knowledge

    Qian Chen, Xiaodan Zhu, Zhen-Hua Ling +2

    cs.CLarXiv:1711.04289v32017
  14. On the Predictive Power of Neural Language Models for Human Real-Time Comprehension Behavior

    Ethan Gotlieb Wilcox, Jon Gauthier, Jennifer Hu +2

    cs.CLarXiv:2006.01912v12020
  15. Recurrent Dropout without Memory Loss

    Stanislau Semeniuta, Aliaksei Severyn, Erhardt Barth

    cs.CLarXiv:1603.05118v22016
  16. Privacy- and Utility-Preserving Textual Analysis via Calibrated Multivariate Perturbations

    Oluwaseyi Feyisetan, Borja Balle, Thomas Drake +1

    cs.LGcs.CLcs.CRarXiv:1910.08902v12019
  17. Improving Entity Linking by Modeling Latent Relations between Mentions

    Phong Le, Ivan Titov

    cs.CLarXiv:1804.10637v12018
  18. MiNER: Fine-Tuned Biomedical Natural Language Processing for Malaria Disease Entity Recognition in Clinical Texts

    V. S. Anoop, Devika N

    cs.AIcs.CLarXiv:2609.00073v12026
  19. An Analysis of Hierarchical Text Classification Using Word Embeddings

    Roger A. Stein, Patricia A. Jaques, Joao F. Valiati

    cs.CLcs.AIcs.LGarXiv:1809.01771v12018
  20. Augmenting Language Models with Long-Term Memory

    Weizhi Wang, Li Dong, Hao Cheng +4

    cs.CLarXiv:2306.07174v12023
  21. Vision-Language Pre-training: Basics, Recent Advances, and Future Trends

    Zhe Gan, Linjie Li, Chunyuan Li +3

    cs.CVcs.CLarXiv:2210.09263v12022
  22. COCO-LM: Correcting and Contrasting Text Sequences for Language Model Pretraining

    Yu Meng, Chenyan Xiong, Payal Bajaj +4

    cs.CLcs.LGarXiv:2102.08473v22021
  23. BERT-of-Theseus: Compressing BERT by Progressive Module Replacing

    Canwen Xu, Wangchunshu Zhou, Tao Ge +2

    cs.CLcs.LGarXiv:2002.02925v42020
  24. Between words and characters: A Brief History of Open-Vocabulary Modeling and Tokenization in NLP

    Sabrina J. Mielke, Zaid Alyafeai, Elizabeth Salesky +8

    cs.CLcs.LGarXiv:2112.10508v12021
  25. Colossal-AI: A Unified Deep Learning System For Large-Scale Parallel Training

    Shenggui Li, Hongxin Liu, Zhengda Bian +5

    cs.LGcs.AIcs.CLarXiv:2110.14883v32021
  26. InCharacter: Evaluating Personality Fidelity in Role-Playing Agents through Psychological Interviews

    Xintao Wang, Yunze Xiao, Jen-tse Huang +10

    cs.CLarXiv:2310.17976v42023
  27. The Language Interpretability Tool: Extensible, Interactive Visualizations and Analysis for NLP Models

    Ian Tenney, James Wexler, Jasmijn Bastings +8

    cs.CLarXiv:2008.05122v12020
  28. Automatically Identifying Words That Can Serve as Labels for Few-Shot Text Classification

    Timo Schick, Helmut Schmid, Hinrich Schütze

    cs.CLcs.AIcs.LGarXiv:2010.13641v12020
  29. Detecting Hidden Behaviors in LLMs via Activation-matched Finetuning

    Robin Haselhorst, Lucie Flek, Florian Mai

    cs.CLcs.AIarXiv:2609.00351v12026
  30. From Distillation to Hard Negative Sampling: Making Sparse Neural IR Models More Effective

    Thibault Formal, Carlos Lassance, Benjamin Piwowarski +1

    cs.IRcs.CLarXiv:2205.04733v22022
  31. Mobile-Agent-v2: Mobile Device Operation Assistant with Effective Navigation via Multi-Agent Collaboration

    Junyang Wang, Haiyang Xu, Haitao Jia +6

    cs.CLcs.CVarXiv:2406.01014v12024
  32. Strategies for Structuring Story Generation

    Angela Fan, Mike Lewis, Yann Dauphin

    cs.CLarXiv:1902.01109v22019
  33. FlashRAG: A Modular Toolkit for Efficient Retrieval-Augmented Generation Research

    Jiajie Jin, Yutao Zhu, Guanting Dong +7

    cs.CLcs.IRarXiv:2405.13576v22024
  34. Multilingual Language Processing From Bytes

    Dan Gillick, Cliff Brunk, Oriol Vinyals +1

    cs.CLarXiv:1512.00103v22015
  35. Open-Domain Targeted Sentiment Analysis via Span-Based Extraction and Classification

    Minghao Hu, Yuxing Peng, Zhen Huang +2

    cs.CLarXiv:1906.03820v12019
  36. Identifying beneficial task relations for multi-task learning in deep neural networks

    Joachim Bingel, Anders Søgaard

    cs.CLarXiv:1702.08303v12017
  37. DEGREE: A Data-Efficient Generation-Based Event Extraction Model

    I-Hung Hsu, Kuan-Hao Huang, Elizabeth Boschee +4

    cs.CLcs.AIarXiv:2108.12724v32021
  38. Connecting the Dots: Document-level Neural Relation Extraction with Edge-oriented Graphs

    Fenia Christopoulou, Makoto Miwa, Sophia Ananiadou

    cs.CLarXiv:1909.00228v12019
  39. Fake News Early Detection: An Interdisciplinary Study

    Xinyi Zhou, Atishay Jain, Vir V. Phoha +1

    cs.CLcs.SIarXiv:1904.11679v22019
  40. Normalized and Geometry-Aware Self-Attention Network for Image Captioning

    Longteng Guo, Jing Liu, Xinxin Zhu +3

    cs.CVcs.CLcs.MMarXiv:2003.08897v12020
  41. CCAligned: A Massive Collection of Cross-Lingual Web-Document Pairs

    Ahmed El-Kishky, Vishrav Chaudhary, Francisco Guzman +1

    cs.CLcs.LGstat.MLarXiv:1911.06154v22019
  42. RENSA: Rich Environment Metadata to Navigate Shared and Distributed Endpoints for Automated Federated SPARQL Query Generation

    Victor Eiti Yamamoto, Takeda Hideaki, Yamamoto Yasunori

    cs.DBcs.CLarXiv:2608.28963v12026
  43. Towards Crafting Text Adversarial Samples

    Suranjana Samanta, Sameep Mehta

    cs.LGcs.AIcs.CLarXiv:1707.02812v12017
  44. Deductive Verification of Chain-of-Thought Reasoning

    Zhan Ling, Yunhao Fang, Xuanlin Li +4

    cs.CLcs.AIcs.LGarXiv:2306.03872v32023
  45. Attention Correctness in Neural Image Captioning

    Chenxi Liu, Junhua Mao, Fei Sha +1

    cs.CVcs.CLcs.LGarXiv:1605.09553v22016
  46. On the Impact of Various Types of Noise on Neural Machine Translation

    Huda Khayrallah, Philipp Koehn

    cs.CLarXiv:1805.12282v12018
  47. Knowledge Matters: Radiology Report Generation with General and Specific Knowledge

    Shuxin Yang, Xian Wu, Shen Ge +2

    eess.IVcs.CLcs.CVarXiv:2112.15009v22021
  48. When Are Tree Structures Necessary for Deep Learning of Representations?

    Jiwei Li, Minh-Thang Luong, Dan Jurafsky +1

    cs.AIcs.CLarXiv:1503.00185v52015
  49. The Parallelism Tradeoff: Limitations of Log-Precision Transformers

    William Merrill, Ashish Sabharwal

    cs.CCcs.CLarXiv:2207.00729v42022
  50. RankT5: Fine-Tuning T5 for Text Ranking with Ranking Losses

    Honglei Zhuang, Zhen Qin, Rolf Jagerman +6

    cs.IRcs.CLarXiv:2210.10634v12022
  51. Rate-Coding Bundle Memory: A Unified Model of Memory and Control for Symbolic Computation in the Brain

    Teun van Gils, Rowan P. Sommers, Markus Ostarek +1

    q-bio.NCcs.AIcs.CLarXiv:2608.29189v12026
  52. BIRD-History: A Benchmark for History-Driven Text-to-SQL with Fine-Grained Knowledge Annotations

    Yunfan Zhou, Qiming Shi, Yizhou Yang +2

    cs.AIcs.CLarXiv:2608.29345v12026
  53. Meta-StyleSpeech : Multi-Speaker Adaptive Text-to-Speech Generation

    Dongchan Min, Dong Bok Lee, Eunho Yang +1

    eess.AScs.CLcs.LGarXiv:2106.03153v32021
  54. Neurosymbolics for Data Engineering: Achieving Long Context Token Reduction Without Finetuning

    Vishvesh Bhat

    cs.CLcs.AIcs.LGarXiv:2609.00367v12026
  55. Concepts and Their Dynamics: A Quantum-Theoretic Modeling of Human Thought

    Diederik Aerts, Liane Gabora, Sandro Sozzo

    cs.AIcs.CLquant-pharXiv:1206.1069v22012
  56. Ultra-Fine Entity Typing

    Eunsol Choi, Omer Levy, Yejin Choi +1

    cs.CLcs.AIcs.LGarXiv:1807.04905v12018
  57. Removable and Irreducible: A Token-Cost Ledger for the Multilingual Tokenization Tax

    Madhulatha Mandarapu, Sandeep Kunkunuru

    cs.CLarXiv:2609.00378v12026
  58. Evaluating the Underlying Gender Bias in Contextualized Word Embeddings

    Christine Basta, Marta R. Costa-jussà, Noe Casas

    cs.CLcs.LGarXiv:1904.08783v12019
  59. Contextual LSTM (CLSTM) models for Large scale NLP tasks

    Shalini Ghosh, Oriol Vinyals, Brian Strope +3

    cs.CLarXiv:1602.06291v22016
  60. A Careful Examination of Large Language Model Performance on Grade School Arithmetic

    Hugh Zhang, Jeff Da, Dean Lee +12

    cs.CLcs.AIcs.LGarXiv:2405.00332v42024