Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

9,781 to 9,840 of 11,254

  1. Repetition over Diversity: High-Signal Data Filtering for Sample-Efficient German Language Modeling

    Ansar Aynetdinov, Patrick Haller, Alan Akbik

    cs.CLcs.AIarXiv:2604.28075v22026
  2. InteractWeb-Bench: Can Multimodal Agent Escape Blind Execution in Interactive Website Generation?

    Qiyao Wang, Haoran Hu, Longze Chen +4

    cs.AIcs.CLarXiv:2604.27419v12026
  3. GoEmotions: A Dataset of Fine-Grained Emotions

    Dorottya Demszky, Dana Movshovitz-Attias, Jeongwoo Ko +3

    cs.CLarXiv:2005.00547v22020
  4. TCDA: Thread-Constrained Discourse-Aware Modeling for Conversational Sentiment Quadruple Analysis

    Xinran Li, Xinze Che, Yifan Lyu +2

    cs.CLcs.AIarXiv:2605.01717v22026
  5. Ask Me Anything: Dynamic Memory Networks for Natural Language Processing

    Ankit Kumar, Ozan Irsoy, Peter Ondruska +6

    cs.CLcs.LGcs.NEarXiv:1506.07285v52015
  6. A Primer on Neural Network Models for Natural Language Processing

    Yoav Goldberg

    cs.CLarXiv:1510.00726v12015
  7. Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

    Hakan Inan, Kartikeya Upasani, Jianfeng Chi +8

    cs.CLcs.AIarXiv:2312.06674v12023
  8. Interleaving Retrieval with Chain-of-Thought Reasoning for Knowledge-Intensive Multi-Step Questions

    Harsh Trivedi, Niranjan Balasubramanian, Tushar Khot +1

    cs.CLarXiv:2212.10509v22022
  9. AdapterFusion: Non-Destructive Task Composition for Transfer Learning

    Jonas Pfeiffer, Aishwarya Kamath, Andreas Rücklé +2

    cs.CLarXiv:2005.00247v32020
  10. Attention is not not Explanation

    Sarah Wiegreffe, Yuval Pinter

    cs.CLarXiv:1908.04626v22019
  11. Plug and Play Language Models: A Simple Approach to Controlled Text Generation

    Sumanth Dathathri, Andrea Madotto, Janice Lan +5

    cs.CLcs.AIcs.LGarXiv:1912.02164v42019
  12. SUPERB: Speech processing Universal PERformance Benchmark

    Shu-wen Yang, Po-Han Chi, Yung-Sung Chuang +17

    cs.CLcs.SDeess.ASarXiv:2105.01051v42021
  13. Representation Engineering: A Top-Down Approach to AI Transparency

    Andy Zou, Long Phan, Sarah Chen +18

    cs.LGcs.AIcs.CLarXiv:2310.01405v42023
  14. OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems

    Chaoqun He, Renjie Luo, Yuzhuo Bai +11

    cs.CLarXiv:2402.14008v22024
  15. Out of One, Many: Using Language Models to Simulate Human Samples

    Lisa P. Argyle, Ethan C. Busby, Nancy Fulda +3

    cs.LGcs.CLarXiv:2209.06899v12022
  16. Towards AI-Complete Question Answering: A Set of Prerequisite Toy Tasks

    Jason Weston, Antoine Bordes, Sumit Chopra +4

    cs.AIcs.CLstat.MLarXiv:1502.05698v102015
  17. Adversarial NLI: A New Benchmark for Natural Language Understanding

    Yixin Nie, Adina Williams, Emily Dinan +3

    cs.CLcs.LGarXiv:1910.14599v22019
  18. End-to-End Relation Extraction using LSTMs on Sequences and Tree Structures

    Makoto Miwa, Mohit Bansal

    cs.CLcs.LGarXiv:1601.00770v32016
  19. The TTS-STT Flywheel: Synthetic Entity-Dense Audio Closes the Indic ASR Gap Where Commercial and Open-Source Systems Fail

    Venkata Pushpak Teja Menta

    cs.CLcs.SDarXiv:2605.03073v12026
  20. WizardLM: Empowering large pre-trained language models to follow complex instructions

    Can Xu, Qingfeng Sun, Kai Zheng +6

    cs.CLcs.AIarXiv:2304.12244v32023
  21. Liberating LLM Capabilities in Full-Duplex Speech Models

    Luoyuan Zhang, Bokai Xu, Junbo Cui +4

    cs.CLcs.AIcs.SDarXiv:2606.07547v12026
  22. Deep Learning for Hate Speech Detection in Tweets

    Pinkesh Badjatiya, Shashank Gupta, Manish Gupta +1

    cs.CLcs.IRarXiv:1706.00188v12017
  23. TabEmbed: Benchmarking and Learning Generalist Embeddings for Tabular Understanding

    Minjie Qiang, Mingming Zhang, Xiaoyi Bao +5

    cs.CLcs.IRarXiv:2605.04962v12026
  24. Deep Captioning with Multimodal Recurrent Neural Networks (m-RNN)

    Junhua Mao, Wei Xu, Yi Yang +3

    cs.CVcs.CLcs.LGarXiv:1412.6632v52014
  25. Kosmos-2: Grounding Multimodal Large Language Models to the World

    Zhiliang Peng, Wenhui Wang, Li Dong +4

    cs.CLcs.CVarXiv:2306.14824v32023
  26. PatRe: A Full-Stage Office Action and Rebuttal Generation Benchmark for Patent Examination

    Qiyao Wang, Xinyi Chen, Longze Chen +4

    cs.CLcs.AIarXiv:2605.03571v12026
  27. Defending Against Neural Fake News

    Rowan Zellers, Ari Holtzman, Hannah Rashkin +4

    cs.CLcs.CYarXiv:1905.12616v32019
  28. The First Token Knows: Single-Decode Confidence for Hallucination Detection

    Mina Gabriel

    cs.CLcs.AIarXiv:2605.05166v12026
  29. Semi-supervised Sequence Learning

    Andrew M. Dai, Quoc V. Le

    cs.LGcs.CLarXiv:1511.01432v12015
  30. Teaching Large Language Models to Self-Debug

    Xinyun Chen, Maxwell Lin, Nathanael Schärli +1

    cs.CLcs.AIarXiv:2304.05128v22023
  31. A Sensitivity Analysis of (and Practitioners' Guide to) Convolutional Neural Networks for Sentence Classification

    Ye Zhang, Byron Wallace

    cs.CLcs.LGcs.NEarXiv:1510.03820v42015
  32. Are NLP Models really able to Solve Simple Math Word Problems?

    Arkil Patel, Satwik Bhattamishra, Navin Goyal

    cs.CLarXiv:2103.07191v22021
  33. VizWiz Grand Challenge: Answering Visual Questions from Blind People

    Danna Gurari, Qing Li, Abigale J. Stangl +5

    cs.CVcs.CLcs.HCarXiv:1802.08218v42018
  34. GLaM: Efficient Scaling of Language Models with Mixture-of-Experts

    Nan Du, Yanping Huang, Andrew M. Dai +24

    cs.CLarXiv:2112.06905v22021
  35. WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents

    Shunyu Yao, Howard Chen, John Yang +1

    cs.CLcs.AIcs.LGarXiv:2207.01206v42022
  36. When No Benchmark Exists: Validating Comparative LLM Safety Scoring Without Ground-Truth Labels

    Sushant Gautam, Finn Schwall, Annika Willoch Olstad +6

    cs.LGcs.AIcs.CLarXiv:2605.06652v12026
  37. Annotation Artifacts in Natural Language Inference Data

    Suchin Gururangan, Swabha Swayamdipta, Omer Levy +3

    cs.CLcs.AIarXiv:1803.02324v22018
  38. Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!

    Xiangyu Qi, Yi Zeng, Tinghao Xie +4

    cs.CLcs.AIcs.CRarXiv:2310.03693v12023
  39. BioTool: A Comprehensive Tool-Calling Dataset for Enhancing Biomedical Capabilities of Large Language Models

    Xin Gao, Ruiyi Zhang, Meixi Du +2

    cs.CLarXiv:2605.05758v12026
  40. Encouraging Divergent Thinking in Large Language Models through Multi-Agent Debate

    Tian Liang, Zhiwei He, Wenxiang Jiao +6

    cs.CLarXiv:2305.19118v42023
  41. The Granularity Axis: A Micro-to-Macro Latent Direction for Social Roles in Language Models

    Chonghan Qin, Xiachong Feng, Ziyun Song +3

    cs.AIcs.CLarXiv:2605.06196v12026
  42. StraTA: Incentivizing Agentic Reinforcement Learning with Strategic Trajectory Abstraction

    Xiangyuan Xue, Yifan Zhou, Zidong Wang +5

    cs.CLcs.AIarXiv:2605.06642v12026
  43. TIDE: Every Layer Knows the Token Beneath the Context

    Ajay Jaiswal, Lauren Hannah, Han-Byul Kim +3

    cs.CLcs.AIcs.LGarXiv:2605.06216v12026
  44. Nonsense Helps: Prompt Space Perturbation Broadens Reasoning Exploration

    Langlin Huang, Chengsong Huang, Jinyuan Li +3

    cs.AIcs.CLcs.LGarXiv:2605.05566v12026
  45. Rethinking RL for LLM Reasoning: It's Sparse Policy Selection, Not Capability Learning

    Ömer Faruk Akgül, Rajgopal Kannan, Willie Neiswanger +1

    cs.CLarXiv:2605.06241v22026
  46. Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers

    Chengyi Wang, Sanyuan Chen, Yu Wu +10

    cs.CLcs.SDeess.ASarXiv:2301.02111v12023
  47. Uncovering Entity Identity Confusion in Multimodal Knowledge Editing

    Shu Wu, Xiaotian Ye, Xinyu Mou +3

    cs.CLcs.CVarXiv:2605.06096v12026
  48. AllenNLP: A Deep Semantic Natural Language Processing Platform

    Matt Gardner, Joel Grus, Mark Neumann +6

    cs.CLarXiv:1803.07640v22018
  49. How Contextual are Contextualized Word Representations? Comparing the Geometry of BERT, ELMo, and GPT-2 Embeddings

    Kawin Ethayarajh

    cs.CLarXiv:1909.00512v12019
  50. Efficient Large-Scale Language Model Training on GPU Clusters Using Megatron-LM

    Deepak Narayanan, Mohammad Shoeybi, Jared Casper +9

    cs.CLcs.DCarXiv:2104.04473v52021
  51. Language-agnostic BERT Sentence Embedding

    Fangxiaoyu Feng, Yinfei Yang, Daniel Cer +2

    cs.CLarXiv:2007.01852v22020
  52. Deep Biaffine Attention for Neural Dependency Parsing

    Timothy Dozat, Christopher D. Manning

    cs.CLcs.NEarXiv:1611.01734v32016
  53. DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

    Dheeru Dua, Yizhong Wang, Pradeep Dasigi +3

    cs.CLarXiv:1903.00161v22019
  54. Quantifying Attention Flow in Transformers

    Samira Abnar, Willem Zuidema

    cs.LGcs.AIcs.CLarXiv:2005.00928v22020
  55. SimLex-999: Evaluating Semantic Models with (Genuine) Similarity Estimation

    Felix Hill, Roi Reichart, Anna Korhonen

    cs.CLarXiv:1408.3456v12014
  56. FastText.zip: Compressing text classification models

    Armand Joulin, Edouard Grave, Piotr Bojanowski +3

    cs.CLcs.LGarXiv:1612.03651v12016
  57. Mind2Web: Towards a Generalist Agent for the Web

    Xiang Deng, Yu Gu, Boyuan Zheng +5

    cs.CLarXiv:2306.06070v32023
  58. Atlas: Few-shot Learning with Retrieval Augmented Language Models

    Gautier Izacard, Patrick Lewis, Maria Lomeli +7

    cs.CLarXiv:2208.03299v32022
  59. SummaRuNNer: A Recurrent Neural Network based Sequence Model for Extractive Summarization of Documents

    Ramesh Nallapati, Feifei Zhai, Bowen Zhou

    cs.CLarXiv:1611.04230v12016
  60. Train Short, Test Long: Attention with Linear Biases Enables Input Length Extrapolation

    Ofir Press, Noah A. Smith, Mike Lewis

    cs.CLarXiv:2108.12409v22021