Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

10,501 to 10,560 of 11,229

  1. Generating Sequences With Recurrent Neural Networks

    Alex Graves

    cs.NEcs.CLarXiv:1308.0850v52013
  2. Open-MOPD: Diagnosing and Fixing Capability Imbalance in Multi-Teacher On-Policy Distillation

    Huan-ang Gao, Haohan Chi, Yong Yan +7

    cs.LGcs.AIcs.CLarXiv:2608.19098v12026
  3. Pointer Sentinel Mixture Models

    Stephen Merity, Caiming Xiong, James Bradbury +1

    cs.CLcs.AIarXiv:1609.07843v12016
  4. Making the V in VQA Matter: Elevating the Role of Image Understanding in Visual Question Answering

    Yash Goyal, Tejas Khot, Douglas Summers-Stay +2

    cs.CVcs.AIcs.CLarXiv:1612.00837v32016
  5. rEDMRec: Distilling Large Language Model Reasoning into an Editable Experience Memory for Recommendation

    Minh Hoang Nguyen, Tung Le, Huy Tien Nguyen

    cs.IRcs.AIcs.CLarXiv:2608.18952v12026
  6. The Curious Case of Neural Text Degeneration

    Ari Holtzman, Jan Buys, Li Du +2

    cs.CLarXiv:1904.09751v22019
  7. Bidirectional LSTM-CRF Models for Sequence Tagging

    Zhiheng Huang, Wei Xu, Kai Yu

    cs.CLarXiv:1508.01991v12015
  8. Tree of Thoughts: Deliberate Problem Solving with Large Language Models

    Shunyu Yao, Dian Yu, Jeffrey Zhao +4

    cs.CLcs.AIcs.LGarXiv:2305.10601v22023
  9. Lost in the Middle: How Language Models Use Long Contexts

    Nelson F. Liu, Kevin Lin, John Hewitt +4

    cs.CLarXiv:2307.03172v32023
  10. Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context

    Zihang Dai, Zhilin Yang, Yiming Yang +3

    cs.LGcs.CLstat.MLarXiv:1901.02860v32019
  11. MedUAG: Unified Understanding and Generation for Medical Multimodal Models

    Zijie Meng, Yuncheng Zhang, Hualiang Wang +8

    cs.CLcs.AIarXiv:2608.18937v12026
  12. Get To The Point: Summarization with Pointer-Generator Networks

    Abigail See, Peter J. Liu, Christopher D. Manning

    cs.CLarXiv:1704.04368v22017
  13. SMTrap: Cost-Effective DoS Attacks Against Large Reasoning Models via SMT Conflict Guidance

    Jian Yang, Zhenqi Feng, Zhaoyang Yu +7

    cs.CLcs.AIarXiv:2608.18921v12026
  14. Identifying Implicit Premises for Logical Reconstruction of Argument Graphs

    Xuyao Feng, Anthony Hunter

    cs.CLcs.AIarXiv:2608.18821v12026
  15. Do Large Language Models Hallucinate Electric Fata Morganas?

    Kristina Šekrst

    cs.CLcs.AIarXiv:2608.18816v12026
  16. Decomposing Wrong-Consensus Agreement in LLM Self-Consistency: A GPT-4.1 Case Study

    Lizhuo Zhang, Mengmeng Tang, Chenfeng Long +2

    cs.CLcs.AIarXiv:2608.18795v12026
  17. ViLBERT: Pretraining Task-Agnostic Visiolinguistic Representations for Vision-and-Language Tasks

    Jiasen Lu, Dhruv Batra, Devi Parikh +1

    cs.CVcs.CLarXiv:1908.02265v12019
  18. SimCSE: Simple Contrastive Learning of Sentence Embeddings

    Tianyu Gao, Xingcheng Yao, Danqi Chen

    cs.CLcs.LGarXiv:2104.08821v42021
  19. HellaSwag: Can a Machine Really Finish Your Sentence?

    Rowan Zellers, Ari Holtzman, Yonatan Bisk +2

    cs.CLarXiv:1905.07830v12019
  20. A Survey of Large Language Models

    Wayne Xin Zhao, Kun Zhou, Junyi Li +19

    cs.CLcs.AIarXiv:2303.18223v192023
  21. Evaluating and Explaining Prompt Sensitivity of LLMs Using Interactions

    Ruiyang Qin, Qingzhuo Wang, Tian Wang +2

    cs.LGcs.AIcs.CLarXiv:2608.18539v12026
  22. OPT: Open Pre-trained Transformer Language Models

    Susan Zhang, Stephen Roller, Naman Goyal +16

    cs.CLcs.LGarXiv:2205.01068v42022
  23. HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units

    Wei-Ning Hsu, Benjamin Bolte, Yao-Hung Hubert Tsai +3

    cs.CLcs.AIcs.LGarXiv:2106.07447v12021
  24. Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

    Peng Wang, Shuai Bai, Sinan Tan +16

    cs.CVcs.AIcs.CLarXiv:2409.12191v22024
  25. GPT-4o System Card

    OpenAI, :, Aaron Hurst +417

    cs.CLcs.AIcs.CVarXiv:2410.21276v12024
  26. DART-SD: Diamond-topology Aware Retrieval and Tuning for Self-Distillation of Multi-Turn Tool-Calling Agents

    Hangrui Xu, Jiarui Wang, Yang Yang +5

    cs.CLcs.AIcs.LGarXiv:2608.18524v12026
  27. Qwen2.5 Technical Report

    Qwen, :, An Yang +41

    cs.CLarXiv:2412.15115v22024
  28. A large annotated corpus for learning natural language inference

    Samuel R. Bowman, Gabor Angeli, Christopher Potts +1

    cs.CLarXiv:1508.05326v12015
  29. Reflexion: Language Agents with Verbal Reinforcement Learning

    Noah Shinn, Federico Cassano, Edward Berman +3

    cs.AIcs.CLcs.LGarXiv:2303.11366v42023
  30. Toolformer: Language Models Can Teach Themselves to Use Tools

    Timo Schick, Jane Dwivedi-Yu, Roberto Dessì +5

    cs.CLarXiv:2302.04761v12023
  31. Attention Amnesia in Hybrid LLMs: When CoT Fine-Tuning Breaks Long-Range Recall, and How to Fix It

    Xinyu Zhou, Boyu Zhu, Yi Xu +4

    cs.CLarXiv:2606.11052v12026
  32. A Broad-Coverage Challenge Corpus for Sentence Understanding through Inference

    Adina Williams, Nikita Nangia, Samuel R. Bowman

    cs.CLarXiv:1704.05426v42017
  33. Pedagogical AI in Mental Health: A Tri-Stream Fine-Tuned LLM Framework for Automated Clinical Supervision and Risk Triage

    Shreeya Sharma, Ravish Gupta, Saket Kumar +1

    cs.CLcs.AIcs.LGarXiv:2608.18438v12026
  34. Selection, Recombination, or a Fresh Solve? A Candidate-Free Control for Single-Pass Test-Time Aggregation

    Guiv Farmanfarmaian

    cs.LGcs.AIcs.CLarXiv:2608.18379v12026
  35. Lius: Translation Model Based Instructional Lingustic Using Continual Instruction Tuning In Kupang Malay

    Joanito Agili Lopo, Yunita Sari, Guntur Budi Herwanto

    cs.CLarXiv:2606.11786v12026
  36. Survey of Hallucination in Natural Language Generation

    Ziwei Ji, Nayeon Lee, Rita Frieske +10

    cs.CLarXiv:2202.03629v72022
  37. Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

    Peter Clark, Isaac Cowhey, Oren Etzioni +4

    cs.AIcs.CLcs.IRarXiv:1803.05457v12018
  38. Accurate Decoding of Natural Sentences from Non-Invasive Brain Recordings

    Mingfang Zhang, Jarod Lévy, Cedric Rommel +9

    cs.CLcs.AIcs.LGarXiv:2608.18114v12026
  39. Improved Baselines with Visual Instruction Tuning

    Haotian Liu, Chunyuan Li, Yuheng Li +1

    cs.CVcs.AIcs.CLarXiv:2310.03744v22023
  40. Finetuned Language Models Are Zero-Shot Learners

    Jason Wei, Maarten Bosma, Vincent Y. Zhao +6

    cs.CLarXiv:2109.01652v52021
  41. Language Models for Portuguese: A Systematic Mapping Study

    Jhessica Silva, Carlos Caetano, Helena Maia +3

    cs.CLcs.AIarXiv:2608.18138v12026
  42. Are LLMs Safe Beyond Text: Do Emojis Expose Gaps in Safety Evaluation

    M P V S Gopinadh

    cs.CLcs.AIcs.CRarXiv:2608.18164v12026
  43. When Do LLMs Actually Help? Evaluating LLMs as Data Quality Annotators

    Praphulla Lal Shrestha

    cs.CLcs.AIarXiv:2608.18158v12026
  44. The Deontic Gap: Large Language Models and the Modal Language of Obligation

    Daniel Hart, Sarah Allred, Joseph Abbas +1

    cs.CLcs.AIarXiv:2608.18144v12026
  45. Same Facts, Different Updates: Inference Setup Shapes LLM Behavior in Medical Allocation

    Spencer Gibson, Tyler Crosse, Magnus Saebo +3

    cs.CLcs.AIcs.HCarXiv:2608.18108v12026
  46. Different Facets of Verbalised Overconfidence: an Interpretability Study

    Davide Mazzaccara, Leonardo Bertolazzi, Raffaella Bernardi

    cs.CLcs.AIarXiv:2608.18106v12026
  47. Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision

    Chao Jia, Yinfei Yang, Ye Xia +7

    cs.CVcs.CLcs.LGarXiv:2102.05918v22021
  48. Prefix-Tuning: Optimizing Continuous Prompts for Generation

    Xiang Lisa Li, Percy Liang

    cs.CLarXiv:2101.00190v12021
  49. StocksTalk: A Voice-Enabled Conversational Agent for Structured Query Generation over Web Data

    Akshat Parmar, Vikranth Udandarao, Abhay Shakya +4

    cs.CLcs.AIcs.LGarXiv:2608.18105v12026
  50. Computational Orientalism: Measuring Structural Discourse Bias in Large Language Models Using the Middle East Cultural Sensitivity Score (MECSS)

    Maha Shahid

    cs.CLcs.AIcs.CYarXiv:2608.18100v12026
  51. Fractional Decay KV-Cache: Ownership-Aware Memory Management for Improved Inference Relevancy in Dialog Systems

    Sukanta Ganguly

    cs.CLcs.AIarXiv:2608.18098v12026
  52. Longformer: The Long-Document Transformer

    Iz Beltagy, Matthew E. Peters, Arman Cohan

    cs.CLarXiv:2004.05150v22020
  53. NE-BERT: A Multilingual Language Model for Nine Northeast Indian Languages

    Badal Nyalang

    cs.CLcs.AIarXiv:2608.18094v12026
  54. Abliteration Mitigation via Refusal Aliases

    Nathan Truong

    cs.CLcs.AIcs.CRarXiv:2608.18093v12026
  55. Nine Emotion Centroids: A Label-Free Valence Axis That Transfers Across Four Modalities

    Yousef Radwan

    cs.CLcs.AIarXiv:2608.18090v12026
  56. SuTRA : Structurally-Unified Tokenization with Root Awareness

    Vaibhav Rathore, Siddhant Gole, Dadhichi Telwadkar +4

    cs.CLcs.AIarXiv:2608.18087v12026
  57. Adaptive Memory and Reflection Multi-Agent System for Medical Question Answering

    Pradeep Murugesan, Luoxiao Yang, Xueli Chen +1

    cs.AIcs.CLcs.MAarXiv:2608.19029v12026
  58. Metrics That Write Themselves: Evolving an Evaluator from Its Own Blind Spots

    Xing Zhang, Yanwei Cui, Guanghui Wang +2

    cs.AIcs.CLcs.SEarXiv:2608.18744v12026
  59. Can a Lightweight Multimodal Model Estimate LLM Reasoning Performance? A Study for Compute-Optimal Document Inference

    Zishan Ahmad, Vishal Vaddina

    cs.AIcs.CLarXiv:2608.18591v12026
  60. Character-level Convolutional Networks for Text Classification

    Xiang Zhang, Junbo Zhao, Yann LeCun

    cs.LGcs.CLarXiv:1509.01626v32015