Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
10,501 to 10,560 of 11,229
Generating Sequences With Recurrent Neural Networks
Alex Graves
cs.NEcs.CLarXiv:1308.0850v52013Open-MOPD: Diagnosing and Fixing Capability Imbalance in Multi-Teacher On-Policy Distillation
Huan-ang Gao, Haohan Chi, Yong Yan +7
cs.LGcs.AIcs.CLarXiv:2608.19098v12026Pointer Sentinel Mixture Models
Stephen Merity, Caiming Xiong, James Bradbury +1
cs.CLcs.AIarXiv:1609.07843v12016Making the V in VQA Matter: Elevating the Role of Image Understanding in Visual Question Answering
Yash Goyal, Tejas Khot, Douglas Summers-Stay +2
cs.CVcs.AIcs.CLarXiv:1612.00837v32016rEDMRec: Distilling Large Language Model Reasoning into an Editable Experience Memory for Recommendation
Minh Hoang Nguyen, Tung Le, Huy Tien Nguyen
cs.IRcs.AIcs.CLarXiv:2608.18952v12026The Curious Case of Neural Text Degeneration
Ari Holtzman, Jan Buys, Li Du +2
cs.CLarXiv:1904.09751v22019Bidirectional LSTM-CRF Models for Sequence Tagging
Zhiheng Huang, Wei Xu, Kai Yu
cs.CLarXiv:1508.01991v12015Tree of Thoughts: Deliberate Problem Solving with Large Language Models
Shunyu Yao, Dian Yu, Jeffrey Zhao +4
cs.CLcs.AIcs.LGarXiv:2305.10601v22023Lost in the Middle: How Language Models Use Long Contexts
Nelson F. Liu, Kevin Lin, John Hewitt +4
cs.CLarXiv:2307.03172v32023Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context
Zihang Dai, Zhilin Yang, Yiming Yang +3
cs.LGcs.CLstat.MLarXiv:1901.02860v32019MedUAG: Unified Understanding and Generation for Medical Multimodal Models
Zijie Meng, Yuncheng Zhang, Hualiang Wang +8
cs.CLcs.AIarXiv:2608.18937v12026Get To The Point: Summarization with Pointer-Generator Networks
Abigail See, Peter J. Liu, Christopher D. Manning
cs.CLarXiv:1704.04368v22017SMTrap: Cost-Effective DoS Attacks Against Large Reasoning Models via SMT Conflict Guidance
Jian Yang, Zhenqi Feng, Zhaoyang Yu +7
cs.CLcs.AIarXiv:2608.18921v12026Identifying Implicit Premises for Logical Reconstruction of Argument Graphs
Xuyao Feng, Anthony Hunter
cs.CLcs.AIarXiv:2608.18821v12026Do Large Language Models Hallucinate Electric Fata Morganas?
Kristina Šekrst
cs.CLcs.AIarXiv:2608.18816v12026Decomposing Wrong-Consensus Agreement in LLM Self-Consistency: A GPT-4.1 Case Study
Lizhuo Zhang, Mengmeng Tang, Chenfeng Long +2
cs.CLcs.AIarXiv:2608.18795v12026ViLBERT: Pretraining Task-Agnostic Visiolinguistic Representations for Vision-and-Language Tasks
Jiasen Lu, Dhruv Batra, Devi Parikh +1
cs.CVcs.CLarXiv:1908.02265v12019SimCSE: Simple Contrastive Learning of Sentence Embeddings
Tianyu Gao, Xingcheng Yao, Danqi Chen
cs.CLcs.LGarXiv:2104.08821v42021HellaSwag: Can a Machine Really Finish Your Sentence?
Rowan Zellers, Ari Holtzman, Yonatan Bisk +2
cs.CLarXiv:1905.07830v12019A Survey of Large Language Models
Wayne Xin Zhao, Kun Zhou, Junyi Li +19
cs.CLcs.AIarXiv:2303.18223v192023Evaluating and Explaining Prompt Sensitivity of LLMs Using Interactions
Ruiyang Qin, Qingzhuo Wang, Tian Wang +2
cs.LGcs.AIcs.CLarXiv:2608.18539v12026OPT: Open Pre-trained Transformer Language Models
Susan Zhang, Stephen Roller, Naman Goyal +16
cs.CLcs.LGarXiv:2205.01068v42022HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Wei-Ning Hsu, Benjamin Bolte, Yao-Hung Hubert Tsai +3
cs.CLcs.AIcs.LGarXiv:2106.07447v12021Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Peng Wang, Shuai Bai, Sinan Tan +16
cs.CVcs.AIcs.CLarXiv:2409.12191v22024GPT-4o System Card
OpenAI, :, Aaron Hurst +417
cs.CLcs.AIcs.CVarXiv:2410.21276v12024DART-SD: Diamond-topology Aware Retrieval and Tuning for Self-Distillation of Multi-Turn Tool-Calling Agents
Hangrui Xu, Jiarui Wang, Yang Yang +5
cs.CLcs.AIcs.LGarXiv:2608.18524v12026Qwen2.5 Technical Report
Qwen, :, An Yang +41
cs.CLarXiv:2412.15115v22024A large annotated corpus for learning natural language inference
Samuel R. Bowman, Gabor Angeli, Christopher Potts +1
cs.CLarXiv:1508.05326v12015Reflexion: Language Agents with Verbal Reinforcement Learning
Noah Shinn, Federico Cassano, Edward Berman +3
cs.AIcs.CLcs.LGarXiv:2303.11366v42023Toolformer: Language Models Can Teach Themselves to Use Tools
Timo Schick, Jane Dwivedi-Yu, Roberto Dessì +5
cs.CLarXiv:2302.04761v12023Attention Amnesia in Hybrid LLMs: When CoT Fine-Tuning Breaks Long-Range Recall, and How to Fix It
Xinyu Zhou, Boyu Zhu, Yi Xu +4
cs.CLarXiv:2606.11052v12026A Broad-Coverage Challenge Corpus for Sentence Understanding through Inference
Adina Williams, Nikita Nangia, Samuel R. Bowman
cs.CLarXiv:1704.05426v42017Pedagogical AI in Mental Health: A Tri-Stream Fine-Tuned LLM Framework for Automated Clinical Supervision and Risk Triage
Shreeya Sharma, Ravish Gupta, Saket Kumar +1
cs.CLcs.AIcs.LGarXiv:2608.18438v12026Selection, Recombination, or a Fresh Solve? A Candidate-Free Control for Single-Pass Test-Time Aggregation
Guiv Farmanfarmaian
cs.LGcs.AIcs.CLarXiv:2608.18379v12026Lius: Translation Model Based Instructional Lingustic Using Continual Instruction Tuning In Kupang Malay
Joanito Agili Lopo, Yunita Sari, Guntur Budi Herwanto
cs.CLarXiv:2606.11786v12026Survey of Hallucination in Natural Language Generation
Ziwei Ji, Nayeon Lee, Rita Frieske +10
cs.CLarXiv:2202.03629v72022Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Peter Clark, Isaac Cowhey, Oren Etzioni +4
cs.AIcs.CLcs.IRarXiv:1803.05457v12018Accurate Decoding of Natural Sentences from Non-Invasive Brain Recordings
Mingfang Zhang, Jarod Lévy, Cedric Rommel +9
cs.CLcs.AIcs.LGarXiv:2608.18114v12026Improved Baselines with Visual Instruction Tuning
Haotian Liu, Chunyuan Li, Yuheng Li +1
cs.CVcs.AIcs.CLarXiv:2310.03744v22023Finetuned Language Models Are Zero-Shot Learners
Jason Wei, Maarten Bosma, Vincent Y. Zhao +6
cs.CLarXiv:2109.01652v52021Language Models for Portuguese: A Systematic Mapping Study
Jhessica Silva, Carlos Caetano, Helena Maia +3
cs.CLcs.AIarXiv:2608.18138v12026Are LLMs Safe Beyond Text: Do Emojis Expose Gaps in Safety Evaluation
M P V S Gopinadh
cs.CLcs.AIcs.CRarXiv:2608.18164v12026When Do LLMs Actually Help? Evaluating LLMs as Data Quality Annotators
Praphulla Lal Shrestha
cs.CLcs.AIarXiv:2608.18158v12026The Deontic Gap: Large Language Models and the Modal Language of Obligation
Daniel Hart, Sarah Allred, Joseph Abbas +1
cs.CLcs.AIarXiv:2608.18144v12026Same Facts, Different Updates: Inference Setup Shapes LLM Behavior in Medical Allocation
Spencer Gibson, Tyler Crosse, Magnus Saebo +3
cs.CLcs.AIcs.HCarXiv:2608.18108v12026Different Facets of Verbalised Overconfidence: an Interpretability Study
Davide Mazzaccara, Leonardo Bertolazzi, Raffaella Bernardi
cs.CLcs.AIarXiv:2608.18106v12026Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision
Chao Jia, Yinfei Yang, Ye Xia +7
cs.CVcs.CLcs.LGarXiv:2102.05918v22021Prefix-Tuning: Optimizing Continuous Prompts for Generation
Xiang Lisa Li, Percy Liang
cs.CLarXiv:2101.00190v12021StocksTalk: A Voice-Enabled Conversational Agent for Structured Query Generation over Web Data
Akshat Parmar, Vikranth Udandarao, Abhay Shakya +4
cs.CLcs.AIcs.LGarXiv:2608.18105v12026Computational Orientalism: Measuring Structural Discourse Bias in Large Language Models Using the Middle East Cultural Sensitivity Score (MECSS)
Maha Shahid
cs.CLcs.AIcs.CYarXiv:2608.18100v12026Fractional Decay KV-Cache: Ownership-Aware Memory Management for Improved Inference Relevancy in Dialog Systems
Sukanta Ganguly
cs.CLcs.AIarXiv:2608.18098v12026Longformer: The Long-Document Transformer
Iz Beltagy, Matthew E. Peters, Arman Cohan
cs.CLarXiv:2004.05150v22020NE-BERT: A Multilingual Language Model for Nine Northeast Indian Languages
Badal Nyalang
cs.CLcs.AIarXiv:2608.18094v12026Abliteration Mitigation via Refusal Aliases
Nathan Truong
cs.CLcs.AIcs.CRarXiv:2608.18093v12026Nine Emotion Centroids: A Label-Free Valence Axis That Transfers Across Four Modalities
Yousef Radwan
cs.CLcs.AIarXiv:2608.18090v12026SuTRA : Structurally-Unified Tokenization with Root Awareness
Vaibhav Rathore, Siddhant Gole, Dadhichi Telwadkar +4
cs.CLcs.AIarXiv:2608.18087v12026Adaptive Memory and Reflection Multi-Agent System for Medical Question Answering
Pradeep Murugesan, Luoxiao Yang, Xueli Chen +1
cs.AIcs.CLcs.MAarXiv:2608.19029v12026Metrics That Write Themselves: Evolving an Evaluator from Its Own Blind Spots
Xing Zhang, Yanwei Cui, Guanghui Wang +2
cs.AIcs.CLcs.SEarXiv:2608.18744v12026Can a Lightweight Multimodal Model Estimate LLM Reasoning Performance? A Study for Compute-Optimal Document Inference
Zishan Ahmad, Vishal Vaddina
cs.AIcs.CLarXiv:2608.18591v12026Character-level Convolutional Networks for Text Classification
Xiang Zhang, Junbo Zhao, Yann LeCun
cs.LGcs.CLarXiv:1509.01626v32015