Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
421 to 480 of 11,252
Dream-RSI: Recursive Self-Improvement through Evolving Worlds
Tong Zheng, Xidong Wu, Zheng Zhang +14
cs.CLarXiv:2609.14858v12026RAGCache: Efficient Knowledge Caching for Retrieval-Augmented Generation
Chao Jin, Zili Zhang, Xuanlin Jiang +4
cs.DCcs.CLcs.LGarXiv:2404.12457v22024Prompting for Multimodal Hateful Meme Classification
Rui Cao, Roy Ka-Wei Lee, Wen-Haw Chong +1
cs.CLcs.IRcs.MMarXiv:2302.04156v12023ScienceBuddy: Recursive-in-Recursive Self-Improvement for Interactive Scientific Agents
Shuhan Xue, Jianyuan Zhong, Ziyuan Nan +10
cs.AIcs.CLarXiv:2609.17523v12026SAS: Simple Attention Sparsification via End-to-End Optimization of Context Ranking
Zhiwei Li, Lei Zhu, Hao Gu +6
cs.CLarXiv:2609.13141v12026AttnLRP: Attention-Aware Layer-Wise Relevance Propagation for Transformers
Reduan Achtibat, Sayed Mohammad Vakilzadeh Hatefi, Maximilian Dreyer +4
cs.CLcs.AIcs.CVarXiv:2402.05602v22024Ghostbuster: Detecting Text Ghostwritten by Large Language Models
Vivek Verma, Eve Fleisig, Nicholas Tomlin +1
cs.CLcs.AIarXiv:2305.15047v32023It Takes Two: Your GRPO Is Secretly DPO
Yihong Wu, Liheng Ma, Lei Ding +9
cs.LGcs.CLarXiv:2510.00977v32025RLAD: Training LLMs to Discover Abstractions for Solving Reasoning Problems
Yuxiao Qu, Anikait Singh, Yoonho Lee +4
cs.AIcs.CLcs.LGarXiv:2510.02263v12025Parallel Loop Transformer for Efficient Test-Time Computation Scaling
Bohong Wu, Mengzhao Chen, Xiang Luo +9
cs.CLarXiv:2510.24824v12025Response Ranking with Deep Matching Networks and External Knowledge in Information-seeking Conversation Systems
Liu Yang, Minghui Qiu, Chen Qu +5
cs.IRcs.CLarXiv:1805.00188v32018Improving Scene Text Recognition for Character-Level Long-Tailed Distribution
Sunghyun Park, Sunghyo Chung, Jungsoo Lee +1
cs.CVcs.CLarXiv:2304.08592v12023TrustJudge: Inconsistencies of LLM-as-a-Judge and How to Alleviate Them
Yidong Wang, Yunze Song, Tingyuan Zhu +11
cs.AIcs.CLarXiv:2509.21117v22025SWIM: Student Writing Simulation via Proficiency-Conditioned Generation
Heejin Do, Jakub Kontak, Mrinmaya Sachan
cs.CLcs.LGarXiv:2609.03215v12026DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads
Guangxuan Xiao, Jiaming Tang, Jingwei Zuo +5
cs.CLarXiv:2410.10819v12024LLM Maybe LongLM: Self-Extend LLM Context Window Without Tuning
Hongye Jin, Xiaotian Han, Jingfeng Yang +5
cs.CLcs.AIcs.LGarXiv:2401.01325v32024Symbolic Discovery of Optimization Algorithms
Xiangning Chen, Chen Liang, Da Huang +9
cs.LGcs.AIcs.CLarXiv:2302.06675v42023Bias Out-of-the-Box: An Empirical Analysis of Intersectional Occupational Biases in Popular Generative Language Models
Hannah Kirk, Yennie Jun, Haider Iqbal +5
cs.CLcs.AIarXiv:2102.04130v32021Transfer Learning for Named-Entity Recognition with Neural Networks
Ji Young Lee, Franck Dernoncourt, Peter Szolovits
cs.CLcs.AIcs.NEarXiv:1705.06273v12017Generating Wikipedia by Summarizing Long Sequences
Peter J. Liu, Mohammad Saleh, Etienne Pot +4
cs.CLarXiv:1801.10198v12018UltraCUA: A Foundation Model for Computer Use Agents with Hybrid Action
Yuhao Yang, Zhen Yang, Zi-Yi Dou +10
cs.CVcs.CLarXiv:2510.17790v32025d$^2$Cache: Accelerating Diffusion-Based LLMs via Dual Adaptive Caching
Yuchu Jiang, Yue Cai, Xiangzhong Luo +4
cs.CLarXiv:2509.23094v22025Multimodal ArXiv: A Dataset for Improving Scientific Comprehension of Large Vision-Language Models
Lei Li, Yuqi Wang, Runxin Xu +4
cs.CVcs.CLarXiv:2403.00231v32024Shaping capabilities with token-level data filtering
Neil Rathi, Alec Radford
cs.LGcs.AIcs.CLarXiv:2601.21571v22026Agentic Reasoning for Large Language Models
Tianxin Wei, Ting-Wei Li, Zhining Liu +26
cs.AIcs.CLarXiv:2601.12538v12026ToolSafe: Enhancing Tool Invocation Safety of LLM-based agents via Proactive Step-level Guardrail and Feedback
Yutao Mou, Zhangchi Xue, Lijun Li +4
cs.CLarXiv:2601.10156v12026ExpSeek: Self-Triggered Experience Seeking for Web Agents
Wenyuan Zhang, Xinghua Zhang, Haiyang Yu +5
cs.CLcs.AIarXiv:2601.08605v22026Double: Breaking the Acceleration Limit via Double Retrieval Speculative Parallelism
Yuhao Shen, Tianyu Liu, Junyi Shen +4
cs.CLarXiv:2601.05524v32026Stronger Normalization-Free Transformers
Mingzhi Chen, Taiming Lu, Jiachen Zhu +2
cs.LGcs.AIcs.CLarXiv:2512.10938v22025Adapting Language Models for Zero-shot Learning by Meta-tuning on Dataset and Prompt Collections
Ruiqi Zhong, Kristy Lee, Zheng Zhang +1
cs.CLcs.AIarXiv:2104.04670v52021SPIRAL: Self-Play on Zero-Sum Games Incentivizes Reasoning via Multi-Agent Multi-Turn Reinforcement Learning
Bo Liu, Leon Guertler, Simon Yu +9
cs.AIcs.CLcs.LGarXiv:2506.24119v32025Improving and Simplifying Pattern Exploiting Training
Derek Tam, Rakesh R Menon, Mohit Bansal +2
cs.CLcs.AIcs.LGarXiv:2103.11955v32021The Geometry of Culture: Analyzing Meaning through Word Embeddings
Austin C. Kozlowski, Matt Taddy, James A. Evans
cs.CLarXiv:1803.09288v12018Measuring Mathematical Problem Solving With the MATH Dataset
Dan Hendrycks, Collin Burns, Saurav Kadavath +5
cs.LGcs.AIcs.CLarXiv:2103.03874v22021Thought Crime: Backdoors and Emergent Misalignment in Reasoning Models
James Chua, Jan Betley, Mia Taylor +1
cs.LGcs.AIcs.CLarXiv:2506.13206v22025Deterministic Non-Autoregressive Neural Sequence Modeling by Iterative Refinement
Jason Lee, Elman Mansimov, Kyunghyun Cho
cs.LGcs.CLstat.MLarXiv:1802.06901v32018Neural Voice Cloning with a Few Samples
Sercan O. Arik, Jitong Chen, Kainan Peng +2
cs.CLcs.LGcs.SDarXiv:1802.06006v32018Structured Prediction as Translation between Augmented Natural Languages
Giovanni Paolini, Ben Athiwaratkun, Jason Krone +6
cs.LGcs.CLarXiv:2101.05779v32021Trankit: A Light-Weight Transformer-based Toolkit for Multilingual Natural Language Processing
Minh Van Nguyen, Viet Dac Lai, Amir Pouran Ben Veyseh +1
cs.CLarXiv:2101.03289v52021LLMs Get Lost In Multi-Turn Conversation
Philippe Laban, Hiroaki Hayashi, Yingbo Zhou +1
cs.CLcs.HCarXiv:2505.06120v12025Reinforcement Learning for Reasoning in Large Language Models with One Training Example
Yiping Wang, Qing Yang, Zhiyuan Zeng +11
cs.LGcs.AIcs.CLarXiv:2504.20571v32025Simple Recurrent Units for Highly Parallelizable Recurrence
Tao Lei, Yu Zhang, Sida I. Wang +2
cs.CLcs.NEarXiv:1709.02755v52017Paper2Code: Automating Code Generation from Scientific Papers in Machine Learning
Minju Seo, Jinheon Baek, Seongyun Lee +1
cs.CLarXiv:2504.17192v52025Keep CALM and Explore: Language Models for Action Generation in Text-based Games
Shunyu Yao, Rohan Rao, Matthew Hausknecht +1
cs.CLarXiv:2010.02903v12020GenPRM: Scaling Test-Time Compute of Process Reward Models via Generative Reasoning
Jian Zhao, Runze Liu, Kaiyan Zhang +8
cs.CLarXiv:2504.00891v22025Aligning AI With Shared Human Values
Dan Hendrycks, Collin Burns, Steven Basart +4
cs.CYcs.AIcs.CLarXiv:2008.02275v62020Measuring Chain of Thought Faithfulness by Unlearning Reasoning Steps
Martin Tutek, Fateme Hashemi Chaleshtori, Ana Marasović +1
cs.CLarXiv:2502.14829v42025A Survey on Large Language Models with some Insights on their Capabilities and Limitations
Andrea Matarazzo, Riccardo Torlone
cs.CLcs.AIcs.LGarXiv:2501.04040v22025Studying Image Tokenizers as Visual Languages in Unified Multimodal Models
Siting Li, Zhengyang Wang, Simon Shaolei Du +2
cs.CVcs.CLarXiv:2609.09143v12026DART: Open-Domain Structured Data Record to Text Generation
Linyong Nan, Dragomir Radev, Rui Zhang +21
cs.CLarXiv:2007.02871v22020wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations
Alexei Baevski, Henry Zhou, Abdelrahman Mohamed +1
cs.CLcs.LGcs.SDarXiv:2006.11477v32020Record Grouping Controls Evidence Weight in Language Models
Zhongxuan Liu, Sicheng Zhou, Hongzhi Wang
cs.CLarXiv:2609.08698v12026Language Models are Few-Shot Learners
Tom B. Brown, Benjamin Mann, Nick Ryder +28
cs.CLarXiv:2005.14165v42020Leveraging Low-Level Symbolic Competences for Unsupervised Grounding in Hallucination Detection
Renato Vukovic, Hsien-chin Lin, Carel van Niekerk +5
cs.CLcs.AIcs.IRarXiv:2609.05025v12026Embodied Agent Interface: Benchmarking LLMs for Embodied Decision Making
Manling Li, Shiyu Zhao, Qineng Wang +12
cs.CLcs.AIcs.LGarXiv:2410.07166v32024A Better Use of Audio-Visual Cues: Dense Video Captioning with Bi-modal Transformer
Vladimir Iashin, Esa Rahtu
cs.CVcs.CLcs.LGarXiv:2005.08271v22020Causal Mediation Analysis for Interpreting Neural NLP: The Case of Gender Bias
Jesse Vig, Sebastian Gehrmann, Yonatan Belinkov +6
cs.CLarXiv:2004.12265v22020We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?
Runqi Qiao, Qiuna Tan, Guanting Dong +15
cs.AIcs.CLcs.CVarXiv:2407.01284v12024PiPMRE: A Pipeline Based on Language Model for Medical Relation Extraction
Jiaxin Duan, Fengyu Lu, Junfei Liu
cs.CLarXiv:2609.02896v12026Large Language Models are Inconsistent and Biased Evaluators
Rickard Stureborg, Dimitris Alikaniotis, Yoshi Suhara
cs.CLcs.AIarXiv:2405.01724v12024