Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
3,901 to 3,960 of 11,325
Path2ST: Hierarchical Cell-Tissue Grounded Cross-Modal Translation for Spatial Transcriptomics
Ruochen Liu, Wei Lou
cs.CVcs.AIcs.CLarXiv:2608.14710v12026KV Cache Compression Through the Lens of Transform Coding
Hannah Laus, Claudio Mayrink Verdun, Hao Wang +2
cs.LGcs.CLeess.SParXiv:2608.14191v12026Jais 2: A Family of Arabic-Centric Open Large Language Models
Mohamed Anwar, Abed Alhakim Freihat, George Ibrahim +57
cs.CLcs.AIarXiv:2608.13580v12026Modular TTT: Rethinking Test-Time Training as Composable Modules
Bohao Tang, Zhen Qin, Yuqi Pan +3
cs.LGcs.CLarXiv:2608.07110v12026Length Penalties Make Chain-of-Thought Less Monitorable
Bryce Little
cs.AIcs.CLcs.LGarXiv:2607.09786v32026STARE: Surprisal-Guided Token-Level Advantage Reweighting for Policy Entropy Stability
Haipeng Luo, Qingfeng Sun, Songli Wu +4
cs.LGcs.AIcs.CLarXiv:2606.19236v12026Exploring the Political Agenda of the European Parliament Using a Dynamic Topic Modeling Approach
Derek Greene, James P. Cross
cs.CLcs.CYarXiv:1607.03055v12016Can Generalist Agents Automate Data Curation?
Feiyang Kang, Hanze Li, Adam Nguyen +5
cs.AIcs.CLcs.CVarXiv:2606.04261v12026What Does an Agentic Software Engineering Benchmark Measure? Profiling Task Demands and Agent Behaviour Beyond What Category Labels Reveal
Radin Shayanfar, Keheliya Gallaba, Ahmed E. Hassan
cs.SEcs.CLarXiv:2609.01271v12026EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle
Rong Wu, Xiaoman Wang, Jianbiao Mei +8
cs.CLcs.AIarXiv:2510.16079v32025CEPO: RLVR Self-Distillation using Contrastive Evidence Policy Optimization
Ahmed Heakl, Abdelrahman M. Shaker, Youssef Mohamed +4
cs.LGcs.CLcs.CVarXiv:2605.19436v12026SDARE-Bench: Evaluating Large Language Models on Conversational Stigma Detection and Response in Dyadic and Group Dialogue
Stephanie Fong, Yiwen Jiang, Zimu Wang +12
cs.CLarXiv:2609.01548v12026Latent Preference Modeling for Cross-Session Personalized Tool Calling
Yejin Yoon, Minseo Kim, Taeuk Kim
cs.CLcs.AIarXiv:2604.17886v12026Cut Your Losses! Learning to Prune Paths Early for Efficient Parallel Reasoning
Jiaxi Bi, Tongxu Luo, Wenyu Du +2
cs.CLcs.LGarXiv:2604.16029v22026ExpArt-KG: Artwork Image Description Generation through Iterative Exploration of Knowledge Graphs
Yuta Kato, Shintaro Ozaki, Kazuki Hayashi +4
cs.CLcs.CVarXiv:2609.00629v12026Live-SWE-agent: Can Software Engineering Agents Self-Evolve on the Fly?
Chunqiu Steven Xia, Zhe Wang, Yan Yang +2
cs.SEcs.AIcs.CLarXiv:2511.13646v32025Scaling Computer-Use Grounding via User Interface Decomposition and Synthesis
Tianbao Xie, Jiaqi Deng, Xiaochuan Li +12
cs.AIcs.CLcs.CVarXiv:2505.13227v32025X-Teaming: Multi-Turn Jailbreaks and Defenses with Adaptive Multi-Agents
Salman Rahman, Liwei Jiang, James Shiffer +7
cs.CRcs.AIcs.CLarXiv:2504.13203v22025What Makes a Good Query? Measuring the Impact of Human-Confusing Linguistic Features on LLM Performance
William Watson, Nicole Cho, Sumitra Ganesh +1
cs.CLcs.AIarXiv:2602.20300v12026Accelerating Scientific Research with Gemini: Case Studies and Common Techniques
David P. Woodruff, Vincent Cohen-Addad, Lalit Jain +33
cs.CLcs.AIarXiv:2602.03837v32026Towards Autonomous Mathematics Research
Tony Feng, Trieu H. Trinh, Garrett Bingham +25
cs.LGcs.AIcs.CLarXiv:2602.10177v32026WebWalker: Benchmarking LLMs in Web Traversal
Jialong Wu, Wenbiao Yin, Yong Jiang +8
cs.CLcs.AIarXiv:2501.07572v32025Exploiting Semantics in Neural Machine Translation with Graph Convolutional Networks
Diego Marcheggiani, Jasmijn Bastings, Ivan Titov
cs.CLarXiv:1804.08313v22018Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification
Aojun Zhou, Ke Wang, Zimu Lu +8
cs.CLcs.AIcs.CVarXiv:2308.07921v12023MedRAG: Enhancing Retrieval-augmented Generation with Knowledge Graph-Elicited Reasoning for Healthcare Copilot
Xuejiao Zhao, Siyan Liu, Su-Yin Yang +1
cs.CLcs.AIcs.IRarXiv:2502.04413v22025Exploring Sparse Autoencoders in Text-Based Causal Confounding Adjustment
Mian Zhong, Katherine A. Keith, Anjalie Field
cs.CLcs.LGarXiv:2609.01322v12026RewardBench 2: Advancing Reward Model Evaluation
Saumya Malik, Valentina Pyatkin, Sander Land +4
cs.CLarXiv:2506.01937v22025MultiChallenge: A Realistic Multi-Turn Conversation Evaluation Benchmark Challenging to Frontier LLMs
Ved Sirdeshmukh, Kaustubh Deshpande, Johannes Mols +7
cs.CLcs.AIarXiv:2501.17399v22025Overfitting Mitigation via Singular Value Decomposition in Minimum Bayes Risk Decoding
Riza Setiawan Soetedjo, Yusuke Sakai, Hidetaka Kamigaito +2
cs.CLarXiv:2609.01135v12026Prompt Programming for Large Language Models: Beyond the Few-Shot Paradigm
Laria Reynolds, Kyle McDonell
cs.CLcs.AIarXiv:2102.07350v12021Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More
Zichen Wen, Yifeng Gao, Shaobo Wang +5
cs.CLcs.CVarXiv:2502.11494v22025A Survey of Evaluation Metrics Used for NLG Systems
Ananya B. Sai, Akash Kumar Mohankumar, Mitesh M. Khapra
cs.CLarXiv:2008.12009v22020Thought Anchors: Which LLM Reasoning Steps Matter?
Paul C. Bogdan, Uzay Macar, Neel Nanda +1
cs.LGcs.AIcs.CLarXiv:2506.19143v42025The Surprising Effectiveness of Negative Reinforcement in LLM Reasoning
Xinyu Zhu, Mengzhou Xia, Zhepei Wei +3
cs.CLcs.LGarXiv:2506.01347v22025Characterizing the Google Books corpus: Strong limits to inferences of socio-cultural and linguistic evolution
Eitan Adam Pechenick, Christopher M. Danforth, Peter Sheridan Dodds
physics.soc-phcond-mat.stat-mechcs.CLarXiv:1501.00960v42015MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Ziyang Ma, Yinghao Ma, Yanqiao Zhu +31
cs.SDcs.CLcs.MMarXiv:2505.13032v12025The Tower of Babel Meets Web 2.0: User-Generated Content and its Applications in a Multilingual Context
B. Hecht, D. Gergle
cs.CLcs.HCarXiv:1904.01689v12019On the Theoretical Limitations of Embedding-Based Retrieval
Orion Weller, Michael Boratko, Iftekhar Naim +1
cs.IRcs.CLcs.LGarXiv:2508.21038v22025Summaries:한국어Step-by-Step: Separating Planning from Realization in Neural Data-to-Text Generation
Amit Moryossef, Yoav Goldberg, Ido Dagan
cs.CLcs.AIarXiv:1904.03396v22019StateSwap: Probing Support-Elimination Hidden States in Multiple-Choice Questions
Chao Gao, Haijiang Liu, Qiyuan Li +3
cs.CLcs.AIarXiv:2609.01081v12026Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains
Yi Su, Dian Yu, Linfeng Song +5
cs.CLarXiv:2503.23829v22025A Text Classification Framework for Simple and Effective Early Depression Detection Over Social Media Streams
Sergio G. Burdisso, Marcelo Errecalde, Manuel Montes-y-Gómez
cs.CYcs.CLcs.IRarXiv:1905.08772v22019Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond
Liang Wen, Yunke Cai, Fenrui Xiao +11
cs.CLcs.LGarXiv:2503.10460v42025Evaluating Commonsense in Pre-trained Language Models
Xuhui Zhou, Yue Zhang, Leyang Cui +1
cs.CLcs.AIarXiv:1911.11931v22019The Landscape of Agentic Reinforcement Learning for LLMs: A Survey
Guibin Zhang, Hejia Geng, Xiaohang Yu +22
cs.AIcs.CLarXiv:2509.02547v52025AgentSpec: Customizable Runtime Enforcement for Safe and Reliable LLM Agents
Haoyu Wang, Christopher M. Poskitt, Jun Sun
cs.AIcs.CLarXiv:2503.18666v32025Language Modeling with Deep Transformers
Kazuki Irie, Albert Zeyer, Ralf Schlüter +1
cs.CLcs.LGarXiv:1905.04226v22019RM-R1: Reward Modeling as Reasoning
Xiusi Chen, Gaotang Li, Ziqi Wang +9
cs.CLcs.AIcs.LGarXiv:2505.02387v42025VIBE-Bench: Evaluating Personalized Large Language Models When Profiles Don't Mean Preferences
Yiwen Jiang, Yang Deng, Stephanie Fong +9
cs.AIcs.CLarXiv:2609.00921v12026Predictors of Loneliness in Older Adults Using Multimodal Analysis of Speech and Language
Vinmay Khandode, Sai Karthik Kosuri, Neil K. R. Sehgal +6
cs.CLarXiv:2609.02606v12026LLaDA-V: Large Language Diffusion Models with Visual Instruction Tuning
Zebin You, Shen Nie, Xiaolu Zhang +5
cs.LGcs.CLcs.CVarXiv:2505.16933v22025Slow to See, Slow to Suppress: Understanding the Effects of Modality in Context-Memory Conflicts
Athulith Paraselli, Etha Tianze Hua, Ellie Pavlick
cs.CLarXiv:2609.00293v12026CoT-Valve: Length-Compressible Chain-of-Thought Tuning
Xinyin Ma, Guangnian Wan, Runpeng Yu +2
cs.AIcs.CLarXiv:2502.09601v12025Evaluating and Improving LLM Self-Modeling
Siqi Zeng, Andre N. Assis, Rowan Wang
cs.CLcs.AIarXiv:2608.30980v12026Summaries:한국어How Prolific Sellers Self-Present: Dissecting the Communication Patterns of 1.6 Million Reverb Listings
David M. Markowitz
cs.CLarXiv:2608.29952v12026Semantic Head Specialization Guides Hybrid ViT Attention for Multimodal LLMs
Chenhong He, Lei Li, Shicheng Li +5
cs.CVcs.CLarXiv:2608.28383v12026Puro-2B: Poor Lab's Qwen2-1.5B Trained on RTX 5090 within $5090
Kairong Luo, Jiarui Cui, Yaorui Yin +8
cs.CLcs.LGarXiv:2608.27370v12026Design and Empirical Characterization of a Hardware-Realized Turing Machine with Automated Card-Based Programming
Agrima Regmi, Jenish Pant, Pratistha Sapkota +2
cs.LOcs.CLarXiv:2608.24742v12026MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Junbo Niu, Zheng Liu, Zhuangcheng Gu +58
cs.CVcs.CLarXiv:2509.22186v22025A Tensorized Transformer for Language Modeling
Xindian Ma, Peng Zhang, Shuai Zhang +4
cs.CLcs.LGarXiv:1906.09777v32019