Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

3,901 to 3,960 of 11,325

  1. Path2ST: Hierarchical Cell-Tissue Grounded Cross-Modal Translation for Spatial Transcriptomics

    Ruochen Liu, Wei Lou

    cs.CVcs.AIcs.CLarXiv:2608.14710v12026
  2. KV Cache Compression Through the Lens of Transform Coding

    Hannah Laus, Claudio Mayrink Verdun, Hao Wang +2

    cs.LGcs.CLeess.SParXiv:2608.14191v12026
  3. Jais 2: A Family of Arabic-Centric Open Large Language Models

    Mohamed Anwar, Abed Alhakim Freihat, George Ibrahim +57

    cs.CLcs.AIarXiv:2608.13580v12026
  4. Modular TTT: Rethinking Test-Time Training as Composable Modules

    Bohao Tang, Zhen Qin, Yuqi Pan +3

    cs.LGcs.CLarXiv:2608.07110v12026
  5. Length Penalties Make Chain-of-Thought Less Monitorable

    Bryce Little

    cs.AIcs.CLcs.LGarXiv:2607.09786v32026
  6. STARE: Surprisal-Guided Token-Level Advantage Reweighting for Policy Entropy Stability

    Haipeng Luo, Qingfeng Sun, Songli Wu +4

    cs.LGcs.AIcs.CLarXiv:2606.19236v12026
  7. Exploring the Political Agenda of the European Parliament Using a Dynamic Topic Modeling Approach

    Derek Greene, James P. Cross

    cs.CLcs.CYarXiv:1607.03055v12016
  8. Can Generalist Agents Automate Data Curation?

    Feiyang Kang, Hanze Li, Adam Nguyen +5

    cs.AIcs.CLcs.CVarXiv:2606.04261v12026
  9. What Does an Agentic Software Engineering Benchmark Measure? Profiling Task Demands and Agent Behaviour Beyond What Category Labels Reveal

    Radin Shayanfar, Keheliya Gallaba, Ahmed E. Hassan

    cs.SEcs.CLarXiv:2609.01271v12026
  10. EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle

    Rong Wu, Xiaoman Wang, Jianbiao Mei +8

    cs.CLcs.AIarXiv:2510.16079v32025
  11. CEPO: RLVR Self-Distillation using Contrastive Evidence Policy Optimization

    Ahmed Heakl, Abdelrahman M. Shaker, Youssef Mohamed +4

    cs.LGcs.CLcs.CVarXiv:2605.19436v12026
  12. SDARE-Bench: Evaluating Large Language Models on Conversational Stigma Detection and Response in Dyadic and Group Dialogue

    Stephanie Fong, Yiwen Jiang, Zimu Wang +12

    cs.CLarXiv:2609.01548v12026
  13. Latent Preference Modeling for Cross-Session Personalized Tool Calling

    Yejin Yoon, Minseo Kim, Taeuk Kim

    cs.CLcs.AIarXiv:2604.17886v12026
  14. Cut Your Losses! Learning to Prune Paths Early for Efficient Parallel Reasoning

    Jiaxi Bi, Tongxu Luo, Wenyu Du +2

    cs.CLcs.LGarXiv:2604.16029v22026
  15. ExpArt-KG: Artwork Image Description Generation through Iterative Exploration of Knowledge Graphs

    Yuta Kato, Shintaro Ozaki, Kazuki Hayashi +4

    cs.CLcs.CVarXiv:2609.00629v12026
  16. Live-SWE-agent: Can Software Engineering Agents Self-Evolve on the Fly?

    Chunqiu Steven Xia, Zhe Wang, Yan Yang +2

    cs.SEcs.AIcs.CLarXiv:2511.13646v32025
  17. Scaling Computer-Use Grounding via User Interface Decomposition and Synthesis

    Tianbao Xie, Jiaqi Deng, Xiaochuan Li +12

    cs.AIcs.CLcs.CVarXiv:2505.13227v32025
  18. X-Teaming: Multi-Turn Jailbreaks and Defenses with Adaptive Multi-Agents

    Salman Rahman, Liwei Jiang, James Shiffer +7

    cs.CRcs.AIcs.CLarXiv:2504.13203v22025
  19. What Makes a Good Query? Measuring the Impact of Human-Confusing Linguistic Features on LLM Performance

    William Watson, Nicole Cho, Sumitra Ganesh +1

    cs.CLcs.AIarXiv:2602.20300v12026
  20. Accelerating Scientific Research with Gemini: Case Studies and Common Techniques

    David P. Woodruff, Vincent Cohen-Addad, Lalit Jain +33

    cs.CLcs.AIarXiv:2602.03837v32026
  21. Towards Autonomous Mathematics Research

    Tony Feng, Trieu H. Trinh, Garrett Bingham +25

    cs.LGcs.AIcs.CLarXiv:2602.10177v32026
  22. WebWalker: Benchmarking LLMs in Web Traversal

    Jialong Wu, Wenbiao Yin, Yong Jiang +8

    cs.CLcs.AIarXiv:2501.07572v32025
  23. Exploiting Semantics in Neural Machine Translation with Graph Convolutional Networks

    Diego Marcheggiani, Jasmijn Bastings, Ivan Titov

    cs.CLarXiv:1804.08313v22018
  24. Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

    Aojun Zhou, Ke Wang, Zimu Lu +8

    cs.CLcs.AIcs.CVarXiv:2308.07921v12023
  25. MedRAG: Enhancing Retrieval-augmented Generation with Knowledge Graph-Elicited Reasoning for Healthcare Copilot

    Xuejiao Zhao, Siyan Liu, Su-Yin Yang +1

    cs.CLcs.AIcs.IRarXiv:2502.04413v22025
  26. Exploring Sparse Autoencoders in Text-Based Causal Confounding Adjustment

    Mian Zhong, Katherine A. Keith, Anjalie Field

    cs.CLcs.LGarXiv:2609.01322v12026
  27. RewardBench 2: Advancing Reward Model Evaluation

    Saumya Malik, Valentina Pyatkin, Sander Land +4

    cs.CLarXiv:2506.01937v22025
  28. MultiChallenge: A Realistic Multi-Turn Conversation Evaluation Benchmark Challenging to Frontier LLMs

    Ved Sirdeshmukh, Kaustubh Deshpande, Johannes Mols +7

    cs.CLcs.AIarXiv:2501.17399v22025
  29. Overfitting Mitigation via Singular Value Decomposition in Minimum Bayes Risk Decoding

    Riza Setiawan Soetedjo, Yusuke Sakai, Hidetaka Kamigaito +2

    cs.CLarXiv:2609.01135v12026
  30. Prompt Programming for Large Language Models: Beyond the Few-Shot Paradigm

    Laria Reynolds, Kyle McDonell

    cs.CLcs.AIarXiv:2102.07350v12021
  31. Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More

    Zichen Wen, Yifeng Gao, Shaobo Wang +5

    cs.CLcs.CVarXiv:2502.11494v22025
  32. A Survey of Evaluation Metrics Used for NLG Systems

    Ananya B. Sai, Akash Kumar Mohankumar, Mitesh M. Khapra

    cs.CLarXiv:2008.12009v22020
  33. Thought Anchors: Which LLM Reasoning Steps Matter?

    Paul C. Bogdan, Uzay Macar, Neel Nanda +1

    cs.LGcs.AIcs.CLarXiv:2506.19143v42025
  34. The Surprising Effectiveness of Negative Reinforcement in LLM Reasoning

    Xinyu Zhu, Mengzhou Xia, Zhepei Wei +3

    cs.CLcs.LGarXiv:2506.01347v22025
  35. Characterizing the Google Books corpus: Strong limits to inferences of socio-cultural and linguistic evolution

    Eitan Adam Pechenick, Christopher M. Danforth, Peter Sheridan Dodds

    physics.soc-phcond-mat.stat-mechcs.CLarXiv:1501.00960v42015
  36. MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix

    Ziyang Ma, Yinghao Ma, Yanqiao Zhu +31

    cs.SDcs.CLcs.MMarXiv:2505.13032v12025
  37. The Tower of Babel Meets Web 2.0: User-Generated Content and its Applications in a Multilingual Context

    B. Hecht, D. Gergle

    cs.CLcs.HCarXiv:1904.01689v12019
  38. On the Theoretical Limitations of Embedding-Based Retrieval

    Orion Weller, Michael Boratko, Iftekhar Naim +1

    cs.IRcs.CLcs.LGarXiv:2508.21038v22025
    Summaries:한국어
  39. Step-by-Step: Separating Planning from Realization in Neural Data-to-Text Generation

    Amit Moryossef, Yoav Goldberg, Ido Dagan

    cs.CLcs.AIarXiv:1904.03396v22019
  40. StateSwap: Probing Support-Elimination Hidden States in Multiple-Choice Questions

    Chao Gao, Haijiang Liu, Qiyuan Li +3

    cs.CLcs.AIarXiv:2609.01081v12026
  41. Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains

    Yi Su, Dian Yu, Linfeng Song +5

    cs.CLarXiv:2503.23829v22025
  42. A Text Classification Framework for Simple and Effective Early Depression Detection Over Social Media Streams

    Sergio G. Burdisso, Marcelo Errecalde, Manuel Montes-y-Gómez

    cs.CYcs.CLcs.IRarXiv:1905.08772v22019
  43. Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond

    Liang Wen, Yunke Cai, Fenrui Xiao +11

    cs.CLcs.LGarXiv:2503.10460v42025
  44. Evaluating Commonsense in Pre-trained Language Models

    Xuhui Zhou, Yue Zhang, Leyang Cui +1

    cs.CLcs.AIarXiv:1911.11931v22019
  45. The Landscape of Agentic Reinforcement Learning for LLMs: A Survey

    Guibin Zhang, Hejia Geng, Xiaohang Yu +22

    cs.AIcs.CLarXiv:2509.02547v52025
  46. AgentSpec: Customizable Runtime Enforcement for Safe and Reliable LLM Agents

    Haoyu Wang, Christopher M. Poskitt, Jun Sun

    cs.AIcs.CLarXiv:2503.18666v32025
  47. Language Modeling with Deep Transformers

    Kazuki Irie, Albert Zeyer, Ralf Schlüter +1

    cs.CLcs.LGarXiv:1905.04226v22019
  48. RM-R1: Reward Modeling as Reasoning

    Xiusi Chen, Gaotang Li, Ziqi Wang +9

    cs.CLcs.AIcs.LGarXiv:2505.02387v42025
  49. VIBE-Bench: Evaluating Personalized Large Language Models When Profiles Don't Mean Preferences

    Yiwen Jiang, Yang Deng, Stephanie Fong +9

    cs.AIcs.CLarXiv:2609.00921v12026
  50. Predictors of Loneliness in Older Adults Using Multimodal Analysis of Speech and Language

    Vinmay Khandode, Sai Karthik Kosuri, Neil K. R. Sehgal +6

    cs.CLarXiv:2609.02606v12026
  51. LLaDA-V: Large Language Diffusion Models with Visual Instruction Tuning

    Zebin You, Shen Nie, Xiaolu Zhang +5

    cs.LGcs.CLcs.CVarXiv:2505.16933v22025
  52. Slow to See, Slow to Suppress: Understanding the Effects of Modality in Context-Memory Conflicts

    Athulith Paraselli, Etha Tianze Hua, Ellie Pavlick

    cs.CLarXiv:2609.00293v12026
  53. CoT-Valve: Length-Compressible Chain-of-Thought Tuning

    Xinyin Ma, Guangnian Wan, Runpeng Yu +2

    cs.AIcs.CLarXiv:2502.09601v12025
  54. Evaluating and Improving LLM Self-Modeling

    Siqi Zeng, Andre N. Assis, Rowan Wang

    cs.CLcs.AIarXiv:2608.30980v12026
    Summaries:한국어
  55. How Prolific Sellers Self-Present: Dissecting the Communication Patterns of 1.6 Million Reverb Listings

    David M. Markowitz

    cs.CLarXiv:2608.29952v12026
  56. Semantic Head Specialization Guides Hybrid ViT Attention for Multimodal LLMs

    Chenhong He, Lei Li, Shicheng Li +5

    cs.CVcs.CLarXiv:2608.28383v12026
  57. Puro-2B: Poor Lab's Qwen2-1.5B Trained on RTX 5090 within $5090

    Kairong Luo, Jiarui Cui, Yaorui Yin +8

    cs.CLcs.LGarXiv:2608.27370v12026
  58. Design and Empirical Characterization of a Hardware-Realized Turing Machine with Automated Card-Based Programming

    Agrima Regmi, Jenish Pant, Pratistha Sapkota +2

    cs.LOcs.CLarXiv:2608.24742v12026
  59. MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing

    Junbo Niu, Zheng Liu, Zhuangcheng Gu +58

    cs.CVcs.CLarXiv:2509.22186v22025
  60. A Tensorized Transformer for Language Modeling

    Xindian Ma, Peng Zhang, Shuai Zhang +4

    cs.CLcs.LGarXiv:1906.09777v32019