Computation and Language

Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,141 to 4,200 of 11,216

  1. The Ubuntu Dialogue Corpus: A Large Dataset for Research in Unstructured Multi-Turn Dialogue Systems

    Ryan Lowe, Nissan Pow, Iulian Serban +1

    cs.CLcs.AIcs.LGarXiv:1506.08909v32015
  2. On-Policy Context Distillation for Language Models

    Tianzhu Ye, Li Dong, Xun Wu +2

    cs.CLarXiv:2602.12275v22026
  3. TRIS: A Tri-Layer Retrieval Integrity Sieve Against Knowledge Poisoning

    Muhaimin Bin Munir, Akib Jawad Ononto, Nazia Shehnaz Joynab +2

    cs.CLcs.CRcs.IRarXiv:2609.00470v12026
  4. WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation

    Yuwei Niu, Munan Ning, Mengren Zheng +9

    cs.CVcs.AIcs.CLarXiv:2503.07265v42025
  5. Scalable Best-of-N Selection for Large Language Models via Self-Certainty

    Zhewei Kang, Xuandong Zhao, Dawn Song

    cs.CLcs.AIcs.LGarXiv:2502.18581v32025
  6. SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution

    Yuxiang Wei, Olivier Duchenne, Jade Copet +6

    cs.SEcs.AIcs.CLarXiv:2502.18449v22025
  7. Instella-MoE Technical Report

    Jiang Liu, Sudhanshu Ranjan, Prakamya Mishra +10

    cs.CLcs.AIarXiv:2609.00791v12026
  8. A Multi-Axis Annotation Scheme for Event Temporal Relations

    Qiang Ning, Hao Wu, Dan Roth

    cs.CLarXiv:1804.07828v22018
  9. Expressing stigma and inappropriate responses prevents LLMs from safely replacing mental health providers

    Jared Moore, Declan Grabb, William Agnew +4

    cs.CLarXiv:2504.18412v12025
  10. d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning

    Siyan Zhao, Devaansh Gupta, Qinqing Zheng +1

    cs.CLcs.LGarXiv:2504.12216v22025
  11. LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

    Yingzhe Peng, Gongrui Zhang, Miaosen Zhang +7

    cs.CLcs.AIarXiv:2503.07536v22025
  12. DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation

    Shansan Gong, Ruixiang Zhang, Huangjie Zheng +4

    cs.CLarXiv:2506.20639v22025
  13. Stochastic Estimation of Transduced Language Models

    Vésteinn Snæbjarnarson, Samuel Kiegeland, Manuel de Prada Corral +2

    cs.CLarXiv:2608.27428v12026
  14. Safety Hacking in Constrained Best-of-$N$ Inference-time Scaling

    Akifumi Wachi, Takumi Tanabe, Youhei Akimoto

    cs.LGcs.AIcs.CLarXiv:2608.22915v12026
  15. Meta-Moderator: Empowering Multi-Agent Debate with Meta-Cognition

    Wentao Hu, Zhuoyue Wan, Jinhao Shen +3

    cs.CLarXiv:2608.23029v12026
  16. Test-Time Scaling in the Wild: Why Exploitation, Not Exploration, Is the Bottleneck

    Davide Romano, Kanak Raj, Jerrod Parker +1

    cs.CLcs.AIarXiv:2608.18931v12026
  17. BERTilda: Explainable Topic Lifecycle Tracking with Split/Merge Detection via Similarity-and-Flow Temporal Graphs

    Cláudia Oliveira, Álvaro Figueira

    cs.CLcs.LGarXiv:2608.18101v12026
  18. Institutional Prestige as Geographic Bias in Large Language Models: Evidence from Three Factorial Experiments with Bootstrap Confidence Intervals

    Maikel Leyva-Vazquez, Florentin Smarandache

    cs.CLcs.AIarXiv:2608.18107v12026
  19. TokEval: A Tokenizer Evaluation Suite

    Clara Meister

    cs.CLcs.LGarXiv:2608.18062v12026
  20. Institution-Specific LLM Prompting Recovers PHI That De-identification Systems and Their Gold Standards Both Miss

    Daniel Palacios, Matthew Brady Neeley, Angel Adetomike Otto +6

    cs.CLcs.AIarXiv:2608.17051v12026
  21. There is No Theoretical Curse of Multilinguality For Embedding Space Structure

    Niyati Bafna, Neha Verma, Vilém Zouhar +2

    cs.CLarXiv:2608.17088v12026
  22. Closing the Affective Loop: Multimodal Speaker-Listener Emotion-Dynamics-Aware Empathetic Social Robots

    Zi Haur Pang, Casey Kennington, Tatsuya Kawahara

    cs.HCcs.CLcs.ROarXiv:2608.16686v12026
  23. Bilingual-GAN: A Step Towards Parallel Text Generation

    Ahmad Rashid, Alan Do-Omri, Md. Akmal Haidar +2

    cs.CLcs.LGarXiv:1904.04742v22019
  24. Using the Mimi codec for metalinguistic representations

    Artem Saloev, Erin Pacquetet, Nicolas Ballier

    cs.CLarXiv:2608.15799v12026
  25. Multi-Modal Generative Fuzzy System: Fuzzy Inference Guided Large Model Interactive Question Answering Framework

    Hailong Yang, Jianqi Wang, Guanjin Wang +1

    cs.CLcs.AIarXiv:2608.14584v12026
  26. QUASAR: Lowering the Loss Floor of Quantization-Aware Training with Loss-Aware Reconstruction

    Vincent Counathe, Ben Athiwaratkun, Christopher De Sa +1

    cs.LGcs.CLstat.MLarXiv:2608.13966v12026
  27. SimpleOPD: Simple Tokenizer-Agnostic On-Policy Distillation for Long-Context Reasoning

    Haonan He, Haodi Lei, Yun Luo +13

    cs.CLcs.AIarXiv:2608.14277v12026
  28. Mental-LLM: Leveraging Large Language Models for Mental Health Prediction via Online Text Data

    Xuhai Xu, Bingsheng Yao, Yuanzhe Dong +6

    cs.CLarXiv:2307.14385v42023
  29. Capacity-Dependent Effects of Data Selection for Reasoning

    Cuong Dang, Hoang Anh Just, Ruoxi Jia

    cs.LGcs.AIcs.CLarXiv:2608.13721v12026
  30. ARC: Fair Relative Advantage Comparison in Open-Ended Real-World Interaction

    Yongqi Tong, Tan Li Hui Faith, Choy Zhen Wen Marcus +5

    cs.AIcs.CLarXiv:2608.13622v12026
  31. OmniScientist: An Omni-Modal Omni-Discipline AI Scientist

    Bobo Li, Hao Fei, Tianjie Ju +2

    cs.AIcs.CLarXiv:2608.13558v12026
    Summaries:한국어
  32. LittleLearner: Language Models Under Pedagogically Controlled Knowledge Exposure

    Fanfei Li, Jana Zeller, Manuel Prada-Corral +4

    cs.CLcs.AIcs.LGarXiv:2608.13545v12026
  33. LycheeMemory V2: Efficient Long-Term Memory for LLM Agents via Semantic Segment-Level Consolidation

    Dongfang Li, Zixuan Liu, Junmai Wang +5

    cs.CLarXiv:2608.12990v12026
  34. Mechanist: AI as a Scientific Instrument for Discovering the Mechanisms of Intelligence

    Mengru Wang, Junfeng Fang, Shuofei Qiao +16

    cs.AIcs.CLcs.HCarXiv:2608.12036v12026
  35. AI4AI at Test-Time: Strong-to-Weak Capability Transfer via Harnesses

    Cheng Qian, Wenting Zhao, Liangwei Yang +6

    cs.LGcs.AIcs.CLarXiv:2608.12307v12026
  36. Reference-Free Post-Training of Open Large Language Models for Multilingual Machine Translation

    Chris Han, Pengzhi Gao, Pei Fu +1

    cs.CLcs.AIarXiv:2608.10812v22026
  37. Claim-Level Reliability Assessment for Efficient Test-Time Reasoning

    Sen Xu, Wei Wang, Shixi Liu +5

    cs.AIcs.CLarXiv:2608.11994v12026
  38. Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill

    Zhuoyang Qian, Biao Wu, Yiran Wang +6

    cs.CLarXiv:2608.11924v12026
  39. Hybrid-Policy Self-Editing for Composable Unstructured Knowledge Editing

    Tianci Liu, Zihan Dong, Tianchun Li +8

    cs.CLcs.AIcs.LGarXiv:2608.11660v12026
  40. Ex-Omni-2D: Expressive Omni-Modal Dialogue Models with Native Visual Presence

    Haoyu Zhang, Zhipeng Li, Xiaoying Tang +2

    cs.AIcs.CLcs.CVarXiv:2608.10720v12026
  41. DistilVDR: A Compact End-to-End Visual Document Retriever via Dual-Student Distillation

    Zhuchenyang Liu, Ziyi Wang, Yao Zhang +1

    cs.IRcs.CLcs.CVarXiv:2608.10636v12026
  42. Power law graph attention: exact generalization of scaled dot-product attention, empirical collapse at inference

    Burc Gokden

    cs.LGcs.CLarXiv:2608.10288v12026
  43. Co-Evolution in Agentic Systems: Toward Self-Directed Evolution Beyond Human Design

    Qing Zong, Jiayu Liu, Junhao Shen +9

    cs.CLarXiv:2608.10299v12026
  44. Multimodal Model Diffing for Feature Discovery and Control

    Hunar Batra, Lachin Naghashyar, Ashkan Khakzar +4

    cs.CVcs.AIcs.CLarXiv:2608.09928v12026
  45. Simplex Relaxation for Discrete Diffusion

    Jinya Sakurai, Patrick Pynadath, Satoshi Hayakawa +4

    cs.CLarXiv:2608.10615v12026
  46. Decoding-Level Taboo: A Diagnostic Stress Test for LLM Robustness

    Tadanobu Chuyo Kamijo, Ori Rottenstreich, Javier Conde +2

    cs.CLarXiv:2608.09900v22026
  47. Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA

    Mind Lab, :, Vin Bo +74

    cs.LGcs.CLarXiv:2608.09819v12026
  48. How Can Rhetoric Reward-Hack AI Reviewers? Dissecting Rhetorical Sensitivity in AI-Based Peer Review

    Ming Li, Chenguang Wang, Xirui Li +5

    cs.CLcs.AIarXiv:2608.08975v12026
    Summaries:한국어
  49. Parameter Exploration for RLVR via Variational Learning

    Vatsal Venkatkrishna, Nico Daheim, Iryna Gurevych

    cs.LGcs.AIcs.CLarXiv:2608.09805v12026
  50. Reading Cognition as Decisions Unfold in Words: A Factorized Inverse Decision Model

    Jiawen Kang, Dongrui Han, Xixin Wu +1

    cs.CLq-bio.NCarXiv:2608.09222v12026
  51. Don't Scroll Back: Missing-Evidence Memory for Streaming Dialogue Summarization

    Hyangsuk Min, Hwanjun Song

    cs.CLcs.AIarXiv:2608.09043v12026
  52. Scaling Inherently Interpretable Language Models

    Guide Labs Team, Andreas Madsen, Aya Abdelsalam Ismail +7

    cs.CLcs.AIarXiv:2608.07594v12026
    Summaries:한국어
  53. VectraYX-Vision-1B: A Sub-2B Spanish/LATAM Cybersecurity Vision-Language Model with Structured Visual Reasoning and Native Tool Use

    Juan S. Santillana

    cs.CLarXiv:2608.08477v12026
  54. Vision-Language Grounding as Bidirectional Concept Correspondence

    Jieyu Zhang, Ziqi Gao, Luke Zettlemoyer +1

    cs.CVcs.AIcs.CLarXiv:2608.07886v12026
  55. CalibForge: Adversarial Solver Calibration for Scaling Learnable Terminal Tasks

    Fanzhe Meng, Guoxin Chen, Jiale Zhao +6

    cs.LGcs.CLarXiv:2608.06352v12026
  56. Kimi K3: Open Frontier Intelligence

    Kimi Team, Tongtong Bai, Yifan Bai +399

    cs.CLcs.LGarXiv:2607.24653v22026
  57. SkillZip: Contract-Preserving Graph Compression for Scalable Agent Skill Libraries

    Xingyu Tan, Xiaoyang Wang, Qing Liu +4

    cs.CLcs.AIarXiv:2608.05604v12026
    Summaries:한국어
  58. Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning

    Yinghui He, Ling Yang, Jiarui Liu +6

    cs.CLcs.LGarXiv:2608.05139v12026
  59. PAST-Bench: Benchmarking the Foundations of Recursive Self-Improvement in Personal Agents

    Shuhan Xue, Zixin Ding, Yichen Shen +6

    cs.CLarXiv:2608.04003v12026
  60. RoMeRL: Balancing Feedback Coverage and the Memory-Reward Trap in Self-Evolving Agent Memory via Reduced-Order Utility States

    Yi Yang, Zhennan Chen, Yihong Zhuang +5

    cs.LGcs.CLarXiv:2608.02508v32026