Computation and Language
Papers filed under cs.CL on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,141 to 4,200 of 11,216
The Ubuntu Dialogue Corpus: A Large Dataset for Research in Unstructured Multi-Turn Dialogue Systems
Ryan Lowe, Nissan Pow, Iulian Serban +1
cs.CLcs.AIcs.LGarXiv:1506.08909v32015On-Policy Context Distillation for Language Models
Tianzhu Ye, Li Dong, Xun Wu +2
cs.CLarXiv:2602.12275v22026TRIS: A Tri-Layer Retrieval Integrity Sieve Against Knowledge Poisoning
Muhaimin Bin Munir, Akib Jawad Ononto, Nazia Shehnaz Joynab +2
cs.CLcs.CRcs.IRarXiv:2609.00470v12026WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation
Yuwei Niu, Munan Ning, Mengren Zheng +9
cs.CVcs.AIcs.CLarXiv:2503.07265v42025Scalable Best-of-N Selection for Large Language Models via Self-Certainty
Zhewei Kang, Xuandong Zhao, Dawn Song
cs.CLcs.AIcs.LGarXiv:2502.18581v32025SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution
Yuxiang Wei, Olivier Duchenne, Jade Copet +6
cs.SEcs.AIcs.CLarXiv:2502.18449v22025Instella-MoE Technical Report
Jiang Liu, Sudhanshu Ranjan, Prakamya Mishra +10
cs.CLcs.AIarXiv:2609.00791v12026A Multi-Axis Annotation Scheme for Event Temporal Relations
Qiang Ning, Hao Wu, Dan Roth
cs.CLarXiv:1804.07828v22018Expressing stigma and inappropriate responses prevents LLMs from safely replacing mental health providers
Jared Moore, Declan Grabb, William Agnew +4
cs.CLarXiv:2504.18412v12025d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Siyan Zhao, Devaansh Gupta, Qinqing Zheng +1
cs.CLcs.LGarXiv:2504.12216v22025LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL
Yingzhe Peng, Gongrui Zhang, Miaosen Zhang +7
cs.CLcs.AIarXiv:2503.07536v22025DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation
Shansan Gong, Ruixiang Zhang, Huangjie Zheng +4
cs.CLarXiv:2506.20639v22025Stochastic Estimation of Transduced Language Models
Vésteinn Snæbjarnarson, Samuel Kiegeland, Manuel de Prada Corral +2
cs.CLarXiv:2608.27428v12026Safety Hacking in Constrained Best-of-$N$ Inference-time Scaling
Akifumi Wachi, Takumi Tanabe, Youhei Akimoto
cs.LGcs.AIcs.CLarXiv:2608.22915v12026Meta-Moderator: Empowering Multi-Agent Debate with Meta-Cognition
Wentao Hu, Zhuoyue Wan, Jinhao Shen +3
cs.CLarXiv:2608.23029v12026Test-Time Scaling in the Wild: Why Exploitation, Not Exploration, Is the Bottleneck
Davide Romano, Kanak Raj, Jerrod Parker +1
cs.CLcs.AIarXiv:2608.18931v12026BERTilda: Explainable Topic Lifecycle Tracking with Split/Merge Detection via Similarity-and-Flow Temporal Graphs
Cláudia Oliveira, Álvaro Figueira
cs.CLcs.LGarXiv:2608.18101v12026Institutional Prestige as Geographic Bias in Large Language Models: Evidence from Three Factorial Experiments with Bootstrap Confidence Intervals
Maikel Leyva-Vazquez, Florentin Smarandache
cs.CLcs.AIarXiv:2608.18107v12026TokEval: A Tokenizer Evaluation Suite
Clara Meister
cs.CLcs.LGarXiv:2608.18062v12026Institution-Specific LLM Prompting Recovers PHI That De-identification Systems and Their Gold Standards Both Miss
Daniel Palacios, Matthew Brady Neeley, Angel Adetomike Otto +6
cs.CLcs.AIarXiv:2608.17051v12026There is No Theoretical Curse of Multilinguality For Embedding Space Structure
Niyati Bafna, Neha Verma, Vilém Zouhar +2
cs.CLarXiv:2608.17088v12026Closing the Affective Loop: Multimodal Speaker-Listener Emotion-Dynamics-Aware Empathetic Social Robots
Zi Haur Pang, Casey Kennington, Tatsuya Kawahara
cs.HCcs.CLcs.ROarXiv:2608.16686v12026Bilingual-GAN: A Step Towards Parallel Text Generation
Ahmad Rashid, Alan Do-Omri, Md. Akmal Haidar +2
cs.CLcs.LGarXiv:1904.04742v22019Using the Mimi codec for metalinguistic representations
Artem Saloev, Erin Pacquetet, Nicolas Ballier
cs.CLarXiv:2608.15799v12026Multi-Modal Generative Fuzzy System: Fuzzy Inference Guided Large Model Interactive Question Answering Framework
Hailong Yang, Jianqi Wang, Guanjin Wang +1
cs.CLcs.AIarXiv:2608.14584v12026QUASAR: Lowering the Loss Floor of Quantization-Aware Training with Loss-Aware Reconstruction
Vincent Counathe, Ben Athiwaratkun, Christopher De Sa +1
cs.LGcs.CLstat.MLarXiv:2608.13966v12026SimpleOPD: Simple Tokenizer-Agnostic On-Policy Distillation for Long-Context Reasoning
Haonan He, Haodi Lei, Yun Luo +13
cs.CLcs.AIarXiv:2608.14277v12026Mental-LLM: Leveraging Large Language Models for Mental Health Prediction via Online Text Data
Xuhai Xu, Bingsheng Yao, Yuanzhe Dong +6
cs.CLarXiv:2307.14385v42023Capacity-Dependent Effects of Data Selection for Reasoning
Cuong Dang, Hoang Anh Just, Ruoxi Jia
cs.LGcs.AIcs.CLarXiv:2608.13721v12026ARC: Fair Relative Advantage Comparison in Open-Ended Real-World Interaction
Yongqi Tong, Tan Li Hui Faith, Choy Zhen Wen Marcus +5
cs.AIcs.CLarXiv:2608.13622v12026OmniScientist: An Omni-Modal Omni-Discipline AI Scientist
Bobo Li, Hao Fei, Tianjie Ju +2
cs.AIcs.CLarXiv:2608.13558v12026Summaries:한국어LittleLearner: Language Models Under Pedagogically Controlled Knowledge Exposure
Fanfei Li, Jana Zeller, Manuel Prada-Corral +4
cs.CLcs.AIcs.LGarXiv:2608.13545v12026LycheeMemory V2: Efficient Long-Term Memory for LLM Agents via Semantic Segment-Level Consolidation
Dongfang Li, Zixuan Liu, Junmai Wang +5
cs.CLarXiv:2608.12990v12026Mechanist: AI as a Scientific Instrument for Discovering the Mechanisms of Intelligence
Mengru Wang, Junfeng Fang, Shuofei Qiao +16
cs.AIcs.CLcs.HCarXiv:2608.12036v12026AI4AI at Test-Time: Strong-to-Weak Capability Transfer via Harnesses
Cheng Qian, Wenting Zhao, Liangwei Yang +6
cs.LGcs.AIcs.CLarXiv:2608.12307v12026Reference-Free Post-Training of Open Large Language Models for Multilingual Machine Translation
Chris Han, Pengzhi Gao, Pei Fu +1
cs.CLcs.AIarXiv:2608.10812v22026Claim-Level Reliability Assessment for Efficient Test-Time Reasoning
Sen Xu, Wei Wang, Shixi Liu +5
cs.AIcs.CLarXiv:2608.11994v12026Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill
Zhuoyang Qian, Biao Wu, Yiran Wang +6
cs.CLarXiv:2608.11924v12026Hybrid-Policy Self-Editing for Composable Unstructured Knowledge Editing
Tianci Liu, Zihan Dong, Tianchun Li +8
cs.CLcs.AIcs.LGarXiv:2608.11660v12026Ex-Omni-2D: Expressive Omni-Modal Dialogue Models with Native Visual Presence
Haoyu Zhang, Zhipeng Li, Xiaoying Tang +2
cs.AIcs.CLcs.CVarXiv:2608.10720v12026DistilVDR: A Compact End-to-End Visual Document Retriever via Dual-Student Distillation
Zhuchenyang Liu, Ziyi Wang, Yao Zhang +1
cs.IRcs.CLcs.CVarXiv:2608.10636v12026Power law graph attention: exact generalization of scaled dot-product attention, empirical collapse at inference
Burc Gokden
cs.LGcs.CLarXiv:2608.10288v12026Co-Evolution in Agentic Systems: Toward Self-Directed Evolution Beyond Human Design
Qing Zong, Jiayu Liu, Junhao Shen +9
cs.CLarXiv:2608.10299v12026Multimodal Model Diffing for Feature Discovery and Control
Hunar Batra, Lachin Naghashyar, Ashkan Khakzar +4
cs.CVcs.AIcs.CLarXiv:2608.09928v12026Simplex Relaxation for Discrete Diffusion
Jinya Sakurai, Patrick Pynadath, Satoshi Hayakawa +4
cs.CLarXiv:2608.10615v12026Decoding-Level Taboo: A Diagnostic Stress Test for LLM Robustness
Tadanobu Chuyo Kamijo, Ori Rottenstreich, Javier Conde +2
cs.CLarXiv:2608.09900v22026Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA
Mind Lab, :, Vin Bo +74
cs.LGcs.CLarXiv:2608.09819v12026How Can Rhetoric Reward-Hack AI Reviewers? Dissecting Rhetorical Sensitivity in AI-Based Peer Review
Ming Li, Chenguang Wang, Xirui Li +5
cs.CLcs.AIarXiv:2608.08975v12026Summaries:한국어Parameter Exploration for RLVR via Variational Learning
Vatsal Venkatkrishna, Nico Daheim, Iryna Gurevych
cs.LGcs.AIcs.CLarXiv:2608.09805v12026Reading Cognition as Decisions Unfold in Words: A Factorized Inverse Decision Model
Jiawen Kang, Dongrui Han, Xixin Wu +1
cs.CLq-bio.NCarXiv:2608.09222v12026Don't Scroll Back: Missing-Evidence Memory for Streaming Dialogue Summarization
Hyangsuk Min, Hwanjun Song
cs.CLcs.AIarXiv:2608.09043v12026Scaling Inherently Interpretable Language Models
Guide Labs Team, Andreas Madsen, Aya Abdelsalam Ismail +7
cs.CLcs.AIarXiv:2608.07594v12026Summaries:한국어VectraYX-Vision-1B: A Sub-2B Spanish/LATAM Cybersecurity Vision-Language Model with Structured Visual Reasoning and Native Tool Use
Juan S. Santillana
cs.CLarXiv:2608.08477v12026Vision-Language Grounding as Bidirectional Concept Correspondence
Jieyu Zhang, Ziqi Gao, Luke Zettlemoyer +1
cs.CVcs.AIcs.CLarXiv:2608.07886v12026CalibForge: Adversarial Solver Calibration for Scaling Learnable Terminal Tasks
Fanzhe Meng, Guoxin Chen, Jiale Zhao +6
cs.LGcs.CLarXiv:2608.06352v12026Kimi K3: Open Frontier Intelligence
Kimi Team, Tongtong Bai, Yifan Bai +399
cs.CLcs.LGarXiv:2607.24653v22026SkillZip: Contract-Preserving Graph Compression for Scalable Agent Skill Libraries
Xingyu Tan, Xiaoyang Wang, Qing Liu +4
cs.CLcs.AIarXiv:2608.05604v12026Summaries:한국어Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning
Yinghui He, Ling Yang, Jiarui Liu +6
cs.CLcs.LGarXiv:2608.05139v12026PAST-Bench: Benchmarking the Foundations of Recursive Self-Improvement in Personal Agents
Shuhan Xue, Zixin Ding, Yichen Shen +6
cs.CLarXiv:2608.04003v12026RoMeRL: Balancing Feedback Coverage and the Memory-Reward Trap in Self-Evolving Agent Memory via Reduced-Order Utility States
Yi Yang, Zhennan Chen, Yihong Zhuang +5
cs.LGcs.CLarXiv:2608.02508v32026