Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
60,721 to 60,780 of 61,271
WorldDiT: A Unified Diffusion Architecture for World and Action Modeling
Sen Wang, R. Gnana Praveen, Bidhan Roy +1
cs.LGcs.ROarXiv:2607.23909v22026Where Should Optimizer State Live? Tiered State Allocation for Memory-Efficient Mixture-of-Experts Training
Nuemaan Malik
cs.LGcs.AIarXiv:2607.19058v22026OPERA: Offline Policy-guided Expert Routing and Adaptation for Universal Biomedical Image Analysis
Zihan Li, Feiyang Liu, Dandan Shan +2
cs.CVcs.AIcs.LGarXiv:2607.25108v12026UltraViT: Latency-Optimized On-device Vision Encoder for Large Vision-Language Models
Ioannis Maniadis Metaxas, Adrian Bulat, Alberto Baldrati +4
cs.CVarXiv:2607.23373v12026LLM-as-a-Coach: Experiential Learning for Non-Verifiable Tasks
Tianzhu Ye, Li Dong, Guanheng Chen +4
cs.LGcs.CLarXiv:2607.18110v12026Reading and Steering Representations of Materials-Science Mechanisms in an Open-Weight Language Model
Markus J. Buehler
cs.AIcond-mat.mes-hallcond-mat.mtrl-sciarXiv:2607.20058v12026LAMAR: An Open Language-Aware Multilingual Alignment Reranker
Seongtae Hong, Youngjoon Jang, Jungseob Lee +2
cs.IRarXiv:2607.22042v22026SecRespond: Benchmarking AI Agents for Real-World Post-Compromise Incident Response
Lehan Wang, Boli Chen, Ruixue Ding +7
cs.CRcs.AIcs.CLarXiv:2607.26791v12026Streaming Multi-Agent Autoregressive Diffusion Model with World State Registers
Sicheng Mo, Yuheng Li, Ziyang Leng +2
cs.CVarXiv:2607.21594v12026Oxygen-TryOn: Fashion-Native Foundation Model for Any-item Virtual Try-On
Yong Liu, Xiaolong Fu, Zihang Xu +10
cs.CVarXiv:2607.21694v12026Distilled Reinforcement Learning for LLM Post-training
Chen Wang, Zhaochun Li, Jionghao Bai +4
cs.LGcs.AIarXiv:2607.17247v12026Tencent WorkBuddy Bench: A Multi-Domain Coding-Agent Benchmark with Contamination-Resistant Task Construction
Tencent WorkBuddy Bench Team, Siqi Cai, Shaopeng Chen +35
cs.CLcs.SEarXiv:2607.20911v12026Leveraging External Knowledge for Historical Document Restoration via Retrieval-Augmented Large Language Models
Gabeen Kim, Kyeongpil Kang
cs.CLarXiv:2607.21936v12026Not All Tokens Deserve Equal Credit: Counterfactual Sensitivity Credit Reallocation for Long-CoT Reasoning
Qiangqiang He, Zhongheng Wu, ZiJian Wang
cs.AIarXiv:2607.27888v12026CriPO: Enhancing Rubric-based RL via Self-Distillation
Mingxuan Xia, Yuhang Yang, Chao Ye +7
cs.LGcs.AIarXiv:2607.18082v32026ShotPlan: Cinematic Video Generation with Learnable Planning Token
Su Guo, Guangce Liu, Haosen Yang +7
cs.CVarXiv:2607.17675v12026WorldCupArena: Fine-Grained Evaluation of Language Models and Deep-Research Agents on Football Forecasting
Zhaokai Wang, Tianlin Gui, Jiayuan Rao +3
cs.AIcs.CLcs.LGarXiv:2607.18084v12026DocOps: A Verifiable Benchmark for Autonomous Agents in Complex Document Operations
Jiazhen Jiang, Boxi Cao, Lingyong Yan +6
cs.AIcs.CLcs.LGarXiv:2607.19865v12026AutoIndex: Learning Representation Programs for Retrieval
Sam O'Nuallain, Nithya Rajkumar, Ramya Narayanasamy +3
cs.IRcs.AIcs.CLarXiv:2607.18603v12026Would You Walk to the Car Wash? Revealing the Salience Bias of Large Language Models in Commonsense Reasoning
Zheng Wu, Chenhao Xue, Shijie Zheng +3
cs.CLarXiv:2607.28478v12026Beyond Feeling Better: Capability-Sustaining Emotional Dialogue as a Longitudinal Research Paradigm
Ming Wang, Jiaqi Wu Young, Wenfang Wu +2
cs.CLcs.HCcs.SIarXiv:2607.27851v12026Codifying the Judge: Scalable Evaluation via Program Distillation
Tzu-Heng Huang, Shengqi Qiu, Frederic Sala
cs.AIcs.LGarXiv:2607.22561v12026OmniDelta: Skill-Driven Budget Allocation for Token Compression in OmniLLMs
Haoyang Huang, Wenjie Huang, Tianqi Xu +14
cs.AIarXiv:2607.25669v12026Towards Robust Reinforcement Learning for Small-Scale Language Model Agents
Md Rezwanul Haque, Md. Milon Islam, Fakhri Karray
cs.AIcs.CLcs.LGarXiv:2607.25091v12026DataPrep-Bench: Benchmarking LLMs as Training Data Preparators
Hao Liang, Qifeng Cai, Yibo Lin +11
cs.LGcs.CLarXiv:2607.20465v12026SceneActBench: Can Agents Act on the 3D Scenes They See?
Yifei Zhao, Xiangxin Zhou, Wenhao Yang +11
cs.AIcs.CVarXiv:2607.22393v12026dRAE: Representation Autoencoder with Hyper-Spherical Codes
Tianren Ma, Lin Long, Chuyan Chen +4
cs.CVcs.AIarXiv:2607.22148v12026Scaling Laws for Hypernetwork-Based Knowledge Injection in Large Language Models
Nischay Dhankhar, Dos Baha, Abulhair Saparov
cs.CLcs.LGarXiv:2607.19604v12026Scaling Native Multimodal Pre-Training From Scratch
Haoyuan Wu, Aoqi Wu, Hai Wang +3
cs.CLcs.CVarXiv:2607.22043v12026Mapping CVEs to MITRE ATT&CK Techniques: A Curated Gold-Set Classifier and the Limits of LLM-Assisted Label Expansion
Cédric Bonhomme, Alexandre Dulaunoy
cs.CRarXiv:2607.25572v12026One Future, Every Robot: Label-Efficient Collective-State Prediction with Decentralized JEPA
Alan-Barsag Gazzaev, Alexey Gavrilov, Sergey Muravyov
cs.ROarXiv:2607.28443v32026TRACE: Business Rule-Grounded Reasoning Curriculum for Knowledge-Preserving Parametric Tool Retrieval in Enterprise LLMs
Sai Shruthi Sistla, Ashutosh Hathidara, Christopher Toukmaji +2
cs.AIarXiv:2607.22639v12026LEDGERMIND: Provenance-Constrained Multimodal Agentic Reasoning with a Structured Evidence Ledger
Enjun Du, Hange Zhou, Chenxu Du +4
cs.LGarXiv:2607.28374v12026CORAL: Curriculum-Optimized Reward Adaptation for LiDAR-Based Goal-Directed Urban Driving
Anisa Saleem, Duksu Kim
cs.ROcs.LGarXiv:2608.14332v12026Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing
Xinjie Zhang, Peng Zhang, Shicheng Zheng +21
cs.CVcs.AIcs.LGarXiv:2607.19064v22026Safeguards Based on Copyable Context Cannot Provide Reliable Safety for LLMs
Pingyu Wu, Lingyao Zhu, Weiming Zhang +1
cs.CRcs.AIarXiv:2607.27951v12026Grading the Narrators: An Isnad-Rijal Framework for Claim-Level Provenance in Multi-Agent Knowledge Systems
Ali Zahid Raja
cs.AIcs.MAarXiv:2607.24117v12026StealthBench: Measuring Operational Stealth in Autonomous Offensive-Security Agents
Ads Dawson, Adrian Wood
cs.CRcs.AIarXiv:2607.26314v12026Chamaileon: Cross-Context Binder Design with Contextualized Modeling and Mixed Sampling
Hengyuan Cao, Shizhuo Cheng, Mingxuan Liu +5
cs.LGq-bio.BMarXiv:2607.23518v12026GPT-Red: Automated Red Teaming via Self-Play at Scale
Eric Wallace, Christopher A. Choquette-Choo, Nikhil Kandpal +15
cs.CRcs.AIcs.CLarXiv:2607.26115v12026StateAct: Program State, before Pixels, for Long-Horizon Computer-Use Agents
Yan Yang, Xiangru Jian, Ziyang Luo +7
cs.SEcs.CVarXiv:2607.22798v12026Text Template Tokens Are Implicit Semantic Registers in Diffusion Transformers
Maohua Li, Qirui Li, Yanke Zhou +10
cs.CVarXiv:2607.19139v22026Filesystem-Based Memory for LLM Agents: Organization, Evolution, and Sustainability
Sizhe Zhou, Sheldon Yu, Hui Wei +8
cs.CLcs.AIarXiv:2607.26637v12026IDEAgent: Agentic Quality-Diversity Search for Research Idea Generation
Varun Gumma, Navonil Majumder, Soumitra Sinhahajari +1
cs.AIarXiv:2607.22375v12026Wonder: Video World Model Done Better
Jiacong Xu, Hanwen Jiang, Zhixin Shu +3
cs.CVcs.GRarXiv:2607.26037v12026Self-Supervised Learning of Structured Dynamics from Videos
Lukas Knobel, Andrew Zisserman, Yuki M. Asano
cs.CVarXiv:2607.21576v12026Human-in-the-Loop Signature Bootstrapping for UAV Hyperspectral PFM-1 Mine Detection
Sagar Lekhak, Prasanna Reddy Pulakurthi, Emmett J. Ientilucci
cs.CVeess.IVarXiv:2607.25310v12026RynnBrain 1.1: Towards More Capable and Generalizable Embodied Foundation Model
Kehan Li, Bohan Hou, Minghao Zhu +28
cs.ROarXiv:2607.17977v22026PerceptionBench: Evaluating Atomic Visual Perception in Multimodal Large Language Models
Zichao Lin, Yifeng Xie, Bowen Qu +30
cs.CVarXiv:2607.24957v12026SpecFirst: Behavioral Specification Elicitation as a First-Class Step in Agent-Based Program Synthesis from Scratch
Yihao Chen, Shi Chang, Feng Lin +4
cs.SEcs.CLarXiv:2607.27167v12026Masked Visual Actions for Unified World Modeling
Hadi Alzayer, Wenlong Huang, Haonan Chen +8
cs.CVcs.ROarXiv:2607.19343v12026Transcription Policy as a Latent Variable: Activating Controllable Verbatim ASR with Word-Level Timing
Laurin Wagner, Mario Zusag, Bernhard Thallinger
cs.CLarXiv:2607.18934v12026Explicit Layer Modeling for Video Object Insertion and Layer Decomposition
Kyujin Han, Seungjoo Shin, Sunghyun Cho
cs.CVarXiv:2607.25802v22026Memory for Large Language Models
Sining Zhoubian, Dan Zhang, Evgeny Kharlamov +1
cs.CLarXiv:2607.25380v12026Novel Claim or Déjà Vu? Rethinking "Contamination-Free'' Dynamic Evaluation for Multimodal Automated Fact-Checking
Haorui He, Xinwen Chen, Dacheng Wen +3
cs.CLcs.AIcs.MMarXiv:2607.23514v12026Stale but Stable: Staleness-Adaptive Trust Regions for Stabilizing Asynchronous Reinforcement Learning
Junyao Yang, Yucheng Shi, Zongxia Li +6
cs.LGcs.CLarXiv:2607.18722v32026SGTP: Sampling-based Game-Theoretic Planning for Real-Time Multi-Vehicle Autonomous Racing
Zhouheng Li, Fangguo Zhao, Mattia Piccinini +6
cs.ROarXiv:2607.25388v12026Shieldstral
Antonia Calvi, Avinash Sooriyarachchi, Giada Pistilli +273
cs.CLcs.CVarXiv:2607.25857v22026Sol-Attn: Accelerating Video Generation Inference via On-the-Fly Attention Sparsification
Haopeng Li, Yitong Li, Junsong Chen +8
cs.CVarXiv:2607.24027v12026Toward Robust and 3D-Aware RGB-NIR Imaging in the Dark
Muyao Niu, Mingze Ma, Yifan Zhan +5
cs.CVarXiv:2607.29684v12026