Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
59,701 to 59,760 of 61,374
MuJoCo-Drones-Gym: A GPU-Accelerated Multi-Drone Simulator for Control and Reinforcement Learning
Manan Tayal
cs.ROarXiv:2606.08039v12026ARISE: An adaptive residual-informed stability ensemble for feature selection in small-sample biomedical omics
Zardad Khan, Amjad Ali, Naz Gul +2
stat.MLcs.LGarXiv:2608.14866v12026MCP-Persona: Benchmarking LLM Agents on Real-World Personal Applications via Environment Simulation
Wenhao Wang, Peizhi Niu, Gongyi Zou +9
cs.AIarXiv:2606.02470v12026VLMs are Good Teachers for Video Reasoning via Adaptive Test-Time Optimization
Junhao Cheng, Liang Hou, Tianxiong Zhong +4
cs.CVarXiv:2606.02564v32026SkillHarm: Lifecycle-Aware Skill-Based Attacks via Automated Construction
Yuting Ning, Zhehao Zhang, Yash Kumar Lal +8
cs.CLarXiv:2606.02540v12026Eliciting Complex Spatial Reasoning in MLLMs through Wide-Baseline Matching
Hao Zhong, Muzhi Zhu, Shenyan Zeng +8
cs.CVarXiv:2606.03577v12026Evaluating Large Language Models in Dynamic Clinical Decision-Making with Standardized Patient Cases
Cheng Liang, Pengcheng Qiu, Ya Zhang +3
cs.CLarXiv:2606.05112v12026Rethinking Continual Experience Internalization for Self-Evolving LLM Agents
Jingwen Chen, Wenkai Yang, Shengda Fan +7
cs.CLcs.LGarXiv:2606.04703v12026VideoKR: Towards Knowledge- and Reasoning-Intensive Video Understanding
Lin Fu, Zheyuan Yang, Yang Wang +3
cs.CVarXiv:2606.05259v12026WaveDiT: Distribution-Aware Wavelet Flow Matching for Efficient 3D Brain MRI Synthesis
Danilo Danese, Angela Lombardi, Giuseppe Fasano +2
cs.CVarXiv:2606.08670v22026Echo-Memory: A Controlled Study of Memory in Action World Models
Wayne King, Zeyue Xue, Yuxuan Bian +13
cs.CVcs.GRcs.LGarXiv:2606.09803v12026Precision Is Not Faithfulness: Coverage-Aware Evaluation of Grounded Generation with a Complete Oracle
Juan S. Santillana
cs.CLarXiv:2606.09376v22026AlloSpatial: Agentic Harness Framework for Spatial Reasoning in Foundation Models
Shouwei Ruan, Bin Wang, Zhenyu Wu +5
cs.AIarXiv:2606.08952v12026Hardening Agent Benchmarks with Adversarial Hacker-Fixer Loops
Ziqian Zhong, Ivgeni Segal, Ivan Bercovich +3
cs.CRcs.AIcs.LGarXiv:2606.08960v12026Routing Divergence Is Not Evidence of Behavioral Influence in Same-Weight MoE Self-Distillation
Cedric Caruzzo, Donggeun Yoo, Tae Soo Kim
cs.LGcs.AIcs.CLarXiv:2608.15787v12026Large Language Models Are Overconfident in Their Own Responses
Mario Sanz-Guerrero, Manuel Mager, Katharina von der Wense
cs.CLarXiv:2606.03437v12026Deep Embedded Multiplicative DMD for Algebra-Preserving Koopman Learning
Kelan Gray, Finlay Brown, Nicolas Boullé +1
cs.LGmath.DSmath.NAarXiv:2606.05131v12026A Cookbook of 3D Vision: Data, Learning Paradigms, and Application
Hongyang Du, Zongxia Li, Dawei Liu +8
cs.CVarXiv:2606.04291v12026Latent Reasoning with Normalizing Flows
Guancheng Tu, Xiangjun Fu, Suhao Yu +5
cs.CLcs.LGarXiv:2606.06447v12026LongLive-RAG: A General Retrieval-Augmented Framework for Long Video Generation
Qixin Hu, Shuai Yang, Wei Huang +2
cs.CVarXiv:2606.02553v12026WorldBench: A Challenging and Visually Diverse Multimodal Reasoning Benchmark
Yida Yin, Harish Krishnakumar, Chung Peng Lee +9
cs.CVarXiv:2606.06538v12026A Geometric Account of Activation Steering through Angle-Norm Decomposition
Georgii Aparin, Tatiana Gaintseva
cs.AIarXiv:2606.06735v22026Where, What, Why, and Importance: Structured Defect Grounding for Text-to-Image Feedback
Huaisong Zhang, Hao Yu, Yuxuan Zhang +7
cs.CVarXiv:2606.06113v22026Re-Centering Humans in LLM Personalization
Lechen Zhang, Jiarui Liu, Tal August
cs.CLcs.AIcs.HCarXiv:2606.06614v12026Synthesizing Feature Extractors: An Agentic Approach for Algorithm Selection
Hai Xia, Carlos Ansótegui, Stefan Szeider
cs.AIarXiv:2608.17170v12026Without journalists, there is no journalism: the social dimension of generative artificial intelligence in the media
Simón Peña-Fernández, Koldobika Meso-Ayerdi, Ainara Larrondo-Ureta +1
cs.CYcs.AIarXiv:2608.17017v12026APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies
Kechun Xu, Zhenjie Zhu, Anzhe Chen +2
cs.ROarXiv:2606.12366v12026ICA Lens: Interpreting Language Models Without Training Another Dictionary
Sida Liu, Feijiang Han
cs.LGcs.AIcs.CLarXiv:2606.11722v12026IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder
Yitong Chen, Zijie Diao, Junke Wang +5
cs.CVarXiv:2606.11096v12026Which Models Are Our Models Built On? Auditing Invisible Dependencies in Modern LLMs
Sanjay Adhikesaven, Haoxiang Sun, Sewon Min
cs.CLarXiv:2606.12385v12026Verifiable Environments Are LEGO Bricks: Recursive Composition for Reasoning Generalization
Hao Xiang, Qiaoyu Tang, Le Yu +8
cs.CLarXiv:2606.12373v12026Reason, Then Re-reason: Cross-view Revisiting Improves Spatial Reasoning
Chaofan Ma, Zhenjie Mao, Yuhuan Yang +5
cs.CVcs.AIarXiv:2606.11683v12026RedAct: Redacting Agent Capability Traces for Procedural Skill Protection
Shuwen Xu, Zhitao He, Yi R. Fung
cs.CRcs.CLarXiv:2606.10813v32026TuneJury: An Open Metric for Improving Music Generation Preference Alignment
Yonghyun Kim, Junwon Lee, Haiwen Xia +5
cs.SDcs.AIcs.LGarXiv:2606.17006v12026TokenPilot: Cache-Efficient Context Management for LLM Agents
Buqiang Xu, Zirui Xue, Dianmou Chen +12
cs.CLcs.AIcs.LGarXiv:2606.17016v12026RODS: Reward-Driven Online Data Synthesis for Multi-Turn Tool-Use Agents
Ruishan Fang, Siyuan Lu, Chenyi Zhuang +1
cs.AIarXiv:2606.19047v12026Native Active Perception as Reasoning for Omni-Modal Understanding
Zhenghao Xing, Ruiyang Xu, Yuxuan Wang +8
cs.CVcs.CLcs.SDarXiv:2606.19341v22026BRAID: Learning Equilibrium Maps in Interdependent Security Games via Weight-Tied Iterative Graph Neural Networks
Elnaz Nowrouzi, Zhiqun Zuo, Xueru Zhang +1
cs.GTcs.LGarXiv:2608.14856v12026How Does Reasoning Flow? Tracing Attention-Induced Information Flow for Targeted RL in LLMs
Zhichen Dong, Yang Li, Yuhan Sun +9
cs.LGcs.CLarXiv:2606.10646v12026Adaptive Multi-Resolution Procedural Knowledge Compression for Large Language Models
Changyue Wang, Weihang Su, Qingyao Ai +5
cs.CLarXiv:2606.12203v12026Time-Series Foundation Model Embeddings for Remaining Useful Life Estimation
Amir El-Ghoussani, Michele De Vita, Ronald Naumann +1
cs.LGcs.AIarXiv:2606.11990v32026InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning
Ziang Yan, Sheng Xia, Jiashuo Yu +10
cs.CVarXiv:2606.12195v12026Pythagoras-Prover: Advancing Efficient Formal Proving via Augmented Lean Formalisation
Joshua Ong Jun Leang, Zheng Zhao, Mihaela Cătălina Stoian +5
cs.AIarXiv:2606.12594v12026From AGI to ASI
Tim Genewein, Matija Franklin, Alexander Lerchner +11
cs.AIcs.CYcs.LGarXiv:2606.12683v12026LLM-Enabled NWDAF: A Step Toward AI-Native 6G Network Intelligence
Henok Daniel, Omar Alhussein, Cheng Li +2
cs.NIarXiv:2606.11877v12026Quickest Detection of Hallucination Onset: Delay Bounds and Learned CUSUM Statistics
Igor Itkin
cs.LGcs.AIcs.CLarXiv:2606.12476v32026HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers
Guozhen Zhang, Xuerui Qiu, Yutao Cui +11
cs.CVcs.AIarXiv:2606.13289v12026LabVLA: Grounding Vision-Language-Action Models in Scientific Laboratories
Baochang Ren, Xinjie Liu, Xi Chen +15
cs.CLcs.AIcs.LGarXiv:2606.13578v22026Dense Supervision, Sparse Updates: On the Sparsity and Geometry of On-Policy Distillation
Guo Yu, Wenlin Liu, Yulan Hu +3
cs.LGarXiv:2606.13657v32026ClinHallu: A Benchmark for Diagnosing Stage-Wise Hallucinations in Medical MLLM Reasoning
Sicheng Yang, Hangjie Yuan, Wenjun Zhang +5
cs.CVcs.AIcs.CLarXiv:2606.14697v12026IndustryBench-MIPU: Benchmarking Multi-Image Attribute Value Extraction for Industrial Products
Haonan Qi, Jin Cao, Yongqi Zhang +9
cs.CVarXiv:2606.14383v22026Retrieve, Don't Retrain: Extending Vision Language Action Models to New Tasks at Test Time
Jeongeun Park, Juhan Park, Taekyung Kim +3
cs.ROcs.AIarXiv:2606.15631v12026Human Universal Grasping
Kevin Yuanbo Wu, Tianxing Zhou, Isaac Tu +5
cs.ROcs.AIcs.CVarXiv:2606.17054v12026GD$^2$PO: Mitigating Multi-Reward Conflicts via Group-Dynamic reward-Decoupled Policy Optimization
Haotian Liu, Yihao Liu, Jingwei Ni +11
cs.LGarXiv:2606.16771v12026VibeThinker-3B: Exploring the Frontier of Verifiable Reasoning in Small Language Models
Sen Xu, Shixi Liu, Wei Wang +6
cs.AIcs.CLarXiv:2606.16140v12026ACE-Ego-0: Unifying Egocentric Human and Robotic Data for VLA Pretraining
Hao Li, Ganlong Zhao, Yufei Liu +8
cs.ROarXiv:2606.17200v12026Speaking the Language of Science: Toward a General-Purpose Generative Foundation Model for the Natural Sciences
Mingyang Li, Yurou Liu, Jieping Ye +3
cs.CLarXiv:2606.16905v12026LoopCoder-v2: Only Loop Once for Efficient Test-Time Computation Scaling
Jian Yang, Shawn Guo, Wei Zhang +16
cs.LGcs.AIarXiv:2606.18023v12026Looped World Models
Hongyuan Adam Lu, Z. L. Victor Wei, Qun Zhang +28
cs.LGcs.AIcs.CLarXiv:2606.18208v12026SAE Interventions are Unreliable: Post-Intervention Recovery of Suppressed Behavior
Mingyue Cui, Linghui Shen, Xingyi Yang
cs.LGcs.AIarXiv:2606.18322v12026