Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

59,701 to 59,760 of 61,374

  1. MuJoCo-Drones-Gym: A GPU-Accelerated Multi-Drone Simulator for Control and Reinforcement Learning

    Manan Tayal

    cs.ROarXiv:2606.08039v12026
  2. ARISE: An adaptive residual-informed stability ensemble for feature selection in small-sample biomedical omics

    Zardad Khan, Amjad Ali, Naz Gul +2

    stat.MLcs.LGarXiv:2608.14866v12026
  3. MCP-Persona: Benchmarking LLM Agents on Real-World Personal Applications via Environment Simulation

    Wenhao Wang, Peizhi Niu, Gongyi Zou +9

    cs.AIarXiv:2606.02470v12026
  4. VLMs are Good Teachers for Video Reasoning via Adaptive Test-Time Optimization

    Junhao Cheng, Liang Hou, Tianxiong Zhong +4

    cs.CVarXiv:2606.02564v32026
  5. SkillHarm: Lifecycle-Aware Skill-Based Attacks via Automated Construction

    Yuting Ning, Zhehao Zhang, Yash Kumar Lal +8

    cs.CLarXiv:2606.02540v12026
  6. Eliciting Complex Spatial Reasoning in MLLMs through Wide-Baseline Matching

    Hao Zhong, Muzhi Zhu, Shenyan Zeng +8

    cs.CVarXiv:2606.03577v12026
  7. Evaluating Large Language Models in Dynamic Clinical Decision-Making with Standardized Patient Cases

    Cheng Liang, Pengcheng Qiu, Ya Zhang +3

    cs.CLarXiv:2606.05112v12026
  8. Rethinking Continual Experience Internalization for Self-Evolving LLM Agents

    Jingwen Chen, Wenkai Yang, Shengda Fan +7

    cs.CLcs.LGarXiv:2606.04703v12026
  9. VideoKR: Towards Knowledge- and Reasoning-Intensive Video Understanding

    Lin Fu, Zheyuan Yang, Yang Wang +3

    cs.CVarXiv:2606.05259v12026
  10. WaveDiT: Distribution-Aware Wavelet Flow Matching for Efficient 3D Brain MRI Synthesis

    Danilo Danese, Angela Lombardi, Giuseppe Fasano +2

    cs.CVarXiv:2606.08670v22026
  11. Echo-Memory: A Controlled Study of Memory in Action World Models

    Wayne King, Zeyue Xue, Yuxuan Bian +13

    cs.CVcs.GRcs.LGarXiv:2606.09803v12026
  12. Precision Is Not Faithfulness: Coverage-Aware Evaluation of Grounded Generation with a Complete Oracle

    Juan S. Santillana

    cs.CLarXiv:2606.09376v22026
  13. AlloSpatial: Agentic Harness Framework for Spatial Reasoning in Foundation Models

    Shouwei Ruan, Bin Wang, Zhenyu Wu +5

    cs.AIarXiv:2606.08952v12026
  14. Hardening Agent Benchmarks with Adversarial Hacker-Fixer Loops

    Ziqian Zhong, Ivgeni Segal, Ivan Bercovich +3

    cs.CRcs.AIcs.LGarXiv:2606.08960v12026
  15. Routing Divergence Is Not Evidence of Behavioral Influence in Same-Weight MoE Self-Distillation

    Cedric Caruzzo, Donggeun Yoo, Tae Soo Kim

    cs.LGcs.AIcs.CLarXiv:2608.15787v12026
  16. Large Language Models Are Overconfident in Their Own Responses

    Mario Sanz-Guerrero, Manuel Mager, Katharina von der Wense

    cs.CLarXiv:2606.03437v12026
  17. Deep Embedded Multiplicative DMD for Algebra-Preserving Koopman Learning

    Kelan Gray, Finlay Brown, Nicolas Boullé +1

    cs.LGmath.DSmath.NAarXiv:2606.05131v12026
  18. A Cookbook of 3D Vision: Data, Learning Paradigms, and Application

    Hongyang Du, Zongxia Li, Dawei Liu +8

    cs.CVarXiv:2606.04291v12026
  19. Latent Reasoning with Normalizing Flows

    Guancheng Tu, Xiangjun Fu, Suhao Yu +5

    cs.CLcs.LGarXiv:2606.06447v12026
  20. LongLive-RAG: A General Retrieval-Augmented Framework for Long Video Generation

    Qixin Hu, Shuai Yang, Wei Huang +2

    cs.CVarXiv:2606.02553v12026
  21. WorldBench: A Challenging and Visually Diverse Multimodal Reasoning Benchmark

    Yida Yin, Harish Krishnakumar, Chung Peng Lee +9

    cs.CVarXiv:2606.06538v12026
  22. A Geometric Account of Activation Steering through Angle-Norm Decomposition

    Georgii Aparin, Tatiana Gaintseva

    cs.AIarXiv:2606.06735v22026
  23. Where, What, Why, and Importance: Structured Defect Grounding for Text-to-Image Feedback

    Huaisong Zhang, Hao Yu, Yuxuan Zhang +7

    cs.CVarXiv:2606.06113v22026
  24. Re-Centering Humans in LLM Personalization

    Lechen Zhang, Jiarui Liu, Tal August

    cs.CLcs.AIcs.HCarXiv:2606.06614v12026
  25. Synthesizing Feature Extractors: An Agentic Approach for Algorithm Selection

    Hai Xia, Carlos Ansótegui, Stefan Szeider

    cs.AIarXiv:2608.17170v12026
  26. Without journalists, there is no journalism: the social dimension of generative artificial intelligence in the media

    Simón Peña-Fernández, Koldobika Meso-Ayerdi, Ainara Larrondo-Ureta +1

    cs.CYcs.AIarXiv:2608.17017v12026
  27. APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies

    Kechun Xu, Zhenjie Zhu, Anzhe Chen +2

    cs.ROarXiv:2606.12366v12026
  28. ICA Lens: Interpreting Language Models Without Training Another Dictionary

    Sida Liu, Feijiang Han

    cs.LGcs.AIcs.CLarXiv:2606.11722v12026
  29. IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder

    Yitong Chen, Zijie Diao, Junke Wang +5

    cs.CVarXiv:2606.11096v12026
  30. Which Models Are Our Models Built On? Auditing Invisible Dependencies in Modern LLMs

    Sanjay Adhikesaven, Haoxiang Sun, Sewon Min

    cs.CLarXiv:2606.12385v12026
  31. Verifiable Environments Are LEGO Bricks: Recursive Composition for Reasoning Generalization

    Hao Xiang, Qiaoyu Tang, Le Yu +8

    cs.CLarXiv:2606.12373v12026
  32. Reason, Then Re-reason: Cross-view Revisiting Improves Spatial Reasoning

    Chaofan Ma, Zhenjie Mao, Yuhuan Yang +5

    cs.CVcs.AIarXiv:2606.11683v12026
  33. RedAct: Redacting Agent Capability Traces for Procedural Skill Protection

    Shuwen Xu, Zhitao He, Yi R. Fung

    cs.CRcs.CLarXiv:2606.10813v32026
  34. TuneJury: An Open Metric for Improving Music Generation Preference Alignment

    Yonghyun Kim, Junwon Lee, Haiwen Xia +5

    cs.SDcs.AIcs.LGarXiv:2606.17006v12026
  35. TokenPilot: Cache-Efficient Context Management for LLM Agents

    Buqiang Xu, Zirui Xue, Dianmou Chen +12

    cs.CLcs.AIcs.LGarXiv:2606.17016v12026
  36. RODS: Reward-Driven Online Data Synthesis for Multi-Turn Tool-Use Agents

    Ruishan Fang, Siyuan Lu, Chenyi Zhuang +1

    cs.AIarXiv:2606.19047v12026
  37. Native Active Perception as Reasoning for Omni-Modal Understanding

    Zhenghao Xing, Ruiyang Xu, Yuxuan Wang +8

    cs.CVcs.CLcs.SDarXiv:2606.19341v22026
  38. BRAID: Learning Equilibrium Maps in Interdependent Security Games via Weight-Tied Iterative Graph Neural Networks

    Elnaz Nowrouzi, Zhiqun Zuo, Xueru Zhang +1

    cs.GTcs.LGarXiv:2608.14856v12026
  39. How Does Reasoning Flow? Tracing Attention-Induced Information Flow for Targeted RL in LLMs

    Zhichen Dong, Yang Li, Yuhan Sun +9

    cs.LGcs.CLarXiv:2606.10646v12026
  40. Adaptive Multi-Resolution Procedural Knowledge Compression for Large Language Models

    Changyue Wang, Weihang Su, Qingyao Ai +5

    cs.CLarXiv:2606.12203v12026
  41. Time-Series Foundation Model Embeddings for Remaining Useful Life Estimation

    Amir El-Ghoussani, Michele De Vita, Ronald Naumann +1

    cs.LGcs.AIarXiv:2606.11990v32026
  42. InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning

    Ziang Yan, Sheng Xia, Jiashuo Yu +10

    cs.CVarXiv:2606.12195v12026
  43. Pythagoras-Prover: Advancing Efficient Formal Proving via Augmented Lean Formalisation

    Joshua Ong Jun Leang, Zheng Zhao, Mihaela Cătălina Stoian +5

    cs.AIarXiv:2606.12594v12026
  44. From AGI to ASI

    Tim Genewein, Matija Franklin, Alexander Lerchner +11

    cs.AIcs.CYcs.LGarXiv:2606.12683v12026
  45. LLM-Enabled NWDAF: A Step Toward AI-Native 6G Network Intelligence

    Henok Daniel, Omar Alhussein, Cheng Li +2

    cs.NIarXiv:2606.11877v12026
  46. Quickest Detection of Hallucination Onset: Delay Bounds and Learned CUSUM Statistics

    Igor Itkin

    cs.LGcs.AIcs.CLarXiv:2606.12476v32026
  47. HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers

    Guozhen Zhang, Xuerui Qiu, Yutao Cui +11

    cs.CVcs.AIarXiv:2606.13289v12026
  48. LabVLA: Grounding Vision-Language-Action Models in Scientific Laboratories

    Baochang Ren, Xinjie Liu, Xi Chen +15

    cs.CLcs.AIcs.LGarXiv:2606.13578v22026
  49. Dense Supervision, Sparse Updates: On the Sparsity and Geometry of On-Policy Distillation

    Guo Yu, Wenlin Liu, Yulan Hu +3

    cs.LGarXiv:2606.13657v32026
  50. ClinHallu: A Benchmark for Diagnosing Stage-Wise Hallucinations in Medical MLLM Reasoning

    Sicheng Yang, Hangjie Yuan, Wenjun Zhang +5

    cs.CVcs.AIcs.CLarXiv:2606.14697v12026
  51. IndustryBench-MIPU: Benchmarking Multi-Image Attribute Value Extraction for Industrial Products

    Haonan Qi, Jin Cao, Yongqi Zhang +9

    cs.CVarXiv:2606.14383v22026
  52. Retrieve, Don't Retrain: Extending Vision Language Action Models to New Tasks at Test Time

    Jeongeun Park, Juhan Park, Taekyung Kim +3

    cs.ROcs.AIarXiv:2606.15631v12026
  53. Human Universal Grasping

    Kevin Yuanbo Wu, Tianxing Zhou, Isaac Tu +5

    cs.ROcs.AIcs.CVarXiv:2606.17054v12026
  54. GD$^2$PO: Mitigating Multi-Reward Conflicts via Group-Dynamic reward-Decoupled Policy Optimization

    Haotian Liu, Yihao Liu, Jingwei Ni +11

    cs.LGarXiv:2606.16771v12026
  55. VibeThinker-3B: Exploring the Frontier of Verifiable Reasoning in Small Language Models

    Sen Xu, Shixi Liu, Wei Wang +6

    cs.AIcs.CLarXiv:2606.16140v12026
  56. ACE-Ego-0: Unifying Egocentric Human and Robotic Data for VLA Pretraining

    Hao Li, Ganlong Zhao, Yufei Liu +8

    cs.ROarXiv:2606.17200v12026
  57. Speaking the Language of Science: Toward a General-Purpose Generative Foundation Model for the Natural Sciences

    Mingyang Li, Yurou Liu, Jieping Ye +3

    cs.CLarXiv:2606.16905v12026
  58. LoopCoder-v2: Only Loop Once for Efficient Test-Time Computation Scaling

    Jian Yang, Shawn Guo, Wei Zhang +16

    cs.LGcs.AIarXiv:2606.18023v12026
  59. Looped World Models

    Hongyuan Adam Lu, Z. L. Victor Wei, Qun Zhang +28

    cs.LGcs.AIcs.CLarXiv:2606.18208v12026
  60. SAE Interventions are Unreliable: Post-Intervention Recovery of Suppressed Behavior

    Mingyue Cui, Linghui Shen, Xingyi Yang

    cs.LGcs.AIarXiv:2606.18322v12026