Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

60,841 to 60,900 of 61,255

  1. ReDesign: Recovering Editable Design Structures from Images via Agentic Decomposition

    Jooyeol Yun, Jintae Park, Hyesu Lim +3

    cs.CVarXiv:2607.25565v12026
  2. Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making

    Amirhosein Ghasemabadi, Ruichen Chen, Bahador Rashidi +1

    cs.CLarXiv:2607.14277v12026
  3. Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model

    Senqiao Yang, Kaichen Zhang, Zhaoyang Jia +20

    cs.CVcs.CLarXiv:2607.24904v12026
  4. Trajectory-aware Cross-view Geo-localization with Sequential Observations

    Tianyi Gao, Jiayu Lin, Danielle Beaulieu +1

    cs.CVarXiv:2607.15491v12026
  5. Nonuniformity Principle in Human-AI Coworking

    An Luo, Jie Ding

    cs.AIstat.MEarXiv:2607.16530v12026
  6. VideoCoCo: Code-as-CoT for Physically-Consistent Video Generation via an Agentic Dual-Engine System

    Haodong Li, Tianfei Ren, Xiaoxiao Ma +25

    cs.CVcs.AIarXiv:2607.27380v22026
  7. AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis

    Bing Yan, Gregory Wolfe, Stefano Martiniani +1

    cs.CLcs.AIcs.IRarXiv:2607.28618v12026
  8. FVAttn: Adaptive Sparse Attention with Runtime Load Balancing for Video Generation

    Hao Liu, Chenghuan Huang, Ye Huang +7

    cs.CVarXiv:2607.16190v12026
  9. An Exam for Active Observers

    Jiarui Zhang, Muzi Tao, Shangshang Wang +3

    cs.CVcs.AIcs.CLarXiv:2607.16165v12026
  10. Interactive Training 2: Auditable Control Plane for Live Model Training

    Wentao Zhang, Xuanhe Pan, Han Zhou +2

    cs.LGarXiv:2607.18314v12026
  11. JoyNexus: Service-Oriented Multi-Tenant Post-Training for VLA Models

    Haoran Sun, Wentao Zhang, Junyang Hua +18

    cs.DCcs.AIcs.SEarXiv:2607.16074v12026
  12. Metis: Memory Foundation Model

    Zeyu Zhang, Ziliang Guo, Yihang Sun +14

    cs.CLcs.LGarXiv:2607.26760v22026
  13. Open-AoE: An Open Egocentric Manipulation Dataset and Toolchain for Embodied Learning

    Zishuo Li, Bowen Yang, Changtao Miao +29

    cs.ROcs.CVarXiv:2607.14183v22026
  14. CADENCE: Closing the Reasoning Gap via Coverage-Adaptive On-Policy Distillation

    Satyam Kumar, Saurabh Jha

    cs.LGcs.AIarXiv:2607.16955v12026
  15. In the Driver's Seat: A Multi-Company Study on the Reality of Autonomous Driving System Testing

    Qunying Song, Yuan Gao, Johannes Betz +3

    cs.SEarXiv:2607.15820v12026
  16. TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM

    Hengyi Xie, Chenfei Yao, Xianjin Wu +7

    cs.CVcs.ROarXiv:2607.27205v12026
  17. DataFlow-Harness: A Grounded Code-Agent Platform for Constructing Editable LLM Data Pipelines

    Runming He, Zhen Hao Wong, Hao Liang +4

    cs.SEcs.AIarXiv:2607.16617v22026
  18. Can Multimodal Large Language Models Understand OCT?

    Baochen Fu, Wenzhi Deng, Baihao Jin +5

    cs.CVcs.CLarXiv:2607.16609v12026
  19. Dataset Distillation by Influence Matching

    Haoru Tan, Wang Wang, Sitong Wu +5

    cs.CVarXiv:2607.16859v12026
  20. Multi-Turn On-Policy Distillation with Prefix Replay

    Baohao Liao, Hanze Dong, Christof Monz +3

    cs.LGcs.AIcs.CLarXiv:2607.04763v32026
  21. Optimal Scheduling of Road Maintenance Jobs Considering Impact on Traffic Flows

    Charitha Nandepu, Lohitha Kalepu, Gabriele Ciavarella +1

    eess.SYcs.AIarXiv:2608.14491v12026
  22. Reflex: Enabling Fast and Predictive Vision-Language-Action Models for Reaction-Critical Manipulation

    Yuxuan Chen, Wanruo Zhang, Xiao Li

    cs.ROcs.AIarXiv:2608.14379v12026
  23. MedClaw: Heuristic Agent Harness for Long-Horizon Surgical Video Reasoning

    Yingying Fan, Penghui Du, Leyan Zhu +10

    cs.CVcs.AIarXiv:2608.14015v12026
  24. Agentic Transaction: Towards ACID-Compliant Agent Systems

    Zhaoyan Sun, Xiaoxiao Wang, Guoliang Li

    cs.DBcs.AIcs.CLarXiv:2608.13900v12026
  25. Engineering Signals of Human-AI Collaboration in the Agentic Coding Era: A Longitudinal Analysis of 33,228 Pull Requests from vLLM and SGLang with Implications for Biomedical AI Agents and Bioinformatics Pipeline Developmen

    Jiada Li, Xuesong Ye, Olamide Olowoniyi

    cs.SEcs.AIcs.ETarXiv:2608.13884v12026
  26. Split the Labor: Separating Evidence Interpretation from Decision Aggregation

    Zhelun Wu

    cs.AIcs.CLcs.LGarXiv:2608.14509v12026
  27. LLMs Don't Pay for the Jump

    Paras Balani, Subhrakanta Panda

    cs.AIcs.CLarXiv:2608.14397v12026
  28. Tripwire: Triggering Aligned Refusal via Statistically Certified Safety Neurons

    Wei Zhao, Zhe Li, Peixin Zhang +1

    cs.AIarXiv:2608.14392v12026
  29. Attributing Preprocessing Invariance in Spectral Foundation Models

    Dongjun Wei, Hongyi Wu, Yinuo Zou

    cs.AIcs.CEcs.LGarXiv:2608.14227v12026
  30. Towards Efficient Multimodal and Multilingual Opinion Extraction for STI: A QLoRA-Based Fine-Tuning Approach

    Sheng Hong, Xuanqi Wang, Jiacheng Wang +1

    cs.AIarXiv:2608.14152v12026
  31. Mandato: Protocol-Level Enforcement of Digitally Signed Mandates on AI Agent Actions with Cryptographically Chained Audit Trails

    Giovanni Racioppi

    cs.AIarXiv:2608.14074v12026
  32. Implementing Computational Law in Wolfram Language for the Governance of Artificial Intelligence

    James K. Wiles

    cs.AIcs.CYcs.LOarXiv:2608.13958v12026
  33. AI Research Preference Models

    Thomas Simon Foster, Bassel Al Omari, Tingchen Fu +30

    cs.AIarXiv:2608.13940v12026
  34. Never the Number: Structural Abstention for AI Systems Whose Answers Are Consumed as Fact

    Zhelun, Wu

    cs.AIcs.CLcs.DBarXiv:2608.13926v12026
  35. Reading Between The Lines: Modeling and Evaluating Behavioral Realism in Legal Simulation

    Divya Vetticaden, Arya Gupta, Julian Nyarko +1

    cs.CYcs.AIcs.CLarXiv:2608.13712v12026
  36. MemoryLake on MemoryArena: A Matched Study of Agent Memory Backends

    Chaoqun Zhan, Qiang Zhou, Guannan Li +2

    cs.AIarXiv:2608.13883v12026
  37. Secret-Stego Dissimilarity as a Design Axis: Invertible Coverless Image Steganography with Diffusion Models

    Hongxin Xu, Jianping Mei, Can Wang +1

    eess.IVcs.AIcs.CRarXiv:2608.13597v12026
  38. Active Perception for Embodied Disambiguation

    Yiwei Liu, Luwei Yang

    cs.AIcs.ROarXiv:2608.13605v12026
  39. What AI Red-Team Evaluations Can and Cannot Prove

    Bandana Kaur

    cs.AIcs.CRarXiv:2607.21735v22026
  40. Seeing or Knowing? Visual Context Sensitivity in Multimodal Large Language Models

    Jiaang Li, Chengzu Li, Zhaochong An +4

    cs.CVarXiv:2607.26326v12026
  41. Evaluating Agentic Learning Harness Capabilities Without Labels via the Scaling Hypothesis

    Aryan Luthra, Kshitij Jain, Siddharth Arya +2

    cs.AIcs.CRcs.LGarXiv:2608.13608v12026
  42. Are the Financial Reasoning from LLMs Credible? A Real World Test over Long-Horizon Statements

    Xinke Tong, Xuanming Zhang, Tianyi Tang +10

    cs.CLarXiv:2607.28661v12026
  43. When Activation Oracles Learn Not to Read: Concept-Specific Blind Spots in Fine-Tuned Oracles

    Tobias Bersia, Tatiana Gaintseva

    cs.CLcs.AIarXiv:2607.23379v12026
  44. Compute Globally, Materialize Locally: The Memory Contract of Sparse Event-KV

    Zefeng Cai, Zerui Cai

    cs.AIarXiv:2607.23693v12026
  45. The Query Knows What to Forget: A Second Erase Direction for Linear Attention

    Dhruman Gupta, Aritra Das, Debayan Gupta

    cs.LGarXiv:2608.13668v12026
  46. Helping Music Co-Creation Agents 'Listen' Well: Hierarchical Self-Supervised World Models for Understanding and Generation

    Scott H. Hawley

    cs.SDcs.LGeess.ASarXiv:2608.04378v12026
  47. Exploring ESC Winners with Nested Diagrams

    Anurag Sharma, Marcel Nöhre, Gerd Stumme

    cs.AIarXiv:2608.13630v12026
  48. Acoustic UAV Detection in Battlefield Scenarios: Handling Noise, Domain Shift, and Weak Labels

    Vadym Vilhurin, Volodymyr Sydorskyi, Andrii Shevtsov

    cs.SDcs.AIcs.CVarXiv:2608.14287v12026
  49. FreeBalance: Pre-Routing Online Moe Load Balancing via Residual Workload Prediction

    Pengfei Chen, Yize Wu, Shouxu Kuang +2

    cs.AIcs.LGarXiv:2608.14205v12026
  50. Removing Temporal Note Redundancy Improves Multimodal Reinforcement Learning for Medicine

    Chenran Weng, Joo Seung Lee, Malini Mahendra +1

    cs.AIcs.LGarXiv:2608.14157v12026
  51. Traj-LeWM: Path-Aware World-Model Planning via Latent Trajectory Cost

    Xiaodi Huang, Ziyi Ding, Jingtian Wan +6

    cs.AIarXiv:2608.14125v12026
  52. Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures

    Harsh Raj, Vipul Gupta, Anas Mahmoud +4

    cs.AIarXiv:2607.28802v12026
  53. Push-Wiper: Toward General-Purpose Robotic Cleaning across Varied Stains and Surfaces with Segmented Pushing Trajectories

    Renhao Lu, Mingxin Wang, Chenyang Cao +5

    cs.ROarXiv:2608.00730v12026
  54. Wnuan: Staged Post-Training for Question Answering over Proprietary Enterprise Knowledge

    Xiaofeng Shi, Xiaosong Qiu, Wenxin Ma +6

    cs.AIarXiv:2608.01862v12026
  55. DRIFT: Derailing Denoising Trajectories of Flow-Matching VLAs with Adversarial Patch Attack

    Hoseong Tae, Jong-Seok Lee

    cs.CVcs.LGarXiv:2608.03207v12026
  56. Simulation-Aware In-Context Policy Improvement for LLM-Aided Analog Layout Refinement

    Bingyang Liu, Ziming Wei, Xiaohan Gao +1

    cs.AIcs.ROarXiv:2608.13767v12026
  57. Adjacency-Based Spectral Proxy Control of Mobile Communication Agents

    Mariana del Castillo, Federico Larroca

    cs.ROcs.LGcs.MAarXiv:2608.13616v12026
  58. Omega-S: A Functional Resilience Index for LLM Fine-Tuning

    Alberto Acedo

    cs.LGcs.NEq-bio.MNarXiv:2608.03887v12026
  59. BM25-Augmented Many-Shot Translation for Low-Resource North-Eastern Indian Languages

    Aashish Dhawan, Christopher Driggers-Ellis, Dzmitry Kasinets +2

    cs.CLarXiv:2608.13722v12026
  60. Agent-Orchestration in Autonomous Chip Design

    Linyang Li

    cs.AIarXiv:2608.14035v12026