Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

60,781 to 60,840 of 61,181

  1. In the Driver's Seat: A Multi-Company Study on the Reality of Autonomous Driving System Testing

    Qunying Song, Yuan Gao, Johannes Betz +3

    cs.SEarXiv:2607.15820v12026
  2. TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM

    Hengyi Xie, Chenfei Yao, Xianjin Wu +7

    cs.CVcs.ROarXiv:2607.27205v12026
  3. DataFlow-Harness: A Grounded Code-Agent Platform for Constructing Editable LLM Data Pipelines

    Runming He, Zhen Hao Wong, Hao Liang +4

    cs.SEcs.AIarXiv:2607.16617v22026
  4. Can Multimodal Large Language Models Understand OCT?

    Baochen Fu, Wenzhi Deng, Baihao Jin +5

    cs.CVcs.CLarXiv:2607.16609v12026
  5. Dataset Distillation by Influence Matching

    Haoru Tan, Wang Wang, Sitong Wu +5

    cs.CVarXiv:2607.16859v12026
  6. Multi-Turn On-Policy Distillation with Prefix Replay

    Baohao Liao, Hanze Dong, Christof Monz +3

    cs.LGcs.AIcs.CLarXiv:2607.04763v32026
  7. Optimal Scheduling of Road Maintenance Jobs Considering Impact on Traffic Flows

    Charitha Nandepu, Lohitha Kalepu, Gabriele Ciavarella +1

    eess.SYcs.AIarXiv:2608.14491v12026
  8. Reflex: Enabling Fast and Predictive Vision-Language-Action Models for Reaction-Critical Manipulation

    Yuxuan Chen, Wanruo Zhang, Xiao Li

    cs.ROcs.AIarXiv:2608.14379v12026
  9. MedClaw: Heuristic Agent Harness for Long-Horizon Surgical Video Reasoning

    Yingying Fan, Penghui Du, Leyan Zhu +10

    cs.CVcs.AIarXiv:2608.14015v12026
  10. Agentic Transaction: Towards ACID-Compliant Agent Systems

    Zhaoyan Sun, Xiaoxiao Wang, Guoliang Li

    cs.DBcs.AIcs.CLarXiv:2608.13900v12026
  11. Engineering Signals of Human-AI Collaboration in the Agentic Coding Era: A Longitudinal Analysis of 33,228 Pull Requests from vLLM and SGLang with Implications for Biomedical AI Agents and Bioinformatics Pipeline Developmen

    Jiada Li, Xuesong Ye, Olamide Olowoniyi

    cs.SEcs.AIcs.ETarXiv:2608.13884v12026
  12. Split the Labor: Separating Evidence Interpretation from Decision Aggregation

    Zhelun Wu

    cs.AIcs.CLcs.LGarXiv:2608.14509v12026
  13. LLMs Don't Pay for the Jump

    Paras Balani, Subhrakanta Panda

    cs.AIcs.CLarXiv:2608.14397v12026
  14. Tripwire: Triggering Aligned Refusal via Statistically Certified Safety Neurons

    Wei Zhao, Zhe Li, Peixin Zhang +1

    cs.AIarXiv:2608.14392v12026
  15. Attributing Preprocessing Invariance in Spectral Foundation Models

    Dongjun Wei, Hongyi Wu, Yinuo Zou

    cs.AIcs.CEcs.LGarXiv:2608.14227v12026
  16. Towards Efficient Multimodal and Multilingual Opinion Extraction for STI: A QLoRA-Based Fine-Tuning Approach

    Sheng Hong, Xuanqi Wang, Jiacheng Wang +1

    cs.AIarXiv:2608.14152v12026
  17. Mandato: Protocol-Level Enforcement of Digitally Signed Mandates on AI Agent Actions with Cryptographically Chained Audit Trails

    Giovanni Racioppi

    cs.AIarXiv:2608.14074v12026
  18. Implementing Computational Law in Wolfram Language for the Governance of Artificial Intelligence

    James K. Wiles

    cs.AIcs.CYcs.LOarXiv:2608.13958v12026
  19. AI Research Preference Models

    Thomas Simon Foster, Bassel Al Omari, Tingchen Fu +30

    cs.AIarXiv:2608.13940v12026
  20. Never the Number: Structural Abstention for AI Systems Whose Answers Are Consumed as Fact

    Zhelun, Wu

    cs.AIcs.CLcs.DBarXiv:2608.13926v12026
  21. Reading Between The Lines: Modeling and Evaluating Behavioral Realism in Legal Simulation

    Divya Vetticaden, Arya Gupta, Julian Nyarko +1

    cs.CYcs.AIcs.CLarXiv:2608.13712v12026
  22. MemoryLake on MemoryArena: A Matched Study of Agent Memory Backends

    Chaoqun Zhan, Qiang Zhou, Guannan Li +2

    cs.AIarXiv:2608.13883v12026
  23. Secret-Stego Dissimilarity as a Design Axis: Invertible Coverless Image Steganography with Diffusion Models

    Hongxin Xu, Jianping Mei, Can Wang +1

    eess.IVcs.AIcs.CRarXiv:2608.13597v12026
  24. Active Perception for Embodied Disambiguation

    Yiwei Liu, Luwei Yang

    cs.AIcs.ROarXiv:2608.13605v12026
  25. What AI Red-Team Evaluations Can and Cannot Prove

    Bandana Kaur

    cs.AIcs.CRarXiv:2607.21735v22026
  26. Seeing or Knowing? Visual Context Sensitivity in Multimodal Large Language Models

    Jiaang Li, Chengzu Li, Zhaochong An +4

    cs.CVarXiv:2607.26326v12026
  27. Evaluating Agentic Learning Harness Capabilities Without Labels via the Scaling Hypothesis

    Aryan Luthra, Kshitij Jain, Siddharth Arya +2

    cs.AIcs.CRcs.LGarXiv:2608.13608v12026
  28. Are the Financial Reasoning from LLMs Credible? A Real World Test over Long-Horizon Statements

    Xinke Tong, Xuanming Zhang, Tianyi Tang +10

    cs.CLarXiv:2607.28661v12026
  29. When Activation Oracles Learn Not to Read: Concept-Specific Blind Spots in Fine-Tuned Oracles

    Tobias Bersia, Tatiana Gaintseva

    cs.CLcs.AIarXiv:2607.23379v12026
  30. Compute Globally, Materialize Locally: The Memory Contract of Sparse Event-KV

    Zefeng Cai, Zerui Cai

    cs.AIarXiv:2607.23693v12026
  31. The Query Knows What to Forget: A Second Erase Direction for Linear Attention

    Dhruman Gupta, Aritra Das, Debayan Gupta

    cs.LGarXiv:2608.13668v12026
  32. Helping Music Co-Creation Agents 'Listen' Well: Hierarchical Self-Supervised World Models for Understanding and Generation

    Scott H. Hawley

    cs.SDcs.LGeess.ASarXiv:2608.04378v12026
  33. Exploring ESC Winners with Nested Diagrams

    Anurag Sharma, Marcel Nöhre, Gerd Stumme

    cs.AIarXiv:2608.13630v12026
  34. Acoustic UAV Detection in Battlefield Scenarios: Handling Noise, Domain Shift, and Weak Labels

    Vadym Vilhurin, Volodymyr Sydorskyi, Andrii Shevtsov

    cs.SDcs.AIcs.CVarXiv:2608.14287v12026
  35. FreeBalance: Pre-Routing Online Moe Load Balancing via Residual Workload Prediction

    Pengfei Chen, Yize Wu, Shouxu Kuang +2

    cs.AIcs.LGarXiv:2608.14205v12026
  36. Removing Temporal Note Redundancy Improves Multimodal Reinforcement Learning for Medicine

    Chenran Weng, Joo Seung Lee, Malini Mahendra +1

    cs.AIcs.LGarXiv:2608.14157v12026
  37. Traj-LeWM: Path-Aware World-Model Planning via Latent Trajectory Cost

    Xiaodi Huang, Ziyi Ding, Jingtian Wan +6

    cs.AIarXiv:2608.14125v12026
  38. Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures

    Harsh Raj, Vipul Gupta, Anas Mahmoud +4

    cs.AIarXiv:2607.28802v12026
  39. Push-Wiper: Toward General-Purpose Robotic Cleaning across Varied Stains and Surfaces with Segmented Pushing Trajectories

    Renhao Lu, Mingxin Wang, Chenyang Cao +5

    cs.ROarXiv:2608.00730v12026
  40. Wnuan: Staged Post-Training for Question Answering over Proprietary Enterprise Knowledge

    Xiaofeng Shi, Xiaosong Qiu, Wenxin Ma +6

    cs.AIarXiv:2608.01862v12026
  41. DRIFT: Derailing Denoising Trajectories of Flow-Matching VLAs with Adversarial Patch Attack

    Hoseong Tae, Jong-Seok Lee

    cs.CVcs.LGarXiv:2608.03207v12026
  42. Simulation-Aware In-Context Policy Improvement for LLM-Aided Analog Layout Refinement

    Bingyang Liu, Ziming Wei, Xiaohan Gao +1

    cs.AIcs.ROarXiv:2608.13767v12026
  43. Adjacency-Based Spectral Proxy Control of Mobile Communication Agents

    Mariana del Castillo, Federico Larroca

    cs.ROcs.LGcs.MAarXiv:2608.13616v12026
  44. Omega-S: A Functional Resilience Index for LLM Fine-Tuning

    Alberto Acedo

    cs.LGcs.NEq-bio.MNarXiv:2608.03887v12026
  45. BM25-Augmented Many-Shot Translation for Low-Resource North-Eastern Indian Languages

    Aashish Dhawan, Christopher Driggers-Ellis, Dzmitry Kasinets +2

    cs.CLarXiv:2608.13722v12026
  46. Agent-Orchestration in Autonomous Chip Design

    Linyang Li

    cs.AIarXiv:2608.14035v12026
  47. Articulated Object Reconstruction from Rest-State Observation

    Daeun Lee, Jaeah Lee, Woosung Kim +2

    cs.CVcs.ROarXiv:2607.27749v12026
  48. FactorJEPA: Factorizing Monolithic Futures into Layout-Agent-Interaction Channels for Crowded and Chaotic Global South Urban Worlds

    Kapil Wanaskar, Gaytri Jena, Aman Chadha +3

    cs.AIcs.CVcs.LGarXiv:2608.01049v12026
  49. Capek 0.5: An Execution-Centric Vision-Language Model for Embodied Intelligence

    Ying Chen, Weizhen Li, Zhe Hu +7

    cs.AIarXiv:2608.06756v12026
  50. GEOID-Flood: A Large-Scale Multi-Modal Benchmark Dataset for Flood Segmentation

    Gaetano Chiriaco, Luca Barco, Andrea Bragagnolo +2

    cs.CVarXiv:2608.02315v12026
  51. Can MLLMs Decode the Creative Leap? Introducing C4 for Cross-Concept Understanding

    Ming Wang, Yuqing Zhang, Tingna Xie +5

    cs.AIcs.CLcs.MMarXiv:2608.06501v12026
  52. Zero-Mem: Zero-Token Memory Operations for LLM Agents

    Yilin Xiao, Zhehan Zhu, Yujing Zhang +8

    cs.CLarXiv:2607.29377v12026
  53. To Add Is Machine, To Delete Is Human: Measuring and Mitigating Deletion Avoidance in LLM Code Editing

    Amir M. Ebrahimi, Mohammed Mehedi Hasan, Aaditya Bhatia +2

    cs.SEcs.AIcs.LGarXiv:2607.28887v12026
  54. DuplexGen: Adaptive Synthesis of Human-AI Turn-Taking Dialogues

    Takyoung Kim, Kang-wook Kim, Sang Hoon Woo +3

    cs.CLarXiv:2607.26178v12026
  55. DreamTraj: Generating 6-DoF Object Trajectories by Reading Unrendered Video Diffusion Latents

    Tongsheng Ding, Zhen Luo, Yixuan Yang +4

    cs.CVarXiv:2608.00486v12026
  56. SIGNPOST-Bench: Benchmarking Text-Vision Conflict Resolution in Multimodal Large Language Models

    Sirun Li, Minghao Liu, Ling Dai +4

    cs.CVcs.CLarXiv:2608.04244v12026
  57. Towards Interpretable Foundation Models for Retinal Fundus Images

    Samuel Ofosu Mensah, Camila Roa, Kerol Djoumessi +1

    cs.CVcs.LGstat.COarXiv:2603.18846v42026
  58. When Lexical Change Misleads: Rethinking Dynamic Topic Model Evaluation with Traditional and LLM-Based Metrics

    Charu Karakkaparambil James

    cs.CLarXiv:2608.13835v12026
  59. Generation as Auxiliary Supervision: Enhancing Visual Understanding at Zero Inference Overhead via Decoupled Embedding Prediction

    Zhongbin Guo, Jiahao Xie, Dongling Xiao +5

    cs.CVarXiv:2608.12209v12026
  60. Poplar: A Scalable Pipeline for Human-Centric Image Dataset Synthesis

    Zhishan Zou

    cs.CVarXiv:2608.00440v12026