Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

18,961 to 19,020 of 61,112

  1. On Memory Construction and Retrieval for Personalized Conversational Agents

    Zhuoshi Pan, Qianhui Wu, Huiqiang Jiang +8

    cs.CLcs.AIarXiv:2502.05589v32025
  2. Streaming Long Video Understanding with Large Language Models

    Rui Qian, Xiaoyi Dong, Pan Zhang +4

    cs.CVarXiv:2405.16009v12024
  3. Improving Object Localization with Fitness NMS and Bounded IoU Loss

    Lachlan Tychsen-Smith, Lars Petersson

    cs.CVarXiv:1711.00164v32017
  4. Hessian-based Analysis of Large Batch Training and Robustness to Adversaries

    Zhewei Yao, Amir Gholami, Qi Lei +2

    cs.CVcs.LGstat.MLarXiv:1802.08241v42018
  5. Closing Cost-Quality Gap in Document VLMs: Difficulty-Aware Data Curation and Quality-Adjusted Deployment Economics

    Maksim Evdokimov, Matvey Ivanov, Dmitrii Tsiupin +3

    cs.CLarXiv:2609.01575v12026
  6. Real or Fake? Learning to Discriminate Machine from Human Generated Text

    Anton Bakhtin, Sam Gross, Myle Ott +3

    cs.LGcs.CLstat.MLarXiv:1906.03351v22019
  7. RepoAudit: An Autonomous LLM-Agent for Repository-Level Code Auditing

    Jinyao Guo, Chengpeng Wang, Xiangzhe Xu +2

    cs.SEcs.PLarXiv:2501.18160v32025
  8. ChipGPT: How far are we from natural language hardware design

    Kaiyan Chang, Ying Wang, Haimeng Ren +5

    cs.AIcs.ARcs.PLarXiv:2305.14019v42023
  9. From Rollouts to Recipes: Self-Contained Post-Training for LLMs

    Yifei Li, Lingling Zhang, Muye Huang +3

    cs.CLarXiv:2609.01422v12026
  10. Mitigating Hallucinations in Large Vision-Language Models via DPO: On-Policy Data Hold the Key

    Zhihe Yang, Xufang Luo, Dongqi Han +2

    cs.CVarXiv:2501.09695v22025
  11. Chebyshev Polynomial-Based Kolmogorov-Arnold Networks: An Efficient Architecture for Nonlinear Function Approximation

    Sidharth SS, Keerthana AR, Gokul R +1

    cs.LGcs.AIarXiv:2405.07200v32024
  12. ClinTraceBench: Source-Verifiable Longitudinal Clinical Reasoning over EHR-Derived Dialogues

    Huimin Wang, Zhengyi Zhao, Yutian Zhao

    cs.CLarXiv:2609.01111v12026
  13. A Survey of Layer-Two Blockchain Protocols

    Ankit Gangwal, Haripriya Ravali Gangavalli, Apoorva Thirupathi

    cs.CRarXiv:2204.08032v32022
  14. MASTER: Multi-Aspect Non-local Network for Scene Text Recognition

    Ning Lu, Wenwen Yu, Xianbiao Qi +4

    cs.CVarXiv:1910.02562v32019
  15. 3DCodeBench: Benchmarking Agentic Procedural 3D Modeling Via Code

    Yipeng Gao, Lei Shu, Genzhi Ye +5

    cs.CVcs.AIcs.GRarXiv:2606.01057v12026
  16. Deepfake-Eval-2024: A Multi-Modal In-the-Wild Benchmark of Deepfakes Circulated in 2024

    Nuria Alina Chandra, Hannah Lee, Ryan Murtfeldt +10

    cs.CVcs.AIcs.CYarXiv:2503.02857v52025
  17. MARS: Modular Agent with Reflective Search for Automated AI Research

    Jiefeng Chen, Bhavana Dalvi Mishra, Jaehyun Nam +3

    cs.AIarXiv:2602.02660v32026
  18. An Evaluation Framework for National AI Regulation

    Kaushik Sanjay Prabhakar, Tarun Adarsh R S, Amal Dhivyan Gregory +3

    cs.CYcs.AIarXiv:2608.15417v12026
  19. LLM-Grounder: Open-Vocabulary 3D Visual Grounding with Large Language Model as an Agent

    Jianing Yang, Xuweiyi Chen, Shengyi Qian +4

    cs.CVcs.AIcs.CLarXiv:2309.12311v12023
  20. StarCoder: may the source be with you!

    Raymond Li, Loubna Ben Allal, Yangtian Zi +64

    cs.CLcs.AIcs.PLarXiv:2305.06161v22023
  21. Forecasting Corn Yield with Machine Learning Ensembles

    Mohsen Shahhosseini, Guiping Hu, Sotirios V. Archontoulis

    stat.APcs.LGstat.MLarXiv:2001.09055v22020
  22. TacoMAS: Test-Time Co-Evolution of Topology and Capability in LLM-based Multi-Agent Systems

    Chen Xu, Yicheng Hu, Ruizi Wang +4

    cs.CLarXiv:2605.09539v12026
  23. Modality Fault Lines: Structural Corruptions Reveal Fragile Omni-Modal Reasoning

    Zhaolu Kang, Meixin Wu, Yu Xue +6

    cs.CLarXiv:2608.29278v12026
  24. Temporal Leakage in Financial News NLP: A Multi-Architecture Audit with a Regime-Specific M&A Signal

    Chenhao Xue, Raslen Guesmi, Siwei Feng +5

    cs.CLcs.LGarXiv:2608.17223v12026
  25. A Trajectory-Based Safety Audit of Clawdbot (OpenClaw)

    Tianyu Chen, Dongrui Liu, Xia Hu +2

    cs.CRcs.AIarXiv:2602.14364v12026
  26. PIVOT: Iterative Visual Prompting Elicits Actionable Knowledge for VLMs

    Soroush Nasiriany, Fei Xia, Wenhao Yu +20

    cs.ROcs.CLcs.CVarXiv:2402.07872v12024
  27. Recursive Multi-Agent Systems

    Jiaru Zou, Rui Pan, Ruizhong Qiu +8

    cs.AIcs.CLcs.LGarXiv:2604.25917v22026
    Summaries:한국어
  28. Diffusion as a Training Curriculum for Timestep-Free Iterative Reasoning

    Mariia Drozdova, Aidan Sirbu, Pietro Miotti +4

    cs.LGarXiv:2609.01449v12026
  29. Building Production-Ready Probes For Gemini

    János Kramár, Joshua Engels, Zheng Wang +4

    cs.LGcs.AIcs.CLarXiv:2601.11516v42026
  30. P1-VL: Bridging Visual Perception and Scientific Reasoning in Physics Olympiads

    Yun Luo, Futing Wang, Qianjia Cheng +28

    cs.AIarXiv:2602.09443v12026
  31. EnvHarness: Awakening Static Worlds for Agent Learning

    Chengsong Huang, Zifeng Wang, Rujun Han +14

    cs.AIcs.CLcs.LGarXiv:2608.19880v12026
    Summaries:简体中文
  32. Think Again or Think Longer? Selective Verification for Budget-Aware Reasoning

    Sajib Acharjee Dip, Dawei Zhou, Liqing Zhang

    cs.AIcs.CLarXiv:2606.19808v12026
  33. Ventor-QTest: Threat-Model-Driven Verification of Vendor-Hosted LLM APIs

    Xiangfan Wu, Zonghao Ying, Huiyu Wu +4

    cs.CRcs.AIarXiv:2608.16391v12026
  34. Pruning and Quantization for Deep Neural Network Acceleration: A Survey

    Tailin Liang, John Glossner, Lei Wang +2

    cs.CVcs.AIarXiv:2101.09671v32021
  35. Memory Intelligence Agent

    Jingyang Qiao, Weicheng Meng, Yu Cheng +6

    cs.AIcs.MAarXiv:2604.04503v42026
  36. ScientistOne: Towards Human-Level Autonomous Research via Chain-of-Evidence

    Rui Meng, Bhavana Dalvi Mishra, Jiefeng Chen +10

    cs.AIcs.CLcs.MAarXiv:2605.26340v12026
  37. Is Mamba Effective for Time Series Forecasting?

    Zihan Wang, Fanheng Kong, Shi Feng +5

    cs.LGarXiv:2403.11144v32024
  38. Co-Director: Agentic Generative Video Storytelling

    Yale Song, Yiwen Song, Nick Losier +13

    cs.AIcs.MAcs.MMarXiv:2604.24842v12026
  39. No Hidden Prompts Needed! You Can Game AI Peer Review with Presentation-Only Revisions

    Xu Yang, Zhizhou Sha, Junbo Li +10

    cs.CLarXiv:2606.13044v12026
  40. Transparency of Deep Neural Networks for Medical Image Analysis: A Review of Interpretability Methods

    Zohaib Salahuddin, Henry C Woodruff, Avishek Chatterjee +1

    eess.IVcs.AIcs.CVarXiv:2111.02398v12021
  41. Data-Efficient Autoregressive-to-Diffusion Language Models via On-Policy Distillation

    Xingyu Su, Jacob Helwig, Shubham Parashar +6

    cs.CLcs.AIarXiv:2606.06712v12026
  42. FAIR1M: A Benchmark Dataset for Fine-grained Object Recognition in High-Resolution Remote Sensing Imagery

    Xian Sun, Peijin Wang, Zhiyuan Yan +11

    cs.CVarXiv:2103.05569v22021
  43. ProEval: Proactive Failure Discovery and Efficient Performance Estimation for Generative AI Evaluation

    Yizheng Huang, Wenjun Zeng, Aditi Kumaresan +1

    cs.LGcs.AIstat.MLarXiv:2604.23099v22026
  44. VID-AD: A Dataset for Image-Level Logical Anomaly Detection under Vision-Induced Distraction

    Hiroto Nakata, Yawen Zou, Shunsuke Sakai +5

    cs.CVarXiv:2603.13964v12026
  45. CLIPO: Contrastive Learning in Policy Optimization Generalizes RLVR

    Sijia Cui, Pengyu Cheng, Jiajun Song +6

    cs.LGcs.AIcs.CLarXiv:2603.10101v12026
  46. The spatial anatomy of urban wildfire vulnerability: a spatially validated GeoAI framework reveals the roles of building density and vegetation moisture in structure loss during the 2025 Palisades Fire

    Parastoo Farajpoor, Mohammadreza Narimani

    physics.geo-phcs.LGeess.IVarXiv:2608.22293v12026
  47. Benchmarking Vision-Language Models for Automated Pathology Diagnosis and Report Generation

    Yumi Lee, Harim Oh, Hyoryung Kim +52

    cs.CVcs.AIarXiv:2609.00866v12026
  48. Video models are zero-shot learners and reasoners

    Thaddäus Wiedemer, Yuxuan Li, Paul Vicol +6

    cs.LGcs.AIcs.CVarXiv:2509.20328v22025
  49. GigaWorld-Policy-0.5: A Faster and Stronger WAM Empowered by AutoResearch

    GigaWorld Team, Angen Ye, Angyuan Ma +26

    cs.ROarXiv:2607.13960v32026
  50. World Simulation with Video Foundation Models for Physical AI

    NVIDIA, :, Arslan Ali +87

    cs.CVcs.AIcs.LGarXiv:2511.00062v22025
  51. VBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning

    Junxiang Xu, Ruisi Wang, Fanyi Pu +49

    cs.CVcs.AIcs.LGarXiv:2608.26105v12026
  52. NatureBench: Can Coding Agents Match the Published SOTA of Nature-Family Papers?

    Yuru Wang, Lejun Cheng, Yuxin Zuo +14

    cs.CLarXiv:2606.24530v22026
  53. Stitched Value Model for Diffusion Alignment

    Hyojun Go, Hyungjin Chung, Prune Truong +8

    cs.CVcs.AIcs.LGarXiv:2605.19804v12026
  54. Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination

    Shuo Liang, Yixing Ma, Pengfei Zhou +32

    cs.CVcs.AIarXiv:2608.14391v12026
  55. Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory

    Rubin Wei, Jiaqi Cao, Jiarui Wang +4

    cs.CLarXiv:2607.27919v12026
  56. Gated DeltaNet-2: Decoupling Erase and Write in Linear Attention

    Ali Hatamizadeh, Yejin Choi, Jan Kautz

    cs.AIarXiv:2605.22791v12026
  57. ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU

    Fan Jiang, Zhaoxu Sun, Mengchao Wang +38

    cs.CVcs.AIcs.LGarXiv:2607.19191v12026
  58. SANA-Streaming: Real-time Streaming Video Editing with Hybrid Diffusion Transformer

    Yuyang Zhao, Yicheng Pan, Qiyuan He +6

    cs.CVcs.AIarXiv:2605.30409v12026
  59. Repo-To-Skill: Distilling GitHub Repositories Into AI4AI Skills

    Jianlyu Chen, Yuyang Hu, Hongjin Qian +8

    cs.AIcs.CLarXiv:2609.02749v12026
  60. Pretraining Large Language Models with NVFP4

    NVIDIA, Felix Abecassis, Anjulie Agrusa +87

    cs.CLcs.AIcs.LGarXiv:2509.25149v22025