Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
18,961 to 19,020 of 61,112
On Memory Construction and Retrieval for Personalized Conversational Agents
Zhuoshi Pan, Qianhui Wu, Huiqiang Jiang +8
cs.CLcs.AIarXiv:2502.05589v32025Streaming Long Video Understanding with Large Language Models
Rui Qian, Xiaoyi Dong, Pan Zhang +4
cs.CVarXiv:2405.16009v12024Improving Object Localization with Fitness NMS and Bounded IoU Loss
Lachlan Tychsen-Smith, Lars Petersson
cs.CVarXiv:1711.00164v32017Hessian-based Analysis of Large Batch Training and Robustness to Adversaries
Zhewei Yao, Amir Gholami, Qi Lei +2
cs.CVcs.LGstat.MLarXiv:1802.08241v42018Closing Cost-Quality Gap in Document VLMs: Difficulty-Aware Data Curation and Quality-Adjusted Deployment Economics
Maksim Evdokimov, Matvey Ivanov, Dmitrii Tsiupin +3
cs.CLarXiv:2609.01575v12026Real or Fake? Learning to Discriminate Machine from Human Generated Text
Anton Bakhtin, Sam Gross, Myle Ott +3
cs.LGcs.CLstat.MLarXiv:1906.03351v22019RepoAudit: An Autonomous LLM-Agent for Repository-Level Code Auditing
Jinyao Guo, Chengpeng Wang, Xiangzhe Xu +2
cs.SEcs.PLarXiv:2501.18160v32025ChipGPT: How far are we from natural language hardware design
Kaiyan Chang, Ying Wang, Haimeng Ren +5
cs.AIcs.ARcs.PLarXiv:2305.14019v42023From Rollouts to Recipes: Self-Contained Post-Training for LLMs
Yifei Li, Lingling Zhang, Muye Huang +3
cs.CLarXiv:2609.01422v12026Mitigating Hallucinations in Large Vision-Language Models via DPO: On-Policy Data Hold the Key
Zhihe Yang, Xufang Luo, Dongqi Han +2
cs.CVarXiv:2501.09695v22025Chebyshev Polynomial-Based Kolmogorov-Arnold Networks: An Efficient Architecture for Nonlinear Function Approximation
Sidharth SS, Keerthana AR, Gokul R +1
cs.LGcs.AIarXiv:2405.07200v32024ClinTraceBench: Source-Verifiable Longitudinal Clinical Reasoning over EHR-Derived Dialogues
Huimin Wang, Zhengyi Zhao, Yutian Zhao
cs.CLarXiv:2609.01111v12026A Survey of Layer-Two Blockchain Protocols
Ankit Gangwal, Haripriya Ravali Gangavalli, Apoorva Thirupathi
cs.CRarXiv:2204.08032v32022MASTER: Multi-Aspect Non-local Network for Scene Text Recognition
Ning Lu, Wenwen Yu, Xianbiao Qi +4
cs.CVarXiv:1910.02562v320193DCodeBench: Benchmarking Agentic Procedural 3D Modeling Via Code
Yipeng Gao, Lei Shu, Genzhi Ye +5
cs.CVcs.AIcs.GRarXiv:2606.01057v12026Deepfake-Eval-2024: A Multi-Modal In-the-Wild Benchmark of Deepfakes Circulated in 2024
Nuria Alina Chandra, Hannah Lee, Ryan Murtfeldt +10
cs.CVcs.AIcs.CYarXiv:2503.02857v52025MARS: Modular Agent with Reflective Search for Automated AI Research
Jiefeng Chen, Bhavana Dalvi Mishra, Jaehyun Nam +3
cs.AIarXiv:2602.02660v32026An Evaluation Framework for National AI Regulation
Kaushik Sanjay Prabhakar, Tarun Adarsh R S, Amal Dhivyan Gregory +3
cs.CYcs.AIarXiv:2608.15417v12026LLM-Grounder: Open-Vocabulary 3D Visual Grounding with Large Language Model as an Agent
Jianing Yang, Xuweiyi Chen, Shengyi Qian +4
cs.CVcs.AIcs.CLarXiv:2309.12311v12023StarCoder: may the source be with you!
Raymond Li, Loubna Ben Allal, Yangtian Zi +64
cs.CLcs.AIcs.PLarXiv:2305.06161v22023Forecasting Corn Yield with Machine Learning Ensembles
Mohsen Shahhosseini, Guiping Hu, Sotirios V. Archontoulis
stat.APcs.LGstat.MLarXiv:2001.09055v22020TacoMAS: Test-Time Co-Evolution of Topology and Capability in LLM-based Multi-Agent Systems
Chen Xu, Yicheng Hu, Ruizi Wang +4
cs.CLarXiv:2605.09539v12026Modality Fault Lines: Structural Corruptions Reveal Fragile Omni-Modal Reasoning
Zhaolu Kang, Meixin Wu, Yu Xue +6
cs.CLarXiv:2608.29278v12026Temporal Leakage in Financial News NLP: A Multi-Architecture Audit with a Regime-Specific M&A Signal
Chenhao Xue, Raslen Guesmi, Siwei Feng +5
cs.CLcs.LGarXiv:2608.17223v12026A Trajectory-Based Safety Audit of Clawdbot (OpenClaw)
Tianyu Chen, Dongrui Liu, Xia Hu +2
cs.CRcs.AIarXiv:2602.14364v12026PIVOT: Iterative Visual Prompting Elicits Actionable Knowledge for VLMs
Soroush Nasiriany, Fei Xia, Wenhao Yu +20
cs.ROcs.CLcs.CVarXiv:2402.07872v12024Recursive Multi-Agent Systems
Jiaru Zou, Rui Pan, Ruizhong Qiu +8
cs.AIcs.CLcs.LGarXiv:2604.25917v22026Summaries:한국어Diffusion as a Training Curriculum for Timestep-Free Iterative Reasoning
Mariia Drozdova, Aidan Sirbu, Pietro Miotti +4
cs.LGarXiv:2609.01449v12026Building Production-Ready Probes For Gemini
János Kramár, Joshua Engels, Zheng Wang +4
cs.LGcs.AIcs.CLarXiv:2601.11516v42026P1-VL: Bridging Visual Perception and Scientific Reasoning in Physics Olympiads
Yun Luo, Futing Wang, Qianjia Cheng +28
cs.AIarXiv:2602.09443v12026EnvHarness: Awakening Static Worlds for Agent Learning
Chengsong Huang, Zifeng Wang, Rujun Han +14
cs.AIcs.CLcs.LGarXiv:2608.19880v12026Summaries:简体中文Think Again or Think Longer? Selective Verification for Budget-Aware Reasoning
Sajib Acharjee Dip, Dawei Zhou, Liqing Zhang
cs.AIcs.CLarXiv:2606.19808v12026Ventor-QTest: Threat-Model-Driven Verification of Vendor-Hosted LLM APIs
Xiangfan Wu, Zonghao Ying, Huiyu Wu +4
cs.CRcs.AIarXiv:2608.16391v12026Pruning and Quantization for Deep Neural Network Acceleration: A Survey
Tailin Liang, John Glossner, Lei Wang +2
cs.CVcs.AIarXiv:2101.09671v32021Memory Intelligence Agent
Jingyang Qiao, Weicheng Meng, Yu Cheng +6
cs.AIcs.MAarXiv:2604.04503v42026ScientistOne: Towards Human-Level Autonomous Research via Chain-of-Evidence
Rui Meng, Bhavana Dalvi Mishra, Jiefeng Chen +10
cs.AIcs.CLcs.MAarXiv:2605.26340v12026Is Mamba Effective for Time Series Forecasting?
Zihan Wang, Fanheng Kong, Shi Feng +5
cs.LGarXiv:2403.11144v32024Co-Director: Agentic Generative Video Storytelling
Yale Song, Yiwen Song, Nick Losier +13
cs.AIcs.MAcs.MMarXiv:2604.24842v12026No Hidden Prompts Needed! You Can Game AI Peer Review with Presentation-Only Revisions
Xu Yang, Zhizhou Sha, Junbo Li +10
cs.CLarXiv:2606.13044v12026Transparency of Deep Neural Networks for Medical Image Analysis: A Review of Interpretability Methods
Zohaib Salahuddin, Henry C Woodruff, Avishek Chatterjee +1
eess.IVcs.AIcs.CVarXiv:2111.02398v12021Data-Efficient Autoregressive-to-Diffusion Language Models via On-Policy Distillation
Xingyu Su, Jacob Helwig, Shubham Parashar +6
cs.CLcs.AIarXiv:2606.06712v12026FAIR1M: A Benchmark Dataset for Fine-grained Object Recognition in High-Resolution Remote Sensing Imagery
Xian Sun, Peijin Wang, Zhiyuan Yan +11
cs.CVarXiv:2103.05569v22021ProEval: Proactive Failure Discovery and Efficient Performance Estimation for Generative AI Evaluation
Yizheng Huang, Wenjun Zeng, Aditi Kumaresan +1
cs.LGcs.AIstat.MLarXiv:2604.23099v22026VID-AD: A Dataset for Image-Level Logical Anomaly Detection under Vision-Induced Distraction
Hiroto Nakata, Yawen Zou, Shunsuke Sakai +5
cs.CVarXiv:2603.13964v12026CLIPO: Contrastive Learning in Policy Optimization Generalizes RLVR
Sijia Cui, Pengyu Cheng, Jiajun Song +6
cs.LGcs.AIcs.CLarXiv:2603.10101v12026The spatial anatomy of urban wildfire vulnerability: a spatially validated GeoAI framework reveals the roles of building density and vegetation moisture in structure loss during the 2025 Palisades Fire
Parastoo Farajpoor, Mohammadreza Narimani
physics.geo-phcs.LGeess.IVarXiv:2608.22293v12026Benchmarking Vision-Language Models for Automated Pathology Diagnosis and Report Generation
Yumi Lee, Harim Oh, Hyoryung Kim +52
cs.CVcs.AIarXiv:2609.00866v12026Video models are zero-shot learners and reasoners
Thaddäus Wiedemer, Yuxuan Li, Paul Vicol +6
cs.LGcs.AIcs.CVarXiv:2509.20328v22025GigaWorld-Policy-0.5: A Faster and Stronger WAM Empowered by AutoResearch
GigaWorld Team, Angen Ye, Angyuan Ma +26
cs.ROarXiv:2607.13960v32026World Simulation with Video Foundation Models for Physical AI
NVIDIA, :, Arslan Ali +87
cs.CVcs.AIcs.LGarXiv:2511.00062v22025VBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning
Junxiang Xu, Ruisi Wang, Fanyi Pu +49
cs.CVcs.AIcs.LGarXiv:2608.26105v12026NatureBench: Can Coding Agents Match the Published SOTA of Nature-Family Papers?
Yuru Wang, Lejun Cheng, Yuxin Zuo +14
cs.CLarXiv:2606.24530v22026Stitched Value Model for Diffusion Alignment
Hyojun Go, Hyungjin Chung, Prune Truong +8
cs.CVcs.AIcs.LGarXiv:2605.19804v12026Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination
Shuo Liang, Yixing Ma, Pengfei Zhou +32
cs.CVcs.AIarXiv:2608.14391v12026Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory
Rubin Wei, Jiaqi Cao, Jiarui Wang +4
cs.CLarXiv:2607.27919v12026Gated DeltaNet-2: Decoupling Erase and Write in Linear Attention
Ali Hatamizadeh, Yejin Choi, Jan Kautz
cs.AIarXiv:2605.22791v12026ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU
Fan Jiang, Zhaoxu Sun, Mengchao Wang +38
cs.CVcs.AIcs.LGarXiv:2607.19191v12026SANA-Streaming: Real-time Streaming Video Editing with Hybrid Diffusion Transformer
Yuyang Zhao, Yicheng Pan, Qiyuan He +6
cs.CVcs.AIarXiv:2605.30409v12026Repo-To-Skill: Distilling GitHub Repositories Into AI4AI Skills
Jianlyu Chen, Yuyang Hu, Hongjin Qian +8
cs.AIcs.CLarXiv:2609.02749v12026Pretraining Large Language Models with NVFP4
NVIDIA, Felix Abecassis, Anjulie Agrusa +87
cs.CLcs.AIcs.LGarXiv:2509.25149v22025