Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

20,821 to 20,880 of 61,306

  1. Conditional Affordance Learning for Driving in Urban Environments

    Axel Sauer, Nikolay Savinov, Andreas Geiger

    cs.ROcs.LGeess.SYarXiv:1806.06498v32018
  2. Multiplicative comparisons of Rényi entropies for weighted Bernoulli sums

    Jiange Li

    math.PRcs.ITmath.COarXiv:2609.01529v12026
  3. Nemotron 3 Nano: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning

    NVIDIA, :, Aaron Blakeman +311

    cs.CLcs.AIcs.LGarXiv:2512.20848v12025
  4. A General Construction of Codes from Drinfeld Modules

    Alessandro Giannoni, Giacomo Micheli, Mihran Papikian

    math.NTcs.ITmath.COarXiv:2609.01484v12026
  5. PaddleOCR-VL: Boosting Multilingual Document Parsing via a 0.9B Ultra-Compact Vision-Language Model

    Cheng Cui, Ting Sun, Suyin Liang +15

    cs.CVarXiv:2510.14528v42025
  6. Dominant Set Clustering and Pooling for Multi-View 3D Object Recognition

    Chu Wang, Marcello Pelillo, Kaleem Siddiqi

    cs.CVarXiv:1906.01592v12019
  7. WebAgent-R1: Training Web Agents via End-to-End Multi-Turn Reinforcement Learning

    Zhepei Wei, Wenlin Yao, Yao Liu +9

    cs.CLcs.LGarXiv:2505.16421v22025
  8. DeepCorr: Strong Flow Correlation Attacks on Tor Using Deep Learning

    Milad Nasr, Alireza Bahramali, Amir Houmansadr

    cs.CRcs.LGarXiv:1808.07285v12018
  9. TiRex: Zero-Shot Forecasting Across Long and Short Horizons with Enhanced In-Context Learning

    Andreas Auer, Patrick Podest, Daniel Klotz +3

    cs.LGarXiv:2505.23719v22025
  10. SparseTSF: Modeling Long-term Time Series Forecasting with 1k Parameters

    Shengsheng Lin, Weiwei Lin, Wentai Wu +2

    cs.LGarXiv:2405.00946v22024
  11. A Survey on Test-Time Scaling in Large Language Models: What, How, Where, and How Well?

    Qiyuan Zhang, Fuyuan Lyu, Zexu Sun +10

    cs.CLcs.AIarXiv:2503.24235v32025
  12. Virtual Element Methods on Meshes with Small Edges or Faces

    Susanne C. Brenner, Li-yeng Sung

    math.NAarXiv:1710.00442v12017
  13. Self-supervised Image-specific Prototype Exploration for Weakly Supervised Semantic Segmentation

    Qi Chen, Lingxiao Yang, Jianhuang Lai +1

    cs.CVarXiv:2203.02909v12022
  14. Attention Convolutional Binary Neural Tree for Fine-Grained Visual Categorization

    Ruyi Ji, Longyin Wen, Libo Zhang +5

    cs.CVarXiv:1909.11378v22019
  15. Sponge Examples: Energy-Latency Attacks on Neural Networks

    Ilia Shumailov, Yiren Zhao, Daniel Bates +3

    cs.LGcs.CLcs.CRarXiv:2006.03463v22020
  16. Training Agents Inside of Scalable World Models

    Danijar Hafner, Wilson Yan, Timothy Lillicrap

    cs.AIcs.LGcs.ROarXiv:2509.24527v12025
  17. Solving (most) of a set of quadratic equalities: Composite optimization for robust phase retrieval

    John C. Duchi, Feng Ruan

    math.STcs.ITmath.OCarXiv:1705.02356v22017
  18. Algorand

    Jing Chen, Silvio Micali

    cs.CRcs.DCarXiv:1607.01341v92016
  19. The Polar Express: Optimal Matrix Sign Methods and Their Application to the Muon Algorithm

    Noah Amsel, David Persson, Christopher Musco +1

    cs.LGcs.AIcs.CLarXiv:2505.16932v52025
  20. SAEBench: A Comprehensive Benchmark for Sparse Autoencoders in Language Model Interpretability

    Adam Karvonen, Can Rager, Johnny Lin +12

    cs.LGcs.CLarXiv:2503.09532v42025
  21. Understanding plasticity in neural networks

    Clare Lyle, Zeyu Zheng, Evgenii Nikishin +3

    cs.LGarXiv:2303.01486v42023
  22. Joint Training Is Not Enough: Conditioned Cross-Granularity Training for Multimodal Document Understanding

    Chengguang Gan, Yunhao Liang, Hanjun Wei +2

    cs.CLarXiv:2609.00756v12026
  23. X-Omni: Reinforcement Learning Makes Discrete Autoregressive Image Generative Models Great Again

    Zigang Geng, Yibing Wang, Yeyao Ma +10

    cs.CVarXiv:2507.22058v12025
  24. Autonomous discovery of new structure-plausibility laws for explainable and rapid crystal diagnosis and screening

    Zhilong Song, Lixue Cheng

    cond-mat.mtrl-scics.AIarXiv:2609.01209v12026
  25. Uncertainty-aware Score Distribution Learning for Action Quality Assessment

    Yansong Tang, Zanlin Ni, Jiahuan Zhou +4

    cs.CVarXiv:2006.07665v12020
  26. Practical Implementation of Spatial Modulation

    N. Serafimovski, A. Younis, R. Mesleh +6

    cs.ITarXiv:1305.0664v22013
  27. Single Image Reflection Removal Exploiting Misaligned Training Data and Network Enhancements

    Kaixuan Wei, Jiaolong Yang, Ying Fu +2

    cs.CVarXiv:1904.00637v12019
  28. SoK: When Safe Agents Fail Together: The Security of Multi Agent LLM Systems

    Rui Yang, Junjie Xu, Zhengyu Liu +4

    cs.CRcs.AIarXiv:2609.00595v12026
  29. Prompt Injection Attack to Tool Selection in LLM Agents

    Jiawen Shi, Zenghui Yuan, Guiyao Tie +3

    cs.CRarXiv:2504.19793v32025
  30. Q-Insight: Understanding Image Quality via Visual Reinforcement Learning

    Weiqi Li, Xuanyu Zhang, Shijie Zhao +4

    cs.CVarXiv:2503.22679v22025
  31. Ultralytics YOLO Evolution: An Overview of YOLO26, YOLO11, YOLOv8 and YOLOv5 Object Detectors for Computer Vision and Pattern Recognition

    Ranjan Sapkota, Manoj Karkee

    cs.CVcs.AIarXiv:2510.09653v32025
  32. Don't Trust the Code, Check Its Effects: Runtime Refinement for Regenerated Systems Code Under an Adversarial Generator

    Jinhao Hu, Ashvin Goel, Laurent Bindschaedler

    cs.CRarXiv:2609.00430v12026
  33. ZebraPose: Coarse to Fine Surface Encoding for 6DoF Object Pose Estimation

    Yongzhi Su, Mahdi Saleh, Torben Fetzer +5

    cs.CVarXiv:2203.09418v22022
  34. Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model

    Guoqing Ma, Haoyang Huang, Kun Yan +112

    cs.CVcs.CLarXiv:2502.10248v32025
  35. Programming Refusal with Conditional Activation Steering

    Bruce W. Lee, Inkit Padhi, Karthikeyan Natesan Ramamurthy +4

    cs.LGcs.AIcs.CLarXiv:2409.05907v32024
  36. Controllable Image Captioning with Prompt-Conditioned Scene Rewards

    Jongyeop Hyun, Taeyoung Kim, Hyounghun Kim

    cs.CVcs.CLcs.LGarXiv:2609.00709v12026
  37. Computing B-Stationary Points of Nonsmooth DC Programs

    Jong-Shi Pang, Meisam Razaviyayn, Alberth Alvarado

    math.OCarXiv:1511.01796v12015
  38. XAttention: Block Sparse Attention with Antidiagonal Scoring

    Ruyi Xu, Guangxuan Xiao, Haofeng Huang +2

    cs.CLcs.CVarXiv:2503.16428v12025
  39. OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

    Anurag Ajay, Aviral Kumar, Pulkit Agrawal +2

    cs.LGarXiv:2010.13611v32020
  40. PersuaRL: Reinforcement Learning-Driven Multi-Expert Selection for Persuasive Dialogue Generation in Insurance

    Rohan Kirti, Akash Ghosh, Aryan Vats +5

    cs.CLarXiv:2609.01188v12026
  41. Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation

    Qianhao Yuan, Jie Lou, Xing Yu +4

    cs.CVcs.AIcs.CLarXiv:2605.18740v42026
  42. A Unified RGB-T Saliency Detection Benchmark: Dataset, Baselines, Analysis and A Novel Approach

    Chenglong Li, Guizhao Wang, Yunpeng Ma +3

    cs.CVarXiv:1701.02829v12017
  43. PixMix: Dreamlike Pictures Comprehensively Improve Safety Measures

    Dan Hendrycks, Andy Zou, Mantas Mazeika +4

    cs.LGcs.CVarXiv:2112.05135v32021
  44. GOOD: A Graph Out-of-Distribution Benchmark

    Shurui Gui, Xiner Li, Limei Wang +1

    cs.LGcs.AIarXiv:2206.08452v22022
  45. Membership Inference in Fine-tuned Diffusion Language Models via Token-level Memorization Asymmetry

    Shengfang Zhai, Leo Marchyok, Yuling Shi +4

    cs.CLcs.CRarXiv:2609.00873v12026
  46. REVE: A Foundation Model for EEG -- Adapting to Any Setup with Large-Scale Pretraining on 25,000 Subjects

    Yassine El Ouahidi, Jonathan Lys, Philipp Thölke +5

    cs.LGq-bio.NCarXiv:2510.21585v12025
  47. Deep Exemplar-based Video Colorization

    Bo Zhang, Mingming He, Jing Liao +4

    cs.CVcs.AIcs.LGarXiv:1906.09909v12019
  48. Infinity-RoPE: Action-Controllable Infinite Video Generation Emerges From Autoregressive Self-Rollout

    Hidir Yesiltepe, Tuna Han Salih Meral, Adil Kaan Akan +2

    cs.CVarXiv:2511.20649v32025
  49. RAID: A Shared Benchmark for Robust Evaluation of Machine-Generated Text Detectors

    Liam Dugan, Alyssa Hwang, Filip Trhlik +5

    cs.CLarXiv:2405.07940v22024
  50. EV-SegNet: Semantic Segmentation for Event-based Cameras

    Iñigo Alonso, Ana C. Murillo

    cs.CVarXiv:1811.12039v12018
  51. Act Only When It Pays: Efficient Reinforcement Learning for LLM Reasoning via Selective Rollouts

    Haizhong Zheng, Yang Zhou, Brian R. Bartoldson +4

    cs.AIcs.LGarXiv:2506.02177v12025
  52. BiTraP: Bi-directional Pedestrian Trajectory Prediction with Multi-modal Goal Estimation

    Yu Yao, Ella Atkins, Matthew Johnson-Roberson +2

    cs.CVcs.ROarXiv:2007.14558v22020
  53. FILM: Following Instructions in Language with Modular Methods

    So Yeon Min, Devendra Singh Chaplot, Pradeep Ravikumar +2

    cs.CLcs.LGarXiv:2110.07342v32021
  54. ArtEmis: Affective Language for Visual Art

    Panos Achlioptas, Maks Ovsjanikov, Kilichbek Haydarov +2

    cs.CVcs.CLarXiv:2101.07396v12021
  55. Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark

    Yunzhuo Hao, Jiawei Gu, Huichen Will Wang +4

    cs.CVarXiv:2501.05444v12025
  56. Go-with-the-Flow: Motion-Controllable Video Diffusion Models Using Real-Time Warped Noise

    Ryan Burgert, Yuancheng Xu, Wenqi Xian +10

    cs.CVarXiv:2501.08331v52025
  57. SPIn-NeRF: Multiview Segmentation and Perceptual Inpainting with Neural Radiance Fields

    Ashkan Mirzaei, Tristan Aumentado-Armstrong, Konstantinos G. Derpanis +4

    cs.CVarXiv:2211.12254v22022
  58. TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate

    Amir Zandieh, Majid Daliri, Majid Hadian +1

    cs.LGcs.AIcs.DBarXiv:2504.19874v12025
  59. Sobolev Norm Learning Rates for Regularized Least-Squares Algorithm

    Simon Fischer, Ingo Steinwart

    stat.MLarXiv:1702.07254v32017
  60. GTA1: GUI Test-time Scaling Agent

    Yan Yang, Dongxu Li, Yutong Dai +12

    cs.AIarXiv:2507.05791v52025