Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
20,821 to 20,880 of 61,306
Conditional Affordance Learning for Driving in Urban Environments
Axel Sauer, Nikolay Savinov, Andreas Geiger
cs.ROcs.LGeess.SYarXiv:1806.06498v32018Multiplicative comparisons of Rényi entropies for weighted Bernoulli sums
Jiange Li
math.PRcs.ITmath.COarXiv:2609.01529v12026Nemotron 3 Nano: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning
NVIDIA, :, Aaron Blakeman +311
cs.CLcs.AIcs.LGarXiv:2512.20848v12025A General Construction of Codes from Drinfeld Modules
Alessandro Giannoni, Giacomo Micheli, Mihran Papikian
math.NTcs.ITmath.COarXiv:2609.01484v12026PaddleOCR-VL: Boosting Multilingual Document Parsing via a 0.9B Ultra-Compact Vision-Language Model
Cheng Cui, Ting Sun, Suyin Liang +15
cs.CVarXiv:2510.14528v42025Dominant Set Clustering and Pooling for Multi-View 3D Object Recognition
Chu Wang, Marcello Pelillo, Kaleem Siddiqi
cs.CVarXiv:1906.01592v12019WebAgent-R1: Training Web Agents via End-to-End Multi-Turn Reinforcement Learning
Zhepei Wei, Wenlin Yao, Yao Liu +9
cs.CLcs.LGarXiv:2505.16421v22025DeepCorr: Strong Flow Correlation Attacks on Tor Using Deep Learning
Milad Nasr, Alireza Bahramali, Amir Houmansadr
cs.CRcs.LGarXiv:1808.07285v12018TiRex: Zero-Shot Forecasting Across Long and Short Horizons with Enhanced In-Context Learning
Andreas Auer, Patrick Podest, Daniel Klotz +3
cs.LGarXiv:2505.23719v22025SparseTSF: Modeling Long-term Time Series Forecasting with 1k Parameters
Shengsheng Lin, Weiwei Lin, Wentai Wu +2
cs.LGarXiv:2405.00946v22024A Survey on Test-Time Scaling in Large Language Models: What, How, Where, and How Well?
Qiyuan Zhang, Fuyuan Lyu, Zexu Sun +10
cs.CLcs.AIarXiv:2503.24235v32025Virtual Element Methods on Meshes with Small Edges or Faces
Susanne C. Brenner, Li-yeng Sung
math.NAarXiv:1710.00442v12017Self-supervised Image-specific Prototype Exploration for Weakly Supervised Semantic Segmentation
Qi Chen, Lingxiao Yang, Jianhuang Lai +1
cs.CVarXiv:2203.02909v12022Attention Convolutional Binary Neural Tree for Fine-Grained Visual Categorization
Ruyi Ji, Longyin Wen, Libo Zhang +5
cs.CVarXiv:1909.11378v22019Sponge Examples: Energy-Latency Attacks on Neural Networks
Ilia Shumailov, Yiren Zhao, Daniel Bates +3
cs.LGcs.CLcs.CRarXiv:2006.03463v22020Training Agents Inside of Scalable World Models
Danijar Hafner, Wilson Yan, Timothy Lillicrap
cs.AIcs.LGcs.ROarXiv:2509.24527v12025Solving (most) of a set of quadratic equalities: Composite optimization for robust phase retrieval
John C. Duchi, Feng Ruan
math.STcs.ITmath.OCarXiv:1705.02356v22017Algorand
Jing Chen, Silvio Micali
cs.CRcs.DCarXiv:1607.01341v92016The Polar Express: Optimal Matrix Sign Methods and Their Application to the Muon Algorithm
Noah Amsel, David Persson, Christopher Musco +1
cs.LGcs.AIcs.CLarXiv:2505.16932v52025SAEBench: A Comprehensive Benchmark for Sparse Autoencoders in Language Model Interpretability
Adam Karvonen, Can Rager, Johnny Lin +12
cs.LGcs.CLarXiv:2503.09532v42025Understanding plasticity in neural networks
Clare Lyle, Zeyu Zheng, Evgenii Nikishin +3
cs.LGarXiv:2303.01486v42023Joint Training Is Not Enough: Conditioned Cross-Granularity Training for Multimodal Document Understanding
Chengguang Gan, Yunhao Liang, Hanjun Wei +2
cs.CLarXiv:2609.00756v12026X-Omni: Reinforcement Learning Makes Discrete Autoregressive Image Generative Models Great Again
Zigang Geng, Yibing Wang, Yeyao Ma +10
cs.CVarXiv:2507.22058v12025Autonomous discovery of new structure-plausibility laws for explainable and rapid crystal diagnosis and screening
Zhilong Song, Lixue Cheng
cond-mat.mtrl-scics.AIarXiv:2609.01209v12026Uncertainty-aware Score Distribution Learning for Action Quality Assessment
Yansong Tang, Zanlin Ni, Jiahuan Zhou +4
cs.CVarXiv:2006.07665v12020Practical Implementation of Spatial Modulation
N. Serafimovski, A. Younis, R. Mesleh +6
cs.ITarXiv:1305.0664v22013Single Image Reflection Removal Exploiting Misaligned Training Data and Network Enhancements
Kaixuan Wei, Jiaolong Yang, Ying Fu +2
cs.CVarXiv:1904.00637v12019SoK: When Safe Agents Fail Together: The Security of Multi Agent LLM Systems
Rui Yang, Junjie Xu, Zhengyu Liu +4
cs.CRcs.AIarXiv:2609.00595v12026Prompt Injection Attack to Tool Selection in LLM Agents
Jiawen Shi, Zenghui Yuan, Guiyao Tie +3
cs.CRarXiv:2504.19793v32025Q-Insight: Understanding Image Quality via Visual Reinforcement Learning
Weiqi Li, Xuanyu Zhang, Shijie Zhao +4
cs.CVarXiv:2503.22679v22025Ultralytics YOLO Evolution: An Overview of YOLO26, YOLO11, YOLOv8 and YOLOv5 Object Detectors for Computer Vision and Pattern Recognition
Ranjan Sapkota, Manoj Karkee
cs.CVcs.AIarXiv:2510.09653v32025Don't Trust the Code, Check Its Effects: Runtime Refinement for Regenerated Systems Code Under an Adversarial Generator
Jinhao Hu, Ashvin Goel, Laurent Bindschaedler
cs.CRarXiv:2609.00430v12026ZebraPose: Coarse to Fine Surface Encoding for 6DoF Object Pose Estimation
Yongzhi Su, Mahdi Saleh, Torben Fetzer +5
cs.CVarXiv:2203.09418v22022Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model
Guoqing Ma, Haoyang Huang, Kun Yan +112
cs.CVcs.CLarXiv:2502.10248v32025Programming Refusal with Conditional Activation Steering
Bruce W. Lee, Inkit Padhi, Karthikeyan Natesan Ramamurthy +4
cs.LGcs.AIcs.CLarXiv:2409.05907v32024Controllable Image Captioning with Prompt-Conditioned Scene Rewards
Jongyeop Hyun, Taeyoung Kim, Hyounghun Kim
cs.CVcs.CLcs.LGarXiv:2609.00709v12026Computing B-Stationary Points of Nonsmooth DC Programs
Jong-Shi Pang, Meisam Razaviyayn, Alberth Alvarado
math.OCarXiv:1511.01796v12015XAttention: Block Sparse Attention with Antidiagonal Scoring
Ruyi Xu, Guangxuan Xiao, Haofeng Huang +2
cs.CLcs.CVarXiv:2503.16428v12025OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning
Anurag Ajay, Aviral Kumar, Pulkit Agrawal +2
cs.LGarXiv:2010.13611v32020PersuaRL: Reinforcement Learning-Driven Multi-Expert Selection for Persuasive Dialogue Generation in Insurance
Rohan Kirti, Akash Ghosh, Aryan Vats +5
cs.CLarXiv:2609.01188v12026Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation
Qianhao Yuan, Jie Lou, Xing Yu +4
cs.CVcs.AIcs.CLarXiv:2605.18740v42026A Unified RGB-T Saliency Detection Benchmark: Dataset, Baselines, Analysis and A Novel Approach
Chenglong Li, Guizhao Wang, Yunpeng Ma +3
cs.CVarXiv:1701.02829v12017PixMix: Dreamlike Pictures Comprehensively Improve Safety Measures
Dan Hendrycks, Andy Zou, Mantas Mazeika +4
cs.LGcs.CVarXiv:2112.05135v32021GOOD: A Graph Out-of-Distribution Benchmark
Shurui Gui, Xiner Li, Limei Wang +1
cs.LGcs.AIarXiv:2206.08452v22022Membership Inference in Fine-tuned Diffusion Language Models via Token-level Memorization Asymmetry
Shengfang Zhai, Leo Marchyok, Yuling Shi +4
cs.CLcs.CRarXiv:2609.00873v12026REVE: A Foundation Model for EEG -- Adapting to Any Setup with Large-Scale Pretraining on 25,000 Subjects
Yassine El Ouahidi, Jonathan Lys, Philipp Thölke +5
cs.LGq-bio.NCarXiv:2510.21585v12025Deep Exemplar-based Video Colorization
Bo Zhang, Mingming He, Jing Liao +4
cs.CVcs.AIcs.LGarXiv:1906.09909v12019Infinity-RoPE: Action-Controllable Infinite Video Generation Emerges From Autoregressive Self-Rollout
Hidir Yesiltepe, Tuna Han Salih Meral, Adil Kaan Akan +2
cs.CVarXiv:2511.20649v32025RAID: A Shared Benchmark for Robust Evaluation of Machine-Generated Text Detectors
Liam Dugan, Alyssa Hwang, Filip Trhlik +5
cs.CLarXiv:2405.07940v22024EV-SegNet: Semantic Segmentation for Event-based Cameras
Iñigo Alonso, Ana C. Murillo
cs.CVarXiv:1811.12039v12018Act Only When It Pays: Efficient Reinforcement Learning for LLM Reasoning via Selective Rollouts
Haizhong Zheng, Yang Zhou, Brian R. Bartoldson +4
cs.AIcs.LGarXiv:2506.02177v12025BiTraP: Bi-directional Pedestrian Trajectory Prediction with Multi-modal Goal Estimation
Yu Yao, Ella Atkins, Matthew Johnson-Roberson +2
cs.CVcs.ROarXiv:2007.14558v22020FILM: Following Instructions in Language with Modular Methods
So Yeon Min, Devendra Singh Chaplot, Pradeep Ravikumar +2
cs.CLcs.LGarXiv:2110.07342v32021ArtEmis: Affective Language for Visual Art
Panos Achlioptas, Maks Ovsjanikov, Kilichbek Haydarov +2
cs.CVcs.CLarXiv:2101.07396v12021Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Yunzhuo Hao, Jiawei Gu, Huichen Will Wang +4
cs.CVarXiv:2501.05444v12025Go-with-the-Flow: Motion-Controllable Video Diffusion Models Using Real-Time Warped Noise
Ryan Burgert, Yuancheng Xu, Wenqi Xian +10
cs.CVarXiv:2501.08331v52025SPIn-NeRF: Multiview Segmentation and Perceptual Inpainting with Neural Radiance Fields
Ashkan Mirzaei, Tristan Aumentado-Armstrong, Konstantinos G. Derpanis +4
cs.CVarXiv:2211.12254v22022TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate
Amir Zandieh, Majid Daliri, Majid Hadian +1
cs.LGcs.AIcs.DBarXiv:2504.19874v12025Sobolev Norm Learning Rates for Regularized Least-Squares Algorithm
Simon Fischer, Ingo Steinwart
stat.MLarXiv:1702.07254v32017GTA1: GUI Test-time Scaling Agent
Yan Yang, Dongxu Li, Yutong Dai +12
cs.AIarXiv:2507.05791v52025