Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
18,421 to 18,480 of 61,393
Dr. Splat: Directly Referring 3D Gaussian Splatting via Direct Language Embedding Registration
Kim Jun-Seong, GeonU Kim, Kim Yu-Ji +3
cs.CVarXiv:2502.16652v12025Multi-fidelity Bayesian Neural Networks: Algorithms and Applications
Xuhui Meng, Hessam Babaee, George Em Karniadakis
cs.LGphysics.comp-pharXiv:2012.13294v12020Validity-Aware Jailbreak Evaluation for Large Language Models
Qilong Wu, Sahil Wadhwa, Pranab Mohanty +2
cs.AIarXiv:2609.00498v12026Reinforcement Learning for Reasoning in Small LLMs: What Works and What Doesn't
Quy-Anh Dang, Chris Ngo
cs.LGcs.CLarXiv:2503.16219v22025Directional Initial Access for Millimeter Wave Cellular Systems
C. Nicolas Barati, S. Amir Hosseini, Marco Mezzavilla +4
cs.ITcs.NIarXiv:1511.06483v22015MotionBench: Benchmarking and Improving Fine-grained Video Motion Understanding for Vision Language Models
Wenyi Hong, Yean Cheng, Zhuoyi Yang +6
cs.CVarXiv:2501.02955v22025Will Metaverse be NextG Internet? Vision, Hype, and Reality
Ruizhi Cheng, Nan Wu, Songqing Chen +1
cs.NIcs.MMcs.SIarXiv:2201.12894v22022Benchmarking Reinforcement Learning Algorithms on Real-World Robots
A. Rupam Mahmood, Dmytro Korenkevych, Gautham Vasan +2
cs.LGcs.AIcs.ROarXiv:1809.07731v12018Actor and Observer: Joint Modeling of First and Third-Person Videos
Gunnar A. Sigurdsson, Abhinav Gupta, Cordelia Schmid +2
cs.CVarXiv:1804.09627v12018SpatialVID: A Large-Scale Video Dataset with Spatial Annotations
Jiahao Wang, Yufeng Yuan, Rujie Zheng +12
cs.CVarXiv:2509.09676v22025Adaptive Depth-Map-Guided Bundle Adjustment for Correspondence-Free Multi-View Point Cloud Registration
Yiran Zhou, Yingyu Wang, Shoudong Huang +1
cs.ROarXiv:2609.01089v12026NoveltyBench: Evaluating Language Models for Humanlike Diversity
Yiming Zhang, Harshita Diddee, Susan Holm +5
cs.CLarXiv:2504.05228v42025Diffusion-4K: Ultra-High-Resolution Image Synthesis with Latent Diffusion Models
Jinjin Zhang, Qiuyu Huang, Junjie Liu +2
cs.CVarXiv:2503.18352v22025DISTAL: Distillation and Self-Supervised Pretraining for Structure-Agnostic Materials Property Prediction
Weiran Wang, Xintong Huo, Yueying Wang +7
cs.LGcs.AIcs.ETarXiv:2609.00059v12026Inverse Optimization with Noisy Data
Anil Aswani, Zuo-Jun Max Shen, Auyon Siddiq
math.OCarXiv:1507.03266v42015FLOWER: Democratizing Generalist Robot Policies with Efficient Vision-Language-Action Flow Policies
Moritz Reuss, Hongyi Zhou, Marcel Rühle +3
cs.ROarXiv:2509.04996v12025ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge
Zhongyi Zhou, Yichen Zhu, Junjie Wen +2
cs.ROcs.AIcs.CVarXiv:2505.21906v22025Generated Faces in the Wild: Quantitative Comparison of Stable Diffusion, Midjourney and DALL-E 2
Ali Borji
cs.CVarXiv:2210.00586v22022VL-JEPA: Joint Embedding Predictive Architecture for Vision-language
Delong Chen, Mustafa Shukor, Theo Moutakanni +7
cs.CVarXiv:2512.10942v22025Terminologies for Reproducible Research
Lorena A. Barba
cs.DLarXiv:1802.03311v12018ART: Automatic multi-step reasoning and tool-use for large language models
Bhargavi Paranjape, Scott Lundberg, Sameer Singh +3
cs.CLarXiv:2303.09014v12023Learning from the Worst: Dynamically Generated Datasets to Improve Online Hate Detection
Bertie Vidgen, Tristan Thrush, Zeerak Waseem +1
cs.CLcs.LGarXiv:2012.15761v22020P3Depth: Monocular Depth Estimation with a Piecewise Planarity Prior
Vaishakh Patil, Christos Sakaridis, Alexander Liniger +1
cs.CVcs.AIcs.LGarXiv:2204.02091v12022Parameterized quantum circuits as machine learning models
Marcello Benedetti, Erika Lloyd, Stefan Sack +1
quant-phcs.LGarXiv:1906.07682v22019Accelerating Unified Multimodal Models with Core-Expansion Routing and Unified Computation Scheduling
Wengyi Zhan, Chenqian Yan, Songwei Liu +2
cs.AIarXiv:2608.29291v32026Cultural Moment Benchmark: Evaluating Video Cultural Reasoning and Grounding in Southeast Asia
Burak Satar, Zhixin Ma, Cheng Yu-Tong +3
cs.CVcs.AIcs.CLarXiv:2608.23065v12026Robust Video Content Alignment and Compensation for Rain Removal in a CNN Framework
Jie Chen, Cheen-Hau Tan, Junhui Hou +2
cs.CVarXiv:1803.10433v12018Simpler but More Accurate Semantic Dependency Parsing
Timothy Dozat, Christopher D. Manning
cs.CLarXiv:1807.01396v12018State-specific protein-ligand complex structure prediction with a multi-scale deep generative model
Zhuoran Qiao, Weili Nie, Arash Vahdat +2
q-bio.QMcs.LGq-bio.BMarXiv:2209.15171v22022Open-environment Machine Learning
Zhi-Hua Zhou
cs.LGarXiv:2206.00423v22022Deep Dynamical Modeling and Control of Unsteady Fluid Flows
Jeremy Morton, Freddie D. Witherden, Antony Jameson +1
cs.CEcs.AIarXiv:1805.07472v22018Robbing the Fed: Directly Obtaining Private Data in Federated Learning with Modified Models
Liam Fowl, Jonas Geiping, Wojtek Czaja +2
cs.LGcs.CRarXiv:2110.13057v22021Increasing Diversity While Maintaining Accuracy: Text Data Generation with Large Language Models and Human Interventions
John Joon Young Chung, Ece Kamar, Saleema Amershi
cs.CLarXiv:2306.04140v12023Neural Embeddings of Graphs in Hyperbolic Space
Benjamin Paul Chamberlain, James Clough, Marc Peter Deisenroth
stat.MLcs.LGarXiv:1705.10359v12017AgentPRM: Process Reward Models for LLM Agents via Step-Wise Promise and Progress
Zhiheng Xi, Chenyang Liao, Guanyu Li +12
cs.CLcs.IRcs.LGarXiv:2511.08325v12025Fine-tuned Language Models are Continual Learners
Thomas Scialom, Tuhin Chakrabarty, Smaranda Muresan
cs.CLarXiv:2205.12393v42022CUDA-Harness: Harnessing Agentic CUDA Kernel Generation and Optimization from Natural Language
Qi Fan, An Zou, Yehan Ma
cs.CLcs.AIcs.MAarXiv:2609.00058v12026Proximity Forest: An effective and scalable distance-based classifier for time series
Benjamin Lucas, Ahmed Shifaz, Charlotte Pelletier +5
cs.LGstat.MLarXiv:1808.10594v22018A-MemGuard: A Proactive Defense Framework for LLM-Based Agent Memory
Qianshan Wei, Tengchao Yang, Yaochen Wang +7
cs.CRcs.AIarXiv:2510.02373v12025Improving Vision-Language-Action Model with Online Reinforcement Learning
Yanjiang Guo, Jianke Zhang, Xiaoyu Chen +4
cs.ROcs.CVcs.LGarXiv:2501.16664v12025SparseFlex: High-Resolution and Arbitrary-Topology 3D Shape Modeling
Xianglong He, Zi-Xin Zou, Chia-Hao Chen +6
cs.CVarXiv:2503.21732v12025ReNFT: Repairing Mode Collapse in Reward Post-Training via Internal Probability-Mass Recalibration
Yuchen Bao, Chao Wen, Haowei Wang +10
cs.LGcs.AIcs.CVarXiv:2609.00061v12026WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL
Zhennan Jiang, Shangqing Zhou, Yutong Jiang +11
cs.ROcs.AIarXiv:2602.13977v22026Uncertainty-Aware Parameter Estimation for Condition Monitoring of Power Converters
Tomas Monopoli, Jiahong Liu, Shuai Zhao
eess.SYarXiv:2609.00218v12026PolaFormer: Polarity-aware Linear Attention for Vision Transformers
Weikang Meng, Yadan Luo, Xin Li +2
cs.CVcs.AIarXiv:2501.15061v22025A soft robot that adapts to environments through shape change
Dylan S. Shah, Joshua P. Powers, Liana G. Tilton +3
cs.ROarXiv:2008.06397v52020OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation
Shenghai Yuan, Xianyi He, Yufan Deng +5
cs.CVcs.AIarXiv:2505.20292v42025Physical Layer Security for Visible Light Communication Systems: A Survey
Mohamed Amine Arfaoui, Mohammad Dehghani Soltani, Iman Tavakkolnia +4
eess.SParXiv:1905.11450v12019An Ensemble 1D-CNN-LSTM-GRU Model with Data Augmentation for Speech Emotion Recognition
Md. Rayhan Ahmed, Salekul Islam, Ph. D +4
cs.SDeess.ASarXiv:2112.05666v22021MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning
Zhaopeng Feng, Shaosheng Cao, Jiahan Ren +7
cs.CLcs.AIcs.LGarXiv:2504.10160v12025Hidden relationships in a document-derived property graph: top-k chunk embeddings and inverse-distance weighting over a dynamically evolving ontology
Bilge Kaan Karamete, Hunter Casten
cs.CEcs.LGarXiv:2609.00387v12026XVAE-WMT: Explainable Wavelet-Temporal Variational Autoencoder for Blind Source Separation of Heart and Lung Sounds
Yasaman Torabi, Shahram Shirani, James P. Reilly
cs.SDeess.SParXiv:2609.00238v12026XNect: Real-time Multi-Person 3D Motion Capture with a Single RGB Camera
Dushyant Mehta, Oleksandr Sotnychenko, Franziska Mueller +7
cs.CVcs.GRarXiv:1907.00837v22019Real-Time Object Detection Meets DINOv3
Shihua Huang, Yongjie Hou, Longfei Liu +2
cs.CVarXiv:2509.20787v42025Enhanced empirical data for the fundamental diagram and the flow through bottlenecks
A. Seyfried, M. Boltes, J. Kähler +6
physics.soc-phphysics.data-anarXiv:0810.1945v12008Dense Weak Hiding: Closing Complexity Gaps in Nonconvex and PL Finite-Sum Optimization under Individual Smoothness
Yuxing Peng, Zhiqing Tang, Weijia Jia
cs.DScs.LGmath.OCarXiv:2609.00045v12026PhaseLink: A Deep Learning Approach to Seismic Phase Association
Zachary E. Ross, Yisong Yue, Men-Andrin Meier +2
cs.LGphysics.geo-phstat.MLarXiv:1809.02880v22018Attentive Relational Networks for Mapping Images to Scene Graphs
Mengshi Qi, Weijian Li, Zhengyuan Yang +2
cs.CVarXiv:1811.10696v22018Experimental Tests of General Relativity: Recent Progress and Future Directions
Slava G. Turyshev
gr-qcarXiv:0809.3730v22008LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion
Jiangran Lyu, Kai Liu, Xuheng Zhang +20
cs.ROarXiv:2602.12215v22026