Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
19,621 to 19,680 of 61,255
Importance Sampling: Intrinsic Dimension and Computational Cost
S. Agapiou, O. Papaspiliopoulos, D. Sanz-Alonso +1
stat.COarXiv:1511.06196v32015Masked Autoencoders Are Effective Tokenizers for Diffusion Models
Hao Chen, Yujin Han, Fangyi Chen +7
cs.CVcs.AIcs.LGarXiv:2502.03444v22025SPICE: Self-Play In Corpus Environments Improves Reasoning
Bo Liu, Chuanyang Jin, Seungone Kim +7
cs.CLarXiv:2510.24684v12025VisuLogic: A Benchmark for Evaluating Visual Reasoning in Multi-modal Large Language Models
Weiye Xu, Jiahao Wang, Weiyun Wang +10
cs.CVarXiv:2504.15279v12025TWIX: a Two-Stage Approach for End-To-End Named Entity Recognition and Relation Extraction
Marco Martinelli, Laura Menotti
cs.CLarXiv:2609.00832v12026AlphaDrive: Unleashing the Power of VLMs in Autonomous Driving via Reinforcement Learning and Reasoning
Bo Jiang, Shaoyu Chen, Qian Zhang +2
cs.CVcs.ROarXiv:2503.07608v12025Semantic Stereo for Incidental Satellite Images
Marc Bosch, Kevin Foster, Gordon Christie +3
cs.CVarXiv:1811.08739v12018Robust Collaborative Nonnegative Matrix Factorization For Hyperspectral Unmixing (R-CoNMF)
Jun Li, Jose M. Bioucas-Dias, Antonio Plaza +1
math.OCarXiv:1506.04870v12015DeepSeekMath-V2: Towards Self-Verifiable Mathematical Reasoning
Zhihong Shao, Yuxiang Luo, Chengda Lu +6
cs.AIcs.CLarXiv:2511.22570v12025MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding
Siwei Han, Peng Xia, Ruiyi Zhang +4
cs.LGarXiv:2503.13964v12025Text-guided flow matching enables sample-efficient crystal structure generation
Wentao Li
cond-mat.mtrl-scics.AIarXiv:2609.01076v12026Efficient Bayesian Phase Estimation
Nathan Wiebe, Christopher E Granade
quant-pharXiv:1508.00869v12015AdaDepth: Unsupervised Content Congruent Adaptation for Depth Estimation
Jogendra Nath Kundu, Phani Krishna Uppala, Anuj Pahuja +1
cs.CVarXiv:1803.01599v22018Panda Diplomacy: Foundation Model Pre-training across Particle Imaging Detectors for High Energy and Nuclear Physics
Samuel Young, César Jesús-Valls, Kazuhiro Terao
hep-excs.CVarXiv:2609.00611v12026TesserAct: Learning 4D Embodied World Models
Haoyu Zhen, Qiao Sun, Hongxin Zhang +4
cs.CVcs.ROarXiv:2504.20995v12025PixelFlow: Pixel-Space Generative Models with Flow
Shoufa Chen, Chongjian Ge, Shilong Zhang +2
cs.CVarXiv:2504.07963v12025Topological Data Analysis of Biological Aggregation Models
Chad M. Topaz, Lori Ziegelmeier, Tom Halverson
q-bio.QMmath.ATnlin.AOarXiv:1412.6430v32014Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tokens
Yiming Qin, Bomin Wei, Jiaxin Ge +4
cs.CVcs.AIcs.LGarXiv:2511.19418v32025Robust Optimization for Deep Regression
Vasileios Belagiannis, Christian Rupprecht, Gustavo Carneiro +1
cs.CVarXiv:1505.06606v22015Candidate-Expanding Routing with Permutation-Stabilized Experts for Mixed-Format Medical VQA
Hai-Dang Nguyen, Huy-Hieu Pham
cs.CVarXiv:2609.00959v12026MVAN: Multi-View Attention Networks for Fake News Detection on Social Media
Shiwen Ni, Jiawen Li, Hung-Yu Kao
cs.CLarXiv:2506.01627v12025A Few Brief Notes on DeepImpact, COIL, and a Conceptual Framework for Information Retrieval Techniques
Jimmy Lin, Xueguang Ma
cs.IRcs.CLarXiv:2106.14807v12021Cooperative Driving at Unsignalized Intersections Using Tree Search
Huile Xu, Yi Zhang, Li Li +1
cs.MAarXiv:1902.01024v12019Ground Slow, Move Fast: A Dual-System Foundation Model for Generalizable Vision-and-Language Navigation
Meng Wei, Chenyang Wan, Jiaqi Peng +8
cs.ROarXiv:2512.08186v12025Goal-Conditioned Reinforcement Learning with Imagined Subgoals
Elliot Chane-Sane, Cordelia Schmid, Ivan Laptev
cs.LGcs.ROarXiv:2107.00541v12021Adversarial Sticker: A Stealthy Attack Method in the Physical World
Xingxing Wei, Ying Guo, Jie Yu
cs.CVarXiv:2104.06728v22021Ad Headline Generation using Self-Critical Masked Language Model
Yashal Shakti Kanungo, Sumit Negi, Aruna Rajan
cs.CLcs.AIcs.LGarXiv:2607.06818v12026GenONet: A Generative operator Network for High-Resolution Precipitation Nowcasting
Mohammad Kian Golkar, Luciano Alves de Oliveira, Mohammad Khanjani
cs.LGphysics.ao-pharXiv:2609.00544v12026G-Memory: Tracing Hierarchical Memory for Multi-Agent Systems
Guibin Zhang, Muxin Fu, Guancheng Wan +3
cs.MAcs.CLcs.LGarXiv:2506.07398v22025Rapid Prediction of Electron-Ionization Mass Spectrometry using Neural Networks
Jennifer N. Wei, David Belanger, Ryan P. Adams +1
physics.chem-phstat.MLarXiv:1811.08545v22018Cloth-Changing Person Re-identification from A Single Image with Gait Prediction and Regularization
Xin Jin, Tianyu He, Kecheng Zheng +7
cs.CVarXiv:2103.15537v42021Does syntax matter? A strong baseline for Aspect-based Sentiment Analysis with RoBERTa
Junqi Dai, Hang Yan, Tianxiang Sun +2
cs.CLarXiv:2104.04986v12021Fast and Flexible Indoor Scene Synthesis via Deep Convolutional Generative Models
Daniel Ritchie, Kai Wang, Yu-an Lin
cs.CVcs.GRarXiv:1811.12463v12018D$^2$NeRF: Self-Supervised Decoupling of Dynamic and Static Objects from a Monocular Video
Tianhao Wu, Fangcheng Zhong, Andrea Tagliasacchi +2
cs.CVarXiv:2205.15838v42022Cryo-CARE: Content-Aware Image Restoration for Cryo-Transmission Electron Microscopy Data
Tim-Oliver Buchholz, Mareike Jordan, Gaia Pigino +1
cs.CVcs.LGarXiv:1810.05420v22018RGB-T Semantic Segmentation with Location, Activation, and Sharpening
Gongyang Li, Yike Wang, Zhi Liu +2
cs.CVarXiv:2210.14530v12022LRW-1000: A Naturally-Distributed Large-Scale Benchmark for Lip Reading in the Wild
Shuang Yang, Yuanhang Zhang, Dalu Feng +6
cs.CVarXiv:1810.06990v62018SCoNE: Selective Context-aware Neuron Editing for Robust Retrieval-Augmented Generation
Chaewon Kim, Seo Yeon Park
cs.CLarXiv:2609.00689v12026Apple Intelligence Foundation Language Models: Tech Report 2025
Ethan Li, Anders Boesen Lindbo Larsen, Chen Zhang +395
cs.LGcs.AIarXiv:2507.13575v32025TempCloze: Can Video-LLMs Identify the Missing Middle?
Wenqi Pei, Henry Hengyuan Zhao, Yilai Liu +4
cs.CVcs.AIarXiv:2609.01515v12026Breaking the Modality Barrier: Universal Embedding Learning with Multimodal LLMs
Tiancheng Gu, Kaicheng Yang, Ziyong Feng +6
cs.CVarXiv:2504.17432v42025Mapping Language to Code in Programmatic Context
Srinivasan Iyer, Ioannis Konstas, Alvin Cheung +1
cs.CLarXiv:1808.09588v12018OFDM-guided Deep Joint Source Channel Coding for Wireless Multipath Fading Channels
Mingyu Yang, Chenghong Bian, Hun-Seok Kim
eess.SParXiv:2109.05194v12021Domino: Discovering Systematic Errors with Cross-Modal Embeddings
Sabri Eyuboglu, Maya Varma, Khaled Saab +5
cs.LGcs.AIarXiv:2203.14960v32022Array Gain for Pinching-Antenna Systems (PASS)
Chongjun Ouyang, Zhaolin Wang, Yuanwei Liu +1
eess.SParXiv:2501.05657v22025Solving 3D Inverse Problems using Pre-trained 2D Diffusion Models
Hyungjin Chung, Dohoon Ryu, Michael T. McCann +2
cs.CVcs.AIcs.LGarXiv:2211.10655v12022Layered Neural Atlases for Consistent Video Editing
Yoni Kasten, Dolev Ofri, Oliver Wang +1
cs.CVcs.GRarXiv:2109.11418v12021MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations
Kaixuan Huang, Jiacheng Guo, Zihao Li +15
cs.LGcs.AIcs.CLarXiv:2502.06453v22025Personalized Transformer for Explainable Recommendation
Lei Li, Yongfeng Zhang, Li Chen
cs.IRcs.AIcs.CLarXiv:2105.11601v22021Adverse Events in Robotic Surgery: A Retrospective Study of 14 Years of FDA Data
Homa Alemzadeh, Ravishankar K. Iyer, Zbigniew Kalbarczyk +2
cs.ROcs.CRarXiv:1507.03518v22015ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific Research
Wanghan Xu, Shuo Li, Tianlin Ye +48
cs.LGcs.AIcs.CLarXiv:2606.07591v52026Likelihood-informed dimension reduction for nonlinear inverse problems
Tiangang Cui, James Martin, Youssef M. Marzouk +2
stat.COmath.NAstat.MEarXiv:1403.4680v22014Closing the Verification Loop: Self-Check Captioning for Long-Paragraph Detailed Audio Captioning
Fengji Ma, Yan Rong, Xu Li +3
cs.SDarXiv:2608.30713v12026Towards End-to-End Automation of AI Research
Yutaro Yamada, Robert Tjarko Lange, Cong Lu +5
cs.AIarXiv:2606.15497v12026SurgSkill-Bench: A Benchmark for Multimodal Surgical Skill Assessment
Chaohui Dang, Zheheng Jiang, James Glasbey +3
cs.CVarXiv:2608.30872v12026Image Hijacks: Adversarial Images can Control Generative Models at Runtime
Luke Bailey, Euan Ong, Stuart Russell +1
cs.LGcs.CLcs.CRarXiv:2309.00236v42023ViSAR: Training-Free Adaptive-$k$ Retrieval for Visual Document Question Answering
Adrien Mialland, Marc Plantevit, Julien Gallois +1
cs.IRcs.AIcs.CLarXiv:2609.02486v12026Narrative-Driven Paper-to-Slide Generation via ArcDeck
Tarik Can Ozden, Sachidanand VS, Furkan Horoz +3
cs.AIarXiv:2604.11969v12026Efficient Reinforcement Finetuning via Adaptive Curriculum Learning
Taiwei Shi, Yiyang Wu, Linxin Song +2
cs.LGcs.CLarXiv:2504.05520v42025NorMuon: Making Muon more efficient and scalable
Zichong Li, Liming Liu, Chen Liang +2
cs.LGcs.CLarXiv:2510.05491v12025