Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
57,301 to 57,360 of 61,306
VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-Training
Zhan Tong, Yibing Song, Jue Wang +1
cs.CVarXiv:2203.12602v32022Streaming Communication in Multi-Agent Reasoning
Zhen Yang, Xiaogang Xu, Wen Wang +3
cs.CLcs.AIcs.MAarXiv:2606.05158v22026Audio Interaction Model
Zhifei Xie, Zihang Liu, Ze An +8
cs.SDcs.AIcs.CLarXiv:2606.05121v12026Measuring Model Robustness via Fisher Information: Spectral Bounds, Theoretical Guarantees, and Practical Algorithms
Chong Zhang, Xiang Li, Jia Wang +2
cs.LGcs.CVarXiv:2606.04767v12026BRepCLIP: Contrastive Multimodal Pretraining on BRep Primitives for CAD Understanding
Muhammad Usama, Didier Stricker, Mohammad Sadil Khan +1
cs.CVarXiv:2606.05515v12026Bridging the Gap Between Anchor-based and Anchor-free Detection via Adaptive Training Sample Selection
Shifeng Zhang, Cheng Chi, Yongqiang Yao +2
cs.CVarXiv:1912.02424v42019TIDE: Proactive Multi-Problem Discovery via Template-Guided Iteration
Soyeong Jeong, Jinheon Baek, Minki Kang +1
cs.CLcs.AIcs.LGarXiv:2606.04743v22026To Explain or to Predict?
Galit Shmueli
stat.MEarXiv:1101.0891v12011CIPER: A Unified Framework for Cross-view Image-retrieval and Pose-estimation
Yurim Jeon, Dongseong Seo, Seung-Woo Seo
cs.CVcs.ROarXiv:2606.05011v12026Weight Uncertainty in Neural Networks
Charles Blundell, Julien Cornebise, Koray Kavukcuoglu +1
stat.MLcs.LGarXiv:1505.05424v22015CrossViT: Cross-Attention Multi-Scale Vision Transformer for Image Classification
Chun-Fu Chen, Quanfu Fan, Rameswar Panda
cs.CVarXiv:2103.14899v22021Open3D: A Modern Library for 3D Data Processing
Qian-Yi Zhou, Jaesik Park, Vladlen Koltun
cs.CVcs.GRcs.ROarXiv:1801.09847v12018What Should Agents Say? Action-state Communication for Efficient Multi-Agent Systems
Chen Huang, Yuhao Wu, Wenxuan Zhang
cs.AIarXiv:2606.05304v12026Designing Network Design Spaces
Ilija Radosavovic, Raj Prateek Kosaraju, Ross Girshick +2
cs.CVcs.LGarXiv:2003.13678v12020Learning Face Representation from Scratch
Dong Yi, Zhen Lei, Shengcai Liao +1
cs.CVarXiv:1411.7923v12014Discrete-WAM: Unified Discrete Vision-Action Token Editing for World-Policy Learning
Ziyang Yao, Haochen Liu, Yuncheng Jiang +10
cs.ROarXiv:2606.05645v22026MS-Celeb-1M: A Dataset and Benchmark for Large-Scale Face Recognition
Yandong Guo, Lei Zhang, Yuxiao Hu +2
cs.CVarXiv:1607.08221v12016Revising Context, Shifting Simulated Stance: Auditing LLM-Based Stance Simulation in Online Discussions
Xinnong Zhang, Wanting Shan, Hanjia Lyu +2
cs.CLcs.MMcs.SIarXiv:2606.06443v22026Self-supervised Learning: Generative or Contrastive
Xiao Liu, Fanjin Zhang, Zhenyu Hou +4
cs.LGstat.MLarXiv:2006.08218v52020Deep metric learning using Triplet network
Elad Hoffer, Nir Ailon
cs.LGcs.CVstat.MLarXiv:1412.6622v42014Towards Truly Multilingual ASR: Generalizing Code-Switching ASR to Unseen Language Pairs
Gio Paik, Hyunseo Shin, Soungmin Lee
cs.CLeess.ASarXiv:2606.05846v22026Deep Learning for Person Re-identification: A Survey and Outlook
Mang Ye, Jianbing Shen, Gaojie Lin +3
cs.CVarXiv:2001.04193v22020Object Detection in Optical Remote Sensing Images: A Survey and A New Benchmark
Ke Li, Gang Wan, Gong Cheng +2
cs.CVarXiv:1909.00133v22019ZOO: Zeroth Order Optimization based Black-box Attacks to Deep Neural Networks without Training Substitute Models
Pin-Yu Chen, Huan Zhang, Yash Sharma +2
stat.MLcs.CRcs.LGarXiv:1708.03999v22017Reluplex: An Efficient SMT Solver for Verifying Deep Neural Networks
Guy Katz, Clark Barrett, David Dill +2
cs.AIcs.LOarXiv:1702.01135v22017Dream to Control: Learning Behaviors by Latent Imagination
Danijar Hafner, Timothy Lillicrap, Jimmy Ba +1
cs.LGcs.AIcs.ROarXiv:1912.01603v32019MICADO: the E-ELT Adaptive Optics Imaging Camera
R. Davies, MICADO Team
astro-ph.IMarXiv:1005.5009v12010SePO: Self-Evolving Prompt Agent for System Prompt Optimization
Wangcheng Tao, Han Wu, Weng-Fai Wong
cs.CLcs.AIarXiv:2606.04465v12026Personal AI Agent for Camera Roll VQA
Thao Nguyen, Krishna Kumar Singh, Donghyun Kim +2
cs.CVcs.AIarXiv:2606.05275v12026A Survey on Explainable Artificial Intelligence (XAI): Towards Medical XAI
Erico Tjoa, Cuntai Guan
cs.LGcs.AIarXiv:1907.07374v52019Aleatoric and Epistemic Uncertainty in Machine Learning: An Introduction to Concepts and Methods
Eyke Hüllermeier, Willem Waegeman
cs.LGstat.MLarXiv:1910.09457v32019GENEB: Why Genomic Models Are Hard to Compare
Daria Ledneva, Mikhail Nuridinov, Denis Kuznetsov
cs.CLcs.LGq-bio.GNarXiv:2606.04525v42026Imaginative Perception Tokens Enhance Spatial Reasoning in Multimodal Language Models
Mahtab Bigverdi, Linjie Li, Weikai Huang +8
cs.AIarXiv:2606.03988v32026SpanBERT: Improving Pre-training by Representing and Predicting Spans
Mandar Joshi, Danqi Chen, Yinhan Liu +3
cs.CLcs.LGarXiv:1907.10529v32019Statistically Reliable LLM-Based Ranking Evaluation via Prediction-Powered Inference
Abhishek Divekar
cs.LGcs.AIcs.CLarXiv:2606.05308v12026Improvements to the APBS biomolecular solvation software suite
Elizabeth Jurrus, Dave Engel, Keith Star +21
q-bio.BMarXiv:1707.00027v22017ForeSci: Evaluating LLM Agents for Forward-Looking AI Research Judgment
Qiuyu Tian, Haojie Yin, Yingce Xia +2
cs.AIarXiv:2606.00644v22026AURA: Intent-Directed Probing for Implicit-Need Surfacing in Situated LLM Agents
Yang Li, Jiaxiang Liu, Jiang Cai +1
cs.CLarXiv:2606.05557v12026Mixed membership stochastic blockmodels
Edoardo M Airoldi, David M Blei, Stephen E Fienberg +1
stat.MEcs.LGmath.STarXiv:0705.4485v12007Supervised Learning of Universal Sentence Representations from Natural Language Inference Data
Alexis Conneau, Douwe Kiela, Holger Schwenk +2
cs.CLarXiv:1705.02364v52017Learning Geometric Representations from Videos for Spatial Intelligent Multimodal Large Language Models
Haibo Wang, Lifu Huang
cs.CVcs.AIarXiv:2606.05833v22026Flower Pollination Algorithm for Global Optimization
Xin-She Yang
math.OCcs.NEnlin.AOarXiv:1312.5673v12013Tabular Data: Deep Learning is Not All You Need
Ravid Shwartz-Ziv, Amitai Armon
cs.LGarXiv:2106.03253v22021LLMs Can Leak Training Data But Do They Want To? A Propensity-Aware Evaluation of Memorization in LLMs
Gianluca Barmina, Peter Schneider-Kamp, Lukas Galke Poech
cs.CLcs.AIarXiv:2606.06286v12026RePaint: Inpainting using Denoising Diffusion Probabilistic Models
Andreas Lugmayr, Martin Danelljan, Andres Romero +3
cs.CVarXiv:2201.09865v42022Oscar: Object-Semantics Aligned Pre-training for Vision-Language Tasks
Xiujun Li, Xi Yin, Chunyuan Li +9
cs.CVcs.CLcs.IRarXiv:2004.06165v52020Kernel methods in machine learning
Thomas Hofmann, Bernhard Schölkopf, Alexander J. Smola
math.STmath.PRarXiv:math/0701907v32007Learning to learn by gradient descent by gradient descent
Marcin Andrychowicz, Misha Denil, Sergio Gomez +5
cs.NEcs.LGarXiv:1606.04474v22016LLM Explainability with Counterfactual Chains and Causal Graphs
Nirit Nussbaum-Hoffer, Nitay Calderon, Liat Ein-Dor +1
cs.LGarXiv:2606.05972v12026AffectNet: A Database for Facial Expression, Valence, and Arousal Computing in the Wild
Ali Mollahosseini, Behzad Hasani, Mohammad H. Mahoor
cs.CVarXiv:1708.03985v42017A Closer Look at Memorization in Deep Networks
Devansh Arpit, Stanisław Jastrzębski, Nicolas Ballas +8
stat.MLcs.LGarXiv:1706.05394v22017Benchmarking Single Image Dehazing and Beyond
Boyi Li, Wenqi Ren, Dengpan Fu +4
cs.CVcs.AIcs.LGarXiv:1712.04143v42017SoCRATES: Towards Reliable Automated Evaluation of Proactive LLM Mediation across Domains and Socio-cognitive Variations
Taewon Yun, Hyeonseong Park, Jeonghwan Choi +3
cs.AIcs.CLarXiv:2606.05563v12026Cosine Misleads: Auxiliary Losses Reshape Vision Language Models, Not Their Latents
XiuYu Zhang, Junfeng Fang, Zhenkai Liang
cs.CVarXiv:2606.05753v12026DRIFT: A Residual Flow Adapter for Decoding Continuous Outputs in Vision-Language Models
Zhuoming Liu, Jinhong Lin, Kwan Man Cheng +3
cs.CVcs.AIcs.LGarXiv:2606.05758v12026European Union regulations on algorithmic decision-making and a "right to explanation"
Bryce Goodman, Seth Flaxman
stat.MLcs.CYcs.LGarXiv:1606.08813v32016In-Context Multiple Instance Learning
Alexander Möllers, Marvin Sextro, Julius Hense +2
cs.LGcs.AIcs.CVarXiv:2606.06458v12026CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
Zhuoyi Yang, Jiayan Teng, Wendi Zheng +15
cs.CVarXiv:2408.06072v32024Complexity-Balanced Diffusion Splitting
Noam Issachar, Dani Lischinski, Raanan Fattal
cs.CVarXiv:2606.06477v12026PCT: Point cloud transformer
Meng-Hao Guo, Jun-Xiong Cai, Zheng-Ning Liu +3
cs.CVarXiv:2012.09688v42020