Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

57,301 to 57,360 of 61,306

  1. VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-Training

    Zhan Tong, Yibing Song, Jue Wang +1

    cs.CVarXiv:2203.12602v32022
  2. Streaming Communication in Multi-Agent Reasoning

    Zhen Yang, Xiaogang Xu, Wen Wang +3

    cs.CLcs.AIcs.MAarXiv:2606.05158v22026
  3. Audio Interaction Model

    Zhifei Xie, Zihang Liu, Ze An +8

    cs.SDcs.AIcs.CLarXiv:2606.05121v12026
  4. Measuring Model Robustness via Fisher Information: Spectral Bounds, Theoretical Guarantees, and Practical Algorithms

    Chong Zhang, Xiang Li, Jia Wang +2

    cs.LGcs.CVarXiv:2606.04767v12026
  5. BRepCLIP: Contrastive Multimodal Pretraining on BRep Primitives for CAD Understanding

    Muhammad Usama, Didier Stricker, Mohammad Sadil Khan +1

    cs.CVarXiv:2606.05515v12026
  6. Bridging the Gap Between Anchor-based and Anchor-free Detection via Adaptive Training Sample Selection

    Shifeng Zhang, Cheng Chi, Yongqiang Yao +2

    cs.CVarXiv:1912.02424v42019
  7. TIDE: Proactive Multi-Problem Discovery via Template-Guided Iteration

    Soyeong Jeong, Jinheon Baek, Minki Kang +1

    cs.CLcs.AIcs.LGarXiv:2606.04743v22026
  8. To Explain or to Predict?

    Galit Shmueli

    stat.MEarXiv:1101.0891v12011
  9. CIPER: A Unified Framework for Cross-view Image-retrieval and Pose-estimation

    Yurim Jeon, Dongseong Seo, Seung-Woo Seo

    cs.CVcs.ROarXiv:2606.05011v12026
  10. Weight Uncertainty in Neural Networks

    Charles Blundell, Julien Cornebise, Koray Kavukcuoglu +1

    stat.MLcs.LGarXiv:1505.05424v22015
  11. CrossViT: Cross-Attention Multi-Scale Vision Transformer for Image Classification

    Chun-Fu Chen, Quanfu Fan, Rameswar Panda

    cs.CVarXiv:2103.14899v22021
  12. Open3D: A Modern Library for 3D Data Processing

    Qian-Yi Zhou, Jaesik Park, Vladlen Koltun

    cs.CVcs.GRcs.ROarXiv:1801.09847v12018
  13. What Should Agents Say? Action-state Communication for Efficient Multi-Agent Systems

    Chen Huang, Yuhao Wu, Wenxuan Zhang

    cs.AIarXiv:2606.05304v12026
  14. Designing Network Design Spaces

    Ilija Radosavovic, Raj Prateek Kosaraju, Ross Girshick +2

    cs.CVcs.LGarXiv:2003.13678v12020
  15. Learning Face Representation from Scratch

    Dong Yi, Zhen Lei, Shengcai Liao +1

    cs.CVarXiv:1411.7923v12014
  16. Discrete-WAM: Unified Discrete Vision-Action Token Editing for World-Policy Learning

    Ziyang Yao, Haochen Liu, Yuncheng Jiang +10

    cs.ROarXiv:2606.05645v22026
  17. MS-Celeb-1M: A Dataset and Benchmark for Large-Scale Face Recognition

    Yandong Guo, Lei Zhang, Yuxiao Hu +2

    cs.CVarXiv:1607.08221v12016
  18. Revising Context, Shifting Simulated Stance: Auditing LLM-Based Stance Simulation in Online Discussions

    Xinnong Zhang, Wanting Shan, Hanjia Lyu +2

    cs.CLcs.MMcs.SIarXiv:2606.06443v22026
  19. Self-supervised Learning: Generative or Contrastive

    Xiao Liu, Fanjin Zhang, Zhenyu Hou +4

    cs.LGstat.MLarXiv:2006.08218v52020
  20. Deep metric learning using Triplet network

    Elad Hoffer, Nir Ailon

    cs.LGcs.CVstat.MLarXiv:1412.6622v42014
  21. Towards Truly Multilingual ASR: Generalizing Code-Switching ASR to Unseen Language Pairs

    Gio Paik, Hyunseo Shin, Soungmin Lee

    cs.CLeess.ASarXiv:2606.05846v22026
  22. Deep Learning for Person Re-identification: A Survey and Outlook

    Mang Ye, Jianbing Shen, Gaojie Lin +3

    cs.CVarXiv:2001.04193v22020
  23. Object Detection in Optical Remote Sensing Images: A Survey and A New Benchmark

    Ke Li, Gang Wan, Gong Cheng +2

    cs.CVarXiv:1909.00133v22019
  24. ZOO: Zeroth Order Optimization based Black-box Attacks to Deep Neural Networks without Training Substitute Models

    Pin-Yu Chen, Huan Zhang, Yash Sharma +2

    stat.MLcs.CRcs.LGarXiv:1708.03999v22017
  25. Reluplex: An Efficient SMT Solver for Verifying Deep Neural Networks

    Guy Katz, Clark Barrett, David Dill +2

    cs.AIcs.LOarXiv:1702.01135v22017
  26. Dream to Control: Learning Behaviors by Latent Imagination

    Danijar Hafner, Timothy Lillicrap, Jimmy Ba +1

    cs.LGcs.AIcs.ROarXiv:1912.01603v32019
  27. MICADO: the E-ELT Adaptive Optics Imaging Camera

    R. Davies, MICADO Team

    astro-ph.IMarXiv:1005.5009v12010
  28. SePO: Self-Evolving Prompt Agent for System Prompt Optimization

    Wangcheng Tao, Han Wu, Weng-Fai Wong

    cs.CLcs.AIarXiv:2606.04465v12026
  29. Personal AI Agent for Camera Roll VQA

    Thao Nguyen, Krishna Kumar Singh, Donghyun Kim +2

    cs.CVcs.AIarXiv:2606.05275v12026
  30. A Survey on Explainable Artificial Intelligence (XAI): Towards Medical XAI

    Erico Tjoa, Cuntai Guan

    cs.LGcs.AIarXiv:1907.07374v52019
  31. Aleatoric and Epistemic Uncertainty in Machine Learning: An Introduction to Concepts and Methods

    Eyke Hüllermeier, Willem Waegeman

    cs.LGstat.MLarXiv:1910.09457v32019
  32. GENEB: Why Genomic Models Are Hard to Compare

    Daria Ledneva, Mikhail Nuridinov, Denis Kuznetsov

    cs.CLcs.LGq-bio.GNarXiv:2606.04525v42026
  33. Imaginative Perception Tokens Enhance Spatial Reasoning in Multimodal Language Models

    Mahtab Bigverdi, Linjie Li, Weikai Huang +8

    cs.AIarXiv:2606.03988v32026
  34. SpanBERT: Improving Pre-training by Representing and Predicting Spans

    Mandar Joshi, Danqi Chen, Yinhan Liu +3

    cs.CLcs.LGarXiv:1907.10529v32019
  35. Statistically Reliable LLM-Based Ranking Evaluation via Prediction-Powered Inference

    Abhishek Divekar

    cs.LGcs.AIcs.CLarXiv:2606.05308v12026
  36. Improvements to the APBS biomolecular solvation software suite

    Elizabeth Jurrus, Dave Engel, Keith Star +21

    q-bio.BMarXiv:1707.00027v22017
  37. ForeSci: Evaluating LLM Agents for Forward-Looking AI Research Judgment

    Qiuyu Tian, Haojie Yin, Yingce Xia +2

    cs.AIarXiv:2606.00644v22026
  38. AURA: Intent-Directed Probing for Implicit-Need Surfacing in Situated LLM Agents

    Yang Li, Jiaxiang Liu, Jiang Cai +1

    cs.CLarXiv:2606.05557v12026
  39. Mixed membership stochastic blockmodels

    Edoardo M Airoldi, David M Blei, Stephen E Fienberg +1

    stat.MEcs.LGmath.STarXiv:0705.4485v12007
  40. Supervised Learning of Universal Sentence Representations from Natural Language Inference Data

    Alexis Conneau, Douwe Kiela, Holger Schwenk +2

    cs.CLarXiv:1705.02364v52017
  41. Learning Geometric Representations from Videos for Spatial Intelligent Multimodal Large Language Models

    Haibo Wang, Lifu Huang

    cs.CVcs.AIarXiv:2606.05833v22026
  42. Flower Pollination Algorithm for Global Optimization

    Xin-She Yang

    math.OCcs.NEnlin.AOarXiv:1312.5673v12013
  43. Tabular Data: Deep Learning is Not All You Need

    Ravid Shwartz-Ziv, Amitai Armon

    cs.LGarXiv:2106.03253v22021
  44. LLMs Can Leak Training Data But Do They Want To? A Propensity-Aware Evaluation of Memorization in LLMs

    Gianluca Barmina, Peter Schneider-Kamp, Lukas Galke Poech

    cs.CLcs.AIarXiv:2606.06286v12026
  45. RePaint: Inpainting using Denoising Diffusion Probabilistic Models

    Andreas Lugmayr, Martin Danelljan, Andres Romero +3

    cs.CVarXiv:2201.09865v42022
  46. Oscar: Object-Semantics Aligned Pre-training for Vision-Language Tasks

    Xiujun Li, Xi Yin, Chunyuan Li +9

    cs.CVcs.CLcs.IRarXiv:2004.06165v52020
  47. Kernel methods in machine learning

    Thomas Hofmann, Bernhard Schölkopf, Alexander J. Smola

    math.STmath.PRarXiv:math/0701907v32007
  48. Learning to learn by gradient descent by gradient descent

    Marcin Andrychowicz, Misha Denil, Sergio Gomez +5

    cs.NEcs.LGarXiv:1606.04474v22016
  49. LLM Explainability with Counterfactual Chains and Causal Graphs

    Nirit Nussbaum-Hoffer, Nitay Calderon, Liat Ein-Dor +1

    cs.LGarXiv:2606.05972v12026
  50. AffectNet: A Database for Facial Expression, Valence, and Arousal Computing in the Wild

    Ali Mollahosseini, Behzad Hasani, Mohammad H. Mahoor

    cs.CVarXiv:1708.03985v42017
  51. A Closer Look at Memorization in Deep Networks

    Devansh Arpit, Stanisław Jastrzębski, Nicolas Ballas +8

    stat.MLcs.LGarXiv:1706.05394v22017
  52. Benchmarking Single Image Dehazing and Beyond

    Boyi Li, Wenqi Ren, Dengpan Fu +4

    cs.CVcs.AIcs.LGarXiv:1712.04143v42017
  53. SoCRATES: Towards Reliable Automated Evaluation of Proactive LLM Mediation across Domains and Socio-cognitive Variations

    Taewon Yun, Hyeonseong Park, Jeonghwan Choi +3

    cs.AIcs.CLarXiv:2606.05563v12026
  54. Cosine Misleads: Auxiliary Losses Reshape Vision Language Models, Not Their Latents

    XiuYu Zhang, Junfeng Fang, Zhenkai Liang

    cs.CVarXiv:2606.05753v12026
  55. DRIFT: A Residual Flow Adapter for Decoding Continuous Outputs in Vision-Language Models

    Zhuoming Liu, Jinhong Lin, Kwan Man Cheng +3

    cs.CVcs.AIcs.LGarXiv:2606.05758v12026
  56. European Union regulations on algorithmic decision-making and a "right to explanation"

    Bryce Goodman, Seth Flaxman

    stat.MLcs.CYcs.LGarXiv:1606.08813v32016
  57. In-Context Multiple Instance Learning

    Alexander Möllers, Marvin Sextro, Julius Hense +2

    cs.LGcs.AIcs.CVarXiv:2606.06458v12026
  58. CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

    Zhuoyi Yang, Jiayan Teng, Wendi Zheng +15

    cs.CVarXiv:2408.06072v32024
  59. Complexity-Balanced Diffusion Splitting

    Noam Issachar, Dani Lischinski, Raanan Fattal

    cs.CVarXiv:2606.06477v12026
  60. PCT: Point cloud transformer

    Meng-Hao Guo, Jun-Xiong Cai, Zheng-Ning Liu +3

    cs.CVarXiv:2012.09688v42020