Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

19,621 to 19,680 of 61,255

  1. Importance Sampling: Intrinsic Dimension and Computational Cost

    S. Agapiou, O. Papaspiliopoulos, D. Sanz-Alonso +1

    stat.COarXiv:1511.06196v32015
  2. Masked Autoencoders Are Effective Tokenizers for Diffusion Models

    Hao Chen, Yujin Han, Fangyi Chen +7

    cs.CVcs.AIcs.LGarXiv:2502.03444v22025
  3. SPICE: Self-Play In Corpus Environments Improves Reasoning

    Bo Liu, Chuanyang Jin, Seungone Kim +7

    cs.CLarXiv:2510.24684v12025
  4. VisuLogic: A Benchmark for Evaluating Visual Reasoning in Multi-modal Large Language Models

    Weiye Xu, Jiahao Wang, Weiyun Wang +10

    cs.CVarXiv:2504.15279v12025
  5. TWIX: a Two-Stage Approach for End-To-End Named Entity Recognition and Relation Extraction

    Marco Martinelli, Laura Menotti

    cs.CLarXiv:2609.00832v12026
  6. AlphaDrive: Unleashing the Power of VLMs in Autonomous Driving via Reinforcement Learning and Reasoning

    Bo Jiang, Shaoyu Chen, Qian Zhang +2

    cs.CVcs.ROarXiv:2503.07608v12025
  7. Semantic Stereo for Incidental Satellite Images

    Marc Bosch, Kevin Foster, Gordon Christie +3

    cs.CVarXiv:1811.08739v12018
  8. Robust Collaborative Nonnegative Matrix Factorization For Hyperspectral Unmixing (R-CoNMF)

    Jun Li, Jose M. Bioucas-Dias, Antonio Plaza +1

    math.OCarXiv:1506.04870v12015
  9. DeepSeekMath-V2: Towards Self-Verifiable Mathematical Reasoning

    Zhihong Shao, Yuxiang Luo, Chengda Lu +6

    cs.AIcs.CLarXiv:2511.22570v12025
  10. MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

    Siwei Han, Peng Xia, Ruiyi Zhang +4

    cs.LGarXiv:2503.13964v12025
  11. Text-guided flow matching enables sample-efficient crystal structure generation

    Wentao Li

    cond-mat.mtrl-scics.AIarXiv:2609.01076v12026
  12. Efficient Bayesian Phase Estimation

    Nathan Wiebe, Christopher E Granade

    quant-pharXiv:1508.00869v12015
  13. AdaDepth: Unsupervised Content Congruent Adaptation for Depth Estimation

    Jogendra Nath Kundu, Phani Krishna Uppala, Anuj Pahuja +1

    cs.CVarXiv:1803.01599v22018
  14. Panda Diplomacy: Foundation Model Pre-training across Particle Imaging Detectors for High Energy and Nuclear Physics

    Samuel Young, César Jesús-Valls, Kazuhiro Terao

    hep-excs.CVarXiv:2609.00611v12026
  15. TesserAct: Learning 4D Embodied World Models

    Haoyu Zhen, Qiao Sun, Hongxin Zhang +4

    cs.CVcs.ROarXiv:2504.20995v12025
  16. PixelFlow: Pixel-Space Generative Models with Flow

    Shoufa Chen, Chongjian Ge, Shilong Zhang +2

    cs.CVarXiv:2504.07963v12025
  17. Topological Data Analysis of Biological Aggregation Models

    Chad M. Topaz, Lori Ziegelmeier, Tom Halverson

    q-bio.QMmath.ATnlin.AOarXiv:1412.6430v32014
  18. Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tokens

    Yiming Qin, Bomin Wei, Jiaxin Ge +4

    cs.CVcs.AIcs.LGarXiv:2511.19418v32025
  19. Robust Optimization for Deep Regression

    Vasileios Belagiannis, Christian Rupprecht, Gustavo Carneiro +1

    cs.CVarXiv:1505.06606v22015
  20. Candidate-Expanding Routing with Permutation-Stabilized Experts for Mixed-Format Medical VQA

    Hai-Dang Nguyen, Huy-Hieu Pham

    cs.CVarXiv:2609.00959v12026
  21. MVAN: Multi-View Attention Networks for Fake News Detection on Social Media

    Shiwen Ni, Jiawen Li, Hung-Yu Kao

    cs.CLarXiv:2506.01627v12025
  22. A Few Brief Notes on DeepImpact, COIL, and a Conceptual Framework for Information Retrieval Techniques

    Jimmy Lin, Xueguang Ma

    cs.IRcs.CLarXiv:2106.14807v12021
  23. Cooperative Driving at Unsignalized Intersections Using Tree Search

    Huile Xu, Yi Zhang, Li Li +1

    cs.MAarXiv:1902.01024v12019
  24. Ground Slow, Move Fast: A Dual-System Foundation Model for Generalizable Vision-and-Language Navigation

    Meng Wei, Chenyang Wan, Jiaqi Peng +8

    cs.ROarXiv:2512.08186v12025
  25. Goal-Conditioned Reinforcement Learning with Imagined Subgoals

    Elliot Chane-Sane, Cordelia Schmid, Ivan Laptev

    cs.LGcs.ROarXiv:2107.00541v12021
  26. Adversarial Sticker: A Stealthy Attack Method in the Physical World

    Xingxing Wei, Ying Guo, Jie Yu

    cs.CVarXiv:2104.06728v22021
  27. Ad Headline Generation using Self-Critical Masked Language Model

    Yashal Shakti Kanungo, Sumit Negi, Aruna Rajan

    cs.CLcs.AIcs.LGarXiv:2607.06818v12026
  28. GenONet: A Generative operator Network for High-Resolution Precipitation Nowcasting

    Mohammad Kian Golkar, Luciano Alves de Oliveira, Mohammad Khanjani

    cs.LGphysics.ao-pharXiv:2609.00544v12026
  29. G-Memory: Tracing Hierarchical Memory for Multi-Agent Systems

    Guibin Zhang, Muxin Fu, Guancheng Wan +3

    cs.MAcs.CLcs.LGarXiv:2506.07398v22025
  30. Rapid Prediction of Electron-Ionization Mass Spectrometry using Neural Networks

    Jennifer N. Wei, David Belanger, Ryan P. Adams +1

    physics.chem-phstat.MLarXiv:1811.08545v22018
  31. Cloth-Changing Person Re-identification from A Single Image with Gait Prediction and Regularization

    Xin Jin, Tianyu He, Kecheng Zheng +7

    cs.CVarXiv:2103.15537v42021
  32. Does syntax matter? A strong baseline for Aspect-based Sentiment Analysis with RoBERTa

    Junqi Dai, Hang Yan, Tianxiang Sun +2

    cs.CLarXiv:2104.04986v12021
  33. Fast and Flexible Indoor Scene Synthesis via Deep Convolutional Generative Models

    Daniel Ritchie, Kai Wang, Yu-an Lin

    cs.CVcs.GRarXiv:1811.12463v12018
  34. D$^2$NeRF: Self-Supervised Decoupling of Dynamic and Static Objects from a Monocular Video

    Tianhao Wu, Fangcheng Zhong, Andrea Tagliasacchi +2

    cs.CVarXiv:2205.15838v42022
  35. Cryo-CARE: Content-Aware Image Restoration for Cryo-Transmission Electron Microscopy Data

    Tim-Oliver Buchholz, Mareike Jordan, Gaia Pigino +1

    cs.CVcs.LGarXiv:1810.05420v22018
  36. RGB-T Semantic Segmentation with Location, Activation, and Sharpening

    Gongyang Li, Yike Wang, Zhi Liu +2

    cs.CVarXiv:2210.14530v12022
  37. LRW-1000: A Naturally-Distributed Large-Scale Benchmark for Lip Reading in the Wild

    Shuang Yang, Yuanhang Zhang, Dalu Feng +6

    cs.CVarXiv:1810.06990v62018
  38. SCoNE: Selective Context-aware Neuron Editing for Robust Retrieval-Augmented Generation

    Chaewon Kim, Seo Yeon Park

    cs.CLarXiv:2609.00689v12026
  39. Apple Intelligence Foundation Language Models: Tech Report 2025

    Ethan Li, Anders Boesen Lindbo Larsen, Chen Zhang +395

    cs.LGcs.AIarXiv:2507.13575v32025
  40. TempCloze: Can Video-LLMs Identify the Missing Middle?

    Wenqi Pei, Henry Hengyuan Zhao, Yilai Liu +4

    cs.CVcs.AIarXiv:2609.01515v12026
  41. Breaking the Modality Barrier: Universal Embedding Learning with Multimodal LLMs

    Tiancheng Gu, Kaicheng Yang, Ziyong Feng +6

    cs.CVarXiv:2504.17432v42025
  42. Mapping Language to Code in Programmatic Context

    Srinivasan Iyer, Ioannis Konstas, Alvin Cheung +1

    cs.CLarXiv:1808.09588v12018
  43. OFDM-guided Deep Joint Source Channel Coding for Wireless Multipath Fading Channels

    Mingyu Yang, Chenghong Bian, Hun-Seok Kim

    eess.SParXiv:2109.05194v12021
  44. Domino: Discovering Systematic Errors with Cross-Modal Embeddings

    Sabri Eyuboglu, Maya Varma, Khaled Saab +5

    cs.LGcs.AIarXiv:2203.14960v32022
  45. Array Gain for Pinching-Antenna Systems (PASS)

    Chongjun Ouyang, Zhaolin Wang, Yuanwei Liu +1

    eess.SParXiv:2501.05657v22025
  46. Solving 3D Inverse Problems using Pre-trained 2D Diffusion Models

    Hyungjin Chung, Dohoon Ryu, Michael T. McCann +2

    cs.CVcs.AIcs.LGarXiv:2211.10655v12022
  47. Layered Neural Atlases for Consistent Video Editing

    Yoni Kasten, Dolev Ofri, Oliver Wang +1

    cs.CVcs.GRarXiv:2109.11418v12021
  48. MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations

    Kaixuan Huang, Jiacheng Guo, Zihao Li +15

    cs.LGcs.AIcs.CLarXiv:2502.06453v22025
  49. Personalized Transformer for Explainable Recommendation

    Lei Li, Yongfeng Zhang, Li Chen

    cs.IRcs.AIcs.CLarXiv:2105.11601v22021
  50. Adverse Events in Robotic Surgery: A Retrospective Study of 14 Years of FDA Data

    Homa Alemzadeh, Ravishankar K. Iyer, Zbigniew Kalbarczyk +2

    cs.ROcs.CRarXiv:1507.03518v22015
  51. ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific Research

    Wanghan Xu, Shuo Li, Tianlin Ye +48

    cs.LGcs.AIcs.CLarXiv:2606.07591v52026
  52. Likelihood-informed dimension reduction for nonlinear inverse problems

    Tiangang Cui, James Martin, Youssef M. Marzouk +2

    stat.COmath.NAstat.MEarXiv:1403.4680v22014
  53. Closing the Verification Loop: Self-Check Captioning for Long-Paragraph Detailed Audio Captioning

    Fengji Ma, Yan Rong, Xu Li +3

    cs.SDarXiv:2608.30713v12026
  54. Towards End-to-End Automation of AI Research

    Yutaro Yamada, Robert Tjarko Lange, Cong Lu +5

    cs.AIarXiv:2606.15497v12026
  55. SurgSkill-Bench: A Benchmark for Multimodal Surgical Skill Assessment

    Chaohui Dang, Zheheng Jiang, James Glasbey +3

    cs.CVarXiv:2608.30872v12026
  56. Image Hijacks: Adversarial Images can Control Generative Models at Runtime

    Luke Bailey, Euan Ong, Stuart Russell +1

    cs.LGcs.CLcs.CRarXiv:2309.00236v42023
  57. ViSAR: Training-Free Adaptive-$k$ Retrieval for Visual Document Question Answering

    Adrien Mialland, Marc Plantevit, Julien Gallois +1

    cs.IRcs.AIcs.CLarXiv:2609.02486v12026
  58. Narrative-Driven Paper-to-Slide Generation via ArcDeck

    Tarik Can Ozden, Sachidanand VS, Furkan Horoz +3

    cs.AIarXiv:2604.11969v12026
  59. Efficient Reinforcement Finetuning via Adaptive Curriculum Learning

    Taiwei Shi, Yiyang Wu, Linxin Song +2

    cs.LGcs.CLarXiv:2504.05520v42025
  60. NorMuon: Making Muon more efficient and scalable

    Zichong Li, Liming Liu, Chen Liang +2

    cs.LGcs.CLarXiv:2510.05491v12025