Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

52,981 to 53,040 of 61,134

  1. Deep Audio-Visual Speech Recognition

    Triantafyllos Afouras, Joon Son Chung, Andrew Senior +2

    cs.CVarXiv:1809.02108v22018
  2. DualPrompt: Complementary Prompting for Rehearsal-free Continual Learning

    Zifeng Wang, Zizhao Zhang, Sayna Ebrahimi +8

    cs.LGcs.CVarXiv:2204.04799v22022
  3. Demystifying Video Reasoning

    Ruisi Wang, Zhongang Cai, Fanyi Pu +11

    cs.CVcs.AIarXiv:2603.16870v32026
  4. Gated Feedback Recurrent Neural Networks

    Junyoung Chung, Caglar Gulcehre, Kyunghyun Cho +1

    cs.NEcs.LGstat.MLarXiv:1502.02367v42015
  5. Towards Automated Kernel Generation in the Era of LLMs

    Yang Yu, Peiyu Zang, Chi Hsu Tsai +11

    cs.LGcs.CLarXiv:2601.15727v32026
  6. ESPNet: Efficient Spatial Pyramid of Dilated Convolutions for Semantic Segmentation

    Sachin Mehta, Mohammad Rastegari, Anat Caspi +2

    cs.CVarXiv:1803.06815v32018
  7. Semantic Autoencoder for Zero-Shot Learning

    Elyor Kodirov, Tao Xiang, Shaogang Gong

    cs.CVarXiv:1704.08345v12017
  8. Video Models Reason Early: Exploiting Plan Commitment for Maze Solving

    Kaleb Newman, Tyler Zhu, Olga Russakovsky

    cs.CVarXiv:2603.30043v12026
  9. Autoregressive Image Generation using Residual Quantization

    Doyup Lee, Chiheon Kim, Saehoon Kim +2

    cs.CVcs.LGarXiv:2203.01941v22022
  10. PhyRPR: Training-Free Physics-Constrained Video Generation

    Yibo Zhao, Hengjia Li, Xiaofei He +1

    cs.CVarXiv:2601.09255v12026
  11. MS-TCN: Multi-Stage Temporal Convolutional Network for Action Segmentation

    Yazan Abu Farha, Juergen Gall

    cs.CVarXiv:1903.01945v22019
  12. Detection and Resolution of Rumours in Social Media: A Survey

    Arkaitz Zubiaga, Ahmet Aker, Kalina Bontcheva +2

    cs.CLcs.HCcs.IRarXiv:1704.00656v32017
  13. EvolVE: Evolutionary Search for LLM-based Verilog Generation and Optimization

    Wei-Po Hsin, Ren-Hao Deng, Yao-Ting Hsieh +2

    cs.AIcs.NEcs.PLarXiv:2601.18067v12026
  14. LumosX: Relate Any Identities with Their Attributes for Personalized Video Generation

    Jiazheng Xing, Fei Du, Hangjie Yuan +7

    cs.CVcs.AIarXiv:2603.20192v12026
  15. Quantum algorithms for supervised and unsupervised machine learning

    Seth Lloyd, Masoud Mohseni, Patrick Rebentrost

    quant-pharXiv:1307.0411v22013
  16. BARF: Bundle-Adjusting Neural Radiance Fields

    Chen-Hsuan Lin, Wei-Chiu Ma, Antonio Torralba +1

    cs.CVcs.GRcs.LGarXiv:2104.06405v22021
  17. DiffusionCLIP: Text-Guided Diffusion Models for Robust Image Manipulation

    Gwanghyun Kim, Taesung Kwon, Jong Chul Ye

    cs.CVcs.AIcs.LGarXiv:2110.02711v62021
  18. Sparsified SGD with Memory

    Sebastian U. Stich, Jean-Baptiste Cordonnier, Martin Jaggi

    cs.LGcs.DCcs.DSarXiv:1809.07599v22018
  19. Bayesian Online Changepoint Detection

    Ryan Prescott Adams, David J. C. MacKay

    stat.MLarXiv:0710.3742v12007
  20. Variational Dropout Sparsifies Deep Neural Networks

    Dmitry Molchanov, Arsenii Ashukha, Dmitry Vetrov

    stat.MLcs.LGarXiv:1701.05369v32017
  21. Multicomponent multisublattice alloys, nonconfigurational entropy and other additions to the Alloy Theoretic Automated Toolkit

    Axel van de Walle

    cond-mat.mtrl-scicond-mat.stat-mecharXiv:0906.1608v22009
  22. FutureOmni: Evaluating Future Forecasting from Omni-Modal Context for Multimodal LLMs

    Qian Chen, Jinlan Fu, Changsong Li +3

    cs.CLcs.CVcs.MMarXiv:2601.13836v22026
  23. Agentic AI and the next intelligence explosion

    James Evans, Benjamin Bratton, Blaise Agüera y Arcas

    cs.AIarXiv:2603.20639v12026
  24. Quantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formatting

    Melanie Sclar, Yejin Choi, Yulia Tsvetkov +1

    cs.CLcs.AIcs.LGarXiv:2310.11324v22023
  25. How Close is ChatGPT to Human Experts? Comparison Corpus, Evaluation, and Detection

    Biyang Guo, Xin Zhang, Ziyuan Wang +5

    cs.CLarXiv:2301.07597v12023
  26. Secure Wireless Communication via Intelligent Reflecting Surface

    Miao Cui, Guangchi Zhang, Rui Zhang

    cs.ITarXiv:1905.10770v12019
  27. One-Shot Learning for Semantic Segmentation

    Amirreza Shaban, Shray Bansal, Zhen Liu +2

    cs.CVarXiv:1709.03410v12017
  28. RAISE: Requirement-Adaptive Evolutionary Refinement for Training-Free Text-to-Image Alignment

    Liyao Jiang, Ruichen Chen, Chao Gao +1

    cs.CVcs.AIarXiv:2603.00483v12026
  29. Manifold-Aware Exploration for Reinforcement Learning in Video Generation

    Mingzhe Zheng, Weijie Kong, Yue Wu +9

    cs.CVcs.AIarXiv:2603.21872v12026
  30. Not-so-supervised: a survey of semi-supervised, multi-instance, and transfer learning in medical image analysis

    Veronika Cheplygina, Marleen de Bruijne, Josien P. W. Pluim

    cs.CVarXiv:1804.06353v22018
  31. Physics-Informed Neural Operator for Learning Partial Differential Equations

    Zongyi Li, Hongkai Zheng, Nikola Kovachki +5

    cs.LGmath.NAarXiv:2111.03794v42021
  32. Reward-free Alignment for Conflicting Objectives

    Peter Chen, Xiaopeng Li, Xi Chen +1

    cs.CLcs.AIcs.LGarXiv:2602.02495v32026
  33. PersonaVLM: Long-Term Personalized Multimodal LLMs

    Chang Nie, Chaoyou Fu, Yifan Zhang +2

    cs.CLcs.CVarXiv:2604.13074v12026
  34. MathQA: Towards Interpretable Math Word Problem Solving with Operation-Based Formalisms

    Aida Amini, Saadia Gabriel, Peter Lin +3

    cs.CLarXiv:1905.13319v12019
  35. ERNIE 2.0: A Continual Pre-training Framework for Language Understanding

    Yu Sun, Shuohuan Wang, Yukun Li +4

    cs.CLarXiv:1907.12412v22019
  36. TuckER: Tensor Factorization for Knowledge Graph Completion

    Ivana Balažević, Carl Allen, Timothy M. Hospedales

    cs.LGstat.MLarXiv:1901.09590v22019
  37. SimRecon: SimReady Compositional Scene Reconstruction from Real Videos

    Chong Xia, Kai Zhu, Zizhuo Wang +3

    cs.CVarXiv:2603.02133v22026
  38. Learning Pixel-level Semantic Affinity with Image-level Supervision for Weakly Supervised Semantic Segmentation

    Jiwoon Ahn, Suha Kwak

    cs.CVarXiv:1803.10464v22018
  39. A Benchmark for Interpretability Methods in Deep Neural Networks

    Sara Hooker, Dumitru Erhan, Pieter-Jan Kindermans +1

    cs.LGcs.AIstat.MLarXiv:1806.10758v32018
  40. TabPFN: A Transformer That Solves Small Tabular Classification Problems in a Second

    Noah Hollmann, Samuel Müller, Katharina Eggensperger +1

    cs.LGstat.MLarXiv:2207.01848v62022
  41. Self-Supervised Pre-Training of Swin Transformers for 3D Medical Image Analysis

    Yucheng Tang, Dong Yang, Wenqi Li +5

    cs.CVcs.AIcs.LGarXiv:2111.14791v22021
  42. PEARL: Personalized Streaming Video Understanding Model

    Yuanhong Zheng, Ruichuan An, Xiaopeng Lin +10

    cs.CVcs.AIcs.IRarXiv:2603.20422v12026
  43. A Speculative Study on 6G

    Faisal Tariq, Muhammad Khandaker, Kai-Kit Wong +3

    cs.NIarXiv:1902.06700v22019
  44. Parseval Networks: Improving Robustness to Adversarial Examples

    Moustapha Cisse, Piotr Bojanowski, Edouard Grave +2

    stat.MLcs.AIcs.CRarXiv:1704.08847v22017
  45. Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Alignment

    Yuming Yang, Mingyoung Lai, Wanxu Zhao +13

    cs.CLarXiv:2601.14249v52026
  46. Recurrent Squeeze-and-Excitation Context Aggregation Net for Single Image Deraining

    Xia Li, Jianlong Wu, Zhouchen Lin +2

    cs.CVarXiv:1807.05698v22018
  47. tttLRM: Test-Time Training for Long Context and Autoregressive 3D Reconstruction

    Chen Wang, Hao Tan, Wang Yifan +6

    cs.CVarXiv:2602.20160v22026
  48. HRank: Filter Pruning using High-Rank Feature Map

    Mingbao Lin, Rongrong Ji, Yan Wang +4

    cs.CVarXiv:2002.10179v22020
  49. Thinking in Frames: How Visual Context and Test-Time Scaling Empower Video Reasoning

    Chengzu Li, Zanyi Wang, Jiaang Li +9

    cs.LGcs.AIcs.CLarXiv:2601.21037v12026
  50. A Comprehensive Survey of Deep Learning for Image Captioning

    Md. Zakir Hossain, Ferdous Sohel, Mohd Fairuz Shiratuddin +1

    cs.CVcs.LGstat.MLarXiv:1810.04020v22018
  51. CAD2RL: Real Single-Image Flight without a Single Real Image

    Fereshteh Sadeghi, Sergey Levine

    cs.LGcs.CVcs.ROarXiv:1611.04201v42016
  52. KAPSO: A Knowledge-grounded framework for Autonomous Program Synthesis and Optimization

    Alireza Nadafian, Alireza Mohammadshahi, Majid Yazdani

    cs.AIcs.CLcs.SEarXiv:2601.21526v22026
  53. Deep Reconstruction-Classification Networks for Unsupervised Domain Adaptation

    Muhammad Ghifary, W. Bastiaan Kleijn, Mengjie Zhang +2

    cs.CVcs.AIcs.LGarXiv:1607.03516v22016
  54. iMAP: Implicit Mapping and Positioning in Real-Time

    Edgar Sucar, Shikun Liu, Joseph Ortiz +1

    cs.CVarXiv:2103.12352v22021
  55. Pose Guided Person Image Generation

    Liqian Ma, Xu Jia, Qianru Sun +3

    cs.CVarXiv:1705.09368v62017
  56. Residual Context Diffusion Language Models

    Yuezhou Hu, Harman Singh, Monishwaran Maheswaran +10

    cs.CLcs.AIarXiv:2601.22954v22026
  57. Dual Path Networks

    Yunpeng Chen, Jianan Li, Huaxin Xiao +3

    cs.CVarXiv:1707.01629v22017
  58. CDDFuse: Correlation-Driven Dual-Branch Feature Decomposition for Multi-Modality Image Fusion

    Zixiang Zhao, Haowen Bai, Jiangshe Zhang +5

    cs.CVarXiv:2211.14461v22022
  59. Neural NILM: Deep Neural Networks Applied to Energy Disaggregation

    Jack Kelly, William Knottenbelt

    cs.NEarXiv:1507.06594v32015
  60. Causal Motion Diffusion Models for Autoregressive Motion Generation

    Qing Yu, Akihisa Watanabe, Kent Fujiwara

    cs.CVarXiv:2602.22594v12026