Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
52,981 to 53,040 of 61,134
Deep Audio-Visual Speech Recognition
Triantafyllos Afouras, Joon Son Chung, Andrew Senior +2
cs.CVarXiv:1809.02108v22018DualPrompt: Complementary Prompting for Rehearsal-free Continual Learning
Zifeng Wang, Zizhao Zhang, Sayna Ebrahimi +8
cs.LGcs.CVarXiv:2204.04799v22022Demystifying Video Reasoning
Ruisi Wang, Zhongang Cai, Fanyi Pu +11
cs.CVcs.AIarXiv:2603.16870v32026Gated Feedback Recurrent Neural Networks
Junyoung Chung, Caglar Gulcehre, Kyunghyun Cho +1
cs.NEcs.LGstat.MLarXiv:1502.02367v42015Towards Automated Kernel Generation in the Era of LLMs
Yang Yu, Peiyu Zang, Chi Hsu Tsai +11
cs.LGcs.CLarXiv:2601.15727v32026ESPNet: Efficient Spatial Pyramid of Dilated Convolutions for Semantic Segmentation
Sachin Mehta, Mohammad Rastegari, Anat Caspi +2
cs.CVarXiv:1803.06815v32018Semantic Autoencoder for Zero-Shot Learning
Elyor Kodirov, Tao Xiang, Shaogang Gong
cs.CVarXiv:1704.08345v12017Video Models Reason Early: Exploiting Plan Commitment for Maze Solving
Kaleb Newman, Tyler Zhu, Olga Russakovsky
cs.CVarXiv:2603.30043v12026Autoregressive Image Generation using Residual Quantization
Doyup Lee, Chiheon Kim, Saehoon Kim +2
cs.CVcs.LGarXiv:2203.01941v22022PhyRPR: Training-Free Physics-Constrained Video Generation
Yibo Zhao, Hengjia Li, Xiaofei He +1
cs.CVarXiv:2601.09255v12026MS-TCN: Multi-Stage Temporal Convolutional Network for Action Segmentation
Yazan Abu Farha, Juergen Gall
cs.CVarXiv:1903.01945v22019Detection and Resolution of Rumours in Social Media: A Survey
Arkaitz Zubiaga, Ahmet Aker, Kalina Bontcheva +2
cs.CLcs.HCcs.IRarXiv:1704.00656v32017EvolVE: Evolutionary Search for LLM-based Verilog Generation and Optimization
Wei-Po Hsin, Ren-Hao Deng, Yao-Ting Hsieh +2
cs.AIcs.NEcs.PLarXiv:2601.18067v12026LumosX: Relate Any Identities with Their Attributes for Personalized Video Generation
Jiazheng Xing, Fei Du, Hangjie Yuan +7
cs.CVcs.AIarXiv:2603.20192v12026Quantum algorithms for supervised and unsupervised machine learning
Seth Lloyd, Masoud Mohseni, Patrick Rebentrost
quant-pharXiv:1307.0411v22013BARF: Bundle-Adjusting Neural Radiance Fields
Chen-Hsuan Lin, Wei-Chiu Ma, Antonio Torralba +1
cs.CVcs.GRcs.LGarXiv:2104.06405v22021DiffusionCLIP: Text-Guided Diffusion Models for Robust Image Manipulation
Gwanghyun Kim, Taesung Kwon, Jong Chul Ye
cs.CVcs.AIcs.LGarXiv:2110.02711v62021Sparsified SGD with Memory
Sebastian U. Stich, Jean-Baptiste Cordonnier, Martin Jaggi
cs.LGcs.DCcs.DSarXiv:1809.07599v22018Bayesian Online Changepoint Detection
Ryan Prescott Adams, David J. C. MacKay
stat.MLarXiv:0710.3742v12007Variational Dropout Sparsifies Deep Neural Networks
Dmitry Molchanov, Arsenii Ashukha, Dmitry Vetrov
stat.MLcs.LGarXiv:1701.05369v32017Multicomponent multisublattice alloys, nonconfigurational entropy and other additions to the Alloy Theoretic Automated Toolkit
Axel van de Walle
cond-mat.mtrl-scicond-mat.stat-mecharXiv:0906.1608v22009FutureOmni: Evaluating Future Forecasting from Omni-Modal Context for Multimodal LLMs
Qian Chen, Jinlan Fu, Changsong Li +3
cs.CLcs.CVcs.MMarXiv:2601.13836v22026Agentic AI and the next intelligence explosion
James Evans, Benjamin Bratton, Blaise Agüera y Arcas
cs.AIarXiv:2603.20639v12026Quantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formatting
Melanie Sclar, Yejin Choi, Yulia Tsvetkov +1
cs.CLcs.AIcs.LGarXiv:2310.11324v22023How Close is ChatGPT to Human Experts? Comparison Corpus, Evaluation, and Detection
Biyang Guo, Xin Zhang, Ziyuan Wang +5
cs.CLarXiv:2301.07597v12023Secure Wireless Communication via Intelligent Reflecting Surface
Miao Cui, Guangchi Zhang, Rui Zhang
cs.ITarXiv:1905.10770v12019One-Shot Learning for Semantic Segmentation
Amirreza Shaban, Shray Bansal, Zhen Liu +2
cs.CVarXiv:1709.03410v12017RAISE: Requirement-Adaptive Evolutionary Refinement for Training-Free Text-to-Image Alignment
Liyao Jiang, Ruichen Chen, Chao Gao +1
cs.CVcs.AIarXiv:2603.00483v12026Manifold-Aware Exploration for Reinforcement Learning in Video Generation
Mingzhe Zheng, Weijie Kong, Yue Wu +9
cs.CVcs.AIarXiv:2603.21872v12026Not-so-supervised: a survey of semi-supervised, multi-instance, and transfer learning in medical image analysis
Veronika Cheplygina, Marleen de Bruijne, Josien P. W. Pluim
cs.CVarXiv:1804.06353v22018Physics-Informed Neural Operator for Learning Partial Differential Equations
Zongyi Li, Hongkai Zheng, Nikola Kovachki +5
cs.LGmath.NAarXiv:2111.03794v42021Reward-free Alignment for Conflicting Objectives
Peter Chen, Xiaopeng Li, Xi Chen +1
cs.CLcs.AIcs.LGarXiv:2602.02495v32026PersonaVLM: Long-Term Personalized Multimodal LLMs
Chang Nie, Chaoyou Fu, Yifan Zhang +2
cs.CLcs.CVarXiv:2604.13074v12026MathQA: Towards Interpretable Math Word Problem Solving with Operation-Based Formalisms
Aida Amini, Saadia Gabriel, Peter Lin +3
cs.CLarXiv:1905.13319v12019ERNIE 2.0: A Continual Pre-training Framework for Language Understanding
Yu Sun, Shuohuan Wang, Yukun Li +4
cs.CLarXiv:1907.12412v22019TuckER: Tensor Factorization for Knowledge Graph Completion
Ivana Balažević, Carl Allen, Timothy M. Hospedales
cs.LGstat.MLarXiv:1901.09590v22019SimRecon: SimReady Compositional Scene Reconstruction from Real Videos
Chong Xia, Kai Zhu, Zizhuo Wang +3
cs.CVarXiv:2603.02133v22026Learning Pixel-level Semantic Affinity with Image-level Supervision for Weakly Supervised Semantic Segmentation
Jiwoon Ahn, Suha Kwak
cs.CVarXiv:1803.10464v22018A Benchmark for Interpretability Methods in Deep Neural Networks
Sara Hooker, Dumitru Erhan, Pieter-Jan Kindermans +1
cs.LGcs.AIstat.MLarXiv:1806.10758v32018TabPFN: A Transformer That Solves Small Tabular Classification Problems in a Second
Noah Hollmann, Samuel Müller, Katharina Eggensperger +1
cs.LGstat.MLarXiv:2207.01848v62022Self-Supervised Pre-Training of Swin Transformers for 3D Medical Image Analysis
Yucheng Tang, Dong Yang, Wenqi Li +5
cs.CVcs.AIcs.LGarXiv:2111.14791v22021PEARL: Personalized Streaming Video Understanding Model
Yuanhong Zheng, Ruichuan An, Xiaopeng Lin +10
cs.CVcs.AIcs.IRarXiv:2603.20422v12026A Speculative Study on 6G
Faisal Tariq, Muhammad Khandaker, Kai-Kit Wong +3
cs.NIarXiv:1902.06700v22019Parseval Networks: Improving Robustness to Adversarial Examples
Moustapha Cisse, Piotr Bojanowski, Edouard Grave +2
stat.MLcs.AIcs.CRarXiv:1704.08847v22017Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Alignment
Yuming Yang, Mingyoung Lai, Wanxu Zhao +13
cs.CLarXiv:2601.14249v52026Recurrent Squeeze-and-Excitation Context Aggregation Net for Single Image Deraining
Xia Li, Jianlong Wu, Zhouchen Lin +2
cs.CVarXiv:1807.05698v22018tttLRM: Test-Time Training for Long Context and Autoregressive 3D Reconstruction
Chen Wang, Hao Tan, Wang Yifan +6
cs.CVarXiv:2602.20160v22026HRank: Filter Pruning using High-Rank Feature Map
Mingbao Lin, Rongrong Ji, Yan Wang +4
cs.CVarXiv:2002.10179v22020Thinking in Frames: How Visual Context and Test-Time Scaling Empower Video Reasoning
Chengzu Li, Zanyi Wang, Jiaang Li +9
cs.LGcs.AIcs.CLarXiv:2601.21037v12026A Comprehensive Survey of Deep Learning for Image Captioning
Md. Zakir Hossain, Ferdous Sohel, Mohd Fairuz Shiratuddin +1
cs.CVcs.LGstat.MLarXiv:1810.04020v22018CAD2RL: Real Single-Image Flight without a Single Real Image
Fereshteh Sadeghi, Sergey Levine
cs.LGcs.CVcs.ROarXiv:1611.04201v42016KAPSO: A Knowledge-grounded framework for Autonomous Program Synthesis and Optimization
Alireza Nadafian, Alireza Mohammadshahi, Majid Yazdani
cs.AIcs.CLcs.SEarXiv:2601.21526v22026Deep Reconstruction-Classification Networks for Unsupervised Domain Adaptation
Muhammad Ghifary, W. Bastiaan Kleijn, Mengjie Zhang +2
cs.CVcs.AIcs.LGarXiv:1607.03516v22016iMAP: Implicit Mapping and Positioning in Real-Time
Edgar Sucar, Shikun Liu, Joseph Ortiz +1
cs.CVarXiv:2103.12352v22021Pose Guided Person Image Generation
Liqian Ma, Xu Jia, Qianru Sun +3
cs.CVarXiv:1705.09368v62017Residual Context Diffusion Language Models
Yuezhou Hu, Harman Singh, Monishwaran Maheswaran +10
cs.CLcs.AIarXiv:2601.22954v22026Dual Path Networks
Yunpeng Chen, Jianan Li, Huaxin Xiao +3
cs.CVarXiv:1707.01629v22017CDDFuse: Correlation-Driven Dual-Branch Feature Decomposition for Multi-Modality Image Fusion
Zixiang Zhao, Haowen Bai, Jiangshe Zhang +5
cs.CVarXiv:2211.14461v22022Neural NILM: Deep Neural Networks Applied to Energy Disaggregation
Jack Kelly, William Knottenbelt
cs.NEarXiv:1507.06594v32015Causal Motion Diffusion Models for Autoregressive Motion Generation
Qing Yu, Akihisa Watanabe, Kent Fujiwara
cs.CVarXiv:2602.22594v12026