Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
53,581 to 53,640 of 61,255
Going Deeper in Facial Expression Recognition using Deep Neural Networks
Ali Mollahosseini, David Chan, Mohammad H. Mahoor
cs.NEcs.CVarXiv:1511.04110v12015Deformable 3D Gaussians for High-Fidelity Monocular Dynamic Scene Reconstruction
Ziyi Yang, Xinyu Gao, Wen Zhou +3
cs.CVarXiv:2309.13101v22023Action100M: A Large-scale Video Action Dataset
Delong Chen, Tejaswi Kasarla, Yejin Bang +6
cs.CVarXiv:2601.10592v12026Network Trimming: A Data-Driven Neuron Pruning Approach towards Efficient Deep Architectures
Hengyuan Hu, Rui Peng, Yu-Wing Tai +1
cs.NEcs.CVcs.LGarXiv:1607.03250v12016Fast Algorithms for Convolutional Neural Networks
Andrew Lavin, Scott Gray
cs.NEcs.LGarXiv:1509.09308v22015Overfitting in adversarially robust deep learning
Leslie Rice, Eric Wong, J. Zico Kolter
cs.LGstat.MLarXiv:2002.11569v22020Integrated photonics on thin-film lithium niobate
Di Zhu, Linbo Shao, Mengjie Yu +12
physics.opticsphysics.app-pharXiv:2102.11956v12021Image-Image Domain Adaptation with Preserved Self-Similarity and Domain-Dissimilarity for Person Re-identification
Weijian Deng, Liang Zheng, Qixiang Ye +3
cs.CVarXiv:1711.07027v32017Omni-Diffusion: Unified Multimodal Understanding and Generation with Masked Discrete Diffusion
Lijiang Li, Zuwei Long, Yunhang Shen +6
cs.CVarXiv:2603.06577v22026UIU-Net: U-Net in U-Net for Infrared Small Object Detection
Xin Wu, Danfeng Hong, Jocelyn Chanussot
cs.CVarXiv:2212.00968v12022THINKSAFE: Self-Generated Safety Alignment for Reasoning Models
Seanie Lee, Sangwoo Park, Yumin Choi +6
cs.AIarXiv:2601.23143v42026The performance of modularity maximization in practical contexts
Benjamin H. Good, Yves-Alexandre de Montjoye, Aaron Clauset
physics.data-ancond-mat.dis-nnphysics.soc-pharXiv:0910.0165v22009ReGuLaR: Variational Latent Reasoning Guided by Rendered Chain-of-Thought
Fanmeng Wang, Haotian Liu, Guojiang Zhao +2
cs.CLarXiv:2601.23184v12026Model-Agnostic Interpretability of Machine Learning
Marco Tulio Ribeiro, Sameer Singh, Carlos Guestrin
stat.MLcs.LGarXiv:1606.05386v12016SiamFC++: Towards Robust and Accurate Visual Tracking with Target Estimation Guidelines
Yinda Xu, Zeyu Wang, Zuoxin Li +2
cs.CVarXiv:1911.06188v42019The Hateful Memes Challenge: Detecting Hate Speech in Multimodal Memes
Douwe Kiela, Hamed Firooz, Aravind Mohan +4
cs.AIcs.CLcs.CVarXiv:2005.04790v32020Terminal Agents Suffice for Enterprise Automation
Patrice Bechard, Orlando Marquez Ayala, Emily Chen +5
cs.SEcs.AIcs.CLarXiv:2604.00073v32026StarCraft II: A New Challenge for Reinforcement Learning
Oriol Vinyals, Timo Ewalds, Sergey Bartunov +22
cs.LGcs.AIarXiv:1708.04782v12017RMA: Rapid Motor Adaptation for Legged Robots
Ashish Kumar, Zipeng Fu, Deepak Pathak +1
cs.LGcs.AIcs.CVarXiv:2107.04034v12021Large Batch Training of Convolutional Networks
Yang You, Igor Gitman, Boris Ginsburg
cs.CVarXiv:1708.03888v32017End-to-end Neural Coreference Resolution
Kenton Lee, Luheng He, Mike Lewis +1
cs.CLarXiv:1707.07045v22017Weakly- and Semi-Supervised Learning of a DCNN for Semantic Image Segmentation
George Papandreou, Liang-Chieh Chen, Kevin Murphy +1
cs.CVarXiv:1502.02734v32015What do you learn from context? Probing for sentence structure in contextualized word representations
Ian Tenney, Patrick Xia, Berlin Chen +8
cs.CLarXiv:1905.06316v12019Exploring Reasoning Reward Model for Agents
Kaixuan Fan, Kaituo Feng, Manyuan Zhang +7
cs.AIcs.CLarXiv:2601.22154v22026A-OKVQA: A Benchmark for Visual Question Answering using World Knowledge
Dustin Schwenk, Apoorv Khandelwal, Christopher Clark +2
cs.CVcs.CLarXiv:2206.01718v12022Semantically Conditioned LSTM-based Natural Language Generation for Spoken Dialogue Systems
Tsung-Hsien Wen, Milica Gasic, Nikola Mrksic +3
cs.CLarXiv:1508.01745v22015Benign Overfitting in Linear Regression
Peter L. Bartlett, Philip M. Long, Gábor Lugosi +1
stat.MLcs.LGmath.STarXiv:1906.11300v32019The Vision Wormhole: Latent-Space Communication in Heterogeneous Multi-Agent Systems
Xiaoze Liu, Ruowang Zhang, Weichen Yu +7
cs.CLcs.CVcs.LGarXiv:2602.15382v22026LiveMedBench: A Contamination-Free Medical Benchmark for LLMs with Automated Rubric Evaluation
Zhiling Yan, Dingjie Song, Zhe Fang +4
cs.AIarXiv:2602.10367v12026Recommendation as Language Processing (RLP): A Unified Pretrain, Personalized Prompt & Predict Paradigm (P5)
Shijie Geng, Shuchang Liu, Zuohui Fu +2
cs.IRcs.AIcs.CLarXiv:2203.13366v72022OPUS: Towards Efficient and Principled Data Selection in Large Language Model Pre-training in Every Iteration
Shaobo Wang, Xuan Ouyang, Tianyi Xu +9
cs.CLarXiv:2602.05400v22026MWM: Mobile World Models for Action-Conditioned Consistent Prediction
Han Yan, Zishang Xiang, Zeyu Zhang +1
cs.CVcs.ROarXiv:2603.07799v12026HiAR: Efficient Autoregressive Long Video Generation via Hierarchical Denoising
Kai Zou, Dian Zheng, Hongbo Liu +3
cs.CVarXiv:2603.08703v12026LIVE: Long-horizon Interactive Video World Modeling
Junchao Huang, Ziyang Ye, Xinting Hu +5
cs.CVarXiv:2602.03747v12026Unity: A General Platform for Intelligent Agents
Arthur Juliani, Vincent-Pierre Berges, Ervin Teng +8
cs.LGcs.AIcs.NEarXiv:1809.02627v22018Temporal Segment Networks for Action Recognition in Videos
Limin Wang, Yuanjun Xiong, Zhe Wang +4
cs.CVarXiv:1705.02953v12017NExT-QA:Next Phase of Question-Answering to Explaining Temporal Actions
Junbin Xiao, Xindi Shang, Angela Yao +1
cs.CVcs.AIarXiv:2105.08276v22021ERNIE 5.0 Technical Report
Haifeng Wang, Hua Wu, Tian Wu +435
cs.CLarXiv:2602.04705v12026pi-GAN: Periodic Implicit Generative Adversarial Networks for 3D-Aware Image Synthesis
Eric R. Chan, Marco Monteiro, Petr Kellnhofer +2
cs.CVcs.GRarXiv:2012.00926v22020Vision2Web: A Hierarchical Benchmark for Visual Website Development with Agent Verification
Zehai He, Wenyi Hong, Zhen Yang +4
cs.SEcs.AIarXiv:2603.26648v32026Medical Image Segmentation Review: The success of U-Net
Reza Azad, Ehsan Khodapanah Aghdam, Amelie Rauland +7
eess.IVcs.CVarXiv:2211.14830v12022LOME: Learning Human-Object Manipulation with Action-Conditioned Egocentric World Model
Quankai Gao, Jiawei Yang, Qiangeng Xu +2
cs.CVarXiv:2603.27449v12026Mobile-GS: Real-time Gaussian Splatting for Mobile Devices
Xiaobiao Du, Yida Wang, Kun Zhan +1
cs.CVarXiv:2603.11531v12026Slither: A Static Analysis Framework For Smart Contracts
Josselin Feist, Gustavo Grieco, Alex Groce
cs.SEcs.CRarXiv:1908.09878v12019A Survey on the Explainability of Supervised Machine Learning
Nadia Burkart, Marco F. Huber
cs.LGcs.AIstat.MLarXiv:2011.07876v12020VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models
Zixuan Wang, Yuxin Chen, Yuqi Liu +6
cs.ROarXiv:2603.22003v32026ILVR: Conditioning Method for Denoising Diffusion Probabilistic Models
Jooyoung Choi, Sungwon Kim, Yonghyun Jeong +2
cs.CVarXiv:2108.02938v22021The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only
Guilherme Penedo, Quentin Malartic, Daniel Hesslow +6
cs.CLcs.AIarXiv:2306.01116v12023AI Gamestore: Scalable, Open-Ended Evaluation of Machine General Intelligence with Human Games
Lance Ying, Ryan Truong, Prafull Sharma +9
cs.AIarXiv:2602.17594v12026Bat Algorithm: Literature Review and Applications
Xin-She Yang
cs.AImath.OCarXiv:1308.3900v12013Understanding data augmentation for classification: when to warp?
Sebastien C. Wong, Adam Gatt, Victor Stamatescu +1
cs.CVarXiv:1609.08764v22016DeepID3: Face Recognition with Very Deep Neural Networks
Yi Sun, Ding Liang, Xiaogang Wang +1
cs.CVarXiv:1502.00873v12015A Survey on Security and Privacy Issues of Bitcoin
Mauro Conti, Sandeep Kumar E, Chhagan Lal +1
cs.CRarXiv:1706.00916v32017ABCNN: Attention-Based Convolutional Neural Network for Modeling Sentence Pairs
Wenpeng Yin, Hinrich Schütze, Bing Xiang +1
cs.CLarXiv:1512.05193v42015Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training
Gen Li, Nan Duan, Yuejian Fang +3
cs.CVarXiv:1908.06066v32019\$OneMillion-Bench: How Far are Language Agents from Human Experts?
Qianyu Yang, Yang Liu, Jiaqi Li +20
cs.LGcs.AIcs.CLarXiv:2603.07980v12026Towards General Text Embeddings with Multi-stage Contrastive Learning
Zehan Li, Xin Zhang, Yanzhao Zhang +3
cs.CLarXiv:2308.03281v12023Understanding disentangling in $β$-VAE
Christopher P. Burgess, Irina Higgins, Arka Pal +4
stat.MLcs.AIcs.LGarXiv:1804.03599v12018Making Convolutional Networks Shift-Invariant Again
Richard Zhang
cs.CVcs.LGarXiv:1904.11486v22019Causal-JEPA: Learning World Models through Object-Level Latent Masking
Heejeong Nam, Quentin Le Lidec, Lucas Maes +2
cs.AIarXiv:2602.11389v22026