Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
56,521 to 56,580 of 61,351
Speeding up Convolutional Neural Networks with Low Rank Expansions
Max Jaderberg, Andrea Vedaldi, Andrew Zisserman
cs.CVarXiv:1405.3866v12014Decision-Based Adversarial Attacks: Reliable Attacks Against Black-Box Machine Learning Models
Wieland Brendel, Jonas Rauber, Matthias Bethge
stat.MLcs.CRcs.CVarXiv:1712.04248v22017CodeGen: An Open Large Language Model for Code with Multi-Turn Program Synthesis
Erik Nijkamp, Bo Pang, Hiroaki Hayashi +5
cs.LGcs.CLcs.PLarXiv:2203.13474v52022iFogSim: A Toolkit for Modeling and Simulation of Resource Management Techniques in Internet of Things, Edge and Fog Computing Environments
Harshit Gupta, Amir Vahid Dastjerdi, Soumya K. Ghosh +1
cs.DCarXiv:1606.02007v12016Learning to See in the Dark
Chen Chen, Qifeng Chen, Jia Xu +1
cs.CVcs.GRcs.LGarXiv:1805.01934v12018Minimalist Visual Inertial Odometry
Francesco Pasti, Jeremy Klotz, Nicola Bellotto +1
cs.ROcs.CVcs.LGarXiv:2605.19990v12026Data-Efficient Image Recognition with Contrastive Predictive Coding
Olivier J. Hénaff, Aravind Srinivas, Jeffrey De Fauw +4
cs.CVcs.LGarXiv:1905.09272v32019Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement Learning
Tabish Rashid, Mikayel Samvelyan, Christian Schroeder de Witt +3
cs.LGcs.MAstat.MLarXiv:2003.08839v22020Platonic Representations in the Human Brain: Unsupervised Recovery of Universal Geometry
Pablo Marcos-Manchón, Rishi Jha, Lluís Fuentemilla
q-bio.NCcs.CVarXiv:2605.20496v12026RETAIN: An Interpretable Predictive Model for Healthcare using Reverse Time Attention Mechanism
Edward Choi, Mohammad Taha Bahadori, Joshua A. Kulas +3
cs.LGcs.AIcs.NEarXiv:1608.05745v42016PCANet: A Simple Deep Learning Baseline for Image Classification?
Tsung-Han Chan, Kui Jia, Shenghua Gao +3
cs.CVcs.LGcs.NEarXiv:1404.3606v22014Perceiver: General Perception with Iterative Attention
Andrew Jaegle, Felix Gimeno, Andrew Brock +3
cs.CVcs.AIcs.LGarXiv:2103.03206v22021Cross-stitch Networks for Multi-task Learning
Ishan Misra, Abhinav Shrivastava, Abhinav Gupta +1
cs.CVcs.LGarXiv:1604.03539v12016Deep & Cross Network for Ad Click Predictions
Ruoxi Wang, Bin Fu, Gang Fu +1
cs.LGstat.MLarXiv:1708.05123v12017CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation
Shuai Lu, Daya Guo, Shuo Ren +19
cs.SEcs.CLarXiv:2102.04664v22021EvolveGCN: Evolving Graph Convolutional Networks for Dynamic Graphs
Aldo Pareja, Giacomo Domeniconi, Jie Chen +6
cs.LGcs.SIstat.MLarXiv:1902.10191v32019Interaction Networks for Learning about Objects, Relations and Physics
Peter W. Battaglia, Razvan Pascanu, Matthew Lai +2
cs.AIcs.LGarXiv:1612.00222v12016Rényi Divergence and Kullback-Leibler Divergence
Tim van Erven, Peter Harremoës
cs.ITmath.STstat.MLarXiv:1206.2459v22012IndusAgent: Reinforcing Open-Vocabulary Industrial Anomaly Detection with Agentic Tools
Rongbin Tan, Fangfang Lin, Zhenlong Yuan +10
cs.CVarXiv:2605.20682v12026FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text Generation
Sewon Min, Kalpesh Krishna, Xinxi Lyu +6
cs.CLcs.AIcs.LGarXiv:2305.14251v22023Deep Reinforcement Learning in a Handful of Trials using Probabilistic Dynamics Models
Kurtland Chua, Roberto Calandra, Rowan McAllister +1
cs.LGcs.AIcs.ROarXiv:1805.12114v22018ERNIE: Enhanced Language Representation with Informative Entities
Zhengyan Zhang, Xu Han, Zhiyuan Liu +3
cs.CLarXiv:1905.07129v32019BranchyNet: Fast Inference via Early Exiting from Deep Neural Networks
Surat Teerapittayanon, Bradley McDanel, H. T. Kung
cs.NEcs.CVcs.LGarXiv:1709.01686v12017Broadening Access to Transportation Safety Data with Generative AI: A Schema-Grounded Framework for Spatial Natural Language Queries
Mahdi Azhdari, Eric J. Gonzales
cs.CLarXiv:2605.21712v12026Analyzing Multi-Head Self-Attention: Specialized Heads Do the Heavy Lifting, the Rest Can Be Pruned
Elena Voita, David Talbot, Fedor Moiseev +2
cs.CLarXiv:1905.09418v22019Generating Videos with Scene Dynamics
Carl Vondrick, Hamed Pirsiavash, Antonio Torralba
cs.CVcs.GRcs.LGarXiv:1609.02612v32016RankJudge: A Multi-Turn LLM-as-a-Judge Synthetic Benchmark Generator
Zhenwei Tang, Zhaoyan Liu, Rasa Hosseinzadeh +3
cs.CLarXiv:2605.21748v12026What Makes for Good Views for Contrastive Learning?
Yonglong Tian, Chen Sun, Ben Poole +3
cs.CVcs.LGarXiv:2005.10243v32020Reflective Prompt Tuning through Language Model Function-Calling
Farima Fatahi Bayat, Moin Aminnaseri, Pouya Pezeshkpour +1
cs.CLarXiv:2605.21781v12026Fine-tuning CNN Image Retrieval with No Human Annotation
Filip Radenović, Giorgos Tolias, Ondřej Chum
cs.CVarXiv:1711.02512v22017Domain Adaptive Faster R-CNN for Object Detection in the Wild
Yuhua Chen, Wen Li, Christos Sakaridis +2
cs.CVarXiv:1803.03243v12018RISE: Randomized Input Sampling for Explanation of Black-box Models
Vitali Petsiuk, Abir Das, Kate Saenko
cs.CVarXiv:1806.07421v32018This Looks Like That: Deep Learning for Interpretable Image Recognition
Chaofan Chen, Oscar Li, Chaofan Tao +3
cs.LGcs.AIcs.CVarXiv:1806.10574v52018How Far Will They Go? Red-Teaming Online Influence with Large Language Models
Daniel C. Ruiz, Anna Serbina, Ashwin Rao +2
cs.CLcs.AIcs.CYarXiv:2605.22880v12026Incorporating Copying Mechanism in Sequence-to-Sequence Learning
Jiatao Gu, Zhengdong Lu, Hang Li +1
cs.CLcs.AIcs.LGarXiv:1603.06393v32016A Dual-Stage Attention-Based Recurrent Neural Network for Time Series Prediction
Yao Qin, Dongjin Song, Haifeng Chen +3
cs.LGstat.MLarXiv:1704.02971v42017Joint 3D Proposal Generation and Object Detection from View Aggregation
Jason Ku, Melissa Mozifian, Jungwook Lee +2
cs.CVarXiv:1712.02294v42017Seq2SQL: Generating Structured Queries from Natural Language using Reinforcement Learning
Victor Zhong, Caiming Xiong, Richard Socher
cs.CLcs.AIarXiv:1709.00103v72017MUSIQ: Multi-scale Image Quality Transformer
Junjie Ke, Qifei Wang, Yilin Wang +2
cs.CVarXiv:2108.05997v12021MotiMotion: Motion-Controlled Video Generation with Visual Reasoning
Lee Hsin-Ying, Hanwen Jiang, Yiqun Mei +3
cs.CVarXiv:2605.22818v12026The Structure and Dynamics of Co-Citation Clusters: A Multiple-Perspective Co-Citation Analysis
Chaomei Chen, Fidelia Ibekwe-SanJuan, Jianhua Hou
cs.CYarXiv:1002.1985v12010DocVQA: A Dataset for VQA on Document Images
Minesh Mathew, Dimosthenis Karatzas, C. V. Jawahar
cs.CVcs.IRarXiv:2007.00398v32020EMMA: Extracting Multiple physical parameters from Multimodal Data
Farhat Shaikh, Ayan Banerjee, Sandeep Gupta
cs.CVarXiv:2605.24047v12026Learning to Simulate Complex Physics with Graph Networks
Alvaro Sanchez-Gonzalez, Jonathan Godwin, Tobias Pfaff +3
cs.LGphysics.comp-phstat.MLarXiv:2002.09405v22020LLMs as Noisy Channels: A Shannon Perspective on Model Capacity and Scaling Laws
Xu Ouyang, Deyi Liu, Yuhang Cai +5
cs.LGcs.AIcs.ITarXiv:2605.23901v12026SpaceDG: Benchmarking Spatial Intelligence under Visual Degradation
Xiaolong Zhou, Yifei Liu, Ziyang Gong +8
cs.CVcs.CLarXiv:2605.22536v22026DecQ: Detail-Condensing Queries for Enhanced Reconstruction and Generation in Representation Autoencoders
Tianhang Wang, Yitong Chen, Wei Song +3
cs.CVarXiv:2605.22777v12026HunyuanVideo: A Systematic Framework For Large Video Generative Models
Weijie Kong, Qi Tian, Zijian Zhang +49
cs.CVarXiv:2412.03603v62024Transformer Feed-Forward Layers Are Key-Value Memories
Mor Geva, Roei Schuster, Jonathan Berant +1
cs.CLarXiv:2012.14913v22020Flat-Pack Bench: Evaluating Spatio-Temporal Understanding in Large Vision-Language Models through Furniture Assembly
Aditya Chetan, Eric Cai, Peeyush Kushwaha +5
cs.CVcs.AIcs.CLarXiv:2605.21625v12026Signed Networks in Social Media
Jure Leskovec, Daniel Huttenlocher, Jon Kleinberg
physics.soc-phcs.CYcs.HCarXiv:1003.2424v12010Large Kernel Matters -- Improve Semantic Segmentation by Global Convolutional Network
Chao Peng, Xiangyu Zhang, Gang Yu +2
cs.CVarXiv:1703.02719v12017Scientific reasoning does not reliably translate into scientific forecasting in frontier AI
Sean Wu, Pan Lu, Yupeng Chen +7
cs.AIarXiv:2605.22681v22026Multimodal Compact Bilinear Pooling for Visual Question Answering and Visual Grounding
Akira Fukui, Dong Huk Park, Daylen Yang +3
cs.CVcs.AIcs.CLarXiv:1606.01847v32016ACC: Compiling Agent Trajectories for Long-Context Training
Qisheng Su, Zhen Fang, Shiting Huang +8
cs.CLcs.AIarXiv:2605.21850v22026Large-Margin Softmax Loss for Convolutional Neural Networks
Weiyang Liu, Yandong Wen, Zhiding Yu +1
stat.MLcs.LGarXiv:1612.02295v42016Perception or Prejudice: Can MLLMs Go Beyond First Impressions of Personality?
Caixin Kang, Tianyu Yan, Sitong Gong +8
cs.AIcs.CVcs.CYarXiv:2605.22109v12026Can AI help in screening Viral and COVID-19 pneumonia?
Muhammad E. H. Chowdhury, Tawsifur Rahman, Amith Khandakar +9
cs.LGcs.CVarXiv:2003.13145v32020TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation
Hanyu Guo, Jiedong Yang, Chao Chen +3
cs.CLcs.AIcs.LGarXiv:2605.22355v12026Cluster-GCN: An Efficient Algorithm for Training Deep and Large Graph Convolutional Networks
Wei-Lin Chiang, Xuanqing Liu, Si Si +3
cs.LGcs.AIstat.MLarXiv:1905.07953v22019