Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

53,581 to 53,640 of 61,255

  1. Going Deeper in Facial Expression Recognition using Deep Neural Networks

    Ali Mollahosseini, David Chan, Mohammad H. Mahoor

    cs.NEcs.CVarXiv:1511.04110v12015
  2. Deformable 3D Gaussians for High-Fidelity Monocular Dynamic Scene Reconstruction

    Ziyi Yang, Xinyu Gao, Wen Zhou +3

    cs.CVarXiv:2309.13101v22023
  3. Action100M: A Large-scale Video Action Dataset

    Delong Chen, Tejaswi Kasarla, Yejin Bang +6

    cs.CVarXiv:2601.10592v12026
  4. Network Trimming: A Data-Driven Neuron Pruning Approach towards Efficient Deep Architectures

    Hengyuan Hu, Rui Peng, Yu-Wing Tai +1

    cs.NEcs.CVcs.LGarXiv:1607.03250v12016
  5. Fast Algorithms for Convolutional Neural Networks

    Andrew Lavin, Scott Gray

    cs.NEcs.LGarXiv:1509.09308v22015
  6. Overfitting in adversarially robust deep learning

    Leslie Rice, Eric Wong, J. Zico Kolter

    cs.LGstat.MLarXiv:2002.11569v22020
  7. Integrated photonics on thin-film lithium niobate

    Di Zhu, Linbo Shao, Mengjie Yu +12

    physics.opticsphysics.app-pharXiv:2102.11956v12021
  8. Image-Image Domain Adaptation with Preserved Self-Similarity and Domain-Dissimilarity for Person Re-identification

    Weijian Deng, Liang Zheng, Qixiang Ye +3

    cs.CVarXiv:1711.07027v32017
  9. Omni-Diffusion: Unified Multimodal Understanding and Generation with Masked Discrete Diffusion

    Lijiang Li, Zuwei Long, Yunhang Shen +6

    cs.CVarXiv:2603.06577v22026
  10. UIU-Net: U-Net in U-Net for Infrared Small Object Detection

    Xin Wu, Danfeng Hong, Jocelyn Chanussot

    cs.CVarXiv:2212.00968v12022
  11. THINKSAFE: Self-Generated Safety Alignment for Reasoning Models

    Seanie Lee, Sangwoo Park, Yumin Choi +6

    cs.AIarXiv:2601.23143v42026
  12. The performance of modularity maximization in practical contexts

    Benjamin H. Good, Yves-Alexandre de Montjoye, Aaron Clauset

    physics.data-ancond-mat.dis-nnphysics.soc-pharXiv:0910.0165v22009
  13. ReGuLaR: Variational Latent Reasoning Guided by Rendered Chain-of-Thought

    Fanmeng Wang, Haotian Liu, Guojiang Zhao +2

    cs.CLarXiv:2601.23184v12026
  14. Model-Agnostic Interpretability of Machine Learning

    Marco Tulio Ribeiro, Sameer Singh, Carlos Guestrin

    stat.MLcs.LGarXiv:1606.05386v12016
  15. SiamFC++: Towards Robust and Accurate Visual Tracking with Target Estimation Guidelines

    Yinda Xu, Zeyu Wang, Zuoxin Li +2

    cs.CVarXiv:1911.06188v42019
  16. The Hateful Memes Challenge: Detecting Hate Speech in Multimodal Memes

    Douwe Kiela, Hamed Firooz, Aravind Mohan +4

    cs.AIcs.CLcs.CVarXiv:2005.04790v32020
  17. Terminal Agents Suffice for Enterprise Automation

    Patrice Bechard, Orlando Marquez Ayala, Emily Chen +5

    cs.SEcs.AIcs.CLarXiv:2604.00073v32026
  18. StarCraft II: A New Challenge for Reinforcement Learning

    Oriol Vinyals, Timo Ewalds, Sergey Bartunov +22

    cs.LGcs.AIarXiv:1708.04782v12017
  19. RMA: Rapid Motor Adaptation for Legged Robots

    Ashish Kumar, Zipeng Fu, Deepak Pathak +1

    cs.LGcs.AIcs.CVarXiv:2107.04034v12021
  20. Large Batch Training of Convolutional Networks

    Yang You, Igor Gitman, Boris Ginsburg

    cs.CVarXiv:1708.03888v32017
  21. End-to-end Neural Coreference Resolution

    Kenton Lee, Luheng He, Mike Lewis +1

    cs.CLarXiv:1707.07045v22017
  22. Weakly- and Semi-Supervised Learning of a DCNN for Semantic Image Segmentation

    George Papandreou, Liang-Chieh Chen, Kevin Murphy +1

    cs.CVarXiv:1502.02734v32015
  23. What do you learn from context? Probing for sentence structure in contextualized word representations

    Ian Tenney, Patrick Xia, Berlin Chen +8

    cs.CLarXiv:1905.06316v12019
  24. Exploring Reasoning Reward Model for Agents

    Kaixuan Fan, Kaituo Feng, Manyuan Zhang +7

    cs.AIcs.CLarXiv:2601.22154v22026
  25. A-OKVQA: A Benchmark for Visual Question Answering using World Knowledge

    Dustin Schwenk, Apoorv Khandelwal, Christopher Clark +2

    cs.CVcs.CLarXiv:2206.01718v12022
  26. Semantically Conditioned LSTM-based Natural Language Generation for Spoken Dialogue Systems

    Tsung-Hsien Wen, Milica Gasic, Nikola Mrksic +3

    cs.CLarXiv:1508.01745v22015
  27. Benign Overfitting in Linear Regression

    Peter L. Bartlett, Philip M. Long, Gábor Lugosi +1

    stat.MLcs.LGmath.STarXiv:1906.11300v32019
  28. The Vision Wormhole: Latent-Space Communication in Heterogeneous Multi-Agent Systems

    Xiaoze Liu, Ruowang Zhang, Weichen Yu +7

    cs.CLcs.CVcs.LGarXiv:2602.15382v22026
  29. LiveMedBench: A Contamination-Free Medical Benchmark for LLMs with Automated Rubric Evaluation

    Zhiling Yan, Dingjie Song, Zhe Fang +4

    cs.AIarXiv:2602.10367v12026
  30. Recommendation as Language Processing (RLP): A Unified Pretrain, Personalized Prompt & Predict Paradigm (P5)

    Shijie Geng, Shuchang Liu, Zuohui Fu +2

    cs.IRcs.AIcs.CLarXiv:2203.13366v72022
  31. OPUS: Towards Efficient and Principled Data Selection in Large Language Model Pre-training in Every Iteration

    Shaobo Wang, Xuan Ouyang, Tianyi Xu +9

    cs.CLarXiv:2602.05400v22026
  32. MWM: Mobile World Models for Action-Conditioned Consistent Prediction

    Han Yan, Zishang Xiang, Zeyu Zhang +1

    cs.CVcs.ROarXiv:2603.07799v12026
  33. HiAR: Efficient Autoregressive Long Video Generation via Hierarchical Denoising

    Kai Zou, Dian Zheng, Hongbo Liu +3

    cs.CVarXiv:2603.08703v12026
  34. LIVE: Long-horizon Interactive Video World Modeling

    Junchao Huang, Ziyang Ye, Xinting Hu +5

    cs.CVarXiv:2602.03747v12026
  35. Unity: A General Platform for Intelligent Agents

    Arthur Juliani, Vincent-Pierre Berges, Ervin Teng +8

    cs.LGcs.AIcs.NEarXiv:1809.02627v22018
  36. Temporal Segment Networks for Action Recognition in Videos

    Limin Wang, Yuanjun Xiong, Zhe Wang +4

    cs.CVarXiv:1705.02953v12017
  37. NExT-QA:Next Phase of Question-Answering to Explaining Temporal Actions

    Junbin Xiao, Xindi Shang, Angela Yao +1

    cs.CVcs.AIarXiv:2105.08276v22021
  38. ERNIE 5.0 Technical Report

    Haifeng Wang, Hua Wu, Tian Wu +435

    cs.CLarXiv:2602.04705v12026
  39. pi-GAN: Periodic Implicit Generative Adversarial Networks for 3D-Aware Image Synthesis

    Eric R. Chan, Marco Monteiro, Petr Kellnhofer +2

    cs.CVcs.GRarXiv:2012.00926v22020
  40. Vision2Web: A Hierarchical Benchmark for Visual Website Development with Agent Verification

    Zehai He, Wenyi Hong, Zhen Yang +4

    cs.SEcs.AIarXiv:2603.26648v32026
  41. Medical Image Segmentation Review: The success of U-Net

    Reza Azad, Ehsan Khodapanah Aghdam, Amelie Rauland +7

    eess.IVcs.CVarXiv:2211.14830v12022
  42. LOME: Learning Human-Object Manipulation with Action-Conditioned Egocentric World Model

    Quankai Gao, Jiawei Yang, Qiangeng Xu +2

    cs.CVarXiv:2603.27449v12026
  43. Mobile-GS: Real-time Gaussian Splatting for Mobile Devices

    Xiaobiao Du, Yida Wang, Kun Zhan +1

    cs.CVarXiv:2603.11531v12026
  44. Slither: A Static Analysis Framework For Smart Contracts

    Josselin Feist, Gustavo Grieco, Alex Groce

    cs.SEcs.CRarXiv:1908.09878v12019
  45. A Survey on the Explainability of Supervised Machine Learning

    Nadia Burkart, Marco F. Huber

    cs.LGcs.AIstat.MLarXiv:2011.07876v12020
  46. VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models

    Zixuan Wang, Yuxin Chen, Yuqi Liu +6

    cs.ROarXiv:2603.22003v32026
  47. ILVR: Conditioning Method for Denoising Diffusion Probabilistic Models

    Jooyoung Choi, Sungwon Kim, Yonghyun Jeong +2

    cs.CVarXiv:2108.02938v22021
  48. The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only

    Guilherme Penedo, Quentin Malartic, Daniel Hesslow +6

    cs.CLcs.AIarXiv:2306.01116v12023
  49. AI Gamestore: Scalable, Open-Ended Evaluation of Machine General Intelligence with Human Games

    Lance Ying, Ryan Truong, Prafull Sharma +9

    cs.AIarXiv:2602.17594v12026
  50. Bat Algorithm: Literature Review and Applications

    Xin-She Yang

    cs.AImath.OCarXiv:1308.3900v12013
  51. Understanding data augmentation for classification: when to warp?

    Sebastien C. Wong, Adam Gatt, Victor Stamatescu +1

    cs.CVarXiv:1609.08764v22016
  52. DeepID3: Face Recognition with Very Deep Neural Networks

    Yi Sun, Ding Liang, Xiaogang Wang +1

    cs.CVarXiv:1502.00873v12015
  53. A Survey on Security and Privacy Issues of Bitcoin

    Mauro Conti, Sandeep Kumar E, Chhagan Lal +1

    cs.CRarXiv:1706.00916v32017
  54. ABCNN: Attention-Based Convolutional Neural Network for Modeling Sentence Pairs

    Wenpeng Yin, Hinrich Schütze, Bing Xiang +1

    cs.CLarXiv:1512.05193v42015
  55. Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training

    Gen Li, Nan Duan, Yuejian Fang +3

    cs.CVarXiv:1908.06066v32019
  56. \$OneMillion-Bench: How Far are Language Agents from Human Experts?

    Qianyu Yang, Yang Liu, Jiaqi Li +20

    cs.LGcs.AIcs.CLarXiv:2603.07980v12026
  57. Towards General Text Embeddings with Multi-stage Contrastive Learning

    Zehan Li, Xin Zhang, Yanzhao Zhang +3

    cs.CLarXiv:2308.03281v12023
  58. Understanding disentangling in $β$-VAE

    Christopher P. Burgess, Irina Higgins, Arka Pal +4

    stat.MLcs.AIcs.LGarXiv:1804.03599v12018
  59. Making Convolutional Networks Shift-Invariant Again

    Richard Zhang

    cs.CVcs.LGarXiv:1904.11486v22019
  60. Causal-JEPA: Learning World Models through Object-Level Latent Masking

    Heejeong Nam, Quentin Le Lidec, Lucas Maes +2

    cs.AIarXiv:2602.11389v22026