Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

56,341 to 56,400 of 61,111

  1. CNN-generated images are surprisingly easy to spot... for now

    Sheng-Yu Wang, Oliver Wang, Richard Zhang +2

    cs.CVarXiv:1912.11035v22019
  2. LaSOT: A High-quality Benchmark for Large-scale Single Object Tracking

    Heng Fan, Liting Lin, Fan Yang +7

    cs.CVarXiv:1809.07845v22018
  3. Joint Training of a Convolutional Network and a Graphical Model for Human Pose Estimation

    Jonathan Tompson, Arjun Jain, Yann LeCun +1

    cs.CVarXiv:1406.2984v22014
  4. E(n) Equivariant Graph Neural Networks

    Victor Garcia Satorras, Emiel Hoogeboom, Max Welling

    cs.LGstat.MLarXiv:2102.09844v32021
  5. Differentially Private Empirical Risk Minimization

    Kamalika Chaudhuri, Claire Monteleoni, Anand D. Sarwate

    cs.LGcs.AIcs.CRarXiv:0912.0071v52009
  6. GenRecon: Bridging Generative Priors for Multi-View 3D Scene Reconstruction

    Katharina Schmid, Nicolas von Lützow, Jozef Hladký +2

    cs.CVarXiv:2605.23888v12026
  7. Transformers in Time Series: A Survey

    Qingsong Wen, Tian Zhou, Chaoli Zhang +4

    cs.LGcs.AIeess.SParXiv:2202.07125v52022
  8. Convex Low-resource Accent-Robust Language Detection in Speech Recognition

    Miria Feng, William Tan, Mert Pilanci

    cs.LGarXiv:2605.23235v12026
  9. Unsupervised Pixel-Level Domain Adaptation with Generative Adversarial Networks

    Konstantinos Bousmalis, Nathan Silberman, David Dohan +2

    cs.CVarXiv:1612.05424v22016
  10. Learning Word Vectors for 157 Languages

    Edouard Grave, Piotr Bojanowski, Prakhar Gupta +2

    cs.CLcs.LGarXiv:1802.06893v22018
  11. ICNet for Real-Time Semantic Segmentation on High-Resolution Images

    Hengshuang Zhao, Xiaojuan Qi, Xiaoyong Shen +2

    cs.CVarXiv:1704.08545v22017
  12. PANNs: Large-Scale Pretrained Audio Neural Networks for Audio Pattern Recognition

    Qiuqiang Kong, Yin Cao, Turab Iqbal +3

    cs.SDeess.ASarXiv:1912.10211v52019
  13. Provable defenses against adversarial examples via the convex outer adversarial polytope

    Eric Wong, J. Zico Kolter

    cs.LGcs.AImath.OCarXiv:1711.00851v32017
  14. The 2017 DAVIS Challenge on Video Object Segmentation

    Jordi Pont-Tuset, Federico Perazzi, Sergi Caelles +3

    cs.CVarXiv:1704.00675v32017
  15. Dense-Captioning Events in Videos

    Ranjay Krishna, Kenji Hata, Frederic Ren +2

    cs.CVarXiv:1705.00754v12017
  16. Injecting Image Guidance into Text-Conditioned Diffusion Models at Inference

    Agata Żywot, Iason Skylitsis, Thijmen Nijdam +4

    cs.CVarXiv:2605.25191v12026
  17. Benchmarking Composed Image Retrieval for Applied Earth Observation

    Bill Psomas, Dionysis Christopoulos, Thanasis Petropoulos +6

    cs.CVarXiv:2605.24442v12026
  18. Bag of Tricks for Image Classification with Convolutional Neural Networks

    Tong He, Zhi Zhang, Hang Zhang +3

    cs.CVarXiv:1812.01187v22018
  19. Directional Alignment Mitigates Reward Hacking in Reinforcement Learning for Language Models

    Wenlong Deng, Jiaji Huang, Kaan Ozkara +4

    cs.LGcs.CLarXiv:2605.25189v12026
  20. The Ethics of AI Ethics -- An Evaluation of Guidelines

    Thilo Hagendorff

    cs.AIcs.CYcs.LGarXiv:1903.03425v22019
  21. Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity

    Yao Lu, Max Bartolo, Alastair Moore +2

    cs.CLcs.AIarXiv:2104.08786v22021
  22. DarkForest: Less Talk, Higher Accuracy for Multi-Agent LLMs

    Yi Li, Songtao Wei, Dongming Jiang +3

    cs.AIarXiv:2605.25188v12026
  23. Growing a Neural Network in Breadth, Depth, and Time

    Eivinas Butkus, Kedar Garzón Gupta, Nikolaus Kriegeskorte

    q-bio.NCcs.LGcs.NEarXiv:2605.25174v12026
  24. Frozen in Time: A Joint Video and Image Encoder for End-to-End Retrieval

    Max Bain, Arsha Nagrani, Gül Varol +1

    cs.CVarXiv:2104.00650v22021
  25. PhotoFlow: Agentic 3D Virtual Photography Missions

    Jiarui Guo, Haojia Wei, Yiming Zhang +5

    cs.CVcs.AIcs.MAarXiv:2605.23771v12026
  26. Domain Separation Networks

    Konstantinos Bousmalis, George Trigeorgis, Nathan Silberman +2

    cs.CVarXiv:1608.06019v12016
  27. Neural Attentive Session-based Recommendation

    Jing Li, Pengjie Ren, Zhumin Chen +2

    cs.IRarXiv:1711.04725v12017
  28. Towards Evaluation Engineering: An Empirical Study of ML Evaluation Harnesses in the Wild

    Zhimin Zhao, Zehao Wang, Abdul Ali Bangash +2

    cs.SEcs.AIcs.LGarXiv:2605.24213v12026
  29. Decoding the Critique Mechanism in Large Reasoning Models

    Hoang Phan, Quang H. Nguyen, Hung T. Q. Le +3

    cs.LGarXiv:2603.16331v22026
  30. Recent Advances in Natural Language Processing via Large Pre-Trained Language Models: A Survey

    Bonan Min, Hayley Ross, Elior Sulem +6

    cs.CLcs.AIcs.LGarXiv:2111.01243v12021
  31. OctNet: Learning Deep 3D Representations at High Resolutions

    Gernot Riegler, Ali Osman Ulusoy, Andreas Geiger

    cs.CVarXiv:1611.05009v42016
  32. Self-Supervised Learning of Pretext-Invariant Representations

    Ishan Misra, Laurens van der Maaten

    cs.CVcs.LGarXiv:1912.01991v12019
  33. Molecular Graph Convolutions: Moving Beyond Fingerprints

    Steven Kearnes, Kevin McCloskey, Marc Berndl +2

    stat.MLcs.LGarXiv:1603.00856v32016
  34. Knowing When to Look: Adaptive Attention via A Visual Sentinel for Image Captioning

    Jiasen Lu, Caiming Xiong, Devi Parikh +1

    cs.CVcs.AIarXiv:1612.01887v22016
  35. VaaWIT: Visual-Aware Adaptation of Large Language Models for Multilingual Web Image Translation

    Bo Li, Ronghao Chen, Ningyuan Deng +3

    cs.CVcs.AIarXiv:2605.24675v12026
  36. From Activation to Specificity: Automating Counterfactual Testing of Visual Representations in the Human Brain

    Yuval Golbari, Navve Wasserman, Matias Cosarinsky +5

    cs.CVarXiv:2605.23895v22026
  37. Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs

    Bo Li, Tianyu Dong, Shaolin Zhu +1

    cs.CLcs.AIarXiv:2605.24681v12026
  38. Can Predicted Dynamics Exist in the Physical World?

    Barak Or

    cs.ROcs.AIarXiv:2606.00089v22026
  39. Reinforcement Learning with Deep Energy-Based Policies

    Tuomas Haarnoja, Haoran Tang, Pieter Abbeel +1

    cs.LGcs.AIarXiv:1702.08165v22017
  40. Silent Failures in Physical AI: A Literature Review of Runtime Action Authorization for Autonomous Systems

    Barak Or

    cs.ROcs.AIarXiv:2606.00090v12026
  41. Efficient Transformers: A Survey

    Yi Tay, Mostafa Dehghani, Dara Bahri +1

    cs.LGcs.AIcs.CLarXiv:2009.06732v32020
  42. Learning to Plan Chemical Syntheses

    Marwin H. S. Segler, Mike Preuss, Mark P. Waller

    cs.AIcs.LGphysics.chem-pharXiv:1708.04202v12017
  43. The Online Laboratory: Conducting Experiments in a Real Labor Market

    John J. Horton, David G. Rand, Richard J. Zeckhauser

    cs.HCarXiv:1004.2931v12010
  44. OSMnx: New Methods for Acquiring, Constructing, Analyzing, and Visualizing Complex Street Networks

    Geoff Boeing

    cs.SIphysics.soc-pharXiv:1611.01890v52016
  45. Your Embedding Model is SMARTer Than You Think

    Jianrui Zhang, Hyun Jung Lee, Sukanta Ganguly +3

    cs.IRcs.AIcs.CVarXiv:2605.24938v12026
  46. Jailbreaking Black Box Large Language Models in Twenty Queries

    Patrick Chao, Alexander Robey, Edgar Dobriban +3

    cs.LGcs.AIarXiv:2310.08419v42023
  47. STREAM: A Data-Centric Framework for Mining High-Value Task-Oriented Dialogues from Streaming Media

    Liang Xue, Haoyu Liu, Cheng Wang +3

    cs.CLcs.AIarXiv:2605.25162v12026
  48. OK-VQA: A Visual Question Answering Benchmark Requiring External Knowledge

    Kenneth Marino, Mohammad Rastegari, Ali Farhadi +1

    cs.CVcs.CLarXiv:1906.00067v22019
  49. Graph-based Anomaly Detection and Description: A Survey

    Leman Akoglu, Hanghang Tong, Danai Koutra

    cs.SIcs.CRarXiv:1404.4679v22014
  50. The Open Images Dataset V4: Unified image classification, object detection, and visual relationship detection at scale

    Alina Kuznetsova, Hassan Rom, Neil Alldrin +9

    cs.CVarXiv:1811.00982v22018
  51. ScaleWoB: Guiding GUI Agents with Coding Agents via Large-Scale Environmental Synthesis

    Guohong Liu, Jialei Ye, Pengzhi Gao +4

    cs.AIarXiv:2605.25160v22026
  52. DeepGCNs: Can GCNs Go as Deep as CNNs?

    Guohao Li, Matthias Müller, Ali Thabet +1

    cs.CVcs.LGarXiv:1904.03751v22019
  53. Hypercolumns for Object Segmentation and Fine-grained Localization

    Bharath Hariharan, Pablo Arbeláez, Ross Girshick +1

    cs.CVarXiv:1411.5752v22014
  54. Don't Guess, Just Ask: Resolving Ambiguity in Referring Segmentation via Multi-turn Clarification

    Yuting Yang, Haichao Jiang, Tianming Liang +2

    cs.CVarXiv:2605.17531v22026
  55. Staple: Complementary Learners for Real-Time Tracking

    Luca Bertinetto, Jack Valmadre, Stuart Golodetz +2

    cs.CVarXiv:1512.01355v22015
  56. Opening the Black Box of Deep Neural Networks via Information

    Ravid Shwartz-Ziv, Naftali Tishby

    cs.LGarXiv:1703.00810v32017
  57. Multi-view Consistent 3D Gaussian Head Avatars 'without' Multi-view Generation

    Aviral Chharia, Fernando De la Torre

    cs.CVcs.GRcs.ROarXiv:2605.25220v12026
  58. Countering Adversarial Images using Input Transformations

    Chuan Guo, Mayank Rana, Moustapha Cisse +1

    cs.CVarXiv:1711.00117v32017
  59. A General-Purpose Machine Learning Framework for Predicting Properties of Inorganic Materials

    Logan Ward, Ankit Agrawal, Alok Choudhary +1

    cond-mat.mtrl-sciarXiv:1606.09551v22016
  60. Exploiting Similarities among Languages for Machine Translation

    Tomas Mikolov, Quoc V. Le, Ilya Sutskever

    cs.CLarXiv:1309.4168v12013