Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

39,361 to 39,420 of 61,134

  1. HybridFlow: A Flexible and Efficient RLHF Framework

    Guangming Sheng, Chi Zhang, Zilingfeng Ye +6

    cs.LGcs.DCarXiv:2409.19256v22024
  2. YOLO-World: Real-Time Open-Vocabulary Object Detection

    Tianheng Cheng, Lin Song, Yixiao Ge +3

    cs.CVarXiv:2401.17270v32024
  3. DUSt3R: Geometric 3D Vision Made Easy

    Shuzhe Wang, Vincent Leroy, Yohann Cabon +2

    cs.CVarXiv:2312.14132v32023
  4. Metric3D: Towards Zero-shot Metric 3D Prediction from A Single Image

    Wei Yin, Chi Zhang, Hao Chen +5

    cs.CVcs.AIarXiv:2307.10984v12023
  5. Generative Agents: Interactive Simulacra of Human Behavior

    Joon Sung Park, Joseph C. O'Brien, Carrie J. Cai +3

    cs.HCcs.AIcs.LGarXiv:2304.03442v22023
  6. EVA-CLIP: Improved Training Techniques for CLIP at Scale

    Quan Sun, Yuxin Fang, Ledell Wu +2

    cs.CVarXiv:2303.15389v12023
  7. Tri-Perspective View for Vision-Based 3D Semantic Occupancy Prediction

    Yuanhui Huang, Wenzhao Zheng, Yunpeng Zhang +2

    cs.CVcs.AIcs.LGarXiv:2302.07817v22023
  8. RTMDet: An Empirical Study of Designing Real-Time Object Detectors

    Chengqi Lyu, Wenwei Zhang, Haian Huang +5

    cs.CVarXiv:2212.07784v22022
  9. Magic3D: High-Resolution Text-to-3D Content Creation

    Chen-Hsuan Lin, Jun Gao, Luming Tang +7

    cs.CVcs.GRcs.LGarXiv:2211.10440v22022
  10. Social Simulacra: Creating Populated Prototypes for Social Computing Systems

    Joon Sung Park, Lindsay Popowski, Carrie J. Cai +3

    cs.HCarXiv:2208.04024v12022
  11. Are Transformers Effective for Time Series Forecasting?

    Ailing Zeng, Muxi Chen, Lei Zhang +1

    cs.AIcs.LGarXiv:2205.13504v32022
  12. ViM: Out-Of-Distribution with Virtual-logit Matching

    Haoqi Wang, Zhizhong Li, Litong Feng +1

    cs.CVarXiv:2203.10807v12022
  13. Point-NeRF: Point-based Neural Radiance Fields

    Qiangeng Xu, Zexiang Xu, Julien Philip +4

    cs.CVarXiv:2201.08845v72022
  14. Vision Transformer with Deformable Attention

    Zhuofan Xia, Xuran Pan, Shiji Song +2

    cs.CVarXiv:2201.00520v32022
  15. Task-Oriented Multi-User Semantic Communications

    Huiqiang Xie, Zhijin Qin, Xiaoming Tao +1

    eess.SParXiv:2112.10255v12021
  16. DyTox: Transformers for Continual Learning with DYnamic TOken eXpansion

    Arthur Douillard, Alexandre Ramé, Guillaume Couairon +1

    cs.CVcs.LGarXiv:2111.11326v32021
  17. Pivotal Tuning for Latent-based Editing of Real Images

    Daniel Roich, Ron Mokady, Amit H. Bermano +1

    cs.CVarXiv:2106.05744v12021
  18. Semi-Supervised Semantic Segmentation with Cross Pseudo Supervision

    Xiaokang Chen, Yuhui Yuan, Gang Zeng +1

    cs.CVarXiv:2106.01226v22021
  19. Skillful Precipitation Nowcasting using Deep Generative Models of Radar

    Suman Ravuri, Karel Lenc, Matthew Willson +17

    cs.LGarXiv:2104.00954v12021
  20. The Surprising Effectiveness of PPO in Cooperative, Multi-Agent Games

    Chao Yu, Akash Velu, Eugene Vinitsky +4

    cs.LGcs.AIcs.MAarXiv:2103.01955v42021
  21. Machine learning accelerated computational fluid dynamics

    Dmitrii Kochkov, Jamie A. Smith, Ayya Alieva +3

    physics.flu-dyncs.LGarXiv:2102.01010v12021
  22. E(3)-Equivariant Graph Neural Networks for Data-Efficient and Accurate Interatomic Potentials

    Simon Batzner, Albert Musaelian, Lixin Sun +6

    physics.comp-phcond-mat.mtrl-scics.LGarXiv:2101.03164v32021
  23. Non-Rigid Neural Radiance Fields: Reconstruction and Novel View Synthesis of a Dynamic Scene From Monocular Video

    Edgar Tretschk, Ayush Tewari, Vladislav Golyanik +3

    cs.CVcs.GRarXiv:2012.12247v42020
  24. TDN: Temporal Difference Networks for Efficient Action Recognition

    Limin Wang, Zhan Tong, Bin Ji +1

    cs.CVarXiv:2012.10071v22020
  25. NeRD: Neural Reflectance Decomposition from Image Collections

    Mark Boss, Raphael Braun, Varun Jampani +3

    cs.CVcs.GRcs.LGarXiv:2012.03918v42020
  26. Align Deep Features for Oriented Object Detection

    Jiaming Han, Jian Ding, Jie Li +1

    cs.CVarXiv:2008.09397v32020
  27. MediaPipe Hands: On-device Real-time Hand Tracking

    Fan Zhang, Valentin Bazarevsky, Andrey Vakunov +4

    cs.CVarXiv:2006.10214v12020
  28. X3D: Expanding Architectures for Efficient Video Recognition

    Christoph Feichtenhofer

    cs.CVarXiv:2004.04730v12020
  29. BlendMask: Top-Down Meets Bottom-Up for Instance Segmentation

    Hao Chen, Kunyang Sun, Zhi Tian +3

    cs.CVarXiv:2001.00309v32020
  30. PVN3D: A Deep Point-wise 3D Keypoints Voting Network for 6DoF Pose Estimation

    Yisheng He, Wei Sun, Haibin Huang +3

    cs.CVcs.ROarXiv:1911.04231v22019
  31. CodeSearchNet Challenge: Evaluating the State of Semantic Code Search

    Hamel Husain, Ho-Hsiang Wu, Tiferet Gazit +2

    cs.LGcs.IRcs.SEarXiv:1909.09436v32019
  32. FSGAN: Subject Agnostic Face Swapping and Reenactment

    Yuval Nirkin, Yosi Keller, Tal Hassner

    cs.CVcs.GRcs.LGarXiv:1908.05932v12019
  33. R3Det: Refined Single-Stage Detector with Feature Refinement for Rotating Object

    Xue Yang, Junchi Yan, Ziming Feng +1

    cs.CVcs.LGeess.IVarXiv:1908.05612v62019
  34. Learning Lightweight Lane Detection CNNs by Self Attention Distillation

    Yuenan Hou, Zheng Ma, Chunxiao Liu +1

    cs.CVarXiv:1908.00821v12019
  35. Deep Generative Modeling for Mechanistic-based Learning and Design of Metamaterial Systems

    Liwei Wang, Yu-Chin Chan, Faez Ahmed +3

    cs.CEcs.LGstat.MLarXiv:2006.15274v22020
  36. Joint Radar and Communication Design: Applications, State-of-the-art, and the Road Ahead

    Fan Liu, Christos Masouros, Athina Petropulu +2

    eess.SParXiv:1906.00789v12019
  37. Model-Based Reinforcement Learning with Value-Targeted Regression

    Alex Ayoub, Zeyu Jia, Csaba Szepesvari +2

    cs.LGstat.MLarXiv:2006.01107v12020
  38. 3D Packing for Self-Supervised Monocular Depth Estimation

    Vitor Guizilini, Rares Ambrus, Sudeep Pillai +2

    cs.CVcs.LGcs.ROarXiv:1905.02693v42019
  39. GA-Net: Guided Aggregation Net for End-to-end Stereo Matching

    Feihu Zhang, Victor Prisacariu, Ruigang Yang +1

    cs.CVarXiv:1904.06587v12019
  40. Progressive Image Deraining Networks: A Better and Simpler Baseline

    Dongwei Ren, Wangmeng Zuo, Qinghua Hu +2

    cs.CVarXiv:1901.09221v32019
  41. DeepSDF: Learning Continuous Signed Distance Functions for Shape Representation

    Jeong Joon Park, Peter Florence, Julian Straub +2

    cs.CVarXiv:1901.05103v12019
  42. Normalized Object Coordinate Space for Category-Level 6D Object Pose and Size Estimation

    He Wang, Srinath Sridhar, Jingwei Huang +3

    cs.CVarXiv:1901.02970v22019
  43. High Quality Monocular Depth Estimation via Transfer Learning

    Ibraheem Alhashim, Peter Wonka

    cs.CVarXiv:1812.11941v22018
  44. Ensemble-based Multi-Filter Feature Selection Method for DDoS Detection in Cloud Computing

    Opeyemi Osanaiye, Kim-Kwang Raymond Choo2, Ali Dehghantanha +2

    cs.CRarXiv:1807.10443v12018
  45. Cable Manipulation with a Tactile-Reactive Gripper

    Yu She, Shaoxiong Wang, Siyuan Dong +3

    cs.ROeess.IVeess.SYarXiv:1910.02860v32019
  46. Strike (with) a Pose: Neural Networks Are Easily Fooled by Strange Poses of Familiar Objects

    Michael A. Alcorn, Qi Li, Zhitao Gong +4

    cs.CVcs.LGarXiv:1811.11553v32018
  47. Benchmark Analysis of Representative Deep Neural Network Architectures

    Simone Bianco, Remi Cadene, Luigi Celona +1

    cs.CVarXiv:1810.00736v22018
  48. Sequential Neural Likelihood: Fast Likelihood-free Inference with Autoregressive Flows

    George Papamakarios, David C. Sterratt, Iain Murray

    stat.MLcs.LGarXiv:1805.07226v22018
  49. Large Language Models are Few-shot Testers: Exploring LLM-based General Bug Reproduction

    Sungmin Kang, Juyeon Yoon, Shin Yoo

    cs.SEarXiv:2209.11515v32022
  50. WHFast: A fast and unbiased implementation of a symplectic Wisdom-Holman integrator for long term gravitational simulations

    Hanno Rein, Daniel Tamayo

    astro-ph.EPastro-ph.IMmath.NAarXiv:1506.01084v12015
  51. DEIM: DETR with Improved Matching for Fast Convergence

    Shihua Huang, Zhichao Lu, Xiaodong Cun +3

    cs.CVcs.AIarXiv:2412.04234v32024
  52. Extending the computational reach of a noisy superconducting quantum processor

    Abhinav Kandala, Kristan Temme, Antonio D. Corcoles +3

    quant-pharXiv:1805.04492v12018
  53. On the Equivalence between Kernel Quadrature Rules and Random Feature Expansions

    Francis Bach

    cs.LGmath.NAstat.MLarXiv:1502.06800v22015
  54. Decoupling Direction and Norm for Efficient Gradient-Based L2 Adversarial Attacks and Defenses

    Jérôme Rony, Luiz G. Hafemann, Luiz S. Oliveira +3

    cs.CVcs.CRcs.LGarXiv:1811.09600v32018
  55. Vision Meets Drones: A Challenge

    Pengfei Zhu, Longyin Wen, Xiao Bian +2

    cs.CVarXiv:1804.07437v22018
  56. Point Convolutional Neural Networks by Extension Operators

    Matan Atzmon, Haggai Maron, Yaron Lipman

    cs.CVarXiv:1803.10091v12018
  57. How2: A Large-scale Dataset for Multimodal Language Understanding

    Ramon Sanabria, Ozan Caglayan, Shruti Palaskar +4

    cs.CLarXiv:1811.00347v22018
  58. Learning Spectral-Spatial-Temporal Features via a Recurrent Convolutional Neural Network for Change Detection in Multispectral Imagery

    Lichao Mou, Lorenzo Bruzzone, Xiao Xiang Zhu

    cs.CVarXiv:1803.02642v12018
  59. AutoML to Date and Beyond: Challenges and Opportunities

    Shubhra Kanti Karmaker Santu, Md. Mahadi Hassan, Micah J. Smith +3

    cs.LGcs.AIarXiv:2010.10777v42020
  60. Recurrent Slice Networks for 3D Segmentation of Point Clouds

    Qiangui Huang, Weiyue Wang, Ulrich Neumann

    cs.CVarXiv:1802.04402v22018