Computer Vision and Pattern Recognition

Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

10,081 to 10,140 of 18,848

  1. Routing Networks: Adaptive Selection of Non-linear Functions for Multi-Task Learning

    Clemens Rosenbaum, Tim Klinger, Matthew Riemer

    cs.LGcs.CVcs.NEarXiv:1711.01239v22017
  2. What is YOLOv5: A deep look into the internal features of the popular object detector

    Rahima Khanam, Muhammad Hussain

    cs.CVarXiv:2407.20892v12024
  3. Sparse Fuse Dense: Towards High Quality 3D Detection with Depth Completion

    Xiaopei Wu, Liang Peng, Honghui Yang +5

    cs.CVarXiv:2203.09780v22022
  4. Causality-inspired Single-source Domain Generalization for Medical Image Segmentation

    Cheng Ouyang, Chen Chen, Surui Li +4

    cs.CVarXiv:2111.12525v52021
  5. MeMViT: Memory-Augmented Multiscale Vision Transformer for Efficient Long-Term Video Recognition

    Chao-Yuan Wu, Yanghao Li, Karttikeya Mangalam +4

    cs.CVarXiv:2201.08383v22022
  6. EPOS: Estimating 6D Pose of Objects with Symmetries

    Tomas Hodan, Daniel Barath, Jiri Matas

    cs.CVcs.LGcs.ROarXiv:2004.00605v12020
  7. Bipartite Graph Network with Adaptive Message Passing for Unbiased Scene Graph Generation

    Rongjie Li, Songyang Zhang, Bo Wan +1

    cs.CVcs.AIarXiv:2104.00308v22021
  8. Mitigate Bias in Face Recognition using Skewness-Aware Reinforcement Learning

    Mei Wang, Weihong Deng

    cs.CVarXiv:1911.10692v12019
  9. Sapiens: Foundation for Human Vision Models

    Rawal Khirodkar, Timur Bagautdinov, Julieta Martinez +5

    cs.CVarXiv:2408.12569v32024
  10. Breast cancer detection using artificial intelligence techniques: A systematic literature review

    Ali Bou Nassif, Manar Abu Talib, Qassim Nasir +2

    eess.IVcs.AIcs.CVarXiv:2203.04308v12022
  11. Structured Prediction of 3D Human Pose with Deep Neural Networks

    Bugra Tekin, Isinsu Katircioglu, Mathieu Salzmann +2

    cs.CVarXiv:1605.05180v12016
  12. DetNet: A Backbone network for Object Detection

    Zeming Li, Chao Peng, Gang Yu +3

    cs.CVarXiv:1804.06215v22018
  13. Dissecting Person Re-identification from the Viewpoint of Viewpoint

    Xiaoxiao Sun, Liang Zheng

    cs.CVarXiv:1812.02162v62018
  14. Occluded Person Re-identification

    Jiaxuan Zhuo, Zeyu Chen, Jianhuang Lai +1

    cs.CVcs.AIcs.MMarXiv:1804.02792v32018
  15. Particle Video Revisited: Tracking Through Occlusions Using Point Trajectories

    Adam W. Harley, Zhaoyuan Fang, Katerina Fragkiadaki

    cs.CVarXiv:2204.04153v22022
  16. M-LVC: Multiple Frames Prediction for Learned Video Compression

    Jianping Lin, Dong Liu, Houqiang Li +1

    eess.IVcs.CVcs.LGarXiv:2004.10290v12020
  17. Fine-tuned CLIP Models are Efficient Video Learners

    Hanoona Rasheed, Muhammad Uzair Khattak, Muhammad Maaz +2

    cs.CVcs.AIarXiv:2212.03640v32022
  18. Learning Motion Patterns in Videos

    Pavel Tokmakov, Karteek Alahari, Cordelia Schmid

    cs.CVarXiv:1612.07217v22016
  19. EPINET: A Fully-Convolutional Neural Network Using Epipolar Geometry for Depth from Light Field Images

    Changha Shin, Hae-Gon Jeon, Youngjin Yoon +2

    cs.CVarXiv:1804.02379v12018
  20. Segment Anything Model (SAM) for Digital Pathology: Assess Zero-shot Segmentation on Whole Slide Imaging

    Ruining Deng, Can Cui, Quan Liu +13

    eess.IVcs.CVarXiv:2304.04155v12023
  21. Learning Visual Predictive Models of Physics for Playing Billiards

    Katerina Fragkiadaki, Pulkit Agrawal, Sergey Levine +1

    cs.CVarXiv:1511.07404v32015
  22. Deep3D: Fully Automatic 2D-to-3D Video Conversion with Deep Convolutional Neural Networks

    Junyuan Xie, Ross Girshick, Ali Farhadi

    cs.CVarXiv:1604.03650v12016
  23. Not All Pixels Are Equal: Difficulty-aware Semantic Segmentation via Deep Layer Cascade

    Xiaoxiao Li, Ziwei Liu, Ping Luo +2

    cs.CVcs.LGarXiv:1704.01344v12017
  24. Smooth PARAFAC Decomposition for Tensor Completion

    Tatsuya Yokota, Qibin Zhao, Andrzej Cichocki

    cs.CVarXiv:1505.06611v32015
  25. Domain Adaptation via Prompt Learning

    Chunjiang Ge, Rui Huang, Mixue Xie +4

    cs.CVarXiv:2202.06687v12022
  26. Fishr: Invariant Gradient Variances for Out-of-Distribution Generalization

    Alexandre Rame, Corentin Dancette, Matthieu Cord

    cs.LGcs.AIcs.CVarXiv:2109.02934v32021
  27. Self-supervised Augmentation Consistency for Adapting Semantic Segmentation

    Nikita Araslanov, Stefan Roth

    cs.CVcs.LGarXiv:2105.00097v12021
  28. SE3-Nets: Learning Rigid Body Motion using Deep Neural Networks

    Arunkumar Byravan, Dieter Fox

    cs.LGcs.AIcs.CVarXiv:1606.02378v32016
  29. Houdini: Fooling Deep Structured Prediction Models

    Moustapha Cisse, Yossi Adi, Natalia Neverova +1

    stat.MLcs.AIcs.CRarXiv:1707.05373v12017
  30. TextField: Learning A Deep Direction Field for Irregular Scene Text Detection

    Yongchao Xu, Yukang Wang, Wei Zhou +3

    cs.CVarXiv:1812.01393v22018
  31. Densely Semantically Aligned Person Re-Identification

    Zhizheng Zhang, Cuiling Lan, Wenjun Zeng +1

    cs.CVarXiv:1812.08967v22018
  32. Micro-Net: A unified model for segmentation of various objects in microscopy images

    Shan E Ahmed Raza, Linda Cheung, Muhammad Shaban +5

    cs.CVarXiv:1804.08145v22018
  33. Training Convolutional Networks with Noisy Labels

    Sainbayar Sukhbaatar, Joan Bruna, Manohar Paluri +2

    cs.CVcs.LGcs.NEarXiv:1406.2080v42014
  34. Feature-metric Loss for Self-supervised Learning of Depth and Egomotion

    Chang Shu, Kun Yu, Zhixiang Duan +1

    cs.CVarXiv:2007.10603v12020
  35. Physics-Guided Flow Matching for CT Image Reconstruction

    Davide Evangelista

    cs.AIcs.CVarXiv:2608.28256v12026
  36. Variants of RMSProp and Adagrad with Logarithmic Regret Bounds

    Mahesh Chandra Mukkamala, Matthias Hein

    cs.LGcs.AIcs.CVarXiv:1706.05507v22017
  37. SPACE: Unsupervised Object-Oriented Scene Representation via Spatial Attention and Decomposition

    Zhixuan Lin, Yi-Fu Wu, Skand Vishwanath Peri +5

    cs.LGcs.CVeess.IVarXiv:2001.02407v32020
  38. Improving the Improved Training of Wasserstein GANs: A Consistency Term and Its Dual Effect

    Xiang Wei, Boqing Gong, Zixia Liu +2

    cs.CVcs.LGstat.MLarXiv:1803.01541v12018
  39. GeoTransformer: Fast and Robust Point Cloud Registration with Geometric Transformer

    Zheng Qin, Hao Yu, Changjian Wang +5

    cs.CVarXiv:2308.03768v12023
  40. Augmented Reality and Robotics: A Survey and Taxonomy for AR-enhanced Human-Robot Interaction and Robotic Interfaces

    Ryo Suzuki, Adnan Karim, Tian Xia +2

    cs.ROcs.CVcs.HCarXiv:2203.03254v12022
  41. Lightweight Deep Learning for Resource-Constrained Environments: A Survey

    Hou-I Liu, Marco Galindo, Hongxia Xie +4

    cs.CVcs.LGarXiv:2404.07236v22024
  42. Fast Alternating Linearization Methods for Minimizing the Sum of Two Convex Functions

    Donald Goldfarb, Shiqian Ma, Katya Scheinberg

    math.OCcs.CVmath.NAarXiv:0912.4571v22009
  43. Beyond Temporal Pooling: Recurrence and Temporal Convolutions for Gesture Recognition in Video

    Lionel Pigou, Aäron van den Oord, Sander Dieleman +2

    cs.CVcs.AIcs.LGarXiv:1506.01911v32015
  44. Locate Anything in Videos: Rethinking Efficient Generative Spatio-Temporal Video Grounding

    Hanoona Rasheed, Haania Siddiqui, Ming-Hsuan Yang +2

    cs.CVarXiv:2608.28192v12026
  45. Monte Carlo Convolution for Learning on Non-Uniformly Sampled Point Clouds

    Pedro Hermosilla, Tobias Ritschel, Pere-Pau Vázquez +2

    cs.CVarXiv:1806.01759v22018
  46. Density Map Guided Object Detection in Aerial Images

    Changlin Li, Taojiannan Yang, Sijie Zhu +2

    cs.CVarXiv:2004.05520v12020
  47. Cross-domain Detection via Graph-induced Prototype Alignment

    Minghao Xu, Hang Wang, Bingbing Ni +2

    cs.CVarXiv:2003.12849v12020
  48. SparseDrive: End-to-End Autonomous Driving via Sparse Scene Representation

    Wenchao Sun, Xuewu Lin, Yining Shi +3

    cs.CVarXiv:2405.19620v22024
  49. Person Re-identification by Contour Sketch under Moderate Clothing Change

    Qize Yang, Ancong Wu, Wei-Shi Zheng

    cs.CVarXiv:2002.02295v12020
  50. Cut and Learn for Unsupervised Object Detection and Instance Segmentation

    Xudong Wang, Rohit Girdhar, Stella X. Yu +1

    cs.CVcs.AIcs.LGarXiv:2301.11320v12023
  51. 3D Dynamic Scene Graphs: Actionable Spatial Perception with Places, Objects, and Humans

    Antoni Rosinol, Arjun Gupta, Marcus Abate +2

    cs.ROcs.AIcs.CVarXiv:2002.06289v22020
  52. Uncertainty-aware multi-view co-training for semi-supervised medical image segmentation and domain adaptation

    Yingda Xia, Dong Yang, Zhiding Yu +7

    cs.CVarXiv:2006.16806v12020
  53. GP-GAN: Towards Realistic High-Resolution Image Blending

    Huikai Wu, Shuai Zheng, Junge Zhang +1

    cs.CVarXiv:1703.07195v32017
  54. A Survey of Vision-Language Pre-Trained Models

    Yifan Du, Zikang Liu, Junyi Li +1

    cs.CVcs.CLcs.LGarXiv:2202.10936v22022
  55. Automatic 3D liver location and segmentation via convolutional neural networks and graph cut

    Fang Lu, Fa Wu, Peijun Hu +2

    cs.CVarXiv:1605.03012v12016
  56. GS-IR: 3D Gaussian Splatting for Inverse Rendering

    Zhihao Liang, Qi Zhang, Ying Feng +2

    cs.CVarXiv:2311.16473v32023
  57. Improving Vision-and-Language Navigation with Image-Text Pairs from the Web

    Arjun Majumdar, Ayush Shrivastava, Stefan Lee +3

    cs.CVcs.AIcs.CLarXiv:2004.14973v22020
  58. Beyond the Pixel-Wise Loss for Topology-Aware Delineation

    Agata Mosinska, Pablo Marquez-Neila, Mateusz Kozinski +1

    cs.CVarXiv:1712.02190v12017
  59. Robust and Generalizable Visual Representation Learning via Random Convolutions

    Zhenlin Xu, Deyi Liu, Junlin Yang +2

    cs.CVcs.LGarXiv:2007.13003v32020
  60. Learning a Deep ConvNet for Multi-label Classification with Partial Labels

    Thibaut Durand, Nazanin Mehrasa, Greg Mori

    cs.CVarXiv:1902.09720v12019