Computer Vision and Pattern Recognition

Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,221 to 8,280 of 18,848

  1. MID-Fusion: Octree-based Object-Level Multi-Instance Dynamic SLAM

    Binbin Xu, Wenbin Li, Dimos Tzoumanikas +3

    cs.ROcs.CVarXiv:1812.07976v42018
  2. Breaking the Dilemma of Medical Image-to-image Translation

    Lingke Kong, Chenyu Lian, Detian Huang +3

    eess.IVcs.CVarXiv:2110.06465v22021
  3. Generative Adversarial Minority Oversampling

    Sankha Subhra Mullick, Shounak Datta, Swagatam Das

    cs.CVcs.LGarXiv:1903.09730v32019
  4. PMP-Net: Point Cloud Completion by Learning Multi-step Point Moving Paths

    Xin Wen, Peng Xiang, Zhizhong Han +4

    cs.CVarXiv:2012.03408v32020
  5. K-Radar: 4D Radar Object Detection for Autonomous Driving in Various Weather Conditions

    Dong-Hee Paek, Seung-Hyun Kong, Kevin Tirta Wijaya

    cs.CVcs.AIarXiv:2206.08171v42022
  6. ALICE: Towards Understanding Adversarial Learning for Joint Distribution Matching

    Chunyuan Li, Hao Liu, Changyou Chen +4

    stat.MLcs.AIcs.CVarXiv:1709.01215v22017
  7. Alpha-Refine: Boosting Tracking Performance by Precise Bounding Box Estimation

    Bin Yan, Xinyu Zhang, Dong Wang +2

    cs.CVarXiv:2012.06815v32020
  8. Repeatability Is Not Enough: Learning Affine Regions via Discriminability

    Dmytro Mishkin, Filip Radenovic, Jiri Matas

    cs.CVcs.NEarXiv:1711.06704v42017
  9. TransVOD: End-to-End Video Object Detection with Spatial-Temporal Transformers

    Qianyu Zhou, Xiangtai Li, Lu He +5

    cs.CVarXiv:2201.05047v42022
  10. An end-to-end TextSpotter with Explicit Alignment and Attention

    Tong He, Zhi Tian, Weilin Huang +3

    cs.CVarXiv:1803.03474v32018
  11. Texture Synthesis with Spatial Generative Adversarial Networks

    Nikolay Jetchev, Urs Bergmann, Roland Vollgraf

    cs.CVstat.MLarXiv:1611.08207v42016
  12. GarmentWeaver: Schema-Aware Structured Synthesis for Multimodal Sewing Patterns

    Yinwen Lu, Weihao Luo, Yueqi Zhong

    cs.AIcs.CVarXiv:2608.30550v12026
  13. Pix2Vox++: Multi-scale Context-aware 3D Object Reconstruction from Single and Multiple Images

    Haozhe Xie, Hongxun Yao, Shengping Zhang +2

    cs.CVarXiv:2006.12250v22020
  14. BRF-GS: Hyperspectral Bidirectional Reflectance Factor Modeling and Image Generation Based on 3D Gaussian Splatting

    Yiling Yao, Wenjuan Zhang, Bowen Wang +3

    cs.CVarXiv:2608.31159v12026
  15. Foreground-aware Pyramid Reconstruction for Alignment-free Occluded Person Re-identification

    Lingxiao He, Yinggang Wang, Wu Liu +4

    cs.CVarXiv:1904.04975v22019
  16. Shape, Light, and Material Decomposition from Images using Monte Carlo Rendering and Denoising

    Jon Hasselgren, Nikolai Hofmann, Jacob Munkberg

    cs.GRcs.CVarXiv:2206.03380v22022
  17. AI-enabled Low-Cost 3D Maize Ear Morphometry Platform at Breeding Scale

    Therin Young, Elijah Rodriguez, Lisa Coffey +4

    cs.CVarXiv:2608.30161v12026
  18. Perceptual Adversarial Robustness: Defense Against Unseen Threat Models

    Cassidy Laidlaw, Sahil Singla, Soheil Feizi

    cs.LGcs.CVstat.MLarXiv:2006.12655v42020
  19. Hybrid-SORT: Weak Cues Matter for Online Multi-Object Tracking

    Mingzhan Yang, Guangxin Han, Bin Yan +4

    cs.CVarXiv:2308.00783v22023
  20. CANVAS: Consistency-Aware Navigation via Visual Adaptive Sampling for Long-Context Text-to-SVG Generation

    Yichen Wu, Haoxuan Qu, Yihang Lou +2

    cs.CVarXiv:2608.30689v12026
  21. FaceSnap: Real-Time Personalized Lightstage Facial Performance Capture

    Rukhshanda Hussain, Noé Artru, Emeline Got +6

    cs.CVarXiv:2608.31033v12026
  22. Proximity3D: Shape from Capacitive Proximity on Sensing Manifold

    Hao Chen, Chenming Wu, Chun Ping Lam +6

    cs.CVcs.CGcs.GRarXiv:2608.30344v22026
  23. Texture image analysis and texture classification methods - A review

    Laleh Armi, Shervan Fekri-Ershad

    cs.CVarXiv:1904.06554v12019
  24. Multi-View Reflective Surface Inspection via Semantic-Saliency Cross-Verification

    Van-Giang Nguyen, Thanh-Tuan Tran, Xuan-Hieu Phan +1

    cs.CVarXiv:2608.30997v12026
  25. RL-CycleGAN: Reinforcement Learning Aware Simulation-To-Real

    Kanishka Rao, Chris Harris, Alex Irpan +3

    cs.ROcs.CVcs.LGarXiv:2006.09001v12020
  26. Segmentation of Glioma Tumors in Brain Using Deep Convolutional Neural Network

    Saddam Hussain, Syed Muhammad Anwar, Muhammad Majid

    cs.CVarXiv:1708.00377v12017
  27. Contrastive Learning from Extremely Augmented Skeleton Sequences for Self-supervised Action Recognition

    Tianyu Guo, Hong Liu, Zhan Chen +3

    cs.CVarXiv:2112.03590v12021
  28. Generative Cooperative Learning for Unsupervised Video Anomaly Detection

    Muhammad Zaigham Zaheer, Arif Mahmood, Muhammad Haris Khan +3

    cs.CVarXiv:2203.03962v12022
  29. GAFT: Geo-Anchored Fine-Tuning for Hazard Identification from Rare Failures

    Yanran Xu, Chuanhang Qiu, Yue Wang +2

    cs.ROcs.CVarXiv:2608.30858v12026
  30. Pose-aware Multi-level Feature Network for Human Object Interaction Detection

    Bo Wan, Desen Zhou, Yongfei Liu +2

    cs.CVarXiv:1909.08453v12019
  31. Gen6D: Generalizable Model-Free 6-DoF Object Pose Estimation from RGB Images

    Yuan Liu, Yilin Wen, Sida Peng +4

    cs.CVarXiv:2204.10776v22022
  32. MMA-Diffusion: MultiModal Attack on Diffusion Models

    Yijun Yang, Ruiyuan Gao, Xiaosen Wang +3

    cs.CRcs.CVarXiv:2311.17516v42023
  33. Overview: Computer vision and machine learning for microstructural characterization and analysis

    Elizabeth A. Holm, Ryan Cohn, Nan Gao +4

    cs.CVcond-mat.mtrl-sciarXiv:2005.14260v12020
  34. MIT Advanced Vehicle Technology Study: Large-Scale Naturalistic Driving Study of Driver Behavior and Interaction with Automation

    Lex Fridman, Daniel E. Brown, Michael Glazer +15

    cs.CYcs.CVcs.HCarXiv:1711.06976v42017
  35. Learning to Anonymize Faces for Privacy Preserving Action Detection

    Zhongzheng Ren, Yong Jae Lee, Michael S. Ryoo

    cs.CVcs.AIcs.CRarXiv:1803.11556v22018
  36. UCL-Dehaze: Towards Real-world Image Dehazing via Unsupervised Contrastive Learning

    Yongzhen Wang, Xuefeng Yan, Fu Lee Wang +4

    cs.CVarXiv:2205.01871v12022
  37. CBNet: A Composite Backbone Network Architecture for Object Detection

    Tingting Liang, Xiaojie Chu, Yudong Liu +5

    cs.CVarXiv:2107.00420v72021
  38. Nested Hierarchical Transformer: Towards Accurate, Data-Efficient and Interpretable Visual Understanding

    Zizhao Zhang, Han Zhang, Long Zhao +3

    cs.CVarXiv:2105.12723v42021
  39. Effective Use of Dilated Convolutions for Segmenting Small Object Instances in Remote Sensing Imagery

    Ryuhei Hamaguchi, Aito Fujita, Keisuke Nemoto +2

    cs.CVarXiv:1709.00179v12017
  40. MatrixCity: A Large-scale City Dataset for City-scale Neural Rendering and Beyond

    Yixuan Li, Lihan Jiang, Linning Xu +4

    cs.CVarXiv:2309.16553v12023
  41. VRSTC: Occlusion-Free Video Person Re-Identification

    Ruibing Hou, Bingpeng Ma, Hong Chang +3

    cs.CVarXiv:1907.08427v12019
  42. Group Fisher Pruning for Practical Network Compression

    Liyang Liu, Shilong Zhang, Zhanghui Kuang +7

    cs.CVcs.LGarXiv:2108.00708v12021
  43. Adapting Segment Anything Model for Change Detection in HR Remote Sensing Images

    Lei Ding, Kun Zhu, Daifeng Peng +3

    cs.CVarXiv:2309.01429v42023
  44. CheXGround: Anatomical Region Tokens for Grounded Longitudinal Chest X-ray Interpretation

    Adonay Demewez Gebremedhin, Wessam Shehieb, Sara Alansari +4

    cs.CVarXiv:2608.30758v12026
  45. Unsupervised Domain Adaptation using Generative Adversarial Networks for Semantic Segmentation of Aerial Images

    Bilel Benjdira, Yakoub Bazi, Anis Koubaa +1

    cs.CVarXiv:1905.03198v12019
  46. UFPR-PEs: A Brazilian Face Recognition Benchmark with Self-Declared Race/Color Labels

    Alexandre Diano, Bernardo Biesseck, Gabriel Polo +4

    cs.CVarXiv:2608.30688v12026
  47. PointGrow: Autoregressively Learned Point Cloud Generation with Self-Attention

    Yongbin Sun, Yue Wang, Ziwei Liu +2

    cs.CVarXiv:1810.05591v32018
  48. diffGrad: An Optimization Method for Convolutional Neural Networks

    Shiv Ram Dubey, Soumendu Chakraborty, Swalpa Kumar Roy +3

    cs.LGcs.CVcs.NEarXiv:1909.11015v42019
  49. TUE-Detector: A Tool-Using Expert MLLM-Based Detector for AI-Generated Videos

    Yichen Wu, Haoxuan Qu, Yongxing Dai +5

    cs.CVarXiv:2608.30704v12026
  50. Combining Local Appearance and Holistic View: Dual-Source Deep Neural Networks for Human Pose Estimation

    Xiaochuan Fan, Kang Zheng, Yuewei Lin +1

    cs.CVarXiv:1504.07159v12015
  51. Prime Sample Attention in Object Detection

    Yuhang Cao, Kai Chen, Chen Change Loy +1

    cs.CVarXiv:1904.04821v22019
  52. Relaxed Transformer Decoders for Direct Action Proposal Generation

    Jing Tan, Jiaqi Tang, Limin Wang +1

    cs.CVarXiv:2102.01894v32021
  53. Swin2SR: SwinV2 Transformer for Compressed Image Super-Resolution and Restoration

    Marcos V. Conde, Ui-Jin Choi, Maxime Burchi +1

    cs.CVeess.IVarXiv:2209.11345v12022
  54. Benchmarking Neural Network Robustness to Common Corruptions and Surface Variations

    Dan Hendrycks, Thomas G. Dietterich

    cs.LGcs.AIcs.CVarXiv:1807.01697v52018
  55. RWF-2000: An Open Large Scale Video Database for Violence Detection

    Ming Cheng, Kunjing Cai, Ming Li

    cs.CVarXiv:1911.05913v32019
  56. AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

    Huawei Wei, Zejun Yang, Zhisheng Wang

    cs.CVcs.GReess.IVarXiv:2403.17694v12024
  57. GasHis-Transformer: A Multi-scale Visual Transformer Approach for Gastric Histopathological Image Detection

    Haoyuan Chen, Chen Li, Ge Wang +9

    cs.CVarXiv:2104.14528v72021
  58. SDM-NET: Deep Generative Network for Structured Deformable Mesh

    Lin Gao, Jie Yang, Tong Wu +4

    cs.GRcs.CVarXiv:1908.04520v22019
  59. SCAFFOLD: A Large-Scale Structured Dataset of Computer Science Research Figures with Diagram QA and Chain-of-Thought Reasoning Traces

    Ranjit Raut, Aarav Subedi, Sagun Rai +1

    cs.AIcs.CVarXiv:2609.00018v12026
  60. Fake it till you make it: Learning transferable representations from synthetic ImageNet clones

    Mert Bulent Sariyildiz, Karteek Alahari, Diane Larlus +1

    cs.CVcs.LGarXiv:2212.08420v22022