Computer Vision and Pattern Recognition
Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,221 to 8,280 of 18,848
MID-Fusion: Octree-based Object-Level Multi-Instance Dynamic SLAM
Binbin Xu, Wenbin Li, Dimos Tzoumanikas +3
cs.ROcs.CVarXiv:1812.07976v42018Breaking the Dilemma of Medical Image-to-image Translation
Lingke Kong, Chenyu Lian, Detian Huang +3
eess.IVcs.CVarXiv:2110.06465v22021Generative Adversarial Minority Oversampling
Sankha Subhra Mullick, Shounak Datta, Swagatam Das
cs.CVcs.LGarXiv:1903.09730v32019PMP-Net: Point Cloud Completion by Learning Multi-step Point Moving Paths
Xin Wen, Peng Xiang, Zhizhong Han +4
cs.CVarXiv:2012.03408v32020K-Radar: 4D Radar Object Detection for Autonomous Driving in Various Weather Conditions
Dong-Hee Paek, Seung-Hyun Kong, Kevin Tirta Wijaya
cs.CVcs.AIarXiv:2206.08171v42022ALICE: Towards Understanding Adversarial Learning for Joint Distribution Matching
Chunyuan Li, Hao Liu, Changyou Chen +4
stat.MLcs.AIcs.CVarXiv:1709.01215v22017Alpha-Refine: Boosting Tracking Performance by Precise Bounding Box Estimation
Bin Yan, Xinyu Zhang, Dong Wang +2
cs.CVarXiv:2012.06815v32020Repeatability Is Not Enough: Learning Affine Regions via Discriminability
Dmytro Mishkin, Filip Radenovic, Jiri Matas
cs.CVcs.NEarXiv:1711.06704v42017TransVOD: End-to-End Video Object Detection with Spatial-Temporal Transformers
Qianyu Zhou, Xiangtai Li, Lu He +5
cs.CVarXiv:2201.05047v42022An end-to-end TextSpotter with Explicit Alignment and Attention
Tong He, Zhi Tian, Weilin Huang +3
cs.CVarXiv:1803.03474v32018Texture Synthesis with Spatial Generative Adversarial Networks
Nikolay Jetchev, Urs Bergmann, Roland Vollgraf
cs.CVstat.MLarXiv:1611.08207v42016GarmentWeaver: Schema-Aware Structured Synthesis for Multimodal Sewing Patterns
Yinwen Lu, Weihao Luo, Yueqi Zhong
cs.AIcs.CVarXiv:2608.30550v12026Pix2Vox++: Multi-scale Context-aware 3D Object Reconstruction from Single and Multiple Images
Haozhe Xie, Hongxun Yao, Shengping Zhang +2
cs.CVarXiv:2006.12250v22020BRF-GS: Hyperspectral Bidirectional Reflectance Factor Modeling and Image Generation Based on 3D Gaussian Splatting
Yiling Yao, Wenjuan Zhang, Bowen Wang +3
cs.CVarXiv:2608.31159v12026Foreground-aware Pyramid Reconstruction for Alignment-free Occluded Person Re-identification
Lingxiao He, Yinggang Wang, Wu Liu +4
cs.CVarXiv:1904.04975v22019Shape, Light, and Material Decomposition from Images using Monte Carlo Rendering and Denoising
Jon Hasselgren, Nikolai Hofmann, Jacob Munkberg
cs.GRcs.CVarXiv:2206.03380v22022AI-enabled Low-Cost 3D Maize Ear Morphometry Platform at Breeding Scale
Therin Young, Elijah Rodriguez, Lisa Coffey +4
cs.CVarXiv:2608.30161v12026Perceptual Adversarial Robustness: Defense Against Unseen Threat Models
Cassidy Laidlaw, Sahil Singla, Soheil Feizi
cs.LGcs.CVstat.MLarXiv:2006.12655v42020Hybrid-SORT: Weak Cues Matter for Online Multi-Object Tracking
Mingzhan Yang, Guangxin Han, Bin Yan +4
cs.CVarXiv:2308.00783v22023CANVAS: Consistency-Aware Navigation via Visual Adaptive Sampling for Long-Context Text-to-SVG Generation
Yichen Wu, Haoxuan Qu, Yihang Lou +2
cs.CVarXiv:2608.30689v12026FaceSnap: Real-Time Personalized Lightstage Facial Performance Capture
Rukhshanda Hussain, Noé Artru, Emeline Got +6
cs.CVarXiv:2608.31033v12026Proximity3D: Shape from Capacitive Proximity on Sensing Manifold
Hao Chen, Chenming Wu, Chun Ping Lam +6
cs.CVcs.CGcs.GRarXiv:2608.30344v22026Texture image analysis and texture classification methods - A review
Laleh Armi, Shervan Fekri-Ershad
cs.CVarXiv:1904.06554v12019Multi-View Reflective Surface Inspection via Semantic-Saliency Cross-Verification
Van-Giang Nguyen, Thanh-Tuan Tran, Xuan-Hieu Phan +1
cs.CVarXiv:2608.30997v12026RL-CycleGAN: Reinforcement Learning Aware Simulation-To-Real
Kanishka Rao, Chris Harris, Alex Irpan +3
cs.ROcs.CVcs.LGarXiv:2006.09001v12020Segmentation of Glioma Tumors in Brain Using Deep Convolutional Neural Network
Saddam Hussain, Syed Muhammad Anwar, Muhammad Majid
cs.CVarXiv:1708.00377v12017Contrastive Learning from Extremely Augmented Skeleton Sequences for Self-supervised Action Recognition
Tianyu Guo, Hong Liu, Zhan Chen +3
cs.CVarXiv:2112.03590v12021Generative Cooperative Learning for Unsupervised Video Anomaly Detection
Muhammad Zaigham Zaheer, Arif Mahmood, Muhammad Haris Khan +3
cs.CVarXiv:2203.03962v12022GAFT: Geo-Anchored Fine-Tuning for Hazard Identification from Rare Failures
Yanran Xu, Chuanhang Qiu, Yue Wang +2
cs.ROcs.CVarXiv:2608.30858v12026Pose-aware Multi-level Feature Network for Human Object Interaction Detection
Bo Wan, Desen Zhou, Yongfei Liu +2
cs.CVarXiv:1909.08453v12019Gen6D: Generalizable Model-Free 6-DoF Object Pose Estimation from RGB Images
Yuan Liu, Yilin Wen, Sida Peng +4
cs.CVarXiv:2204.10776v22022MMA-Diffusion: MultiModal Attack on Diffusion Models
Yijun Yang, Ruiyuan Gao, Xiaosen Wang +3
cs.CRcs.CVarXiv:2311.17516v42023Overview: Computer vision and machine learning for microstructural characterization and analysis
Elizabeth A. Holm, Ryan Cohn, Nan Gao +4
cs.CVcond-mat.mtrl-sciarXiv:2005.14260v12020MIT Advanced Vehicle Technology Study: Large-Scale Naturalistic Driving Study of Driver Behavior and Interaction with Automation
Lex Fridman, Daniel E. Brown, Michael Glazer +15
cs.CYcs.CVcs.HCarXiv:1711.06976v42017Learning to Anonymize Faces for Privacy Preserving Action Detection
Zhongzheng Ren, Yong Jae Lee, Michael S. Ryoo
cs.CVcs.AIcs.CRarXiv:1803.11556v22018UCL-Dehaze: Towards Real-world Image Dehazing via Unsupervised Contrastive Learning
Yongzhen Wang, Xuefeng Yan, Fu Lee Wang +4
cs.CVarXiv:2205.01871v12022CBNet: A Composite Backbone Network Architecture for Object Detection
Tingting Liang, Xiaojie Chu, Yudong Liu +5
cs.CVarXiv:2107.00420v72021Nested Hierarchical Transformer: Towards Accurate, Data-Efficient and Interpretable Visual Understanding
Zizhao Zhang, Han Zhang, Long Zhao +3
cs.CVarXiv:2105.12723v42021Effective Use of Dilated Convolutions for Segmenting Small Object Instances in Remote Sensing Imagery
Ryuhei Hamaguchi, Aito Fujita, Keisuke Nemoto +2
cs.CVarXiv:1709.00179v12017MatrixCity: A Large-scale City Dataset for City-scale Neural Rendering and Beyond
Yixuan Li, Lihan Jiang, Linning Xu +4
cs.CVarXiv:2309.16553v12023VRSTC: Occlusion-Free Video Person Re-Identification
Ruibing Hou, Bingpeng Ma, Hong Chang +3
cs.CVarXiv:1907.08427v12019Group Fisher Pruning for Practical Network Compression
Liyang Liu, Shilong Zhang, Zhanghui Kuang +7
cs.CVcs.LGarXiv:2108.00708v12021Adapting Segment Anything Model for Change Detection in HR Remote Sensing Images
Lei Ding, Kun Zhu, Daifeng Peng +3
cs.CVarXiv:2309.01429v42023CheXGround: Anatomical Region Tokens for Grounded Longitudinal Chest X-ray Interpretation
Adonay Demewez Gebremedhin, Wessam Shehieb, Sara Alansari +4
cs.CVarXiv:2608.30758v12026Unsupervised Domain Adaptation using Generative Adversarial Networks for Semantic Segmentation of Aerial Images
Bilel Benjdira, Yakoub Bazi, Anis Koubaa +1
cs.CVarXiv:1905.03198v12019UFPR-PEs: A Brazilian Face Recognition Benchmark with Self-Declared Race/Color Labels
Alexandre Diano, Bernardo Biesseck, Gabriel Polo +4
cs.CVarXiv:2608.30688v12026PointGrow: Autoregressively Learned Point Cloud Generation with Self-Attention
Yongbin Sun, Yue Wang, Ziwei Liu +2
cs.CVarXiv:1810.05591v32018diffGrad: An Optimization Method for Convolutional Neural Networks
Shiv Ram Dubey, Soumendu Chakraborty, Swalpa Kumar Roy +3
cs.LGcs.CVcs.NEarXiv:1909.11015v42019TUE-Detector: A Tool-Using Expert MLLM-Based Detector for AI-Generated Videos
Yichen Wu, Haoxuan Qu, Yongxing Dai +5
cs.CVarXiv:2608.30704v12026Combining Local Appearance and Holistic View: Dual-Source Deep Neural Networks for Human Pose Estimation
Xiaochuan Fan, Kang Zheng, Yuewei Lin +1
cs.CVarXiv:1504.07159v12015Prime Sample Attention in Object Detection
Yuhang Cao, Kai Chen, Chen Change Loy +1
cs.CVarXiv:1904.04821v22019Relaxed Transformer Decoders for Direct Action Proposal Generation
Jing Tan, Jiaqi Tang, Limin Wang +1
cs.CVarXiv:2102.01894v32021Swin2SR: SwinV2 Transformer for Compressed Image Super-Resolution and Restoration
Marcos V. Conde, Ui-Jin Choi, Maxime Burchi +1
cs.CVeess.IVarXiv:2209.11345v12022Benchmarking Neural Network Robustness to Common Corruptions and Surface Variations
Dan Hendrycks, Thomas G. Dietterich
cs.LGcs.AIcs.CVarXiv:1807.01697v52018RWF-2000: An Open Large Scale Video Database for Violence Detection
Ming Cheng, Kunjing Cai, Ming Li
cs.CVarXiv:1911.05913v32019AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation
Huawei Wei, Zejun Yang, Zhisheng Wang
cs.CVcs.GReess.IVarXiv:2403.17694v12024GasHis-Transformer: A Multi-scale Visual Transformer Approach for Gastric Histopathological Image Detection
Haoyuan Chen, Chen Li, Ge Wang +9
cs.CVarXiv:2104.14528v72021SDM-NET: Deep Generative Network for Structured Deformable Mesh
Lin Gao, Jie Yang, Tong Wu +4
cs.GRcs.CVarXiv:1908.04520v22019SCAFFOLD: A Large-Scale Structured Dataset of Computer Science Research Figures with Diagram QA and Chain-of-Thought Reasoning Traces
Ranjit Raut, Aarav Subedi, Sagun Rai +1
cs.AIcs.CVarXiv:2609.00018v12026Fake it till you make it: Learning transferable representations from synthetic ImageNet clones
Mert Bulent Sariyildiz, Karteek Alahari, Diane Larlus +1
cs.CVcs.LGarXiv:2212.08420v22022