Computer Vision and Pattern Recognition
Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,821 to 8,880 of 18,866
Sparse and Imperceivable Adversarial Attacks
Francesco Croce, Matthias Hein
cs.LGcs.CRcs.CVarXiv:1909.05040v12019Dynamic Graph Enhanced Contrastive Learning for Chest X-ray Report Generation
Mingjie Li, Bingqian Lin, Zicong Chen +3
cs.CVarXiv:2303.10323v12023Precise Detection in Densely Packed Scenes
Eran Goldman, Roei Herzig, Aviv Eisenschtat +4
cs.CVarXiv:1904.00853v32019Max-Sliced Wasserstein Distance and its use for GANs
Ishan Deshpande, Yuan-Ting Hu, Ruoyu Sun +6
cs.LGcs.CVstat.MLarXiv:1904.05877v22019Context-Aware Interpretable Representations for Retrieval and Graph Convolutional Network Classification
Thiago César Castilho Almeida, Gustavo Rosseto Letício, Vinicius Atsushi Sato Kawai +1
cs.LGcs.CVcs.IRarXiv:2608.29004v12026A Dual-Path Model With Adaptive Attention For Vehicle Re-Identification
Pirazh Khorramshahi, Amit Kumar, Neehar Peri +3
cs.CVarXiv:1905.03397v32019Robust Invisible Video Watermarking with Attention
Kevin Alex Zhang, Lei Xu, Alfredo Cuesta-Infante +1
cs.MMcs.CVarXiv:1909.01285v12019$\mathcal{N}_0$-Foundation: Towards the Age of Tactile Intelligence
NeoteAI Team, Fudan TEAI Team
cs.ROcs.CVcs.LGarXiv:2608.29601v12026Video-based Remote Physiological Measurement via Cross-verified Feature Disentangling
Xuesong Niu, Zitong Yu, Hu Han +3
cs.CVarXiv:2007.08213v12020Learn From All: Erasing Attention Consistency for Noisy Label Facial Expression Recognition
Yuhang Zhang, Chengrui Wang, Xu Ling +1
cs.CVarXiv:2207.10299v22022Shadow Removal via Shadow Image Decomposition
Hieu Le, Dimitris Samaras
cs.CVarXiv:1908.08628v12019ChessQueries: Toward Better Chess Board Recognition
Joël Seytre
cs.CVarXiv:2608.30762v12026LO-Net: Deep Real-time Lidar Odometry
Qing Li, Shaoyang Chen, Cheng Wang +4
cs.CVarXiv:1904.08242v22019Score-Based Point Cloud Denoising
Shitong Luo, Wei Hu
cs.CVarXiv:2107.10981v52021Lost and Found: Detecting Small Road Hazards for Self-Driving Vehicles
Peter Pinggera, Sebastian Ramos, Stefan Gehrig +3
cs.CVcs.ROarXiv:1609.04653v12016Vibration-Based Damage Detection in Wind Turbine Blades using Phase-Based Motion Estimation and Motion Magnification
Aral Sarrafi, Zhu Mao, Christopher Niezrecki +1
eess.IVcs.CVarXiv:1804.00558v120183D Human Action Representation Learning via Cross-View Consistency Pursuit
Linguo Li, Minsi Wang, Bingbing Ni +3
cs.CVarXiv:2104.14466v22021Do We Need More Training Data?
Xiangxin Zhu, Carl Vondrick, Charless Fowlkes +1
cs.CVarXiv:1503.01508v12015Simultaneous Spectral-Spatial Feature Selection and Extraction for Hyperspectral Images
Lefei Zhang, Qian Zhang, Bo Du +3
cs.CVarXiv:1904.03982v12019Open-World Object Manipulation using Pre-trained Vision-Language Models
Austin Stone, Ted Xiao, Yao Lu +9
cs.ROcs.AIcs.CVarXiv:2303.00905v22023Explainability of deep vision-based autonomous driving systems: Review and challenges
Éloi Zablocki, Hédi Ben-Younes, Patrick Pérez +1
cs.CVcs.AIcs.LGarXiv:2101.05307v22021TGFuse: An Infrared and Visible Image Fusion Approach Based on Transformer and Generative Adversarial Network
Dongyu Rao, Xiao-Jun Wu, Tianyang Xu
cs.CVarXiv:2201.10147v22022A LiDAR Point Cloud Generator: from a Virtual World to Autonomous Driving
Xiangyu Yue, Bichen Wu, Sanjit A. Seshia +2
cs.CVarXiv:1804.00103v120186-DOF Grasping for Target-driven Object Manipulation in Clutter
Adithyavairavan Murali, Arsalan Mousavian, Clemens Eppner +2
cs.ROcs.CVarXiv:1912.03628v22019ARCH++: Animation-Ready Clothed Human Reconstruction Revisited
Tong He, Yuanlu Xu, Shunsuke Saito +2
cs.CVcs.GRarXiv:2108.07845v42021Regularized Discrete Optimal Transport
Sira Ferradans, Nicolas Papadakis, Gabriel Peyré +1
cs.CVcs.DMmath.OCarXiv:1307.5551v12013Say As You Wish: Fine-grained Control of Image Caption Generation with Abstract Scene Graphs
Shizhe Chen, Qin Jin, Peng Wang +1
cs.CVcs.AIarXiv:2003.00387v12020Boosting Crowd Counting via Multifaceted Attention
Hui Lin, Zhiheng Ma, Rongrong Ji +2
cs.CVarXiv:2203.02636v12022Plug-and-Play Algorithms for Large-scale Snapshot Compressive Imaging
Xin Yuan, Yang Liu, Jinli Suo +1
eess.IVcs.CVarXiv:2003.13654v22020Meta Faster R-CNN: Towards Accurate Few-Shot Object Detection with Attentive Feature Alignment
Guangxing Han, Shiyuan Huang, Jiawei Ma +2
cs.CVcs.AIcs.MMarXiv:2104.07719v42021PoseTrack: Joint Multi-Person Pose Estimation and Tracking
Umar Iqbal, Anton Milan, Juergen Gall
cs.CVarXiv:1611.07727v32016UniDexGrasp: Universal Robotic Dexterous Grasping via Learning Diverse Proposal Generation and Goal-Conditioned Policy
Yinzhen Xu, Weikang Wan, Jialiang Zhang +10
cs.ROcs.CVarXiv:2303.00938v22023LaneRCNN: Distributed Representations for Graph-Centric Motion Forecasting
Wenyuan Zeng, Ming Liang, Renjie Liao +1
cs.CVcs.ROarXiv:2101.06653v12021H3DNet: 3D Object Detection Using Hybrid Geometric Primitives
Zaiwei Zhang, Bo Sun, Haitao Yang +1
cs.CVarXiv:2006.05682v32020Object Detection in Video with Spatiotemporal Sampling Networks
Gedas Bertasius, Lorenzo Torresani, Jianbo Shi
cs.CVarXiv:1803.05549v22018DeepRED: Deep Image Prior Powered by RED
Gary Mataev, Michael Elad, Peyman Milanfar
cs.CVeess.IVarXiv:1903.10176v32019Point-Bind & Point-LLM: Aligning Point Cloud with Multi-modality for 3D Understanding, Generation, and Instruction Following
Ziyu Guo, Renrui Zhang, Xiangyang Zhu +8
cs.CVcs.AIcs.CLarXiv:2309.00615v12023PnPNet: End-to-End Perception and Prediction with Tracking in the Loop
Ming Liang, Bin Yang, Wenyuan Zeng +4
cs.CVcs.ROarXiv:2005.14711v22020The Topology ToolKit
Julien Tierny, Guillaume Favelier, Joshua A. Levine +2
cs.GRcs.CGcs.CVarXiv:1805.09110v22018Pixelwise Instance Segmentation with a Dynamically Instantiated Network
Anurag Arnab, Philip H. S Torr
cs.CVarXiv:1704.02386v12017RegionCache: Semantic-Aware Region Reuse for Efficient Multi-Turn Image Generation
Peizheng Li, Xin Ai, Hanyuan Liu +2
cs.CVarXiv:2608.29809v12026Interpolated Convolutional Networks for 3D Point Cloud Understanding
Jiageng Mao, Xiaogang Wang, Hongsheng Li
cs.CVcs.CGeess.IVarXiv:1908.04512v12019Automatic Handgun Detection Alarm in Videos Using Deep Learning
Roberto Olmos, Siham Tabik, Francisco Herrera
cs.CVarXiv:1702.05147v12017BMBC:Bilateral Motion Estimation with Bilateral Cost Volume for Video Interpolation
Junheum Park, Keunsoo Ko, Chul Lee +1
cs.CVarXiv:2007.12622v12020CrossNet: An End-to-end Reference-based Super Resolution Network using Cross-scale Warping
Haitian Zheng, Mengqi Ji, Haoqian Wang +2
cs.CVarXiv:1807.10547v12018Recurrent Attentional Networks for Saliency Detection
Jason Kuen, Zhenhua Wang, Gang Wang
cs.CVcs.LGstat.MLarXiv:1604.03227v12016Visual-Inertial Mapping with Non-Linear Factor Recovery
Vladyslav Usenko, Nikolaus Demmel, David Schubert +2
cs.CVcs.ROarXiv:1904.06504v32019Explicit Visual Prompting for Low-Level Structure Segmentations
Weihuang Liu, Xi Shen, Chi-Man Pun +1
cs.CVarXiv:2303.10883v22023Cross-X Learning for Fine-Grained Visual Categorization
Wei Luo, Xitong Yang, Xianjie Mo +5
cs.CVarXiv:1909.04412v12019ERASOR: Egocentric Ratio of Pseudo Occupancy-based Dynamic Object Removal for Static 3D Point Cloud Map Building
Hyungtae Lim, Sungwon Hwang, Hyun Myung
cs.CVcs.ROarXiv:2103.04316v12021CelebA-Spoof: Large-Scale Face Anti-Spoofing Dataset with Rich Annotations
Yuanhan Zhang, Zhenfei Yin, Yidong Li +4
cs.CVarXiv:2007.12342v32020Bi-Temporal Semantic Reasoning for the Semantic Change Detection in HR Remote Sensing Images
Lei Ding, Haitao Guo, Sicong Liu +3
cs.CVeess.IVarXiv:2108.06103v42021Long-Term Cloth-Changing Person Re-identification
Xuelin Qian, Wenxuan Wang, Li Zhang +5
cs.CVarXiv:2005.12633v32020Tracking of Fingertips and Centres of Palm using KINECT
J. L. Raheja, A. Chaudhary, K Singal
cs.CVarXiv:1304.4662v12013Delving into the Devils of Bird's-eye-view Perception: A Review, Evaluation and Recipe
Hongyang Li, Chonghao Sima, Jifeng Dai +19
cs.CVcs.LGcs.ROarXiv:2209.05324v42022PD-GAN: Probabilistic Diverse GAN for Image Inpainting
Hongyu Liu, Ziyu Wan, Wei Huang +3
cs.CVarXiv:2105.02201v12021ST-GAN: Spatial Transformer Generative Adversarial Networks for Image Compositing
Chen-Hsuan Lin, Ersin Yumer, Oliver Wang +2
cs.CVcs.LGarXiv:1803.01837v12018Synthesizing Programs for Images using Reinforced Adversarial Learning
Yaroslav Ganin, Tejas Kulkarni, Igor Babuschkin +2
cs.CVcs.LGstat.MLarXiv:1804.01118v12018DeepWrinkles: Accurate and Realistic Clothing Modeling
Zorah Laehner, Daniel Cremers, Tony Tung
cs.CVarXiv:1808.03417v12018SpanCalib-VLM: Calibrated Hallucination Span Detection in Vision-Language Models
Amanuel Gizachew Abebe, Yasmin Moslem
cs.CVcs.CLarXiv:2608.29974v12026