Computer Vision and Pattern Recognition

Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,821 to 8,880 of 18,866

  1. Sparse and Imperceivable Adversarial Attacks

    Francesco Croce, Matthias Hein

    cs.LGcs.CRcs.CVarXiv:1909.05040v12019
  2. Dynamic Graph Enhanced Contrastive Learning for Chest X-ray Report Generation

    Mingjie Li, Bingqian Lin, Zicong Chen +3

    cs.CVarXiv:2303.10323v12023
  3. Precise Detection in Densely Packed Scenes

    Eran Goldman, Roei Herzig, Aviv Eisenschtat +4

    cs.CVarXiv:1904.00853v32019
  4. Max-Sliced Wasserstein Distance and its use for GANs

    Ishan Deshpande, Yuan-Ting Hu, Ruoyu Sun +6

    cs.LGcs.CVstat.MLarXiv:1904.05877v22019
  5. Context-Aware Interpretable Representations for Retrieval and Graph Convolutional Network Classification

    Thiago César Castilho Almeida, Gustavo Rosseto Letício, Vinicius Atsushi Sato Kawai +1

    cs.LGcs.CVcs.IRarXiv:2608.29004v12026
  6. A Dual-Path Model With Adaptive Attention For Vehicle Re-Identification

    Pirazh Khorramshahi, Amit Kumar, Neehar Peri +3

    cs.CVarXiv:1905.03397v32019
  7. Robust Invisible Video Watermarking with Attention

    Kevin Alex Zhang, Lei Xu, Alfredo Cuesta-Infante +1

    cs.MMcs.CVarXiv:1909.01285v12019
  8. $\mathcal{N}_0$-Foundation: Towards the Age of Tactile Intelligence

    NeoteAI Team, Fudan TEAI Team

    cs.ROcs.CVcs.LGarXiv:2608.29601v12026
  9. Video-based Remote Physiological Measurement via Cross-verified Feature Disentangling

    Xuesong Niu, Zitong Yu, Hu Han +3

    cs.CVarXiv:2007.08213v12020
  10. Learn From All: Erasing Attention Consistency for Noisy Label Facial Expression Recognition

    Yuhang Zhang, Chengrui Wang, Xu Ling +1

    cs.CVarXiv:2207.10299v22022
  11. Shadow Removal via Shadow Image Decomposition

    Hieu Le, Dimitris Samaras

    cs.CVarXiv:1908.08628v12019
  12. ChessQueries: Toward Better Chess Board Recognition

    Joël Seytre

    cs.CVarXiv:2608.30762v12026
  13. LO-Net: Deep Real-time Lidar Odometry

    Qing Li, Shaoyang Chen, Cheng Wang +4

    cs.CVarXiv:1904.08242v22019
  14. Score-Based Point Cloud Denoising

    Shitong Luo, Wei Hu

    cs.CVarXiv:2107.10981v52021
  15. Lost and Found: Detecting Small Road Hazards for Self-Driving Vehicles

    Peter Pinggera, Sebastian Ramos, Stefan Gehrig +3

    cs.CVcs.ROarXiv:1609.04653v12016
  16. Vibration-Based Damage Detection in Wind Turbine Blades using Phase-Based Motion Estimation and Motion Magnification

    Aral Sarrafi, Zhu Mao, Christopher Niezrecki +1

    eess.IVcs.CVarXiv:1804.00558v12018
  17. 3D Human Action Representation Learning via Cross-View Consistency Pursuit

    Linguo Li, Minsi Wang, Bingbing Ni +3

    cs.CVarXiv:2104.14466v22021
  18. Do We Need More Training Data?

    Xiangxin Zhu, Carl Vondrick, Charless Fowlkes +1

    cs.CVarXiv:1503.01508v12015
  19. Simultaneous Spectral-Spatial Feature Selection and Extraction for Hyperspectral Images

    Lefei Zhang, Qian Zhang, Bo Du +3

    cs.CVarXiv:1904.03982v12019
  20. Open-World Object Manipulation using Pre-trained Vision-Language Models

    Austin Stone, Ted Xiao, Yao Lu +9

    cs.ROcs.AIcs.CVarXiv:2303.00905v22023
  21. Explainability of deep vision-based autonomous driving systems: Review and challenges

    Éloi Zablocki, Hédi Ben-Younes, Patrick Pérez +1

    cs.CVcs.AIcs.LGarXiv:2101.05307v22021
  22. TGFuse: An Infrared and Visible Image Fusion Approach Based on Transformer and Generative Adversarial Network

    Dongyu Rao, Xiao-Jun Wu, Tianyang Xu

    cs.CVarXiv:2201.10147v22022
  23. A LiDAR Point Cloud Generator: from a Virtual World to Autonomous Driving

    Xiangyu Yue, Bichen Wu, Sanjit A. Seshia +2

    cs.CVarXiv:1804.00103v12018
  24. 6-DOF Grasping for Target-driven Object Manipulation in Clutter

    Adithyavairavan Murali, Arsalan Mousavian, Clemens Eppner +2

    cs.ROcs.CVarXiv:1912.03628v22019
  25. ARCH++: Animation-Ready Clothed Human Reconstruction Revisited

    Tong He, Yuanlu Xu, Shunsuke Saito +2

    cs.CVcs.GRarXiv:2108.07845v42021
  26. Regularized Discrete Optimal Transport

    Sira Ferradans, Nicolas Papadakis, Gabriel Peyré +1

    cs.CVcs.DMmath.OCarXiv:1307.5551v12013
  27. Say As You Wish: Fine-grained Control of Image Caption Generation with Abstract Scene Graphs

    Shizhe Chen, Qin Jin, Peng Wang +1

    cs.CVcs.AIarXiv:2003.00387v12020
  28. Boosting Crowd Counting via Multifaceted Attention

    Hui Lin, Zhiheng Ma, Rongrong Ji +2

    cs.CVarXiv:2203.02636v12022
  29. Plug-and-Play Algorithms for Large-scale Snapshot Compressive Imaging

    Xin Yuan, Yang Liu, Jinli Suo +1

    eess.IVcs.CVarXiv:2003.13654v22020
  30. Meta Faster R-CNN: Towards Accurate Few-Shot Object Detection with Attentive Feature Alignment

    Guangxing Han, Shiyuan Huang, Jiawei Ma +2

    cs.CVcs.AIcs.MMarXiv:2104.07719v42021
  31. PoseTrack: Joint Multi-Person Pose Estimation and Tracking

    Umar Iqbal, Anton Milan, Juergen Gall

    cs.CVarXiv:1611.07727v32016
  32. UniDexGrasp: Universal Robotic Dexterous Grasping via Learning Diverse Proposal Generation and Goal-Conditioned Policy

    Yinzhen Xu, Weikang Wan, Jialiang Zhang +10

    cs.ROcs.CVarXiv:2303.00938v22023
  33. LaneRCNN: Distributed Representations for Graph-Centric Motion Forecasting

    Wenyuan Zeng, Ming Liang, Renjie Liao +1

    cs.CVcs.ROarXiv:2101.06653v12021
  34. H3DNet: 3D Object Detection Using Hybrid Geometric Primitives

    Zaiwei Zhang, Bo Sun, Haitao Yang +1

    cs.CVarXiv:2006.05682v32020
  35. Object Detection in Video with Spatiotemporal Sampling Networks

    Gedas Bertasius, Lorenzo Torresani, Jianbo Shi

    cs.CVarXiv:1803.05549v22018
  36. DeepRED: Deep Image Prior Powered by RED

    Gary Mataev, Michael Elad, Peyman Milanfar

    cs.CVeess.IVarXiv:1903.10176v32019
  37. Point-Bind & Point-LLM: Aligning Point Cloud with Multi-modality for 3D Understanding, Generation, and Instruction Following

    Ziyu Guo, Renrui Zhang, Xiangyang Zhu +8

    cs.CVcs.AIcs.CLarXiv:2309.00615v12023
  38. PnPNet: End-to-End Perception and Prediction with Tracking in the Loop

    Ming Liang, Bin Yang, Wenyuan Zeng +4

    cs.CVcs.ROarXiv:2005.14711v22020
  39. The Topology ToolKit

    Julien Tierny, Guillaume Favelier, Joshua A. Levine +2

    cs.GRcs.CGcs.CVarXiv:1805.09110v22018
  40. Pixelwise Instance Segmentation with a Dynamically Instantiated Network

    Anurag Arnab, Philip H. S Torr

    cs.CVarXiv:1704.02386v12017
  41. RegionCache: Semantic-Aware Region Reuse for Efficient Multi-Turn Image Generation

    Peizheng Li, Xin Ai, Hanyuan Liu +2

    cs.CVarXiv:2608.29809v12026
  42. Interpolated Convolutional Networks for 3D Point Cloud Understanding

    Jiageng Mao, Xiaogang Wang, Hongsheng Li

    cs.CVcs.CGeess.IVarXiv:1908.04512v12019
  43. Automatic Handgun Detection Alarm in Videos Using Deep Learning

    Roberto Olmos, Siham Tabik, Francisco Herrera

    cs.CVarXiv:1702.05147v12017
  44. BMBC:Bilateral Motion Estimation with Bilateral Cost Volume for Video Interpolation

    Junheum Park, Keunsoo Ko, Chul Lee +1

    cs.CVarXiv:2007.12622v12020
  45. CrossNet: An End-to-end Reference-based Super Resolution Network using Cross-scale Warping

    Haitian Zheng, Mengqi Ji, Haoqian Wang +2

    cs.CVarXiv:1807.10547v12018
  46. Recurrent Attentional Networks for Saliency Detection

    Jason Kuen, Zhenhua Wang, Gang Wang

    cs.CVcs.LGstat.MLarXiv:1604.03227v12016
  47. Visual-Inertial Mapping with Non-Linear Factor Recovery

    Vladyslav Usenko, Nikolaus Demmel, David Schubert +2

    cs.CVcs.ROarXiv:1904.06504v32019
  48. Explicit Visual Prompting for Low-Level Structure Segmentations

    Weihuang Liu, Xi Shen, Chi-Man Pun +1

    cs.CVarXiv:2303.10883v22023
  49. Cross-X Learning for Fine-Grained Visual Categorization

    Wei Luo, Xitong Yang, Xianjie Mo +5

    cs.CVarXiv:1909.04412v12019
  50. ERASOR: Egocentric Ratio of Pseudo Occupancy-based Dynamic Object Removal for Static 3D Point Cloud Map Building

    Hyungtae Lim, Sungwon Hwang, Hyun Myung

    cs.CVcs.ROarXiv:2103.04316v12021
  51. CelebA-Spoof: Large-Scale Face Anti-Spoofing Dataset with Rich Annotations

    Yuanhan Zhang, Zhenfei Yin, Yidong Li +4

    cs.CVarXiv:2007.12342v32020
  52. Bi-Temporal Semantic Reasoning for the Semantic Change Detection in HR Remote Sensing Images

    Lei Ding, Haitao Guo, Sicong Liu +3

    cs.CVeess.IVarXiv:2108.06103v42021
  53. Long-Term Cloth-Changing Person Re-identification

    Xuelin Qian, Wenxuan Wang, Li Zhang +5

    cs.CVarXiv:2005.12633v32020
  54. Tracking of Fingertips and Centres of Palm using KINECT

    J. L. Raheja, A. Chaudhary, K Singal

    cs.CVarXiv:1304.4662v12013
  55. Delving into the Devils of Bird's-eye-view Perception: A Review, Evaluation and Recipe

    Hongyang Li, Chonghao Sima, Jifeng Dai +19

    cs.CVcs.LGcs.ROarXiv:2209.05324v42022
  56. PD-GAN: Probabilistic Diverse GAN for Image Inpainting

    Hongyu Liu, Ziyu Wan, Wei Huang +3

    cs.CVarXiv:2105.02201v12021
  57. ST-GAN: Spatial Transformer Generative Adversarial Networks for Image Compositing

    Chen-Hsuan Lin, Ersin Yumer, Oliver Wang +2

    cs.CVcs.LGarXiv:1803.01837v12018
  58. Synthesizing Programs for Images using Reinforced Adversarial Learning

    Yaroslav Ganin, Tejas Kulkarni, Igor Babuschkin +2

    cs.CVcs.LGstat.MLarXiv:1804.01118v12018
  59. DeepWrinkles: Accurate and Realistic Clothing Modeling

    Zorah Laehner, Daniel Cremers, Tony Tung

    cs.CVarXiv:1808.03417v12018
  60. SpanCalib-VLM: Calibrated Hallucination Span Detection in Vision-Language Models

    Amanuel Gizachew Abebe, Yasmin Moslem

    cs.CVcs.CLarXiv:2608.29974v12026