Skip to main content

Deep Residual Learning for Image Recognition

Deep Residual Learning for Image Recognition (2016) has been cited 316,706 times according to Google Scholar. CitationMap has resolved 105 citing papers from institutions across 8 countries.

IEEE Conference on Computer Vision and Pattern Recognition (CVPR)2016View paper

Authors: Kaiming He (Microsoft Research Asia), Jian Sun (Microsoft Research Asia), Xiangyu Zhang (Microsoft Research Asia), Shaoqing Ren (Microsoft Research Asia)

See X. Zhang's full citation map →

Where this paper is cited

China · 12United States · 5United Kingdom · 2Mexico · 1South Korea · 1Philippines · 1Bulgaria · 1Malaysia · 1

Top citing institutions

  • Department of Radiology Massachusetts General Hospital Harvard Medical School Boston Massachusetts USA. (3)
  • Hemodialysis Center, Mianyang Central Hospital, Mianyang, Sichuan, China. (2)
  • Department of Geography, University at Buffalo, The State University of New York, Buffalo, NY 14261, United States. (2)
  • Texas Tech University (2)
  • Beijing Electronic Science and Technology Institute (2)
  • (Nanyang Technological University) (2)
  • Instituto Politécnico Nacional (1)
  • Universidad Autónoma de Querétaro (1)
  • University of Huddersfield (1)
  • Harvard Medical School (1)
  • Brigham and Women's Hospital (1)
  • Brigham and Women's Hospital, Harvard Medical School (1)

Papers citing this work (105 resolved)

  1. · Juan Terven, Julio-Alejandro Romero-González, Diana-Margarita Córdova-Esparza

  2. · Muhammad Hussain

  3. · Bowen Chen, Faisal Mahmood, Sharifa Sahai, Guillaume Jaume +24 more

  4. A Survey on Large Language Models for Code Generation

    ACM Transactions on Software Engineering and Methodology (TOSEM) · 2026 · Sunghun Kim, Sungju Kim, Jiasi Shen, Juyong Jiang +1 more

  5. · Lianghui Zhu, Bencheng Liao, Qian Zhang, Xinggang Wang +2 more

  6. Deep Learning in Mechanical Metamaterials: From Prediction and Generation to Inverse Design.

    · Ikumu Watanabe, Ta-Te Chen, Ta‐Te Chen, Xiaoyang Zheng +2 more

  7. A survey of deep learning techniques for autonomous driving

    · Bogdan Trasnea, Bogdan Trăsnea, G. Măceșanu, Gigel Macesanu +9 more

  8. Explainable artificial intelligence: an analytical review

    · Eduardo A. Soares, Nicholas I. Arnold, Peter M. Atkinson, Plamen P. Angelov +9 more

  9. Hyperspectral imaging

    · Antoni Femenias, Antonio J. Plaza, Antonio Plaza, Bing Zhang +13 more

  10. Multimodal Unsupervised Image-to-image Translation

    · Jan Kautz, Ming-Yu Liu, Ming-Yu Liu 0001, Serge Belongie +3 more

  11. CBAM: Convolutional Block Attention Module

    · In So Kweon, Jongchan Park, Joon-Young Lee, Sanghyun Woo +1 more

  12. Image super-resolution using very deep residual channel attention networks

    · Bineng Zhong, Kai Li, Kunpeng Li, Lichen Wang +7 more

  13. BiSeNet: Bilateral Segmentation Network for Real-time Semantic Segmentation

    · Changqian Yu, Changxin Gao, Chao Peng, Chao Peng 0001 +6 more

  14. ShuffleNet V2: Practical Guidelines for Efficient CNN Architecture Design

    · Hai-Tao Zheng, Hai-Tao Zheng 0002, Jian Sun, Jian Sun 0001 +4 more

  15. ESRGAN: Enhanced Super-Resolution Generative Adversarial Networks

    · Chao Dong, Chen Change Loy, Jin Jin Gu, Jinjin Gu +13 more

  16. End-to-End Object Detection with Transformers

    · Alexander Kirillov, Francisco Massa, Gabriel Synnaeve, Gang He +12 more

  17. Big Transfer (BiT): General Visual Representation Learning

    · Alexander Kolesnikov, Alexander Kolesnikov 0003, J. Puigcerver, Jessica Yung +7 more

  18. Lift, Splat, Shoot: Encoding Images From Arbitrary Camera Rigs by Implicitly Unprojecting to 3D

    · Jonah Philion, S. Fidler, Sanja Fidler

  19. Contrastive Multiview Coding

    · Dilip Krishnan, Phillip Isola, Yonglong Tian

  20. Towards real-time multi-object tracking

    · Liang Zheng, Liang Zheng 0001, Shengjin Wang, Yali Li +4 more

  21. Deep Leakage from Gradients

    · Ligeng Zhu, Song Han, Yaqiong Mu, Zhijian Liu +1 more

  22. Padim: a patch distribution modeling framework for anomaly detection and localization

    · Aleksandr Setkov, Angelique Loesch, Angélique Loesch, Romaric Audigier +1 more

  23. Medical Transformer: Gated Axial-Attention for Medical Image Segmentation

    · I. Hacihaliloglu, Ilker Hacihaliloglu, Jeya Maria Jose Valanarasu, Poojan Oza +1 more

  24. Graph Attention Networks

    · Adriana Romero, Arantxa Casanova, Guillem Cucurull, Petar Veličković +2 more

  25. ActionFormer: Localizing Moments of Actions with Transformers

    · Chen-Lin Zhang, Jianxin Wu, Jianxin Wu 0001, Yin Li +1 more

  26. TinyViT: Fast Pretraining Distillation for Small Vision Transformers

    · Bin Xiao, Bin Xiao 0004, Houwen Peng, Jianlong Fu +4 more

  27. FOSTER: Feature Boosting and Compression for Class-Incremental Learning

    · Da-Wei Zhou, Da-Wei Zhou 0001, De-Chuan Zhan, Fu-Yun Wang +1 more

  28. DualPrompt: Complementary Prompting for Rehearsal-Free Continual Learning

    · Zifeng Wang 0002, Zizhao Zhang, Chen-Yu Lee, Guolong Su +11 more

  29. Petr: Position embedding transformation for multi-view 3d object detection

    · Jian Sun 0001, Tiancai Wang, Xiangyu Zhang 0005, Yingfei Liu

  30. Motr: End-to-end multiple-object tracking with transformer

    · Bin Dong, Cheng Chen, Fan-Yi Zeng, Fangao Zeng +7 more

  31. Xmem: Long-term video object segmentation with an atkinson-shiffrin memory model

    · A. Schwing, Ho Kei Cheng, Alexander G. Schwing

  32. Visual Prompt Tuning

    · Bharath Hariharan, Bor-Chun Chen, Borchun Chen, Claire Cardie +5 more

  33. Tip-Adapter: Training-free Adaption of CLIP for Few-shot Classification

    · Hongsheng Li, Hongsheng Li 0001, Jifeng Dai, Kunchang Li +11 more

  34. St-p3: End-to-end vision-based autonomous driving via spatial-temporal feature learning

    · Dacheng Tao, Hongyang Li 0001, Junchi Yan, Li Chen 0008 +2 more

  35. Joint Feature Learning and Relation Modeling for Tracking: A One-Stream Framework

    · B. Ma, Bingpeng Ma, Botao Ye, Hong Chang +6 more

  36. Aiatrack: Attention in attention for transformer visual tracking

    · Chao Ma, Chunluan Zhou, Junsong Yuan, Junsong Yuan 0001 +3 more

  37. MaxViT: Multi-Axis Vision Transformer

    · Alan Bovik, Feng Yang, Han Zhang, Hossein Talebi +7 more

  38. DeiT III: Revenge of the ViT

    · Hervé Jégou, Hugo Touvron, M. Cord, Matthieu Cord

  39. SPot-the-Difference Self-Supervised Pre-training for Anomaly Detection and Segmentation

    · Dongqing Zhang, Jongheon Jeong, Latha Pemula, Onkar Dabeer +1 more

  40. Simple Baselines for Image Restoration

    · Jian Sun, Liangyu Chen, Xiangyu Zhang, Xiaojie Chu +3 more

  41. A-OKVQA: A Benchmark for Visual Question Answering using World Knowledge

    · Apoorv Khandelwal, Christopher Clark, Dustin Schwenk, Kenneth Marino +2 more

  42. Exploring Plain Vision Transformer Backbones for Object Detection

    · Hanzi Mao, Kaiming He, Ross B. Girshick, Ross Girshick +1 more

  43. Simple open-vocabulary object detection

    · A. Gritsenko, Alexey Dosovitskiy, Anurag Arnab, Aravindh Mahendran +15 more

  44. No More Strided Convolutions or Pooling: A New CNN Building Block for Low-Resolution Images and Small Objects

    · Raja Sunkara, Tie Luo, Tie Luo 0001

  45. Mednext: transformer-driven scaling of convnets for medical image segmentation

    · Constantin Ulrich, Fabian Isensee, Gregor Koehler, Gregor Köhler +6 more

  46. Pmc-clip: Contrastive language-image pre-training using biomedical documents

    · Ya Zhang, Chaoyi Wu, Weidi Xie, Weixiong Lin +4 more

  47. OccWorld: Learning a 3D Occupancy World Model for Autonomous Driving

    · Borui Zhang, Jiwen Lu, Weiliang Chen, Wenzhao Zheng +3 more

  48. Frequency-spatial entanglement learning for camouflaged object detection

    · Chunyan Xu, Hanyu Xuan, Jian Yang 0003, Lei Luo 0001 +1 more

  49. MVSplat: Efficient 3D Gaussian Splatting from Sparse Multi-View Images

    · Andreas Geiger, Andreas Geiger 0001, Bohan Zhuang, Chuanxia Zheng +8 more

  50. GS-LRM: Large Reconstruction Model for 3D Gaussian Splatting

    · Hao Tan, Kai Zhang, Kalyan Sunkavalli, Nanxuan Zhao +5 more

  51. Rotary Position Embedding for Vision Transformer

    · Byeongho Heo, Dongyoon Han, Sangdoo Yun, Song Park

  52. YOLOv9: Learning What You Want to Learn Using Programmable Gradient Information

    · Chien-Yao Wang, Hong-Yuan Mark Liao, I-Hau Yeh

  53. Wavelet Convolutions for Large Receptive Fields

    · Eran Treister, Oren Freifeld, Roy Amoyal, Shahaf E. Finder

  54. Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

    · Chunyuan Li, Feng Li, Hang Su, Hao Zhang +14 more

  55. Llava-uhd: an lmm perceiving any aspect ratio and high-resolution images

    · Chunjiang Ge, Gao Huang, Gao Huang 0001, Junbo Cui +9 more

  56. LGM: Large Multi-view Gaussian Model for High-resolution 3D Content Creation

    · Gang Zeng, Jiaxiang Tang, Tengfei Wang, Xiaokang Chen +5 more

  57. Sapiens: Foundation for human vision models

    · Austin James, Julieta Martinez, Peter Selednik, Rawal Khirodkar +6 more

  58. Pixel-Aware Stable Diffusion for Realistic Image Super-Resolution and Personalized Stylization

    · Lei Zhang, Peiran Ren, Rongyuan Wu, T. Yang +4 more

  59. Mm1: methods, analysis and insights from multimodal llm pre-training

    · Alexander Toshev, Anton Belyi, Aonan Zhang, Bowen Zhang +16 more

  60. Genad: Generative end-to-end autonomous driving

    · Long Chen, Ruiqi Song, Wenzhao Zheng, Xianda Guo +2 more

  61. LocalMamba: Visual State Space Model with Windowed Selective Scan

    · Chang Xu, Chang Xu 0002, Chen Qian, Chen Qian 0006 +6 more

  62. InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks

    · Bin Li, Guo Chen, Jiannan Wu, Jifeng Dai +12 more

  63. Deep learning

    · Aaron Courville, Charu C. Aggarwal, Geoffrey E. Hinton, Geoffrey Hinton +4 more

  64. Perceptual losses for real-time style transfer and super-resolution

    · Alexandre Alahi, Justin Johnson, Li Fei-Fei, Justin Johnson 0001 +1 more

  65. XNOR-Net: ImageNet Classification Using Binary Convolutional Neural Networks

    · Ali Farhadi, J. Redmon, Joseph Redmon, Mohammad Rastegari +2 more

  66. Fully-Convolutional Siamese Networks for Object Tracking

    · A. Vedaldi, Andrea Vedaldi, Jack Valmadre, João F. Henriques +3 more

  67. A modified Adam algorithm for deep neural network optimization

    · Amany M Sarhan, Amany M. Sarhan, M. Arafa, Mohamed Reyad +5 more

  68. Deep Learning Techniques for Medical Image Segmentation: Achievements and Challenges.

    · M. H. Hesamian, Mohammad Hesam Hesamian, Paul J. Kennedy, Paul Kennedy +3 more

  69. A Survey of the Recent Architectures of Deep Convolutional Neural Networks

    · A. Sohail, Anabia Sohail, Aqsa Saeed Qureshi, Asifullah Khan +1 more

  70. Artificial intelligence in the creative industries: a review

    · D. Bull, David Bull, David Bull 0001, David R. Bull +2 more

  71. AI-big data analytics for building automation and management systems: a survey, actual challenges and future perspectives

    · A. Amira, Abbes Amira, F. Bensaali, F. Fadli +11 more

  72. Medical image data augmentation: techniques, comparisons and interpretations

    · Evgin Goceri, Evgin Göçeri

  73. Automated machine learning: past, present and future

    · Can Wang, Holger H. Hoos, Holger Hoos, J. N. V. Rijn +7 more

  74. Deepfake video detection: challenges and opportunities

    · A. N. Hoshyar, Achhardeep Kaur, Azadeh Noori Hoshyar, Feng Xia +3 more

  75. A comprehensive survey of deep learning-based lightweight object detection models for edge devices

    · Payal Mittal

  76. A comprehensive survey of loss functions and metrics in deep learning

    · Alfonso Ramirez-Pedraza, Alfonso Ramírez-Pedraza, Diana-Margarita Cordova-Esparza, Diana-Margarita Córdova-Esparza +7 more

  77. Classification of COVID-19 in chest X-ray images using DeTraC deep convolutional neural network

    · Asmaa Abbas, M. Abdelsamea, M. Gaber, Mohamed Medhat Gaber +2 more

  78. Deep learning for time series classification: a review

    · G. Forestier, Germain Forestier, Hassan Ismail Fawaz, J. Weber +4 more

  79. InceptionTime: Finding AlexNet for time series classification

    · Benjamin Lucas, Charlotte Pelletier, Daniel F. Schmidt, Franccois Petitjean +11 more

  80. Object detection using YOLO: challenges, architectural successors, datasets and applications.

    · G Anirudh, G. Anirudh, Jitendra V Tembhurne, Jitendra V. Tembhurne +2 more

  81. Group Normalization

    · Kaiming He, Yuxin Wu, Yuxin Wu 0004

  82. Grad-CAM: Visual Explanations from Deep Networks via Gradient-Based Localization

    · Abhishek Das, Devi Parikh, Dhruv Batra, Michael Cogswell +3 more

  83. Knowledge Distillation: A Survey

    · Jianping Gou, S. Maybank, Stephen J. Maybank, Stephen John Maybank +4 more

  84. BiSeNet V2: Bilateral Network with Guided Aggregation for Real-Time Semantic Segmentation

    · Changqian Yu, Changxin Gao, Chunhua Shen, Gang Yu +5 more

  85. Learning to Prompt for Vision-Language Models

    · C. C. Loy, Chen Change Loy, J. Yang, Jingkang Yang +5 more

  86. 3D Object Detection for Autonomous Driving: A Comprehensive Survey

    · Hongsheng Li, Hongsheng Li 0001, Jiageng Mao, Shaoshuai Shi +2 more

  87. CLIP-Adapter: Better Vision-Language Models with Feature Adapters

    · Hongsheng Li, Hongsheng Li 0001, Peng Gao, Peng Gao 0007 +8 more

  88. Generalized Out-of-Distribution Detection: A Survey

    · Jingkang Yang, Kaiyang Zhou, Yixuan Li, Yixuan Li 0001 +2 more

  89. A Comprehensive Survey on Test-Time Adaptation under Distribution Shifts

    · Jian Liang, Jian Liang 0001, R. He, Ran He +5 more

  90. Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives

    · A. Gebreselasie, A. Southerland, Abrham Gebreselasie, Andrew Westbury +16 more

  91. Perceptual video quality assessment: A survey

    · Guangtao Zhai, Huiyu Duan, Wei Sun, Wei Sun 0029 +2 more

  92. Slim-neck by GSConv: A better design paradigm of detector architectures for autonomous vehicles

    · Hanbing Wei, Hulin Li, Jun Li, Qiliang Ren +3 more

  93. ELA: efficient location attention for deep convolution neural networks

    · Wei Xu, Weina Zhao, Yi Wan

  94. Deep industrial image anomaly detection: A survey

    · Chengjie Wang, Chengjie Wang 0001, Feng Zheng, Feng Zheng 0001 +9 more

  95. Review of Lightweight Deep Convolutional Neural Networks

    · Fanghui Chen, Fengyuan Ren, Jiale Han, Shouliang Li +1 more

  96. A Review on Data-Driven Constitutive Laws for Solids

    · B. Bahmani, Bahador Bahmani, G. Padmanabha, Govinda Anantha Padmanabha +13 more

  97. A Comprehensive Survey on Pretrained Foundation Models: A History from BERT to ChatGPT

    · B. University, Caiming Xiong, Ce Zhou, Chen Li +16 more

  98. Convolutional neural networks: an overview and application in radiology

    · K. Togashi, Kaori Togashi, M. Nishio, Mizuho Nishio +5 more

  99. Attention mechanisms in computer vision: A survey

    · J.-J. Liu, Jiang-Jiang Liu, Jiangjiang Liu, M.-H. Guo +16 more

  100. PVT v2: Improved baselines with Pyramid Vision Transformer

    · Deng-Ping Fan, Ding Liang, Enze Xie, Kaipeng Song +13 more

Showing the top 100 of 105 resolved citing papers — see the full interactive list on X. Zhang's profile.

Map your own citations

CitationMap turns any Google Scholar profile into an interactive world map of citing institutions — free, no sign-up. Used for EB-1A / O-1 / NIW visa evidence, tenure files, and grant applications.

Create your citation map →