Vision Models¶
Resources¶
[lilianweng.github.io] Object Detection for Dummies
[arxiv.org] ResNet: Deep Residual Learning for Image Recognition
[arxiv.org] ViT: An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
[arxiv.org] SimCLR: A Simple Framework for Contrastive Learning of Visual Representations
[arxiv.org] MoCo: Momentum Contrast for Unsupervised Visual Representation Learning
[arxiv.org] YOLO: A Comprehensive Review of YOLO Architectures in Computer Vision
[arxiv.org] DETR: End-to-End Object Detection with Transformers
[arxiv.org] ImageGAN: Image-to-Image Translation with Conditional Adversarial Networks