Advanced Computer Vision
Object detection, segmentation, YOLO, image generation.
Course Duration: 12 hours
What You'll Learn
- Object detection architectures
- Image segmentation
- YOLO family
- Pose estimation
- Face recognition
- OCR and document analysis
Prerequisites
- CNN fundamentals
- TensorFlow or PyTorch
Course Modules
- Object Detection Overview
- R-CNN Family
- YOLO v5/v8
- SSD and RetinaNet
- Semantic Segmentation
- Instance Segmentation (Mask R-CNN)
- Panoptic Segmentation
- Pose Estimation
- Face Detection and Recognition
- OCR with Tesseract
- Vision Transformers
- Project: Real-time Object Detection
Architectures Covered
- YOLO v5/v8
- Faster R-CNN
- Mask R-CNN
- U-Net
- DeepLab