基於深度學習之視覺辨識專論(英文授課)
Selected Topics in Visual Recognition using Deep Learning
| 節 | 週四 |
|---|---|
3 10:10–11:00 | 基於深度學習之視覺辨識專論(英文授課) ED305 3 節連堂 |
4 11:10–12:00 | |
N 12:20–13:10 |
* 根據陽明交大上課時間表所列
Visual recognition aims to enable computers to see, understand, and interpret the world like human visual systems. Deep learning technologies are at the core of the current machine vision revolution. Large-scale annotated data and affordable GPU hardware jointly allow the training of deep learning models with hundreds of layers and millions of parameters, which greatly improve the performance of various machine vision applications and even initiate new vision applications. In the course, I will first introduce some deep learning technologies that are widely used in visual recognition research, including convolutional neural networks, generative adversarial networks, and recurrent neural networks. Then, I will cover some important vision applications such as object recognition, detection, and segmentation, and the corresponding advanced deep learning algorithms.
1. Basic knowledge of linear algebra and calculus. 2. Programming experience such as Python (preferred) and C++.
4 homework assignments 72% Final project 28%
| 週次 | 主題 |
|---|---|
| 第 1 週 | Introduction to Visual Recognition |
| 第 2 週 | Conventional Machine Learning vs. Deep Learning for Visual Recognition |
| 第 3 週 | Convolutional Neural Networks (CNN) |
| 第 4 週 | Representative CNN Architectures |
| 第 5 週 | National Day. No lecture. |
| 第 6 週 | Generative Adversarial Learning |
| 第 7 週 | Object Detection |
| 第 8 週 | Guest Lectures I: Prof. Ching-Chun Huang and Prof. Wei-Chen Chiu in CS, NCTU |
| 第 9 週 | Object Detection / Semantic Segmentation |
| 第 10 週 | Semantic Segmentation |
| 第 11 週 | Image Super-resolution |
| 第 12 週 | Image Matching and Alignment |
| 第 13 週 | Action and Gesture Recognition, 3D Point Classification and Segmentation |
| 第 14 週 | Guest Lectures II: Dr. Joe Yeh at aetherAI |
| 第 15 週 | Image Style Transfer, Video Frame Interpolation, and Video Synthesis |
| 第 16 週 | Final Project Presentation I |
| 第 17 週 | Final Project Presentation II |
Ian Goodfellow, Yoshua Bengio, and Aaron Courville, Deep Learning, MIT Press, 2016 Richard Szeliski, Computer Vision: Algorithms and Applications, Springer Verlag London, 2011.
- 地點
- EC118
- 時間
- Thursday 3:00 pm ~ 4:00 pm
- 聯絡方式
- Contact Instructor Yen-Yu Lin at lin@cs.nctu.edu.tw Contact TA Jimmy Yang at d08922002@ntu.edu.tw