Neural Networks and Deep Learning: Hardware Acceleration
Overview
The course is a module of the Neural Networks and Deep Learning course. The objective of this course is to present practical and implementation issues useful to deploy neural networks on a variety of embedded platforms using different languages and development environments. Topics covered in the course include common development frameworks, networks optimization for embedded platforms, acceleration of deep networks on GPGPUs and FPGA platforms.
Program:
- Programming frameworks for deep learning
- Modeling DNNs in Tensorflow and PyTorch
- DNN optimization for embedded platforms
- The Nvidia TensorRT inference framework
- Accelerating deep networks on FPGA
- GPU programming in CUDA
- Accelerating deep networks on GPGPUs
- The AMD/Xilinx Deep Learning Processing Unit
Format and exam
- Lectures (20 hours): Lectures will be given online over Microsoft Teams.
- Exam (2 CFU): Project work and oral discussion.
Schedule
TBD
Course material
TBD