Back to Optimize TensorFlow Models For Deployment with TensorRT
Coursera

Optimize TensorFlow Models For Deployment with TensorRT

This is a hands-on, guided project on optimizing your TensorFlow models for inference with NVIDIA's TensorRT. By the end of this 1.5 hour long project, you will be able to optimize Tensorflow models using the TensorFlow integration of NVIDIA's TensorRT (TF-TRT), use TF-TRT to optimize several deep learning models at FP32, FP16, and INT8 precision, and observe how tuning TF-TRT parameters affects performance and inference throughput. Prerequisites: In order to successfully complete this project, you should be competent in Python programming, understand deep learning and what inference is, and have experience building deep learning models in TensorFlow and its Keras API. Note: This course works best for learners who are based in the North America region. We’re currently working on providing the same experience in other regions.

Status: Performance Tuning
Status: Model Optimization
IntermediateGuided Project2 hours

Featured reviews

AA

Reviewed Mar 14, 2022

T​he first to introduce such a rare and important topic.

LS

Reviewed Jun 3, 2021

G​reat workshop, all the concepts were very well explained.

All reviews

Showing: 20 of 20

דמיטרי
5.0
Reviewed May 16, 2023
Awais
5.0
Reviewed Mar 28, 2021
Deleted
4.0
Reviewed Jun 15, 2023
Dmytro
2.0
Reviewed Aug 3, 2023
Jorge
3.0
Reviewed Feb 25, 2021
Hleb
5.0
Reviewed May 25, 2023
Luis
5.0
Reviewed Jun 4, 2021
Abdelrahman
5.0
Reviewed Mar 15, 2022
Fabian
5.0
Reviewed Apr 20, 2021
Nusrat
5.0
Reviewed Apr 16, 2021
Chandra
5.0
Reviewed Dec 13, 2020
Fangwen
5.0
Reviewed Sep 12, 2023
Maftuna
5.0
Reviewed Sep 10, 2020
Yushi
4.0
Reviewed Jul 4, 2024
Vignesh
4.0
Reviewed Jul 8, 2021
Yilber
4.0
Reviewed Oct 1, 2020
Amrith
3.0
Reviewed Dec 29, 2022
Nikolai
2.0
Reviewed Apr 29, 2026
Dhanabal
1.0
Reviewed Aug 13, 2025
Rayan
1.0
Reviewed Aug 12, 2023