Back to Optimize TensorFlow Models For Deployment with TensorRT
Coursera

Optimize TensorFlow Models For Deployment with TensorRT

This is a hands-on, guided project on optimizing your TensorFlow models for inference with NVIDIA's TensorRT. By the end of this 1.5 hour long project, you will be able to optimize Tensorflow models using the TensorFlow integration of NVIDIA's TensorRT (TF-TRT), use TF-TRT to optimize several deep learning models at FP32, FP16, and INT8 precision, and observe how tuning TF-TRT parameters affects performance and inference throughput. Prerequisites: In order to successfully complete this project, you should be competent in Python programming, understand deep learning and what inference is, and have experience building deep learning models in TensorFlow and its Keras API. Note: This course works best for learners who are based in the North America region. We’re currently working on providing the same experience in other regions.

Status: Tensorflow
Status: Python Programming
IntermediateGuided Project2 hours

Featured reviews

AA

Reviewed Mar 14, 2022

T​he first to introduce such a rare and important topic.

LS

Reviewed Jun 3, 2021

G​reat workshop, all the concepts were very well explained.

All reviews

Showing: 20 of 20

דמיטרי
Reviewed May 16, 2023
Awais
Reviewed Mar 28, 2021
Deleted
Reviewed Jun 15, 2023
Dmytro
Reviewed Aug 3, 2023
Jorge
Reviewed Feb 25, 2021
Hleb
Reviewed May 25, 2023
Luis
Reviewed Jun 4, 2021
Abdelrahman
Reviewed Mar 15, 2022
Fabian
Reviewed Apr 20, 2021
Nusrat
Reviewed Apr 16, 2021
Chandra
Reviewed Dec 13, 2020
Fangwen
Reviewed Sep 12, 2023
Maftuna
Reviewed Sep 10, 2020
Yushi
Reviewed Jul 4, 2024
Vignesh
Reviewed Jul 8, 2021
Yilber
Reviewed Oct 1, 2020
Amrith
Reviewed Dec 29, 2022
Nikolai
Reviewed Apr 29, 2026
Dhanabal
Reviewed Aug 13, 2025
Rayan
Reviewed Aug 12, 2023