Resource
This notebook demonstrates how to optimize a fine-tuned BERT TF checkpoint to TensorRT and then how to deploy it for inference using Triton inference server on OpenShift cluster.
Use the NGC CLI to download:
Copied!
| Name | Size | Updated | Actions |
|---|---|---|---|
bert_squad_tf_finetuning.ipynbView Notebook | 28.3 KB | November 10, 2020 UTC |