Serving TensorFlow Models with a Custom Environment

Build a custom serving environment to run TensorFlow models with extra dependencies.

August 13, 2023 · 4 min · 688 words · hoangph3

TensorFlow Serving: Handling Content Types

Explore how TensorFlow Serving handles different request content types for inference.

July 29, 2023 · 8 min · 1565 words · hoangph3

Serving TensorFlow Models with JSON Requests

Send and parse JSON prediction requests against a TensorFlow Serving model.

July 27, 2023 · 4 min · 742 words · hoangph3

Serving Custom TensorFlow Models

Package and serve a custom TensorFlow model signature with TensorFlow Serving.

July 26, 2023 · 2 min · 421 words · hoangph3

Serving Multiple TensorFlow Models with Monitoring

Serve multiple TensorFlow models simultaneously and monitor them with Prometheus and Swagger.

July 26, 2023 · 5 min · 953 words · hoangph3

Deploying TensorFlow model as Canary Releases Kubernetes and Istio

Use Istio with Minikube and TensorFlow Serving to create canary deployments of TensorFlow machine learning models!

December 6, 2022 · 11 min · 2240 words · hoangph3

Prepare and Deploy a TensorFlow Model to Tensorflow Serving

Serve tensorflow model efficiently with customized model signatures

December 6, 2022 · 9 min · 1871 words · hoangph3

Integrating Seldon Core with Kafka

Stream predictions in and out of Seldon Core model deployments using Kafka.

December 2, 2022 · 2 min · 362 words · hoangph3

Building Inference Graphs with Seldon Core

Compose multi-step inference graphs (transformers, combiners, routers) with Seldon Core.

December 1, 2022 · 7 min · 1374 words · hoangph3

Deploying Your First Model with Seldon Core

Deploy a first machine learning model on Kubernetes using Seldon Core, including authentication basics.

December 1, 2022 · 2 min · 311 words · hoangph3