Course Introduction
This episode introduces NVIDIA NIM (NVIDIA Inference Microservices), which packages language, vision, and speech models as standardized deployable inference services. It covers NIM containers, model catalogs, APIs, GPU runtime requirements, and deployment, and explains how optimized inference engines, scaling, and consistent service interfaces reduce the engineering cost of moving enterprise AI models into production.
Total recording duration: 21 min 37 sec.
Companion Git repository: https://qytgit.qytang.com/qytadmin/2026-11-Nvidia-NIM
COURSE RECORDINGS
Course Video
Watch the full recording here or open it on the original video platform.
BILIBILI01
FULL SESSION
2026 Evolution Ep. 11: NVIDIA NIM AI Inference Microservices
22 MIN
Watch on the Original Platform