NVIDIA NIM

Prebuilt, optimised inference microservices for NVIDIA GPUs.

Containerised model servers with TensorRT-LLM optimisations and OpenAI-compatible APIs, runnable on any NVIDIA GPU or tried for free on build.nvidia.com.

Vendor
NVIDIA
Category
Inference & model hosting
Pricing
Enterprise
Open source
No
Platforms
Cloud, On-prem, Kubernetes
Launched
2024
Website
build.nvidia.com

Features

Best for

Works with

More inference & model hosting

Links