AI TraceTrace Foundation, Inc.
Productivity AutomationVerified

Reviewed and published by trentmaziarz, April 21, 2026. Discovered and drafted by our automated research pipeline.

NVIDIA offers NIM (NVIDIA Inference Microservices), a set of prepackaged AI model containers that let developers and enterprises deploy large language models and other AI models in minutes rather than weeks, across clouds, data centers, and local workstations. NIM became generally available to developers in June 2024.

Details

NIM microservices provide pre-optimized AI model containers built on inference engines including NVIDIA TensorRT, TensorRT-LLM, vLLM, and SGLang, along with industry-standard APIs, runtime dependencies, and enterprise-grade support. Developers can access NIM for models from NVIDIA and partners through the NVIDIA Developer Program for free research and development, or in production via the NVIDIA AI Enterprise software platform. NIM reduces model deployment times from weeks to minutes and has been embedded into platforms from Amazon SageMaker, Microsoft Azure AI, and over 150 ecosystem partners.

Products affected

NVIDIA NIMNVIDIA AI EnterpriseNVIDIA Developer Program

Sources & Evidence

Cite this record

Trace Foundation. (2026). NVIDIA: NVIDIA offers NIM (NVIDIA Inference Microservices), a set of prepackaged AI model containers that let developers and enterprises deploy large language models and other AI models in minutes rather than weeks, across clouds, data centers, and local workstations. NIM became generally available to developers in June 2024 (data as of 2026-07-22) [Data set record]. AI Trace. https://www.aitrace.org/r/practice/2bdf35d9-6016-49b6-9fba-ab1bab03d88c. Accessed September 6, 2026.

Stable link
https://www.aitrace.org/r/practice/2bdf35d9-6016-49b6-9fba-ab1bab03d88c
Data as of
July 22, 2026
Last verified
April 21, 2026

How to cite AI Trace

Other practices by NVIDIA

Have evidence about NVIDIA's AI practices? Submit a report.

Report a Sighting →