Details
NIM microservices provide pre-optimized AI model containers built on inference engines including NVIDIA TensorRT, TensorRT-LLM, vLLM, and SGLang, along with industry-standard APIs, runtime dependencies, and enterprise-grade support. Developers can access NIM for models from NVIDIA and partners through the NVIDIA Developer Program for free research and development, or in production via the NVIDIA AI Enterprise software platform. NIM reduces model deployment times from weeks to minutes and has been embedded into platforms from Amazon SageMaker, Microsoft Azure AI, and over 150 ecosystem partners.
Have evidence about NVIDIA's AI practices? Submit a report.
Report a Sighting →