Meet the new standard for high-performance, low-cost inference: NVIDIA Dynamo 1.0 is now available to DigitalOcean customers.
Article excerpt
Meet the new standard for high-performance, low-cost inference: NVIDIA Dynamo 1.0 is now available to DigitalOcean customers. Table of contents. NVIDIA Dynamo 1.0, which was released on Monday at NVIDIA GTC, is now available to DigitalOcean customers to help drive performance enhancements and cost efficiency. NVIDIA Dynamo 1.0 offers a 7x inference performance increase on NVIDIA GB200 NVL systems, and by pairing it with DigitalOcean's Agentic Inference Cloud, customers can achieve higher performance at lower costs while benefiting from seamless deployment. Working together, DigitalOcean's optimizations with NVIDIA have already achieved a 67% cost savings for customers like Workato, and this new generation of Dynamo can unlock even greater gains for businesses who run production-grade agentic workflows. DigitalOcean customers can get access to NVIDIA Dynamo 1.0 as a container image that can be run on a Droplet or can deploy directly on DigitalOcean Kubernetes with an inference runtime (vLLM, SGlang, TensorRT). NVIDIA Dynamo is a cutting-edge, high-performance inference service framework specifically designed to accelerate and optimize large-scale generative AI and inference models. Dynamo is an orchestration layer that sits above engines like vLLM, SGLang, and NVIDIA TensorRT-LLM. Think of it as the distributed traffic controller for your GPU fleet, seamlessly...
Keep reading with a free account
The rest of this article, and every signal for DigitalOcean, is in your free account.
Extracted from this sentence
Recently, DigitalOcean, Inc. partnered with Workato's AI Research Lab to scale agentic AI capabilities across its platform, which processes over 1 trillion automated workloads.
.png)