Parasail to Combine NVIDIA AI Infrastructure with d-Matrix Accelerators to Achieve 10x Faster Token Generation
Article excerpt
Highlighted: the sentence this signal was extracted from
Inference service provider to deliver faster, more cost-efficient tokens by pairing NVIDIA Hopper and Blackwell GPUs with d-Matrix Corsair accelerators SAN FRANCISCO and SANTA CLARA, Calif., July 8, 2026 /PRNewswire/ -- Parasail, the inference cloud for AI-native startups, and d-Matrix, a pioneer in low-latency AI inference compute platforms for data centers, today announced Parasail is deploying d-Matrix Corsair inference accelerators alongside the NVIDIA Hopper and NVIDIA Blackwell architectures to deliver up to 10x faster, more cost-efficient inference services to its customers. Parasail's Corsair deployment marks one of the first commercial-scale examples of heterogeneous disaggregated inference in production, where NVIDIA AI infrastructure and d-Matrix's purpose-built Corsair inference accelerators will operate in concert, each performing computing tasks in which they excel. Using this approach, Parasail aims to improve inference economics for select workloads by combining NVIDIA GPUs for compute-intensive prefill with d-Matrix Corsair accelerators for latency-sensitive decode. With approvals and construction buildouts of new data centers a multi-year process, Parasail is leaning into a heterogeneous compute approach with d-Matrix to obtain even more value from the NVIDIA AI infrastructure already in its data centers. By pairing Corsair with its Hopper and Blackwell...
Keep reading with a free account
The rest of this article, and every signal for d-Matrix, is in your free account.
