US patent US12737356
Efficient data delivery for artificial intelligence systems in distributed storage environments
Abstract
A method is disclosed for managing transformed datasets in a compute cluster environment. The method includes identifying, based on one or more machine learning models to be executed on a compute cluster comprising a plurality of GPU servers, one or more transformations to apply to a dataset. The method further includes generating a transformed dataset based on the one or more transformations, storing the transformed dataset, receiving a request to transmit the transformed dataset to at least one GPU server of the plurality of GPU servers, and, responsive to the request, transmitting the stored transformed dataset to the at least one GPU server without re-performing the one or more transformations on the dataset.