Hey,
Thank you for the code. I'm trying to run it on my university's cluster, but the GPUs have memory limits, so I was thinking of splitting the workload across multiple GPUs. Would it take a lot of effort to convert this code to a multi-node setup?
Hey,
Thank you for the code. I'm trying to run it on my university's cluster, but the GPUs have memory limits, so I was thinking of splitting the workload across multiple GPUs. Would it take a lot of effort to convert this code to a multi-node setup?