Skip to content

Folders and files

NameName
Last commit message
Last commit date

Latest commit

 

History

3 Commits
 
 
 
 

Repository files navigation

MLServe

Scaling Machine Learning Workloads on GPU Clusters

  • Configured an NVIDIA GPU cluster with CUDA, cuDNN, and Jaxlib.
  • Used Alpa to leverage model parallelism and statistical multiplexing to scale inference workloads across GPUs using Ray framework.

About

Scaling Machine Learning Workloads on GPU Clusters

Resources

Stars

0 stars

Watchers

1 watching

Forks

Releases

Packages

Contributors