Video by CNCF [Cloud Native Computing Foundation] via YouTube

Join Pavan Madduri of W.W. Grainger at KubeCon + CloudNativeCon North America, November 9-12 in Salt Lake City, Utah.
In “Scaling AI Model Distribution with Dragonfly: From HuggingFace to Production GPU Clusters,” Pavan and Wenbo Qi of ByteDance will explore how cloud native teams can efficiently deliver AI models that span hundreds of gigabytes across GPU infrastructure.
Learn about Dragonfly’s HuggingFace and ModelScope integrations, transparent container acceleration with dragonfly-injector, and patterns for distributing large models at scale.
Explore the session:
https://bit.ly/4hNVUbp
Register:
https://bit.ly/3BDI5XL