Kubernetes Bytes is a podcast bringing you the latest from the world of cloud native data management. Hosts Ryan Wallner and Bhavin Shah come to you from Boston, Massachusetts with experienced backgrounds in cloud-native tech. They'll be sharing their thoughts on recent cloud native news and talking to industry experts about their experiences and challenges managing the wealth of data in today's cloud-native ecosystem.
All content for Kubernetes Bytes is the property of Ryan Wallner & Bhavin Shah and is served directly from their servers
with no modification, redirects, or rehosting. The podcast is not affiliated with or endorsed by Podjoint in any way.
Kubernetes Bytes is a podcast bringing you the latest from the world of cloud native data management. Hosts Ryan Wallner and Bhavin Shah come to you from Boston, Massachusetts with experienced backgrounds in cloud-native tech. They'll be sharing their thoughts on recent cloud native news and talking to industry experts about their experiences and challenges managing the wealth of data in today's cloud-native ecosystem.
Training Machine Learning (ML) models on Kubernetes
Kubernetes Bytes
55 minutes
1 year ago
Training Machine Learning (ML) models on Kubernetes
In this episode of the Kubernetes Bytes podcast, Bhavin sits down with Bernie Wu, VP Strategic Partnerships and AI/CXL/Kubernetes Initiatives at Memverge. They discuss about how Kubernetes is the most popular platform to run AI model training and model inferencing jobs. The discussion dives into model training, talking about different phases of a DAG, and then talk about how Memverge can help users with efficient and cost-effective model checkpoints. The discussion goes into topics like saving costs by using spot instances, hot restart of training jobs, reclaiming unused GPU resources, etc.
Check out our website at https://kubernetesbytes.com/ [https://www.youtube.com/redirect?event=video_description&redir_token=QUFFLUhqbllTN0VpRHpuZWZiem9iWWc2bDF5NF9SMGhod3xBQ3Jtc0tuMnNGTmhwN01xdC01ZEFqRjFERnZBLUlUSTFCRWVta0tYQU4wWndITWhfV3RnbHh1cFVtbUc5NGZRandUcG1ocjQ5Q2R2cG1tQmpibkhJMVNRMHIycE1SYVJNdlVWMGFCa2xwTkhIOEFZUFRqVG5sdw&q=https%3A%2F%2Fkubernetesbytes.com%2F&v=Jf059PFn6l0]
Episode Sponsor: Nethopper
* Learn more about KAOPS: @nethopper.io
* For a supported-demo: info@nethopper.io
* Try the free version of KAOPS now! https://mynethopper.com/auth [https://www.youtube.com/redirect?event=video_description&redir_token=QUFFLUhqa2c2UWw4bFZISFN2X0hIWUtuWHlpWGk3WC1RZ3xBQ3Jtc0trQ3FpMlZDeXUwcVdZZGJ1d2R5RVREc1V0SzNPeXBzZGVCeWdzYjNsRndTVzRGcTYtUVFRUnI2bXEtWE96QVpNNThCaENiWXpqM3dzSmF6cml1c1BEYUk4bnlEQ3lqc1d0cXdTVWxuVE5YZWV3elNITQ&q=https%3A%2F%2Fmynethopper.com%2Fauth&v=Jf059PFn6l0]
Cloud Native News:
* https://www.aquasec.com/blog/linguistic-lumberjack-understanding-cve-2024-4323-in-fluent-bit/
* https://kubernetes.io/blog/2024/05/20/completing-cloud-provider-migration/
* https://thenewstack.io/introducing-aks-automatic-managed-kubernetes-for-developers/
* https://www.harness.io/blog/harness-to-acquire-split
Show Links:
* https://www.linkedin.com/in/berniewu/
* https://criu.org/Main_Page
* https://memverge.com/
* https://youtu.be/tY8YOMRuqWI?si=yB3hHqLUpYPZ-KWN
* https://youtu.be/ND4seSKpJHI?si=shh0iuA9qC-dO6eb
Timestamps:
* 01:04 [https://www.youtube.com/watch?v=Jf059PFn6l0&t=64s] Cloud Native News
* 08:47 [https://www.youtube.com/watch?v=Jf059PFn6l0&t=527s] Interview with Bernie
* 51:40 [https://www.youtube.com/watch?v=Jf059PFn6l0&t=3100s] Key takeaways
Kubernetes Bytes
Kubernetes Bytes is a podcast bringing you the latest from the world of cloud native data management. Hosts Ryan Wallner and Bhavin Shah come to you from Boston, Massachusetts with experienced backgrounds in cloud-native tech. They'll be sharing their thoughts on recent cloud native news and talking to industry experts about their experiences and challenges managing the wealth of data in today's cloud-native ecosystem.