Generative AI on Kubernetes

https://is1-ssl.mzstatic.com/image/thumb/Podcasts122/v4/15/c5/13/15c513e6-e1cc-2c72-d7a3-fbbf854fbc44/mza_13914227192634054396.jpg/600x600bb.jpg

Kubernetes Bytes

Ryan Wallner & Bhavin Shah

89 episodes

7 months ago

Kubernetes Bytes is a podcast bringing you the latest from the world of cloud native data management. Hosts Ryan Wallner and Bhavin Shah come to you from Boston, Massachusetts with experienced backgrounds in cloud-native tech. They'll be sharing their thoughts on recent cloud native news and talking to industry experts about their experiences and challenges managing the wealth of data in today's cloud-native ecosystem.

Technology

RSS

All content for Kubernetes Bytes is the property of Ryan Wallner & Bhavin Shah and is served directly from their servers with no modification, redirects, or rehosting. The podcast is not affiliated with or endorsed by Podjoint in any way.

Technology

https://media.zencastr.com/image-files/60f9ac743534330029a39c99/1b90b3a1-2fb2-4593-8f86-0c3a21f68ffc.jpg

Generative AI on Kubernetes

Kubernetes Bytes

1 hour 15 minutes

1 year ago

Generative AI on Kubernetes

In this episode of the Kubernetes Bytes podcast, Ryan and Bhavin sit down with Janakiram MSV - an advisor, analyst and architect to talk about how users can run Generative AI models on Kubernetes. The discussion revolves around Jani's home lab and his experimentation with different LLM models and how to get them running on NVIDIA GPUs. Jani has spent the past year becoming a subject matter expert in GenAI, and this discussion highlights all the different challenges he faced and what lessons he learnt from them. Check out our website at https://kubernetesbytes.com/ Episode Sponsor: Elotl * https://elotl.co/luna * https://www.elotl.co/luna-free-trial Timestamps: * 02:02 Cloud Native News * 15:31 Interview with Jani * 01:11:00 Key takeaways Cloud Native News: * https://www.techerati.com/press-release/octopus-deploy-acquires-codefresh-to-boost-kubernetes-and-cloud-native-delivery/ * https://www.civo.com/blog/kubefirst-joins-civo * https://cast.ai/kubernetes-cost-benchmark * https://www.techradar.com/pro/vmware-customers-are-jumping-ship-as-broadcom-sales-continue-heres-where-theyre-moving-to * https://cloudonair.withgoogle.com/events/techbyte-making-ai-ml-scalable-cost-effective-gke * https://dok.community/dok-events/dok-day-kubecon-paris/ * https://training.linuxfoundation.org/certification/certified-argo-project-associate-capa Show Links: * https://www.youtube.com/janakirammsv * https://www.linkedin.com/in/janakiramm/ * - NVIDIA Container Toolkit - https://docs.nvidia.com/datacenter/cloud-native/container-toolkit/latest/index.html * NVIDIA Device Plugin - https://github.com/NVIDIA/k8s-device-plugin * NVIDIA Feature Discovery - https://github.com/NVIDIA/gpu-feature-discovery * Hugging Face Text Gen Inference - https://huggingface.co/docs/text-generation-inference/index * Hugging Face Text Embeddings Inference - https://huggingface.co/docs/text-embeddings-inference/index * ChromaDB - https://www.trychroma.com/