32 subscribers
התחל במצב לא מקוון עם האפליקציה Player FM !
פודקאסטים ששווה להאזין
בחסות


1 When Killers Realize It's Over: Raw Police Interrogation Murderer Reaction Compilation 1:22:29
Inference in Action: Scaling Al Smarter with Inferless
Manage episode 446638123 series 3332465
In this episode, we sit down with Nilesh Agarwal, co-founder of Inferless, a platform designed to streamline serverless GPU inference. We’ll cover the evolving landscape of model deployment, explore open-source tools like KServe and Knative, and discuss how Inferless solves common bottlenecks, such as cold starts and scaling issues. We also take a closer look at real-world examples like CleanLab, who saved 90% on GPU costs using Inferless.
Whether you’re a developer, DevOps engineer, or tech enthusiast curious about the latest in AI infrastructure, this podcast offers insights into Kubernetes-based model deployment, efficient updates, and the future of serverless ML. Tune in to hear Nilesh's journey from Amazon to founding Inferless and how his platform is transforming the way companies deploy machine learning models.
Subscribe now for more episodes!
Show Links:
- OpenShift 4.17 is GA https://www.youtube.com/live/DvKHwz-c11c?si=6Zap6hk_GsQfdX2m
- Policy SBOM from Styra: https://www.styra.com/blog/introducing-policy-sbom/
- NVIDIA GEForce NOW runs on KubeVirt https://thenewstack.io/now-nvidia-scaled-its-cloud-services-with-kubevirt/
- CBT feedback https://thenewstack.io/kubernetes-advances-cloud-native-data-protection-share-feedback
- CNCF KUBEEDGE Grad https://www.devopsdigest.com/cncf-announces-kubeedge-graduation?utm_source=tldrdevops
- Palumi Operator 2.0 https://www.pulumi.com/blog/pulumi-kubernetes-operator-2-0
Inferless LInks:
- https://www.inferless.com/blog/cleanlab-saves-90-on-gpu-costs-with-inferless-serverless-inference
- https://www.inferless.com/blog/how-spoofsense-scaled-their-ai-inference-with-inferless-dynamic-batching-autoscaling
- https://www.inferless.com/
- https://docs.inferless.com/introduction/introduction
- LinkedIn - https://www.linkedin.com/in/nilesh-agarwal/
- X- https://x.com/nilesh_agarwal2
- Medium Blog https://nilesh-agarwal.medium.com/
88 פרקים
Manage episode 446638123 series 3332465
In this episode, we sit down with Nilesh Agarwal, co-founder of Inferless, a platform designed to streamline serverless GPU inference. We’ll cover the evolving landscape of model deployment, explore open-source tools like KServe and Knative, and discuss how Inferless solves common bottlenecks, such as cold starts and scaling issues. We also take a closer look at real-world examples like CleanLab, who saved 90% on GPU costs using Inferless.
Whether you’re a developer, DevOps engineer, or tech enthusiast curious about the latest in AI infrastructure, this podcast offers insights into Kubernetes-based model deployment, efficient updates, and the future of serverless ML. Tune in to hear Nilesh's journey from Amazon to founding Inferless and how his platform is transforming the way companies deploy machine learning models.
Subscribe now for more episodes!
Show Links:
- OpenShift 4.17 is GA https://www.youtube.com/live/DvKHwz-c11c?si=6Zap6hk_GsQfdX2m
- Policy SBOM from Styra: https://www.styra.com/blog/introducing-policy-sbom/
- NVIDIA GEForce NOW runs on KubeVirt https://thenewstack.io/now-nvidia-scaled-its-cloud-services-with-kubevirt/
- CBT feedback https://thenewstack.io/kubernetes-advances-cloud-native-data-protection-share-feedback
- CNCF KUBEEDGE Grad https://www.devopsdigest.com/cncf-announces-kubeedge-graduation?utm_source=tldrdevops
- Palumi Operator 2.0 https://www.pulumi.com/blog/pulumi-kubernetes-operator-2-0
Inferless LInks:
- https://www.inferless.com/blog/cleanlab-saves-90-on-gpu-costs-with-inferless-serverless-inference
- https://www.inferless.com/blog/how-spoofsense-scaled-their-ai-inference-with-inferless-dynamic-batching-autoscaling
- https://www.inferless.com/
- https://docs.inferless.com/introduction/introduction
- LinkedIn - https://www.linkedin.com/in/nilesh-agarwal/
- X- https://x.com/nilesh_agarwal2
- Medium Blog https://nilesh-agarwal.medium.com/
88 פרקים
כל הפרקים
×
1 Database as a service with Percona Everest 1:02:44


1 Increasing AI adoption using Kubernetes 52:03

1 Monolith to Microservices using Kubernetes at Guidewire 1:06:28

1 Inference in Action: Scaling Al Smarter with Inferless 55:17

1 Container security with Wiz 1:02:33

1 Dagger.io Deep Dive with Co-Founder Sam Alba 1:06:24

1 Running Ray on Kubernetes with KubeRay 53:06

1 Building scalable data platforms using Data on EKS 1:02:20

1 Deploy and fine-tune LLM models on Kubernetes using KAITO 44:17

1 The business case for cloud-native and Kubernetes 54:24

1 Building the AI Hyperscaler with Kubernetes 54:56

1 Shifting Minds: Exploring OpenShift's AI Landscape 1:05:07

1 Training Machine Learning (ML) models on Kubernetes 55:29

1 The evolution of service mesh technologies 1:08:00



1 Open Policy Agent (OPA) 101 1:07:20

1 Ops Ops Hooray! Navigating IDPs from an Ops perspective 58:17

1 Generative AI on Kubernetes 1:15:56

1 IDPs Unveiled: Accelerating Deployment on Kubernetes 59:52

1 Running Kubernetes at the Edge using K3s 53:51

1 Running multi-tenant Kubernetes clusters using vCluster 57:58


1 Byte-sized: Exploring the Basics of AI in Plain English 1:00:18

1 Kubecon North America 2023: Highlights, Themes and Key Takeaways 57:41

1 Universal Control Planes for Kubernetes and Beyond 59:36

1 DevOpsDays Boston - Helping developers be more productive in a multi-cloud world 35:02

1 DevOpsDays Boston - Platform Engineering and Internal Developer Platforms 31:12

1 DevOpsDays Boston - Real value of community 42:41
ברוכים הבאים אל Player FM!
Player FM סורק את האינטרנט עבור פודקאסטים באיכות גבוהה בשבילכם כדי שתהנו מהם כרגע. זה יישום הפודקאסט הטוב ביותר והוא עובד על אנדרואיד, iPhone ואינטרנט. הירשמו לסנכרון מנויים במכשירים שונים.