BEGIN:VCALENDAR
VERSION:2.0
PRODID:jquery.icalendar
BEGIN:VEVENT
ORGANIZER:MAILTO:nur@eventhandler.co.il
TITLE:Building a Production-Grade AI/ML Inference Platform on Kubernetes
DTSTART:20251211T083500Z
DTEND:20251211T091500Z
SUMMARY:Building a Production-Grade AI/ML Inference Platform on Kubernetes
DESCRIPTION:Building an AI/ML platform that can analyze and summarize massive volumes of unstructured medical data in real time requires more than just powerful models, it demands a production-grade infrastructure capable of handling complex inference workloads at scale.This session dives into how to design and operate AI inference workloads on Kubernetes through an Amazon EKS example, for real-world production environments. It will cover model handoff processes from data science teams, validation and readiness assessments, deployment architecture, scheduling and scaling strategies, and inference traffic management. The session will also explore observability, performance tuning, and optimization techniques that are essential for maintaining reliability under heavy load.Walk away with a deep understanding of the operational challenges, trade-offs, and engineering solutions involved in running high-performance AI/ML inference systems efficiently and responsibly at scale.
LOCATION:Track 2 | A2 A3
END:VEVENT
END:VCALENDAR