Naar hoofdinhoud gaan

Exploring Kaito to streamline AI inference model deployment in Azure Kubernetes

juli 16, 2024 | 9:00 - 10:00 p.m. (UTC) Gecoördineerde Universele Tijd

Klaar om aan de slag te gaan met AI en de nieuwste technologieën? Microsoft Reactor biedt evenementen, training en communitybronnen om ontwikkelaars, ondernemers en startups te helpen bouwen op AI-technologie en meer. Kom kijken.

Exploring Kaito to streamline AI inference model deployment in Azure Kubernetes

juli 16, 2024 | 9:00 - 10:00 p.m. (UTC) Gecoördineerde Universele Tijd

Klaar om aan de slag te gaan met AI en de nieuwste technologieën? Microsoft Reactor biedt evenementen, training en communitybronnen om ontwikkelaars, ondernemers en startups te helpen bouwen op AI-technologie en meer. Kom kijken.

Terug

Exploring Kaito to streamline AI inference model deployment in Azure Kubernetes

juli 16, 2024 | 9:00 - 10:00 p.m. (UTC) Gecoördineerde Universele Tijd

  • Notatie:
  • alt##LivestreamLivestream

Onderwerp: Infrastructuur voor AI

Taal: Engels

About this session:
Roy Kim will be presenting Kaito, an operator streamlining AI/ML inference model deployment in Kubernetes. Discover how Kaito simplifies deployment of large open-source inference models like Falcon and LLAMA2. Learn its unique features: managing large model files with container images, preset GPU configurations, auto-provisioning GPU nodes, and hosting on Microsoft Container Registry (MCR). See how Kaito simplifies the workflow of onboarding large AI inference models in Kubernetes.

Learn more and develop your skills in Azure Kubernetes Service with this Microsoft Learn training module:
https://aka.ms/IntroToAKSLearn1

Sprekers

Delen van deze pagina kunnen machinaal of door AI vertaald zijn.