Naar hoofdinhoud gaan

Build a multi-LLM chat application with Azure Container Apps

april 11, 2024 | 3:30 - 3:45 p.m. (UTC) Gecoördineerde Universele Tijd

Klaar om aan de slag te gaan met AI en de nieuwste technologieën? Microsoft Reactor biedt evenementen, training en communitybronnen om ontwikkelaars, ondernemers en startups te helpen bouwen op AI-technologie en meer. Kom kijken.

Build a multi-LLM chat application with Azure Container Apps

april 11, 2024 | 3:30 - 3:45 p.m. (UTC) Gecoördineerde Universele Tijd

Klaar om aan de slag te gaan met AI en de nieuwste technologieën? Microsoft Reactor biedt evenementen, training en communitybronnen om ontwikkelaars, ondernemers en startups te helpen bouwen op AI-technologie en meer. Kom kijken.

Terug

Build a multi-LLM chat application with Azure Container Apps

april 11, 2024 | 3:30 - 3:45 p.m. (UTC) Gecoördineerde Universele Tijd

  • Notatie:
  • alt##LivestreamLivestream

Onderwerp: Codering, Talen en Frameworks

Taal: Engels

In this demo, explore how to leverage GPU workload profiles in ACA to run your own model backend, and easily switch, compare, and speed up your inference times. You will also explore how to leverage LlamaIndex (https://github.com/run-llama/llama_index) to ingest data on-demand, and host models using Ollama (https://github.com/ollama/ollama). Then finally, decompose the application as a set of microservices written in Python, deployed on ACA.

  • Azure

Sprekers

Delen van deze pagina kunnen machinaal of door AI vertaald zijn.