Artwork

Innhold levert av Conviction. Alt podcastinnhold, inkludert episoder, grafikk og podcastbeskrivelser, lastes opp og leveres direkte av Conviction eller deres podcastplattformpartner. Hvis du tror at noen bruker det opphavsrettsbeskyttede verket ditt uten din tillatelse, kan du følge prosessen skissert her https://no.player.fm/legal.
Player FM - Podcast-app
Gå frakoblet med Player FM -appen!

The marketplace for AI compute with Jared Quincy Davis from Foundry

43:12
 
Del
 

Manage episode 435548496 series 3444082
Innhold levert av Conviction. Alt podcastinnhold, inkludert episoder, grafikk og podcastbeskrivelser, lastes opp og leveres direkte av Conviction eller deres podcastplattformpartner. Hvis du tror at noen bruker det opphavsrettsbeskyttede verket ditt uten din tillatelse, kan du følge prosessen skissert her https://no.player.fm/legal.

In this episode of No Priors, hosts Sarah and Elad are joined by Jared Quincy Davis, former DeepMind researcher and the Founder and CEO of Foundry, a new AI cloud computing service provider. They discuss the research problems that led him to starting Foundry, the current state of GPU cloud utilization, and Foundry's approach to improving cloud economics for AI workloads. Jared also touches on his predictions for the GPU market and the thinking behind his recent paper on designing compound AI systems.

Sign up for new podcasts every week. Email feedback to show@no-priors.com

Follow us on Twitter: @NoPriorsPod | @Saranormous | @EladGil | @jaredq_

Show Notes:

(00:00) Introduction

(02:42) Foundry background

(03:57) GPU utilization for large models

(07:29) Systems to run a large model

(09:54) Historical value proposition of the cloud

(14:45) Sharing cloud compute to increase efficiency

(19:17) Foundry’s new releases

(23:54) The current state of GPU capacity

(29:50) GPU market dynamics

(36:28) Compound systems design

(40:27) Improving open-ended tasks

  continue reading

88 episoder

Artwork
iconDel
 
Manage episode 435548496 series 3444082
Innhold levert av Conviction. Alt podcastinnhold, inkludert episoder, grafikk og podcastbeskrivelser, lastes opp og leveres direkte av Conviction eller deres podcastplattformpartner. Hvis du tror at noen bruker det opphavsrettsbeskyttede verket ditt uten din tillatelse, kan du følge prosessen skissert her https://no.player.fm/legal.

In this episode of No Priors, hosts Sarah and Elad are joined by Jared Quincy Davis, former DeepMind researcher and the Founder and CEO of Foundry, a new AI cloud computing service provider. They discuss the research problems that led him to starting Foundry, the current state of GPU cloud utilization, and Foundry's approach to improving cloud economics for AI workloads. Jared also touches on his predictions for the GPU market and the thinking behind his recent paper on designing compound AI systems.

Sign up for new podcasts every week. Email feedback to show@no-priors.com

Follow us on Twitter: @NoPriorsPod | @Saranormous | @EladGil | @jaredq_

Show Notes:

(00:00) Introduction

(02:42) Foundry background

(03:57) GPU utilization for large models

(07:29) Systems to run a large model

(09:54) Historical value proposition of the cloud

(14:45) Sharing cloud compute to increase efficiency

(19:17) Foundry’s new releases

(23:54) The current state of GPU capacity

(29:50) GPU market dynamics

(36:28) Compound systems design

(40:27) Improving open-ended tasks

  continue reading

88 episoder

Alle episoder

×
 
Loading …

Velkommen til Player FM!

Player FM scanner netter for høykvalitets podcaster som du kan nyte nå. Det er den beste podcastappen og fungerer på Android, iPhone og internett. Registrer deg for å synkronisere abonnement på flere enheter.

 

Hurtigreferanseguide

Copyright 2024 | Sitemap | Personvern | Vilkår for bruk | | opphavsrett