Artwork

Innehåll tillhandahållet av The Gradient. Allt poddinnehåll inklusive avsnitt, grafik och podcastbeskrivningar laddas upp och tillhandahålls direkt av The Gradient eller deras podcastplattformspartner. Om du tror att någon använder ditt upphovsrättsskyddade verk utan din tillåtelse kan du följa processen som beskrivs här https://sv.player.fm/legal.
Player FM - Podcast-app
Gå offline med appen Player FM !

Suhail Doshi: The Future of Computer Vision

1:08:07
 
Dela
 

Manage episode 418564085 series 2975159
Innehåll tillhandahållet av The Gradient. Allt poddinnehåll inklusive avsnitt, grafik och podcastbeskrivningar laddas upp och tillhandahålls direkt av The Gradient eller deras podcastplattformspartner. Om du tror att någon använder ditt upphovsrättsskyddade verk utan din tillåtelse kan du följa processen som beskrivs här https://sv.player.fm/legal.

Episode 123

I spoke with Suhail Doshi about:

* Why benchmarks aren’t prepared for tomorrow’s AI models

* How he thinks about artists in a world with advanced AI tools

* Building a unified computer vision model that can generate, edit, and understand pixels.

Suhail is a software engineer and entrepreneur known for founding Mixpanel, Mighty Computing, and Playground AI (they’re hiring!).

Reach me at editor@thegradient.pub for feedback, ideas, guest suggestions.

Subscribe to The Gradient Podcast: Apple Podcasts | Spotify | Pocket Casts | RSSFollow The Gradient on Twitter

Outline:

* (00:00) Intro

* (00:54) Ad read — MLOps conference

* (01:30) Suhail is *not* in pivot hell but he *is* all-in on 50% AI-generated music

* (03:45) AI and music, similarities to Playground

* (07:50) Skill vs. creative capacity in art

* (12:43) What we look for in music and art

* (15:30) Enabling creative expression

* (18:22) Building a unified computer vision model, underinvestment in computer vision

* (23:14) Enhancing the aesthetic quality of images: color and contrast, benchmarks vs user desires

* (29:05) “Benchmarks are not prepared for how powerful these models will become”

* (31:56) Personalized models and personalized benchmarks

* (36:39) Engaging users and benchmark development

* (39:27) What a foundation model for graphics requires

* (45:33) Text-to-image is insufficient

* (46:38) DALL-E 2 and Imagen comparisons, FID

* (49:40) Compositionality

* (50:37) Why Playground focuses on images vs. 3d, video, etc.

* (54:11) Open source and Playground’s strategy

* (57:18) When to stop open-sourcing?

* (1:03:38) Suhail’s thoughts on AGI discourse

* (1:07:56) Outro

Links:

* Playground homepage

* Suhail on Twitter


Get full access to The Gradient at thegradientpub.substack.com/subscribe
  continue reading

135 episoder

Artwork
iconDela
 
Manage episode 418564085 series 2975159
Innehåll tillhandahållet av The Gradient. Allt poddinnehåll inklusive avsnitt, grafik och podcastbeskrivningar laddas upp och tillhandahålls direkt av The Gradient eller deras podcastplattformspartner. Om du tror att någon använder ditt upphovsrättsskyddade verk utan din tillåtelse kan du följa processen som beskrivs här https://sv.player.fm/legal.

Episode 123

I spoke with Suhail Doshi about:

* Why benchmarks aren’t prepared for tomorrow’s AI models

* How he thinks about artists in a world with advanced AI tools

* Building a unified computer vision model that can generate, edit, and understand pixels.

Suhail is a software engineer and entrepreneur known for founding Mixpanel, Mighty Computing, and Playground AI (they’re hiring!).

Reach me at editor@thegradient.pub for feedback, ideas, guest suggestions.

Subscribe to The Gradient Podcast: Apple Podcasts | Spotify | Pocket Casts | RSSFollow The Gradient on Twitter

Outline:

* (00:00) Intro

* (00:54) Ad read — MLOps conference

* (01:30) Suhail is *not* in pivot hell but he *is* all-in on 50% AI-generated music

* (03:45) AI and music, similarities to Playground

* (07:50) Skill vs. creative capacity in art

* (12:43) What we look for in music and art

* (15:30) Enabling creative expression

* (18:22) Building a unified computer vision model, underinvestment in computer vision

* (23:14) Enhancing the aesthetic quality of images: color and contrast, benchmarks vs user desires

* (29:05) “Benchmarks are not prepared for how powerful these models will become”

* (31:56) Personalized models and personalized benchmarks

* (36:39) Engaging users and benchmark development

* (39:27) What a foundation model for graphics requires

* (45:33) Text-to-image is insufficient

* (46:38) DALL-E 2 and Imagen comparisons, FID

* (49:40) Compositionality

* (50:37) Why Playground focuses on images vs. 3d, video, etc.

* (54:11) Open source and Playground’s strategy

* (57:18) When to stop open-sourcing?

* (1:03:38) Suhail’s thoughts on AGI discourse

* (1:07:56) Outro

Links:

* Playground homepage

* Suhail on Twitter


Get full access to The Gradient at thegradientpub.substack.com/subscribe
  continue reading

135 episoder

Alla avsnitt

×
 
Loading …

Välkommen till Player FM

Player FM scannar webben för högkvalitativa podcasts för dig att njuta av nu direkt. Den är den bästa podcast-appen och den fungerar med Android, Iphone och webben. Bli medlem för att synka prenumerationer mellan enheter.

 

Snabbguide