Challenges and Solutions in Machine Learning Workflows

This chapter discusses the importance of velocity in ML workflows, the use of experiment trackers, and the difficulties faced in testing and validating ideas early on to avoid failures after deployment.

Play episode from 56:53

chevron_right

Transcript

chevron_right

Transcript

Episode notes

In episode 89 of The Gradient Podcast, Daniel Bashir speaks to Shreya Shankar.

Shreya is a computer scientist pursuing her PhD in databases at UC Berkeley. Her research interest is in building end-to-end systems for people to develop production-grade machine learning applications. She was previously the first ML engineer at Viaduct, did research at Google Brain, and software engineering at Facebook. She graduated from Stanford with a B.S. and M.S. in computer science with concentrations in systems and artificial intelligence. At Stanford, helped run SHE++, an organization that helps empower underrepresented minorities in technology.

Have suggestions for future podcast guests (or other feedback)? Let us know here or reach us at editor@thegradient.pub

Subscribe to The Gradient Podcast: Apple Podcasts | Spotify | Pocket Casts | RSSFollow The Gradient on Twitter

Outline:

* (00:00) Intro

* (02:22) Shreya’s background and journey into ML / MLOps

* (04:51) ML advances in 2013-2016

* (05:45) Shift in Stanford undergrad class ecosystems, accessibility of deep learning research

* (09:10) Why Shreya left her job as an ML engineer

* (13:30) How Shreya became interested in databases, data quality in ML

* (14:50) Daniel complains about things

* (16:00) What makes ML engineering uniquely difficult

* (16:50) Being a “historian of the craft” of ML engineering