56 subscribers
با برنامه Player FM !
Systems Engineer Navigating the World of ML // Andrew Dye // MLOps Podcast #136
Manage episode 349591732 series 3241972
MLOps Coffee Sessions #136 with Andrew Dye, Systems Engineer Navigating the World of ML co-hosted by David Aponte.
// Abstract
We don't hear that much about working at a very low level on this podcast but they are still very valid. Andrew is able to give us his take on why and what you need to keep in mind when you are working at these low levels and why it is very important when you are a Machine Learning Engineer and how the two can play together nicely.
Most MLOps teams are formed using existing people and exitsing engineers. More often than not you have to blend these various disciplines and it works well when there's a common goal.
// Bio
Andrew is a software engineer at Union and contributor to Flyte, a production grade data and ML orchestration platform. Prior to that he was a tech lead for ML Infrastructure at Meta, where he focused on ML training reliability.
// MLOps Jobs board
https://mlops.pallet.xyz/jobs
// MLOps Swag/Merch
https://mlops-community.myshopify.com/
// Related Links
--------------- ✌️Connect With Us ✌️ -------------
Join our slack community: https://go.mlops.community/slack
Follow us on Twitter: @mlopscommunity
Sign up for the next meetup: https://go.mlops.community/register
Catch all episodes, blogs, newsletters, and more: https://mlops.community/
Connect with Demetrios on LinkedIn: https://www.linkedin.com/in/dpbrinkm/
Connect with David on LinkedIn: https://www.linkedin.com/in/aponteanalytics/
Connect with Andrew on LinkedIn: https://www.linkedin.com/in/andrewwdye
Timestamps:
[00:00] Andrew's preferred coffee
[03:30] Introduction to Andrew Dye
[03:33] Takeaways
[07:32] Huge shoutout to our sponsors UnionML and UnionAI!
[07:48] Andrew's background
[10:08] Andrew's learning curve
[11:10] Bridging the gap between firmware space and MLOps
[12:18] In connection with Pytorch team
[12:54] Things that should have learned sooner
[14:54] Type of scale Andrew works on
[17:42] Distributed training at Meta
[19:55] Managing the huge search space
[22:18] Execution patterns programs
[23:20] Non-ML engineers dealing with ML engineers having the same skill set
[26:44] Pace rapid change adoptation
[29:18] Consensus challenges
[32:26] Abstractions making sense now
[34:53] Comparing to others
[39:21] General principles in UnionAI tooling
[41:54] Seeing the future
[43:54] Inter-task checkpointing
[44:52] Combining functionality with use cases
[46:17] Wrap up
448 قسمت
Manage episode 349591732 series 3241972
MLOps Coffee Sessions #136 with Andrew Dye, Systems Engineer Navigating the World of ML co-hosted by David Aponte.
// Abstract
We don't hear that much about working at a very low level on this podcast but they are still very valid. Andrew is able to give us his take on why and what you need to keep in mind when you are working at these low levels and why it is very important when you are a Machine Learning Engineer and how the two can play together nicely.
Most MLOps teams are formed using existing people and exitsing engineers. More often than not you have to blend these various disciplines and it works well when there's a common goal.
// Bio
Andrew is a software engineer at Union and contributor to Flyte, a production grade data and ML orchestration platform. Prior to that he was a tech lead for ML Infrastructure at Meta, where he focused on ML training reliability.
// MLOps Jobs board
https://mlops.pallet.xyz/jobs
// MLOps Swag/Merch
https://mlops-community.myshopify.com/
// Related Links
--------------- ✌️Connect With Us ✌️ -------------
Join our slack community: https://go.mlops.community/slack
Follow us on Twitter: @mlopscommunity
Sign up for the next meetup: https://go.mlops.community/register
Catch all episodes, blogs, newsletters, and more: https://mlops.community/
Connect with Demetrios on LinkedIn: https://www.linkedin.com/in/dpbrinkm/
Connect with David on LinkedIn: https://www.linkedin.com/in/aponteanalytics/
Connect with Andrew on LinkedIn: https://www.linkedin.com/in/andrewwdye
Timestamps:
[00:00] Andrew's preferred coffee
[03:30] Introduction to Andrew Dye
[03:33] Takeaways
[07:32] Huge shoutout to our sponsors UnionML and UnionAI!
[07:48] Andrew's background
[10:08] Andrew's learning curve
[11:10] Bridging the gap between firmware space and MLOps
[12:18] In connection with Pytorch team
[12:54] Things that should have learned sooner
[14:54] Type of scale Andrew works on
[17:42] Distributed training at Meta
[19:55] Managing the huge search space
[22:18] Execution patterns programs
[23:20] Non-ML engineers dealing with ML engineers having the same skill set
[26:44] Pace rapid change adoptation
[29:18] Consensus challenges
[32:26] Abstractions making sense now
[34:53] Comparing to others
[39:21] General principles in UnionAI tooling
[41:54] Seeing the future
[43:54] Inter-task checkpointing
[44:52] Combining functionality with use cases
[46:17] Wrap up
448 قسمت
همه قسمت ها
×
1 AI Reliability, Spark, Observability, SLAs and Starting an AI Infra Company 1:37:22

1 The Creator of FastAPI’s Next Chapter // Sebastián Ramírez // #324 1:09:37

1 A Candid Conversation Around MCP and A2A // Rahul Parundekar and Sam Partee // #316 SF Live 1:04:42

1 Making AI Reliable is the Greatest Challenge of the 2020s // Alon Bochman // #312 1:01:37

1 Behavior Modeling, Secondary AI Effects, Bias Reduction & Synthetic Data // Devansh Devansh // #311 1:01:35

1 GraphBI: Expanding Analytics to All Data Through the Combination of GenAI, Graph, & Visual Analytics // Paco Nathan & Weidong Yang // #310 1:14:01

1 I Am Once Again Asking "What is MLOps?" // Oleksandr Stasyk // #308 1:07:22

1 Agents of Innovation: AI-Powered Product Ideation with Synthetic Consumer Testing // Luca Fiaschi // #306 1:02:23

1 We're All Finetuning Incorrectly // Tanmay Chopra // #304 1:00:30



1 From Rules to Reasoning Engines // George Mathew // #296 1:05:26

1 GenAI Traffic: Why API Infrastructure Must Evolve... Again // Erica Hughberg // #296 1:06:24

1 Future of Software, Agents in the Enterprise, and Inception Stage Company Building // Eliot Durbin // #293 54:26

1 The Agent Landscape - Lessons Learned Putting Agents Into Production 1:08:40

1 Evolving Workflow Orchestration // Alex Milowski // #291 1:14:34


به Player FM خوش آمدید!
Player FM در سراسر وب را برای یافتن پادکست های با کیفیت اسکن می کند تا همین الان لذت ببرید. این بهترین برنامه ی پادکست است که در اندروید، آیفون و وب کار می کند. ثبت نام کنید تا اشتراک های شما در بین دستگاه های مختلف همگام سازی شود.