#390: High performance and the lowest cost machine learning inference in the cloud with AWS Inferentia - podcast episode cover

#390: High performance and the lowest cost machine learning inference in the cloud with AWS Inferentia

Sep 06, 202020 min
--:--
--:--
Download Metacast podcast app
Listen to this episode in Metacast mobile app
Don't just listen to podcasts. Learn from them with transcripts, summaries, and chapters for every episode. Skim, search, and bookmark insights. Learn more

Episode description

AWS Inferentia is custom built by AWS to provide high performance and lowest cost machine learning inference in the cloud. Amazon EC2 Inf1 instances, powered by AWS Inferentia, provide up to 3x higher throughput and up to 40% lower cost per inference over comparable GPU-based instances. In this podcast, learn more about AWS Inferentia, Inf1 instances, and how to get started with Inf1 instances. https://aws.amazon.com/machine-learning/inferentia/
For the best experience, listen in Metacast app for iOS or Android