Skip to content

Multi-Cloud Webinar Series: Cloud-Native Model Training on Distributed Data

Photo of Bin Fan
Hosted By
Bin F.
Multi-Cloud Webinar Series: Cloud-Native Model Training on Distributed Data

Details

Register Here:
https://us06web.zoom.us/webinar/register/9016993969420/WN_h056GZPpSAGHt9eTBs1neQ
***
Join this live webinar for a chance to win a $50 Amazon Gift Card!

Cloud-native model training jobs require fast data access to achieve shorter training cycles. Accessing data can be challenging when your datasets are distributed across different regions and clouds. Additionally, as GPUs remain scarce and expensive resources, it becomes more common to set up remote training clusters from where data resides. This multi-region/cloud scenario introduces the challenges of losing data locality, resulting in operational overhead, latency and expensive cloud costs.
In the third webinar of the multi-cloud webinar series, Chanchan and Shawn will dive deep into:

  • The data locality challenges in the multi-region/cloud ML pipeline
  • Using a cloud-native distributed caching system to overcome these challenges
  • The architecture and integration of PyTorch/Ray+Alluxio+S3 using POSIX or RESTful APIs
  • ​Live demo with ResNet and BERT benchmark results showing performance gains and cost savings analysis

The multi-cloud webinar series provides you with insights into the latest trends in large-scale analytics and AI on multi-cloud. Missed the previous two webinars? View the slides and recordings here:

***
Register Here:
https://us06web.zoom.us/webinar/register/9016993969420/WN_h056GZPpSAGHt9eTBs1neQ

Photo of Silicon Valley Data Engineering group
Silicon Valley Data Engineering
See more events
Needs a location