Cover Image for Alluxio Monthly Webinar: Cloud-Native Model Training on Distributed Data
Cover Image for Alluxio Monthly Webinar: Cloud-Native Model Training on Distributed Data
Avatar for Alluxio
Presented by
Alluxio
Hosted By
18 Went

Alluxio Monthly Webinar: Cloud-Native Model Training on Distributed Data

Virtual
Registration Closed
This event is not currently taking registrations. You may contact the host or subscribe to receive updates.
About Event

You're invited to Alluxio's Monthly Webinar - Multi Cloud Series!

PLEASE REGISTER ON ZOOM: https://us06web.zoom.us/webinar/register/6817078655619/WN_h056GZPpSAGHt9eTBs1neQ

Join the Alluxio Slack Community at alluxio.io/slack to engage with our speaker and send us your burning questions 🔥


Cloud-Native Model Training on Distributed Data

Cloud-native model training jobs require fast data access to achieve shorter training cycles. Accessing data can be challenging when your datasets are distributed across different regions and clouds. Additionally, as GPUs remain scarce and expensive resources, it becomes more common to set up remote training clusters from where data resides. This multi-region/cloud scenario introduces the challenges of losing data locality, resulting in operational overhead, latency and expensive cloud costs.

In the third webinar of the multi-cloud webinar series, ChanChan and Shawn will dive deep into:

  • The data locality challenges in the multi-region/cloud ML pipeline

  • Using a cloud-native distributed caching system to overcome these challenges

  • The architecture and integration of PyTorch/Ray+Alluxio+S3 using POSIX or RESTful APIs

  • ​Live demo with ResNet and BERT benchmark results showing performance gains and cost savings analysis

The multi-cloud webinar series provides you with insights into the latest trends in large-scale analytics and AI on multi-cloud. Missed the previous two webinars? View the slides and recordings here:


Speaker: ChanChan Mao

ChanChan Mao is a Developer Advocate at Alluxio. She holds a Bachelor’s degree in Computer Science from UC Santa Barbara and has turned her technical background towards growing and supporting the open source community. She focuses on fostering relationships with open source users and raising awareness of Alluxio’s brand and technology through maintaining Alluxio’s Slack community, producing short form video content, and organizing events with adjacent ecosystem communities.

Speaker: Shawn Sun

Shawn Sun is a Tech Lead of Cloud Native at Alluxio. He is an open-source contributor of Alluxio and a committer of Fluid. He is currently working on containerization of Alluxio, including the integration of Alluxio and docker, Kubernetes, and CSI. Before joining Alluxio, he received his Master’s degree in Computer Science from Duke University.


About Alluxio

Alluxio solves data challenges by providing a platform between your existing compute engines and storage systems, optimizing data access at every step of your data pipeline to accelerate analytics and AI workloads, on-prem, in the cloud, or both.

If you're interested in learning more and discovering how Alluxio can fit into your existing tech stack, schedule a demo here.

Avatar for Alluxio
Presented by
Alluxio
Hosted By
18 Went