Products
On-Demand Videos
video
AI/ML Infra Meetup | Open Source Michelangelo: Uber's Predictive to Generative end to end ML Lifecycle management platform

In this talk, Eric Wang, Senior Staff Software Engineer introduces Uber’s open-source generative end-to-end ML lifecycle management platform: Michelangelo.
video
AI/ML Infra Meetup | Unlock the Future of Generative AI: TorchTitan's Latest Breakthroughs

In this talk, Jiani Wang, Software Engineer Meta's Pytorch Team, dives into the overview and the latest advancements in TorchTitan.
video
AI/ML Infra Meetup | Bringing Data to GPUs Anywhere + Get Low-Latency on Object Store with Alluxio

In this talk, Bin Fan, VP of Technology at Alluxio, explores how to enable efficient data access across distributed GPU infrastructure, achieving low-latency performance for feature stores and RAG workloads.
.png)
Thank you! Your submission has been received!
Oops! Something went wrong while submitting the form.
video
The power of data orchestration: Storage Acceleration and Servitization at Shopee
ALLUXIO DAY XII 2022
April 28, 2022
Shopee is the leading e-commerce platform in SouthEast Asia. In this presentation, Tianbao Ding and Haoning Sun from Shopee will share their Data Infra team’s recent project on acceleration with Presto and storage servitization. They will share the details on how Shopee leverages Alluxio to accelerate Presto query and provide standardized method of accessing data through Alluxio-Fuse and Alluxio-S3.
Large Scale Analytics Acceleration
Data Platform Modernization
video
Alluxio on Kubernetes – Powering training through Container Storage Interface plugin
ALLUXIO DAY XII 2022
April 28, 2022
Shawn Sun from Alluxio will present the journey of using Alluxio as the storage system for Kubernetes through Container Storage Interface (CSI) plugin and Alluxio CSI driver. This talk will cover the challenges we are facing with traditional setup in the AI/ML training jobs, and how Alluxio CSI driver manages to address them. It will also talk about a recent change to the driver that made it more sturdy and robust.
Large Scale Analytics Acceleration
Model Training Acceleration
video
Securing Your Open Source Project
ALLUXIO DAY XII 2022
April 28, 2022
This talk will discuss the process and technical details behind a responsible vulnerability disclosure of an issue detected in Alluxio recently. I will share some of the lessons I’ve learned as a security researcher dealing with multiple open-source vendors and my thoughts about the actions organizations and projects should take to ensure successful vulnerability management and disclosure programs. Learn more about creating more secure software.
No items found.
video
Geo-distributed Analytics with NetApp StorageGRID and Alluxio
This presentation will include information about how Alluxio and NetApp StorageGRID helps enterprises accelerate the adoption of cloud and optimize their resource spend on a modern hybrid big data architecture. The conversation will cover use case and architecture info from a variety of enterprises and some of the high level technical details of how these business solutions are constructed.
Hybrid Multi-Cloud
Data Platform Modernization
Large Scale Analytics Acceleration
video
Speed Up Uber’s Presto with Alluxio
ALLUXIO DAY X 2022
March 3, 2022
Chen Liang from Uber and Beinan Wang from Alluxio will present the practical problems and interesting findings during the launch of Alluxio Local Cache. Their talk covers how Uber’s Presto team implements the cache invalidation and dashboard for Alluxio’s Local Cache. Chen Liang will also share his experience using a customized cache filter to resolve the performance degradation due to a large working set.
Large Scale Analytics Acceleration
video
Alluxio Journal Evolution – Towards high availability and fault tolerance
ALLUXIO DAY X 2022
March 3, 2022
Within Alluxio, the master processes keep track of global metadata for the file system. This includes file system metadata, block cache metadata, and worker metadata. When a client interacts with the filesystem it must first query or update the metadata on the master processes. Given their central role in the system, master processes can be backed by a highly available, fault tolerant replicated journal. This talk will introduce and compare the two available implementations of this journal in Alluxio, the first using Zookeeper and the more recent version using Raft.
Large Scale Analytics Acceleration
video
Building an Efficient AI Training Platform at bilibili with Alluxio
ALLUXIO DAY X 2022
March 3, 2022
In this talk, Lei Li and Zifan Ni share the experience of applying Alluxio in their AI platform to increase training efficiency at bilibili. The talk also includes technical architecture and specific issues addressed.
Model Training Acceleration
Model Distribution
Cloud Cost Savings
Data Platform Modernization
video
Architecting a Heterogeneous Data Platform Across Clusters, Regions, and Clouds
Data platform teams are increasingly challenged with accessing multiple data stores that are separated from compute engines, such as Spark, Presto, TensorFlow or PyTorch. Whether your data is distributed across multiple datacenters and/or clouds, a successful heterogeneous data platform requires efficient data access. Alluxio enables you to embrace the separation of storage from compute and use Alluxio data orchestration to simplify adoption of the data lake and data mesh paradigms for analytics and AI/ML workloads.
Join Alluxio’s Sr. Product Mgr., Adit Madan, to learn:
- Key challenges with architecting a successful heterogeneous data platform
- How data orchestration can overcome data access challenges in a distributed, heterogeneous environment
- How to identify ways to use Alluxio to meet the needs of your own data environment and workload requirements
Large Scale Analytics Acceleration
Model Training Acceleration
Hybrid Multi-Cloud
Data Platform Modernization
Data Migration
video
Industrial Bank’s Alluxio Deployment
ALLUXIO DAY IX 2022
January 21, 2022
Video: Presentation Slides: Industrial Bank's Alluxio Deployment from Alluxio, Inc.
Large Scale Analytics Acceleration
video
Vipshop Offline Data Cache Acceleration System – Alluxio Integration
ALLUXIO DAY IX 2022
January 21, 2022
Large Scale Analytics Acceleration
video
The Evolution of an Open Data Platform with Alluxio
ALLUXIO DAY IX 2022
January 21, 2022
Data Platform Modernization
Hybrid Multi-Cloud
Large Scale Analytics Acceleration
video
Alluxio + Spark: Accelerating Auto Data Tagging in WeRide
ALLUXIO DAY VIII 2021
December 14, 2021
Feifei Cai & Hao Zhu from WeRide provide an overview of Alluxio + Spark use case, which has been deployed and running in production to accelerate auto data tagging in the autonomous driving development.
Large Scale Analytics Acceleration