alluxio day Archives | Page 5 of 7

RaptorX: Building a 10X Faster Presto with hierarchical cache

June 24, 2021

RaptorX is an internal project name aiming to boost query latency significantly beyond what vanilla Presto is capable of. For this session, we introduce the hierarchical cache work including Alluxio data cache, fragment result cache, etc. Cache is the key building block for RaptorX.

Tags: alluxio day, disaggregated storage, facebook, presto, raptorx

Accelerating analytics workloads with Alluxio data orchestration and Intel® Optane™ persistent memory

June 24, 2021

Today’s analytics workloads demand real-time access to expansive amounts of data. This session demonstrates how Alluxio’s data orchestration platform, running on Intel Optane persistent memory, accelerates access to this data and uncovers its valuable business insights faster.

Tags: alluxio day, analytics, data orchestration, intel, optane persistent memory

Alluxio for Machine Learning Workloads

June 24, 2021

Driven by strong interests from our open-source community, the core team of Alluxio started to re-design an efficient and transparent way for users to leverage data orchestration through the POSIX interface.

Tags: alluxio day, data orchestration, fuse, machine learning, POSIX

Alluxio Day 4

Community Virtual Event * June 24, 2021

Join us for our 4th Alluxio Day community virtual event featuring speakers from Facebook, TikTok,
Tencent, and Intel.

Advancing GPU Analytics with RAPIDS Accelerator for Spark and Alluxio

April 27, 2021

RAPIDS is a set of open source libraries enabling GPU aware scheduling and memory representation for analytics and AI. Spark 3.0 uses RAPIDS for GPU computing to accelerate various jobs including SQL and DataFrame. With compute acceleration from massive parallelism on GPUs, there is a need for accelerating data access and this is what Alluxio enables for compute in any cloud. In this talk, you will learn how to use Alluxio and Spark with RAPIDS Accelerator on NVIDIA GPUs without any application changes.

Tags: alluxio day, analytics, gpu, NVIDIA, RAPIDS, spark

Alluxio Data Orchestration for Machine Learning

April 27, 2021

Alluxio’s capabilities as a Data Orchestration framework have encouraged users to onboard more of their data-driven applications to an Alluxio powered data access layer. Driven by strong interests from our open-source community, the core team of Alluxio started to re-design an efficient and transparent way for users to leverage data orchestration through the POSIX interface.

Tags: alluxio day, alluxio engineering, data orchestration, fuse, machine learning, POSIX

Alluxio-FUSE as a data access layer for Dask

April 27, 2021

At Aspect Analytics we intend to use Dask, a distributed computation library for Python, to deal with MSI data stored as large tensors. In this talk we explore using Alluxio and Alluxio FUSE as a data consolidation and caching layer for some of our bioinformatics workflows.

Tags: alluxio day, aspect analytics, caching, dask, fuse

Building a high-performance data lake analytics engine at Alibaba Cloud with Presto+Alluxio

April 27, 2021

Data Lake Analytics(DLA) is a large scale serverless data federation service on Alibaba Cloud. One of its serverless analytics engine is based on Presto. The DLA Presto engine supports a variety of data sources and is widely used in different application scenarios in the cloud. In this session, we will talk about the system architecture of DLA Presto engine, as well as the challenges and solutions. In particular, we will introduce the use of alluxio local cache to solve performance issues on OSS data sources caused by access delay and OSS bandwidth limitation. We will discuss the principle of alluxio local cache and some improvements we have made.

Tags: alibaba, alluxio day, data lake analytics, local cache, performance, presto

Tag: alluxio day