Ray Data
PulseAugur coverage of Ray Data — every cluster mentioning Ray Data across labs, papers, and developer communities, ranked by signal.
- 2026-08-25 product_launch Anyscale released new GPU-native operators for Ray Data, integrating cuDF and RapidsMPF. source
-
Anyscale revamps Ray Data shuffle engine for improved speed and stability
Anyscale has introduced Shuffle V2 for its Ray Data framework, a significant redesign of its shuffle engine. This new version addresses limitations in the previous Shuffle V1 by materializing shuffle intermediates in th…
-
Anyscale boosts Ray Data with GPU-native operators for AI workloads
Anyscale has enhanced its Ray Data engine with GPU-native operators, collaborating with NVIDIA to integrate cuDF and RapidsMPF. These updates allow data processing tasks to execute directly on GPUs, offering significant…
-
Anyscale boosts Ray performance for massive AI training clusters
Anyscale has significantly enhanced its Ray framework to better support large-scale AI workloads. Recent improvements address bottlenecks in driver performance and actor scheduling, leading to substantial speedups for b…
-
Anyscale cuts AI training data latency 20x with Alluxio cache
Anyscale has demonstrated a significant speedup in AI training data reads by integrating Alluxio, a distributed caching layer, with its Ray platform. By deploying Alluxio on NVMe SSDs colocated with Ray clusters, cross-…
-
Anyscale details Ray Data for scaling multimodal AI data pipelines
Anyscale's blog post details challenges in scaling multimodal AI data pipelines, where preprocessing often starves GPUs, leading to underutilization. The article explains that traditional staged batch execution, which i…
-
Anyscale launches persistent dashboards for Ray workload monitoring
Anyscale has launched new Cluster and Actor Dashboards for its Ray platform, providing fully persisted monitoring and debugging tools. These dashboards address limitations of the previous ephemeral data, enabling histor…