DIFF.BLOG
New Following Discover Jobs
More
Top Writers Suggest a blog Upvotes plugin
Report bug Contact About
Sign up
Topics
Follow your own topics →
Menu
New Following Discover Jobs Top Writers
More
Suggest a blog Upvotes plugin Report bug Contact About
Sign up
The home for great developer writing.
We surface the best developer writing from thousands of independent blogs, updated daily.
Join Diff.blog
TOPICS

Shuffle V2 in Ray Data: Faster, Fault-Tolerant Joins and Aggregations

· Anyscale · Aug. 25, 2026, 7:10 p.m.
Ray Data data-processing JOINs Aggregations
Summary
This post discusses Shuffle V2 in Ray Data, highlighting improved performance for joins and aggregations, aiming at developers interested in efficient data processing.
Read full post on anyscale.com →
MORE POSTS LIKE THIS
GPU-Native Operators in Ray Data
Anyscale · Aug 25, 2026
GPU-Native Operators Ray Data
A Tale of Two Flink Autoscalers
Netflix, Inc. · Aug 21, 2026
Stream Processing Autoscaling
Polars for Machine Learning: Zero-Copy to PyTorch and XGBoost
Ahmed Nabil · Jul 15, 2026
Machine Learning Data Science
Processing Data Larger Than RAM: The Polars Streaming Engine (sink_parquet)
Ahmed Nabil · Jun 26, 2026
big data Data Science
Kafka Share Groups - Pathological fetch waits with record_limit
Jack Vanlightly · Jun 24, 2026
kafka kafka
Next Generation DB Ingestion at Pinterest
Pinterest · Feb 5, 2026
engineering spark
Discover more posts →
AUTHOR
RECENT POSTS FROM THE AUTHOR
Choose how you want to continue.
Continue with GitHub Continue with Google