#
DIFF.BLOG
New
Following
Discover
Jobs
More
Top Writers
Suggest a blog
Upvotes plugin
Report bug
Contact
About
Sign up
The home for great developer writing.
We surface the best developer writing from thousands of independent blogs, updated daily.
Join Diff.blog
TOPICS
Shuffle V2 in Ray Data: Faster, Fault-Tolerant Joins and Aggregations
·
Anyscale
·
Aug. 25, 2026, 7:10 p.m.
Ray Data
data-processing
JOINs
Aggregations
Summary
This post discusses Shuffle V2 in Ray Data, highlighting improved performance for joins and aggregations, aiming at developers interested in efficient data processing.
Read full post on anyscale.com →
MORE POSTS LIKE THIS
GPU-Native Operators in Ray Data
Anyscale ·
Aug 25, 2026
GPU-Native Operators
Ray Data
A Tale of Two Flink Autoscalers
Netflix, Inc. ·
Aug 21, 2026
Stream Processing
Autoscaling
Polars for Machine Learning: Zero-Copy to PyTorch and XGBoost
Ahmed Nabil ·
Jul 15, 2026
Machine Learning
Data Science
Processing Data Larger Than RAM: The Polars Streaming Engine (sink_parquet)
Ahmed Nabil ·
Jun 26, 2026
big data
Data Science
Kafka Share Groups - Pathological fetch waits with record_limit
Jack Vanlightly ·
Jun 24, 2026
kafka
kafka
Next Generation DB Ingestion at Pinterest
Pinterest ·
Feb 5, 2026
engineering
spark
Discover more posts →
AUTHOR
RECENT POSTS FROM THE AUTHOR
Choose how you want to continue.
Continue with GitHub
Continue with Google