#
DIFF.BLOG
New
Following
Discover
Jobs
More
Top Writers
Suggest a blog
Upvotes plugin
Report bug
Contact
About
Sign up
The home for great developer writing.
We surface the best developer writing from thousands of independent blogs, updated daily.
Join Diff.blog
TOPICS
Big Data in Polars: Reading and Writing Partitioned Parquet Files
·
Ahmed Nabil
·
June 27, 2026, 2:20 p.m.
big data
Data Science
data engineering
partitioning
big data
polars
parquet files
data-management
Summary
The blog post discusses managing large datasets using Polars, a fast DataFrame library in Python, focusing on the importance of partitioning Parquet files when dealing with substantial amounts of data (such as 1TB).
Read full post on pythonprohub.com →
MORE POSTS LIKE THIS
Polars Structs: How to Pack and Unpack Multiple Columns
Ahmed Nabil ·
Jul 11, 2026
Data Science
data-structures
Visualizing Millions of Rows: Polars + Datashader (Big Data Plotting)
Ahmed Nabil ·
Jul 6, 2026
big data
Data Science
Big Ticks and Small Ticks in Equity Microstructure
Dean ·
Aug 23, 2026
Python
equity
condense-json 1.0
simonw ·
Aug 3, 2026
Python
json
Scaling Grab's Data Lake: Our journey to Apache Iceberg adoption
Grab ·
Jul 10, 2026
engineering
spark
The revolt of the reader
The Observation Deck ·
Sep 5, 2026
LLM-authored content
Reader trust
Discover more posts →
AUTHOR
Sponsored
Zulip
Organized team chat for people who take work seriously. Topic-based threading keeps conversations focused.
Try Zulip
Become a sponsor →
RECENT POSTS FROM THE AUTHOR
Choose how you want to continue.
Continue with GitHub
Continue with Google