Topics
Follow your own topics →
DIFF.BLOG
New Following Discover Jobs
More
Top Writers Suggest a blog Upvotes plugin
Report bug Contact About
Sign up
Menu
New Following Discover Jobs Top Writers
More
Suggest a blog Upvotes plugin Report bug Contact About
Sign up
The home for great developer writing.
We surface the best developer writing from thousands of independent blogs, updated daily.
Join now → Learn more
TOPICS

Big Data in Polars: Reading and Writing Partitioned Parquet Files

1 · Ahmed Nabil · June 27, 2026, 2:20 p.m.
Data Science 2026 big data data engineering big data polars parquet files data-management
Summary
The blog post discusses managing large datasets using Polars, a fast DataFrame library in Python, focusing on the importance of partitioning Parquet files when dealing with substantial amounts of data (such as 1TB).
Read full post on pythonprohub.com →
MORE POSTS LIKE THIS
Polars Structs: How to Pack and Unpack Multiple Columns
Ahmed Nabil · Jul 11, 2026
Data Science 2026
Visualizing Millions of Rows: Polars + Datashader (Big Data Plotting)
Ahmed Nabil · Jul 6, 2026
Data Science 2026
Scaling Grab's Data Lake: Our journey to Apache Iceberg adoption
Grab · Jul 10, 2026
Data database
We have proof automation now
Adam Langley · Jul 26, 2026
dependently-typed languages proof automation
Thoughts about the Leiden Declaration
Gowers Wordpress · Jul 26, 2026
AI and maths AI
PostgreSQL's MVCC is bad. So is everyone else's.
Radim Marek · Jul 27, 2026
postgresql database-design
Discover more posts →
AUTHOR
RECENT POSTS FROM THE AUTHOR
Choose how you want to continue.
Continue with GitHub Continue with Google