AWS Glue 6.0: 30% Cheaper, Iceberg v3, Spark 4.1
AWS Glue 6.0 is GA with 30% lower pricing, full Apache Iceberg v3, and a Spark 4.1 runtime. Here is what changed and how to migrate your ETL jobs safely.
Tag archive
AWS Glue 6.0 is GA with 30% lower pricing, full Apache Iceberg v3, and a Spark 4.1 runtime. Here is what changed and how to migrate your ETL jobs safely.
Why this project I built this repo because I didn't have one of this kind yet and,...
\n In 2025, enterprises wasted $4.2B on underoptimized ETL pipelines, with 68% of teams picking...
While studying the AWS Data Associate certification guide and designing a data pipeline, you will be...
Securing AWS Glue: A Guide to Identifying and Fixing Python Package Vulnerabilities ...
Analyzing data directly from Amazon DynamoDB can be tricky since it doesn’t come with built-in...
Access Glue Iceberg tables via the Iceberg Rest Api AWS Released silently Iceberg REST-API...
Introduction In this blog, we will provide a brief introduction to data governance and...
This post covers a use case of accessing data held in a Snowflake database hosted in GCP within an...
With the evolution of technology, we can see data is growing exponentially, new sources of data,...

What is hot off the press 🪄 from AWS re: Invent 2023? I am still buzzing from attending...

In this post, we'll look at how to use the new automatic compaction feature in AWS Glue and how it can help optimize your Iceberg tables. Use my handy helper script to create an Iceberg table, load sample data into it, and test this new feature hands-on.