The query succeeded. Which table state did it read?
Why two successful queries can count different rows when they resolve different Delta or Iceberg table states. Two queries return different counts f…
Tech news from the best sources
Why two successful queries can count different rows when they resolve different Delta or Iceberg table states. Two queries return different counts f…
TL;DR — DuckDB, DataFusion, Polars, and Postgres each compute and store table statistics their own way, so a histogram built in your ELT pipeline is…
Every time this comes up, someone credits the same setup: Athena on Iceberg is where "the code is the spec" — where you open Git and read the whole…
Original Japanese article : Iceberg REST Catalogを直接叩いて、Glue Data CatalogとS3 Tablesの違いを理解する Introduction I'm Aki, an AWS Community Builder ( @jitepen…
Quick Recap: What We Built in Part 1 In Part 1 , we built a metadata catalog on Apache Iceberg (S3 Tables) that makes unstructured files on FSx for…
What Works Now vs What Requires Validation This article separates verified AWS-native capabilities from cross-platform paths that still require vali…
Original Japanese article : データの主導権から考えるAWSとSnowflakeのレイクハウスアーキテクチャ Introduction I'm Aki, an AWS Community Builder ( @jitepengin ). When designing a…
Introduction Apache Iceberg is the table format that turns a pile of Parquet files in object storage into something that behaves like a warehouse ta…
Why this project I built this repo because I didn't have one of this kind yet and, having worked on data ingestion with Glue for a while, I wanted t…
The future of Kafka is diskless topics + native Apache Iceberg / Delta Lake integration Intro Ursa is a relatively new engine that's being fitted in…