AWS lakehouse twin
The closing iteration of my serverless lakehouse: an AWS twin of the same platform, built from the same data and the same models. The Yandex Cloud original keeps running beside it. S3, Spark on AWS Glue, Iceberg tables, the Glue Data Catalog, Athena, Secrets Manager, Budgets and Cost Anomaly Detection, all in Terraform. I learned AWS by building this. The story is what changed moving from DuckDB and pointer files to Glue and Iceberg.