Skip to main content

aws-lakehouse

스타11
포크4
업데이트2026년 6월 28일 15:44

S3-based lakehouse on Apache Iceberg — Amazon S3 Tables (managed Iceberg table buckets with auto compaction/snapshot maintenance), self-managed Iceberg on plain S3 (Glue/REST catalogs), querying with Athena (engine v3, MERGE/UPDATE/DELETE), and processing with Spark on EMR/Glue/EMR Serverless. Covers catalog choice, Iceberg V3 (deletion vectors, row lineage) and its Athena incompatibility, Spark engine/language speed (Scala vs PySpark vs Kotlin), native accelerators (Comet/Gluten/Photon), and S3 Tables cost pitfalls. Use when building or querying a data lake on S3, choosing between S3 Tables and self-managed Iceberg, picking a query engine, or tuning Spark performance. Triggers: S3 Tables, table bucket, s3tables, Apache Iceberg, Iceberg REST catalog, Glue Data Catalog, Athena Iceberg, MERGE INTO, time travel, deletion vectors, Iceberg V3, EMR Spark, Glue ETL, PySpark slow, Spark accelerator, Comet, Gluten, lakehouse, partition projection.

설치

Codex 또는 Claude로 설치 이 Prompt를 복사해 Codex, Claude 또는 다른 어시스턴트에 붙여 넣으면 Skill 페이지를 검토하고 설치를 진행할 수 있습니다.

SKILL.md
readonly