Generates Terraform for data lake infrastructure from an infrastructure manifest. Produces reusable modules, per-environment application stacks, tfvars, and backend configs for S3 buckets, Glue jobs, IAM, KMS, and supporting services. Use when provisioning or…
aws-samples/sample-datalake-build-agent
SkillsMP has collected 7 skills from aws-samples/sample-datalake-build-agent. Open a skill to review its source and details.
- Latest recorded source activity
- SkillsMP catalog refreshed
- skills collected
- 7
- GitHub stars
- 0
- GitHub forks
- 0
Skills in this repository
Showing 7 of 7 collected skills.
Reference for the Amazon SageMaker Lakehouse architecture — unified data access across S3 data lakes and Redshift warehouses using Apache Iceberg, AWS Glue Data Catalog, and Lake Formation governance. Use when designing pipelines that target the lakehouse,…
Guides agents through data lake architecture design. Use when defining raw, refined, curated, or publish layers; storage organization; retention; and operational boundaries for a data lake.
Guides agents through lakehouse table design and open table format decisions. Use when designing or changing Iceberg, Delta, Hudi, partitioning, schema evolution, compaction, or batch and streaming interoperability.
Guides agents through domain-oriented data product and data mesh design. Use when organizing ownership, domain boundaries, federated governance, and shared platform responsibilities across multiple teams.
Guides agents through data quality design using AWS Glue Data Quality and DQDL. Use when defining quality rules, thresholds, anomaly detection, quarantine patterns, or integrating DQ checks into Glue ETL pipelines and the Data Catalog.
Guides agents through AWS data catalog and lake governance workflows. Use when designing or reviewing Glue Data Catalog, Lake Formation permissions, governed sharing, metadata quality, and access boundaries for S3, Athena, Redshift, EMR, or Glue pipelines.