Job Description

Develop 20 data table ingestion pipelines end-to-end (source extraction → S3 landing → Glue ETL → Lake Formation curated zone)

Implement ETL transformations per data mapping specifications

Build data quality validation rules (Great Expectations / custom Glue checks) - completeness, schema conformance, referential integrity

Configure error handling and dead-letter patterns for failed ingestion records

Register all datasets in AWS Glue Data Catalog with standardised metadata tags (owner, classification, freshness SLA)

Write and maintain IaC (CDK/Terraform) for pipeline resources - Glue jobs, crawlers, S3 buckets, IAM roles

Execute unit testing (per-transform logic) and integration testing (end-to-end flow with sample data)

Support UAT with Agency A data owners - validate output tables match expected schema and row counts

Document pipeline configurations, runbooks, and data flow diagrams for handover

Part...

Ready to Apply?

Take the next step in your AI career. Submit your application to onebyzero pte. ltd. today.

Submit Application