Data Engineer - AWS Data Lakehouse (Public Sector)
At Xtremax, we are looking for a Data Engineer to design and build a scalable data lakehouse platform on AWS. You will work across data architecture, pipeline engineering, data quality, governance, and cloud deployment to deliver production-ready solutions that form the foundation of a modern data platform.
This role is suited to an experienced data engineer who enjoys solving complex data engineering challenges at scale. You will work hands-on with AWS services such as Glue, Step Functions, Lambda, and S3, while applying modern lakehouse technologies including Apache Iceberg or S3 Tables. You will also contribute to automated data quality frameworks, CI/CD, infrastructure as code, and production deployments.
You will collaborate with cross-functional teams to translate technical designs into reliable solutions, communicate architecture clearly to non-technical stakeholders, and produce documentation that enables a smooth transition to Day 2 operations.
Responsibilities
- Design and implement scalable ETL/ELT pipelines using AWS Glue, Step Functions, Lambda, S3, and related AWS services
- Architect and build data lakehouse solutions using Apache Iceberg or S3 Tables, including schema evolution, partition evolution, and ACID transactions
- Optimise data pipelines and lakehouse workloads for performance, cost efficiency, scalability, and reliability
- Design and implement automated data quality validation