Lead Data Migration Engineer
NTT
In this Lead Data Migration Engineer role, you will design, build, and validate end-to-end data migration pipelines from legacy warehouses to AWS-based Data Lakehouse architectures. You will work closely with architects, engineers, and analysts to deliver secure, scalable migration solutions while ensuring data quality and performance through robust testing. You’ll leverage AWS Glue, Apache Iceberg, Python/PySpark, SQL, and YAML configurations to enable efficient modernization. This client-facing position emphasizes collaboration, delivery excellence, and driving data modernization initiatives.
Pay / Benefits- flexible work options
- learning and development opportunities
- tailored benefits
- employer commitment to diversity and inclusion
- support for wellbeing and career growth
- Design, build and optimise data migration pipelines from legacy warehouses to AWS-based Data Lakehouse platforms
- Develop and maintain ETL/ELT pipelines using AWS Glue, Python/PySpark, SQL, YAML-driven configurations
- Implement bulk data migrations followed by incremental/delta loads
- Support pipeline repointing and replatforming to target cloud architectures
- Test end-to-end ETL solutions on AWS and validate migrations to new Data Lakehouse architectures
- Create automated testing and validation frameworks for data integrity and reconciliation
- Collaborate with Solution Architects, Data Engineers, Analysts, and QA to promote best practices and reusable components
- Contribute to migration accelerators and reusable frameworks
- Proven experience in data engineering and data migration delivery in cloud environments
- Strong focus on data pipeline testing, validation, and quality assurance
- Experience across the full data lifecycle with emphasis on migration and transformation
- Strong analytical, problem-solving and communication skills
- Client-facing and delivery-focused experience
- Ability to mentor junior engineers and contribute to team delivery
- Hands-on experience with AWS Glue, Python/PySpark, SQL, YAML configuration
- collaboration
- communication
- mentoring
- AWS Glue
- Python
- PySpark
Reference: WJ-747_30162077