Data Migration Engineer
NTT DATA
In this role you will design, build and validate data migration pipelines to move data from legacy systems to modern AWS-based data lakehouse platforms. You will work with cross-functional teams to ensure data quality, integrity, and performance while leveraging AWS Glue, Python/PySpark, and Iceberg. You’ll focus on delivering scalable migration solutions and robust validation in a collaborative, delivery-focused environment. This is an opportunity to advance cloud data engineering practices within a global consultancy.
Pay / Benefits- tailored benefits
- continuous learning and development
- flexible work options
- Support delivery within data migration programmes and key workstreams
- Collaborate with architects, engineers, and stakeholders to implement migration solutions
- Plan and execute data migration tasks and deliverables
- Build and maintain data migration pipelines from legacy warehouses to AWS platforms
- Develop ETL/ELT pipelines using AWS Glue, Python/PySpark, SQL, YAML configurations
- Support bulk and incremental data migrations and cloud migration repointing
- Test ETL/ELT pipelines on AWS services (Glue, Iceberg) and validate migrations
- Validate data quality, completeness and transformation outputs in Data Lakehouse architectures
- Contribute to improving pipeline performance and reliability in AWS Data Platforms
- Apply data transformation rules and prepare data for target-state models
- Collaborate with Solution Architects, Data Engineers, Analysts and QA teams
- Follow engineering standards and contribute to documentation and reusable components
- Support data quality, governance and security, including GDPR and public sector standards
- Experience in data engineering or data migration delivery
- Strong focus on testing, validation, and data quality assurance
- Ability to work across data pipelines and transformation workflows
- Good analytical and problem-solving skills
- Effective communication and teamwork skills
- Willingness to learn and develop in data migration and cloud technologies
- Hands-on experience with AWS (especially AWS Glue)
- Python / PySpark programming
- SQL querying and validation
- YAML configuration (desirable)
- Familiarity with Data lake / Lakehouse concepts (Apache Iceberg)
- Experience with Spark and distributed processing frameworks
- Understanding of ETL vs ELT approaches
- Exposure to version control and CI/CD tools (desirable)
- collaborative
- delivery-focused
- problem-solving
- AWS Glue
- Python
- PySpark
Reference: WJ-747_30231763