Data Platform Modernization Engineer - Junior (Databricks)
NTT DATA
NTT DATA
Join NTT DATA as a Junior Data Platform Modernization Engineer specializing in Databricks. This role is instrumental in an enterprise data platform modernization initiative, focusing on transitioning legacy workloads to a scalable, cloud-native environment. You will be part of an innovative team dedicated to accelerating client success and driving positive societal impact through advanced technology solutions.
This position offers an opportunity to work with cutting-edge technologies and contribute to significant digital transformation projects. As a forward-thinking organization, NTT DATA fosters an inclusive and adaptable culture where your growth and contributions are highly valued. We are committed to leveraging AI and digital infrastructure to help organizations confidently navigate the future.
Key responsibilities include analyzing existing Informatica and AWS Glue ETL processes to understand data flows and transformations.
You will develop Databricks notebooks and data pipelines to migrate these workloads, implementing Bronze, Silver, and Gold layers. Developing transformations using Python, PySpark, and SQL, alongside building batch and incremental pipelines with Delta Lake, will be crucial. The role involves converting Informatica logic, migrating AWS Glue jobs, and supporting various testing phases, including unit, integration, and user acceptance testing.
Further duties encompass troubleshooting data and pipeline issues, adhering to coding standards, participating in code reviews, and preparing technical documentation. You will also assist with deployment, monitoring, and production stabilization activities, collaborating with senior team members to resolve technical challenges.
We are looking for a Data Engineer with 3-6 years of experience in data engineering or ETL development, possessing strong hands-on skills in Databricks and Apache Spark. Proficiency in Python/PySpark and SQL is essential.
Essential experience includes working with Databricks notebooks, Spark SQL/Databricks SQL, Delta Lake, and Databricks Jobs/Workflows. Familiarity with batch or incremental ETL/ELT pipeline development and at least one ETL technology like Informatica PowerCenter or AWS Glue is required. A solid understanding of data transformation, source-to-target mapping, unit testing, and data validation is also a must.
Candidates should possess a Bachelor's degree in Computer Science, IT, Engineering, or equivalent experience. Excellent analytical, problem-solving, and communication skills are necessary. Experience with layered data architectures (Bronze/Silver/Gold) and the ability to work within established frameworks are key to success in this role.
NTT DATA
IT Consulting