Data Platform Modernization Engineer - Junior (Databricks)

NTT DATA

3–6 yrs Bengaluru Full Time Hybrid (office + remote)
NTT DATA logo
Posted : yesterday
Actively hiring

Job description

Join NTT DATA as a Junior Data Platform Modernization Engineer specializing in Databricks. This role is instrumental in an enterprise data platform modernization initiative, focusing on transitioning legacy workloads to a scalable, cloud-native environment. You will be part of an innovative team dedicated to accelerating client success and driving positive societal impact through advanced technology solutions.

This position offers an opportunity to work with cutting-edge technologies and contribute to significant digital transformation projects. As a forward-thinking organization, NTT DATA fosters an inclusive and adaptable culture where your growth and contributions are highly valued. We are committed to leveraging AI and digital infrastructure to help organizations confidently navigate the future.

Responsibilities

Key responsibilities include analyzing existing Informatica and AWS Glue ETL processes to understand data flows and transformations.

You will develop Databricks notebooks and data pipelines to migrate these workloads, implementing Bronze, Silver, and Gold layers. Developing transformations using Python, PySpark, and SQL, alongside building batch and incremental pipelines with Delta Lake, will be crucial. The role involves converting Informatica logic, migrating AWS Glue jobs, and supporting various testing phases, including unit, integration, and user acceptance testing.

Further duties encompass troubleshooting data and pipeline issues, adhering to coding standards, participating in code reviews, and preparing technical documentation. You will also assist with deployment, monitoring, and production stabilization activities, collaborating with senior team members to resolve technical challenges.

Qualifications

We are looking for a Data Engineer with 3-6 years of experience in data engineering or ETL development, possessing strong hands-on skills in Databricks and Apache Spark. Proficiency in Python/PySpark and SQL is essential.

Essential experience includes working with Databricks notebooks, Spark SQL/Databricks SQL, Delta Lake, and Databricks Jobs/Workflows. Familiarity with batch or incremental ETL/ELT pipeline development and at least one ETL technology like Informatica PowerCenter or AWS Glue is required. A solid understanding of data transformation, source-to-target mapping, unit testing, and data validation is also a must.

Candidates should possess a Bachelor's degree in Computer Science, IT, Engineering, or equivalent experience. Excellent analytical, problem-solving, and communication skills are necessary. Experience with layered data architectures (Bronze/Silver/Gold) and the ability to work within established frameworks are key to success in this role.

Essential Skills

PySparkPythonSQLDatabricksETL DevelopmentData TransformationTestingTroubleshootingDatabricks NotebooksSpark SQLDatabricks SQLDelta LakeDatabricks Jobs/WorkflowsBatch ProcessingIncremental ProcessingInformatica PowerCenterAWS GlueRelational Databases

Good to Have

Databricks CertificationLakeflow Declarative PipelinesDelta Live TablesUnity CatalogDatabricks Asset BundlesCI/CDAWS S3AWS LambdaAWS RedshiftAuto LoaderStreaming IngestionCDCSCD Type 1SCD Type 2Data Quality FrameworksAudit FrameworksError Handling FrameworksSpark Performance OptimizationDelta Lake Performance OptimizationCloud Data ModernizationData Migration

Highlights

  • Actively hiring

More Details

RoleData Platform Modernization Engineer - Junior (Databricks)
DepartmentData Engineering
Employment TypeFull Time, Hybrid (office + remote)

About the Company

NTT DATA logo

NTT DATA

IT Consulting

Data Platform Modernization Engineer - Junior (Databricks) at NTT DATA | SkillMX | SkillMX