Data Platform Modernization Engineer - Junior (Databricks)

NTT DATA

3–6 yrs Bengaluru Full Time Hybrid (office + remote)
NTT DATA logo
Posted : yesterday
Actively hiring

Job description

Join NTT DATA as a Junior Data Platform Modernization Engineer specializing in Databricks. This role is crucial for migrating existing ETL workloads to the Databricks Lakehouse, leveraging modern data technologies.

We are seeking innovative individuals passionate about growing with us in an inclusive and forward-thinking organization. This position offers an opportunity to work on impactful data initiatives.

This is a full-time hybrid role based in Bangalore, India.

Responsibilities

Analyze and understand current Informatica workflows and AWS Glue jobs, including source-to-target mappings and transformation logic.

Develop Databricks notebooks and data pipelines for migrating ETL workloads, utilizing Python, PySpark, Spark SQL, and Databricks SQL.

Implement data processing layers (Bronze, Silver, Gold) following project architecture and coding standards.

Build robust batch and incremental processing pipelines with Delta Lake and establish watermark/control mechanisms.

Develop and maintain Databricks Jobs/Workflows and Lakeflow Declarative Pipelines, ensuring code reusability for logging, auditing, and data quality.

Convert Informatica transformations and business rules into Databricks equivalents, preserving business logic.

Support the migration of AWS Glue jobs and database processing into Databricks.

Conduct unit testing, data reconciliation, and troubleshoot pipeline execution issues.

Participate in integration testing, UAT, and production validation, collaborating with senior developers on technical challenges.

Adhere to established coding, performance, and security standards, actively contributing to code reviews and documentation.

Qualifications

This role requires 3-6 years of experience in data engineering, ETL development, or application/data platform development.

Proficiency with Databricks and Apache Spark is essential, along with strong development skills in Python/PySpark and SQL.

Experience with Databricks notebooks, Spark SQL/Databricks SQL, Delta Lake, and Databricks Jobs/Workflows is mandatory.

You should have experience developing batch or incremental ETL/ELT pipelines and familiarity with ETL technologies like Informatica PowerCenter, AWS Glue, or SSIS.

Understanding of relational databases, SQL-based processing, data transformation, source-to-target mapping, and unit testing is expected.

A Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent experience is required.

Excellent analytical, problem-solving, and communication skills are necessary.

Essential Skills

DatabricksApache SparkPythonPySparkSQLSpark SQLDatabricks SQLDelta LakeDatabricks Jobs/WorkflowsETL/ELTInformatica PowerCenterAWS GlueSSISRelational DatabasesData TransformationSource-to-Target MappingUnit TestingData ValidationTroubleshootingBronze/Silver/Gold Architecture

Good to Have

Databricks CertificationLakeflow Declarative PipelinesDelta Live TablesUnity CatalogDatabricks Asset BundlesCI/CDAWS S3AWS RedshiftAuto LoaderCDCSCD Type 1 / Type 2Data Quality FrameworksAudit FrameworksError Handling FrameworksAWS LambdaSpark Performance OptimizationDelta Performance OptimizationCloud Data ModernizationCloud Data Migration

Highlights

  • Actively hiring

More Details

RoleData Platform Modernization Engineer - Junior (Databricks)
Employment TypeFull Time, Hybrid (office + remote)

About the Company

NTT DATA logo

NTT DATA

IT Consulting

Data Platform Modernization Engineer - Junior (Databricks) at NTT DATA | SkillMX | SkillMX