Azure Subcontractor

Birlasoft

Fresher Hyderabad Full Time Work from office
Birlasoft logo
Posted : 1 week ago
Actively hiring

Job description

Join our team as an experienced Databricks Developer, focusing on crafting and maintaining robust, scalable data pipelines and analytics solutions. You will leverage your deep understanding of Apache Spark, Python, SQL, and cloud data engineering technologies to drive enterprise data integration and analytics initiatives forward.

This role is integral to designing, developing, and optimizing data solutions within a collaborative environment. We seek a proactive individual adept at handling complex data challenges and contributing to best-in-class data engineering practices.

Responsibilities

Your primary duties will involve architecting and implementing ETL/ELT pipelines within Databricks, alongside developing high-performance data processing frameworks using Apache Spark (PySpark). You will build and refine both batch and streaming data pipelines, seamlessly integrating data from diverse sources including databases, APIs, and cloud storage.

Key tasks include performing data transformation, cleansing, and validation, as well as creating and optimizing Delta Lake tables for efficient data storage. You will also be responsible for developing reusable Databricks notebooks, workflows, and jobs, while actively monitoring and troubleshooting pipeline performance and failures. Collaboration with stakeholders to understand data requirements and adherence to coding standards, version control, and CI/CD practices are also essential.

Qualifications

A strong foundation in Databricks, Apache Spark/PySpark, Python, and SQL is mandatory for this position. Demonstrated experience in ETL/ELT development and working with Delta Lake is crucial, complemented by a solid understanding of data warehousing concepts. Familiarity with Git and version control is also required.

Preferred qualifications include experience with major cloud platforms such as Microsoft Azure (Azure Databricks, ADLS Gen2, Azure Data Factory), AWS (S3, Glue, EMR), or Google Cloud (BigQuery, Cloud Storage). Experience with streaming technologies like Spark Structured Streaming and orchestration tools such as Airflow is beneficial. An understanding of data governance and security principles is also a plus.

Essential Skills

DatabricksApache SparkPySparkPythonSQLETL/ELTDelta LakeData WarehousingGitVersion Control

Good to Have

Azure DatabricksADLS Gen2Azure Data FactoryAWS S3AWS GlueAWS EMRGoogle BigQueryGoogle Cloud StorageSpark Structured StreamingAirflowData GovernanceData SecurityMedallion ArchitectureUnity CatalogDatabricks Performance TuningDatabricks Cost OptimizationDevOpsCI/CD

Highlights

  • Actively hiring

More Details

RoleAzure Subcontractor
DepartmentData Engineering
Employment TypeFull Time, Work from office

About the Company

Birlasoft logo

Birlasoft

IT Consulting

Azure Subcontractor at Birlasoft | SkillMX | SkillMX