Senior Associate | Site Reliability Engineer (SRE) | Bengaluru | Engineering as a Service/ Operate (Bengaluru, IN)
Deloitte
Deloitte
Join our pioneering Infrastructure Reliability Engineering (IRE) team, dedicated to fortifying the stability, performance, and scalability of our company's infrastructure. We are a forward-thinking group of engineers building advanced tools for infrastructure verification, automating issue resolution, and developing self-healing systems. Our contributions are fundamental to delivering an exceptional and dependable customer experience.
As a Senior Site Reliability Engineer, you will play a pivotal role in shaping and maintaining the software and tools essential for testing and validating our infrastructure's performance and reliability. You will collaborate closely with fellow engineers to proactively identify potential issues, engineer automated testing frameworks, and implement robust solutions that guarantee system resilience. This role offers the opportunity to take ownership of your solutions and partner with cross-functional teams in a dynamic environment to ensure successful deployments and continuous service improvement through automation.
As a Senior Site Reliability Engineer, your responsibilities will encompass designing and developing innovative software tools for infrastructure testing and verification. You will be instrumental in creating and maintaining automated testing frameworks, and will collaborate with infrastructure and development teams to pinpoint and resolve reliability challenges.
You will actively participate in code reviews, upholding stringent team standards for software quality. Furthermore, you'll be tasked with continuously enhancing service resilience via automation, identifying and rectifying performance bottlenecks, and conducting comprehensive capacity planning. Your role extends to participating in incident reviews, assisting with root cause analysis, and deploying SRE solutions across a global, multi-cloud hybrid environment (AWS, GCP, and On-prem) to ensure unparalleled uptime and Quality of Service (QoS) for our internal users.
We are seeking candidates with a Bachelor’s degree in computer science or a related discipline, or equivalent practical experience. A minimum of 8 years of extensive software development experience is required.
Essential technical proficiencies include expertise in at least one programming language such as Python, Go, or Java. You should possess strong command over Kubernetes administration, modern CI/CD methodologies, and Infrastructure as Code (IaC) principles. A deep understanding of Linux operating systems and fundamental TCP/IP concepts is crucial, alongside experience with software testing methodologies and tools.
Proficiency in utilizing monitoring, metrics gathering, Application Performance Monitoring (APM), container management, and log collection tools is expected. We value creative problem-solvers with exceptional debugging capabilities and strong documentation skills. Preferred qualifications include experience with performance and chaos engineering, CI/CD pipelines, a solid grasp of complex system architectures, and a genuine passion for automation, scalability, and building highly reliable systems from the ground up. Familiarity with cloud infrastructure providers like AWS, Azure, and GCP is also highly beneficial.
Deloitte
IT Consulting