Join a dynamic Site Reliability Engineering (SRE) team where software and systems engineering converge to build and maintain robust, large-scale, and fault-tolerant systems. This role is instrumental in ensuring the reliability and optimal performance of critical Google Cloud services.
Embrace the unique challenges of scale inherent in Google Cloud, leveraging your expertise in coding, algorithms, and system design. Our SRE culture thrives on intellectual curiosity, collaborative problem-solving, and a commitment to innovation in a supportive, blame-free environment that fosters growth and impactful project ownership.
Design, develop, and implement solutions to enhance the reliability of essential enterprise applications.
Analyze and optimize system performance, capacity, and overall reliability.
Automate infrastructure tasks to improve efficiency and reduce manual intervention.
Troubleshoot and resolve complex issues in large-scale distributed systems.
A Bachelor’s degree in Computer Science, a related technical field, or equivalent practical experience is required.
Demonstrate at least 1 year of experience in software development using one or more programming languages.
Possess 1 year of experience working with data structures and algorithms.
Expertise in Unix/Linux environments, IP networking, and diagnosing performance and application-related issues.
Proven ability to solve complex problems and analyze intricate enterprise systems.
Proficiency in navigating and managing enterprise software and workloads is essential.
IT Consulting