Project Lead-App Development
Birlasoft
Birlasoft
Serve as the technical authority for designing, implementing, and operating enterprise monitoring and observability solutions across hybrid IT environments. This role involves hands-on work with platforms like HPE OpsRamp and SolarWinds to deliver comprehensive monitoring, including onboarding, discovery, service mapping, and alert correlation. The goal is to enhance reliability, reduce alert noise, and accelerate Mean Time To Detect (MTTD) and Mean Time To Recover (MTTR) for clients.
Additionally, you will mentor junior technical teams and act as the central point of expertise for all monitoring delivery initiatives, ensuring seamless operations and continuous improvement of our observability capabilities.
Design, deploy, and integrate observability solutions across on-prem, cloud, and edge infrastructures, utilizing tools such as HPE OpsRamp and SolarWinds. Analyze client environments to create tailored monitoring strategies and develop essential documentation, including architecture blueprints and operational runbooks.
Lead critical processes like onboarding, discovery, and service mapping for new and existing engagements. Manage the daily operations of observability platforms, ensuring high availability, performance, and data accuracy. Handle event management by correlating alerts, reducing noise, and routing them effectively.
Troubleshoot complex incidents escalated from L1/L2 teams, conducting root-cause analysis and translating telemetry into actionable insights for improved reliability. Establish and refine standards, thresholds, and SLA-based response models, while maintaining data integrity. Integrate monitoring tools with ITSM platforms like ServiceNow and cloud environments such as AWS, Azure, and GCP.
A minimum of 6 to 8 years of direct experience in deploying, implementing, and supporting IT monitoring tools is essential, along with strong report-building capabilities. You must possess deep hands-on expertise with HPE OpsRamp and SolarWinds, with familiarity in similar tools being a significant advantage. A working understanding of broader observability platforms like Dynatrace, AppDynamics, Datadog, Prometheus, Grafana, ELK, Splunk, and OpenTelemetry concepts is required.
Proven experience in Managed Services or NOC delivery, coupled with ITSM-integrated event management, is crucial. A solid grasp of server, network, virtualization, storage, and hybrid cloud environments is necessary. Hands-on experience integrating with ITSM tools, particularly ServiceNow, and cloud platforms (AWS, Azure, GCP) is essential.
Proficiency in scripting languages such as Python, Bash, or PowerShell for automation and integration is a must. Familiarity with containerization technologies like Docker and Kubernetes is beneficial. A sound understanding of ITIL processes, including Incident, Problem, Change, and Configuration Management, is also a key requirement for this role.
Birlasoft
IT Consulting