Job Summary
We are seeking a Tools Engineer to drive the implementation and support for monitoring and observability tools. In this role, you will provide technical hands on and manage the lifecycle of monitoring platforms including Grafana, Django, and Prometheus. You will partner with cross-functional teams to implement scalable observability solutions, support dashboards and establish engineering standards, and accelerate automation adoption to enhance operational resilience.
Key Responsibilities
Skills & Experience • Education: o Bachelor’s or master’s in computer science, Engineering, or related field (or equivalent experience). o Preferred tools related OEM certification • Technical Skills: o Strong experience in enterprise monitoring, observability, and automation platform engineering/leadership. o Hands-on expertise with Ansible Automation Platform (AAP), Grafana, Prometheus, and Django. o Strong understanding of alerting strategy, operational monitoring, event handling, remediation orchestration, and platform governance. o Experience in enterprise-scale deployment, administration, and optimization of monitoring and automation tools. o Experience working across infrastructure, servers, endpoints, hybrid/cloud, and operational support teams. o Strong scripting/automation skills in one or more of: Python, PowerShell, Bash, YAML/Ansible.