Position: Dynatrace Observability Lead
Location: Remote
Duration: Long term contract
Job Summary:
We are looking for an experienced Dynatrace Observability Lead with strong hands-on expertise in Dynatrace monitoring, observability engineering, automation, and AIOps. The candidate should have proven technical leadership skills, strong troubleshooting capabilities, and excellent client communication skills.
Key Responsibilities:
- Lead end-to-end observability implementation across applications, infrastructure, and cloud environments using Dynatrace.
- Configure and manage dashboards, alerts, metrics, logs, distributed tracing, synthetic monitoring, and RUM.
- Implement automation, self-healing, runbook automation, and automated incident triage.
- Leverage AIOps for anomaly detection, event correlation, predictive alerting, and root-cause analysis.
- Define and optimize SLIs, SLOs, alerting strategies, MTTD, and MTTR.
- Integrate observability tools with CI/CD pipelines, cloud platforms, and ServiceNow.
- Lead incident management, mentor teams, and drive observability best practices.
- Collaborate with clients and stakeholders to improve system reliability and operational efficiency.
Required Technical Skills:
Mandatory: Strong hands-on experience with Dynatrace monitoring, configuration, alerting, and troubleshooting.
Good to Have: Splunk, Grafana, and OpenTelemetry.
Cloud: AWS, Google Cloud Platform, and cloud-native architectures.
Automation: Python and Terraform.
SRE: AIOps, SLIs, SLOs, incident management, MTTD, and MTTR optimization.
Integration: CI/CD pipelines and ServiceNow.
Additional Requirements:
- 7+ years of experience in observability engineering, monitoring, SRE, or related domains.
- Strong technical leadership, problem-solving, and client communication skills.
- Experience driving observability and automation initiatives.
- Minimum 90% weekly usage of enterprise-approved AI tools, such as GitHub Copilot and Microsoft 365 Copilot, for coding, documentation, analysis, and productivity.