Principal Observability Engineer

Remote • Posted 2 hours ago • Updated 2 hours ago
Full Time
Remote
Fitment

Dice Job Match Score™

👤 Reviewing your profile...

Job Details

Skills

  • Systems Architecture
  • Decision-making
  • Roadmaps
  • Strategic Thinking
  • Optimization
  • Kubernetes
  • Microservices
  • Continuous Improvement
  • Dashboard
  • Operational Excellence
  • Mentorship
  • Technical Direction
  • Enterprise Architecture
  • Computer Science
  • Information Systems
  • SaaS
  • Instrumentation
  • Cloud Computing
  • Technical Drafting
  • Communication
  • Collaboration
  • Amazon Web Services
  • Microsoft Azure
  • Google Cloud Platform
  • Google Cloud
  • Reliability Engineering
  • Incident Management
  • Orchestration
  • Workflow
  • Dynatrace
  • Scripting
  • Programming Languages
  • Python
  • Bash
  • JavaScript

Summary

The Principal Observability Engineer is responsible for defining, leading, and advancing enterprise observability strategy, architecture, and implementation across applications, platforms, infrastructure, and operational services. This role requires deep hands-on expertise with Dynatrace SaaS and recent Dynatrace platform capabilities and innovations, as well as strong experience with OpenTelemetry, cloud-native observability, AIOps, and agentic operations.

This individual will serve as a principal-level technical leader, applying strategic thinking, systems architecture expertise, and sound decision-making to establish modern observability standards and scalable telemetry practices. Working in a fast-paced, highly collaborative technology environment, this role partners closely with engineering, site reliability, platform, security, infrastructure, operations, and architecture teams to guide the organization beyond traditional monitoring toward more intelligent and adaptive operational models.

Essential Duties & Responsibilities:

  • Define and evolve our enterprise observability vision, standards, principles, and roadmap, using strategic thinking and sound judgment to align technical direction with business needs.

  • Lead the implementation, optimization, and adoption of Dynatrace SaaS and its latest platform capabilities across the enterprise.

  • Establish scalable telemetry architecture and OpenTelemetry standards that improve consistency, interoperability, and long-term flexibility.

  • Design observability solutions for Kubernetes, containers, microservices, distributed applications, and public cloud environments.

  • Apply AIOps and agentic operations capabilities to strengthen detection, event correlation, diagnosis, automation, operational response, and continuous improvement.

  • Develop and maintain service health models, SLOs, dashboards, alerting strategies, telemetry governance, and observability best practices that support operational excellence and service reliability.

  • Collaborate across engineering, site reliability, platform, security, infrastructure, operations, and architecture functions to expand observability adoption and maturity.

  • Guide the transition from legacy monitoring practices to modern, adaptive, and outcome-focused observability models, including initiatives involving production systems and enterprise platform transformation.

  • Mentor engineers, influence technical direction, solve complex problems, and help shape enterprise architecture and engineering standards.

Required Qualifications :

  • Master's degree from an accredited college or university in Computer Science, Information Systems, Engineering, or a related technical field.

  • 8+ years of experience in observability, monitoring, site reliability engineering, platform engineering, infrastructure engineering, or related technical disciplines.

  • Deep hands-on experience with Dynatrace SaaS and recent Dynatrace platform capabilities and innovations.

  • Strong experience with OpenTelemetry and telemetry instrumentation, collection, and architecture practices.

  • Experience designing and implementing observability solutions for cloud-native infrastructure, distributed systems, and modern application environments.

  • Hands-on experience implementing AIOps and/or agentic operations capabilities.

  • Strong understanding of metrics, logs, traces, event correlation, service health, alerting models, and operational intelligence.

  • Demonstrated ability to apply lessons from traditional monitoring approaches to modern observability strategy and technical design.

  • Proven ability to lead technical initiatives, influence architectural decisions, and drive adoption across multiple teams and stakeholders.

  • Strong verbal and written communication skills, with the ability to collaborate across functions and influence technical and non-technical stakeholders.

Preferred Qualifications:

  • Experience leading enterprise-scale observability transformation initiatives.

  • Experience with AWS, Azure, and/or Google Cloud Platform.

  • Strong knowledge of site reliability engineering principles, including SLIs, SLOs, incident response, and operational resilience.

  • Experience with automation, orchestration, and remediation workflows.

  • Familiarity with additional observability platforms, frameworks, or ecosystems beyond Dynatrace.

  • Experience with scripting or programming languages such as Python, Go, Bash, or JavaScript.
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 10206981
  • Position Id: 307f78bc6feff8cf0f8fef63df2c1ffd
  • Posted 2 hours ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Remote or Texas

Today

Full-time

depends on experience

Remote or Raleigh, North Carolina

Today

Full-time

depends on experience

Remote or Arizona

Today

Full-time

depends on experience

Remote or Raleigh, North Carolina

Today

Full-time

depends on experience

Search all similar jobs