Senior DevOps Engineer

• Posted 30+ days ago • Updated 8 hours ago
Full Time
USD 165,000.00 per year
Fitment

Dice Job Match Score™

📊 Calculating match score...

Job Details

Skills

  • Cyber Security
  • Threat Analysis
  • Clarity
  • Investments
  • Incident Management
  • Roadmaps
  • Leadership
  • Mentorship
  • Accountability
  • Workflow
  • Knowledge Sharing
  • DevOps
  • Terraform
  • Orchestration
  • Kubernetes
  • Microsoft Azure
  • Docker
  • Meta-data Management
  • Caching
  • Optimization
  • Nexus
  • Hardening
  • Supply Chain Management
  • Amazon EC2
  • Computer Networking
  • Amazon Route 53
  • Amazon CloudFront
  • Storage
  • Remote Desktop Services
  • Amazon RDS
  • Amazon S3
  • OIDC
  • Authentication
  • Amazon Web Services
  • Cloud Computing
  • Operational Excellence
  • Lifecycle Management
  • Debugging
  • Artificial Intelligence
  • Python
  • Golang
  • Management
  • Migration
  • Bitbucket
  • Jenkins
  • GitHub
  • Continuous Integration
  • Continuous Delivery
  • Apache Velocity
  • Business Cases
  • Communication
  • Collaboration
  • Documentation
  • Genetics
  • SAP BASIS
  • PASS

Summary

At Dragos, the mission is personal. The systems we protect deliver the water you drink, power your home, and keep the hospitals your community depends on running. Those critical infrastructure systems that power our civilization around the world are under attack every day by adversaries. When those systems fail, people are immediately at risk. We are the global leader in xOT cybersecurity, combining technology, threat intelligence, and expert services. The people here chose this work because they understand what is at stake. Here, you will find a remote-first mission-driven team across North America, Europe, the Middle East, and APAC built on authenticity, transparency, and trust. If safeguarding the systems that protect your family, friends, and community is the kind of work that matters to you, you are in the right place.

About the Role:
The DevOps team sits at the heart of Dragos's ability to deliver world-class security solutions that protect critical OT infrastructure around the world. We build, maintain, and evolve systems and CI/CD practices and related infrastructure to enable and accelerate engineering org-wide.

We are looking for an experienced DevOps engineer to help us deepen the reliability, observability, and operational excellence of engineering at Dragos. You will partner closely with all engineering teams to ensure end-to-end reliability and availability of build, deployment, and test pipeline infrastructure. Your work will impact the developer experience at Dragos in every way and be relied upon by critical security infrastructure globally. You bring a builder mindset, thrive in ambiguity and fast-moving environments, create clarity, and balance long-term platform investments with pragmatic execution.

Responsibilities:
  • Set organizational standards for cloud reliability, observability, and operational excellence, including SLO/SLI frameworks, incident management practices, and the platform tooling that underpins them
  • Own the developer experience roadmap, identifying and closing systemic gaps in self-service infrastructure, CI/CD workflows, and operational visibility across engineering
  • Proactively surface systemic risks in our AWS infrastructure, IaC practices, and delivery pipelines, and drive organizational action before issues become incidents
  • Define & execute the long-term technical vision for CI at Dragos, advising leadership on architectural strategy, investment priorities, and systemic risk
  • Refactor, consolidate, and improve an existing, sprawling CI infrastructure into a single, reliable solution for use throughout engineering at Dragos.
  • Mentor engineers across levels and model the culture of engineering rigor, documentation, and cross-functional accountability expected across the organization
  • Create and maintain comprehensive documentation of all infrastructure components, processes, and deployment workflows to facilitate knowledge sharing and continuity

Qualifications:
  • 5-10 years of relevant experience in DevOps, site reliability, or infrastructure-related roles, including building and operating large-scale distributed systems or platform infrastructure; a background in security is a strong plus
  • Proficiency in AWS cloud services, Terraform, and container orchestration using ECS or EKS (Kubernetes/AKS and a second cloud such as Azure a plus)
  • Deep Docker expertise - building and managing container images and their metadata, mastery of layer and cache optimization, and hands-on experience operating ECR (or comparable container registries).
  • Experience operating artifact registries (Artifactory, Nexus) and hardening the software supply chain - artifact signing, SBOM generation, and build provenance/integrity
  • Expert AWS knowledge across compute (EC2, ECS, Lambda), networking (VPCs, Transit Gateway, Route 53, CloudFront), storage (RDS, S3), and security (IAM, Secrets Manager, Systems Manager), with a track record of high-stakes architectural decisions.
  • Deep understanding of OIDC and workload-identity federation - using OIDC for keyless CI to-cloud authentication (e.g. GitHub Actions to AWS IAM roles) to eliminate long-lived credentials
  • Deep expertise in cloud reliability and operational excellence: SLO/SLI design, centralized observability, alerting strategy, and incident lifecycle management at production scale using metrics, logs, and traces to debug and optimize complex systems
  • Experience with modern AI tooling, security best practices, and building and integrating agentic solutions to improve existing engineering processes
  • Proficient in Python, with GoLang being a nice-to-have for additional development capabilities
  • Strong experience in building and managing bulletproof CI/CD pipelines in GitHub Actions (Buildkite a plus). Experience with migrating BitBucket or Jenkins pipelines to GitHub Actions a big plus
  • Experience setting CI/CD governance and delivery standards at an organizational level defining the self-service patterns teams build on, and driving platform adoption and architectural alignment across multiple teams to reduce toil and improve engineering velocity
  • Data-driven approach to platform work - quantifying toil, lead time, and reliability, and building the business case for infrastructure investment with before/after metrics Clear and concise communication skills, both written and verbal, for effective collaboration and documentation

Compensation:
  • Salary: $165,000
  • Competitive Equity Package
  • Comprehensive Benefits Plan

#LI-NH1 #LI-REMOTE

Dragos is an Equal Opportunity Employer and considers applicants for employment without regard to race, color, religion, sex, orientation, national origin, age, disability, genetics, or any other basis forbidden under federal, state, or local laws. All new hires must pass a background check as a condition of employment.
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: RTX1e15bb
  • Position Id: 7fe442681df10b586eb554d521e1068
  • Posted 30+ days ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Chicago, Illinois

Today

Full-time

USD 120,000.00 - 160,000.00 per year

Philadelphia, Pennsylvania

Today

Full-time

Plano, Texas

25d ago

Full-time

USD 124,000.00 - 163,900.00 per year

Remote or Hybrid in Plano, Texas

Today

Full-time

USD 124,000.00 - 163,900.00 per year

Search all similar jobs