Role: Senior Site Reliability Engineer
Location: Remote- 100%
Duration: 12+Months
Must Have: FedRAMP, DoD IL5, FIPS, GovCloud, vulnerability
management, AWS, Kubernetes, Terraform, CI/CD systems, Unix
Role Summary
Platform Services is looking for a Senior Site Reliability Engineer to help build
and operate the platform capabilities that enable FedRAMP High and IL5
environments. This role is for an independent senior engineer who can own large
features or bounded platform systems with minimal guidance.
You will design and build centralized platform APIs, reusable CI components,
continuous delivery capabilities, federal promotion workflows, and paved-road
service onboarding patterns. You should be comfortable identifying the right
technical approach when the problem is known but the solution is unclear, raising
reliability and security standards, and acting as a trusted resource for engineers
with less experience.
What You ll Do
Build and harden platform services required for FedRAMP High and IL5,
including centralized platform APIs, reusable CI components, and
continuous delivery capabilities.
Help move federal promotion workflows from manual operations toward
automated, gated, auditable deployments.
Support controlled rollout paths for federal staging and production
environments, including validation gates, rollback safety, and operational
readiness.
Build paved-road platform patterns that help teams onboard services into
federal environments safely and consistently.
Partner with other teams to unblock federal build-out and ensure platform
capabilities are secure, reliable, operable, and easy to adopt.
Build infrastructure and automation that improves reliability, security,
repeatability, and operator experience. Participate in production support and incident response for platform
services, including issues that primarily affect federal clusters.
Contribute documentation, runbooks, dashboards, and operational
handoffs so federal systems can be supported sustainably.
Provide guidance and technical feedback to engineers with less
experience. You Are an Ideal Candidate If You Have
8+ years of experience in SRE, infrastructure engineering, DevOps,
platform engineering, or backend systems engineering.
Strong coding or scripting experience in Python, Go, Ruby, Bash, or similar
languages.
Experience with AWS, Kubernetes, Terraform, CI/CD systems,
deployment automation, or service orchestration.
Comfort operating production systems with observability, alerting, incident
response, and post-incident follow-through.
Experience building secure systems with auditability, change control, and
compliance requirements in mind.
Good judgment in distributed systems, deployment safety, reliability
tradeoffs, and operational risk.
Strong written communication and a habit of leaving clear runbooks,
design notes, and implementation plans behind.
Preferred
Experience with FedRAMP, DoD IL5, FIPS, GovCloud, vulnerability
management, or regulated SaaS environments.
Experience with platform APIs, CI components, continuous delivery
systems, artifact promotion, or progressive delivery.
Experience turning manual operational workflows into automated,
supportable systems.