Level: L1
Location: Remote, prefer PST hours
Visa Requirements: UC Citizenship
About the Team
The SRE Platform Engineering team builds and operates the infrastructure that powers our cloud. We focus on delivering reliable, scalable, and simple platforms that enable product teams to move quickly while meeting the requirements of regulated environments such as FedRAMP High and DoD IL5.
About the Role
We’re looking for a Site Reliability Engineer to support the development and operation of our Kubernetes-based platform in regulated environments. In this role, you will work closely with senior engineers and technical leaders to improve reliability, scalability, and compliance across the platform.
This is a hands-on engineering role where you’ll contribute to key systems, build infrastructure to support production operations, and help implement solutions that improve the overall health and performance of the platform.
What You Will Do
Contribute to the design, implementation, and operation of Kubernetes platforms in FedRAMP High / IL5 environments
Support day-to-day reliability and performance of platform services, including monitoring and alerting
Cisco Confidential
Implement automation and tooling to improve operational efficiency and reduce manual effort
Work with senior engineers to define and track SLIs, SLOs, and error budgets
Assist in maintaining compliance and security requirements, including support for audits and continuous monitoring
Contribute to infrastructure as code and CI/CD pipeline improvements
Collaborate with cross-functional teams (Security, Platform, Application teams) to resolve issues and deliver platform capabilities
Participate in on-call rotations supporting customer requests and paging alerts
What You Bring
4–6 years of experience in SRE, DevOps, or platform engineering roles
Experience with Kubernetes in production environments
Familiarity with cloud platforms (AWS, Azure, or similar; GovCloud experience a plus)
Solid understanding of Linux systems, networking, and containerization
Experience with Infrastructure as Code (e.g., Terraform)
Proficiency in scripting or programming (e.g., Python, Go)
Exposure to observability tools (Prometheus, Grafana, logging systems)
Nice to Have
Experience working in FedRAMP High or DoD IL5 environments
Exposure to CI/CD systems and deployment automation (e.g., ArgoCD)
Familiarity with container security practices and tools
Experience supporting regulated or audited systems
How You Work
You take ownership of your work while seeking guidance when needed
You collaborate effectively with more senior engineers and technical leaders
You focus on building reliable, maintainable solutions
You are comfortable working in structured, compliance-driven environments
You are proactive about learning and improving systems