Jr. Site Reliability Engineer – FedRAMP Vulnerability Management
Job Overview The SRE Compliance & Security Initiatives team is seeking a Jr. Site Reliability Engineer to support the stability, scalability, and security of our FedRAMP cloud platform. Operating with a high degree of autonomy, you will focus on vulnerability management, automating security processes, and helping our infrastructure adapt to evolving regulatory requirements.
Key Responsibilities
Vulnerability Management: Own findings from intake through validated remediation and closure. Triage and prioritize host, OS-package, container, base-image, and application-dependency findings based on exploitability, asset criticality, and deadlines.
Remediation & Analysis: Identify the true source of vulnerabilities and coordinate fixes (dependency updates, image rebuilds, host changes, deviations, or false-positive corrections) using Ansible, GitLab CI, package managers, and container builds.
Validation: Verify fixes using package versions, build artifacts, image manifests, deployed-host evidence, and scanner rescans before closing tickets.
Automation & Tooling: Develop and maintain automation solutions using Ansible to improve infrastructure reliability, scalability, and security compliance.
CI/CD & Deployment: Design and enhance deployment pipelines, testing frameworks, and operational tooling to support rapid scaling across global environments.
Documentation & Tracking: Maintain Jira evidence, ownership escalations, backlog metrics, runbooks, and knowledge transfer to streamline processes.
Minimum Qualifications
2+ years of experience in Site Reliability Engineering, DevOps, Infrastructure Engineering, or a related role supporting cloud-based production environments.
Practical vulnerability-management experience using tools like Qualys, JFrog Xray, SCA/SBOM, or equivalent.
Working knowledge of dependency management for Ruby/Bundler, Python/pip, and Go modules (including direct vs. transitive dependencies, version constraints, lockfiles, go.mod/go.sum).
Experience developing and maintaining infrastructure automation using Ansible.
Strong experience administering and troubleshooting Linux-based systems and distributed infrastructure environments.
Experience designing, implementing, and maintaining CI/CD pipelines, including GitLab CI.
Experience supporting large-scale infrastructure environments consisting of hundreds or thousands of systems.
Availability to be online daily from 9:00 AM to 4:00 PM PST.
Preferred Qualifications
Familiarity with AWS or other public cloud platforms and hybrid infrastructure environments.
Knowledge of monitoring, observability, and reliability engineering practices.
Familiarity with Kubernetes and containerized application platforms.
Experience leveraging AI-assisted development tools to improve automation and engineering productivity.
Experience managing fleet-wide software deployments and providing priority incident triage.