Job Summary Client is seeking a Senior DevOps Engineer for a 3-month contingent engagement supporting the High Performance Computing (HPC) and Electronic Design Automation (EDA) infrastructure team. This role operates within the IT Datacenter (ITDC) organization and requires independent senior-level execution with minimal ramp-up time. The ideal candidate brings strong hands-on experience with Linux HPC environments, infrastructure automation, SLURM workload management, datacenter migration, and enterprise identity integration. Key Responsibilities HPC / EDA Platform Operations Support SLURM HPC environments including partition configuration and migration planning Plan and execute datacenter migrations for compute and storage infrastructure Develop migration strategies and author MOPs/runbooks for infrastructure changes Coordinate cross-functionally with EDA, storage, and IDAM teams Verify service continuity following migrations Define HPC storage tiers and gather workload requirements Linux Systems Engineering & OS Deployment Administer SLES systems in production HPC environments Build custom OS images using Kiwi NG Enable bare-metal provisioning via RackN/Digital Rebar Provision Provision VMware vSphere VMs for HPC workloads Troubleshoot Linux HPC services and system daemons Automation & Infrastructure as Code Develop Ansible playbooks for Linux system setup and configuration Ensure multi-version compatibility across SLES versions Manage Git repositories and migrate artifacts to Artifactory Contribute to GitHub repositories and conduct reviews Drive changes via ServiceNow workflows Identity & Access Management Integrate enterprise identity systems for Linux/HPC environments Audit UID/GID data across domains Validate authentication methods and extend SSSD authentication Monitoring, Logging & Operational Readiness Implement log management strategies including Splunk integration Investigate production issues in Linux services Produce technical documentation in Confluence Required Qualifications 5+ years of experience in DevOps, Platform Engineering, or Linux Systems Engineering Hands-on HPC cluster administration experience with SLURM or equivalent workload managers Experience supporting EDA or scientific computing environments Strong Ansible automation skills with production-grade playbook development Experience with bare-metal provisioning tools (RackN, Cobbler, or equivalent) Proven ability to plan and execute datacenter migrations with minimal disruption Familiarity with enterprise Linux identity/authentication stacks (SSSD, LDAP, AD, NIS, Okta) Experience with NetApp or comparable enterprise storage platforms Ability to author technical documentation (MOPs, runbooks, diagrams) Strong written and verbal communication skills Preferred Qualifications Experience with SUSE Linux Enterprise Server (SLES) 12/15 in enterprise environments Familiarity with RackN/Digital Rebar Provision for OS deployment Hands-on experience with Kiwi NG for OS image creation Knowledge of VMware vSphere for HPC VM provisioning Experience migrating artifacts to Artifactory Background in semiconductor, storage, or high-tech manufacturing IT environments Certifications Relevant Linux, DevOps, or HPC certifications preferred (e.g., RHCSA, RHCE, VMware, Ansible) Education: Bachelors Degree
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
- Dice Id: compun
- Position Id: SINDC5868060
- Posted 6 hours ago