What are the top 3 skills required for this role?
1.Compute - Linux
2.AWS
3. Terraform
Job Description/ Responsibilities
* Serve as the highest level technical escalation point (L3 SME) for Linux/Unix infrastructure.
* Manage and support large-scale Red Hat Enterprise Linux (RHEL), Oracle Linux, CentOS, SUSE, and Unix environments.
* Perform advanced troubleshooting of OS, kernel, filesystem, storage, networking, and performance issues.
* Lead OS patching, upgrades, vulnerability remediation, and lifecycle management.
* Ensure system availability, stability, scalability, and operational compliance.
* Lead resolution of critical incidents and major outages.
* Perform root cause analysis (RCA) and implement preventive measures.
* Review and approve changes related to compute infrastructure.
* Drive continuous service improvement initiatives.
* Conduct capacity planning and resource optimization.
* Analyze system utilization, bottlenecks, and trends.
* Implement proactive monitoring and self-healing mechanisms.
* Drive toil reduction through automation and operational innovation.
* Develop and maintain shell scripting and automation solutions.
* Automate provisioning, patching, compliance checks, and operational tasks.
* Contribute to Infrastructure as Code practices using Terraform.
* Support CI/CD integration for infrastructure deployment activities.
Years of Experience: 18.00 Years of Experience