W2 Position - SRE & Production Support Engineer - Local to CA, WA, NY. NJ

Hybrid in New York, NY, US • Posted 1 day ago • Updated 1 day ago
Contract Independent
Contract Corp To Corp
Contract W2
12 Months
No Travel Required
Hybrid
Depends on Experience
Company Branding Image
Fitment

Dice Job Match Score™

📋 Comparing job requirements...

Job Details

Skills

  • SRE
  • Production Support
  • DevOps
  • AI DevOps
  • MLOps
  • Platform Engineering
  • Kubernetes
  • Docker
  • Helm
  • AWS
  • Azure
  • GCP
  • CI/CD
  • Jenkins
  • GitHub Actions
  • GitLab CI
  • Azure DevOps
  • Terraform
  • Ansible
  • Infrastructure as Code (IaC)
  • GitOps
  • Python
  • Bash
  • SQL
  • Go
  • Java
  • Prometheus
  • Grafana
  • Splunk
  • Datadog
  • OpenTelemetry
  • Incident Management
  • Root Cause Analysis (RCA)
  • Monitoring
  • Observability
  • Automation
  • Cloud Security
  • DevSecOps
  • Machine Learning
  • Model Deployment
  • Model Monitoring
  • Distributed Systems
  • Linux.

Summary

W2 position and local to CA, WA, NY, NJ

Position: SRE & Production Support Engineer
Location: CA / WA / NY / NJ (Hybrid)

Requirement:
We are hiring an experienced SRE & Production Support Engineer to support large-scale cloud-native and AI/ML platforms for a leading enterprise client. The ideal candidate will have a strong background in Site Reliability Engineering (SRE), Production Support, DevOps/MLOps, automation, and cloud infrastructure. This role focuses on improving platform reliability, observability, incident response, and operational excellence.


Key Skills:

  • 6+ years of experience in SRE, DevOps, Platform Engineering, or Production Support
  • Strong experience with Kubernetes, Docker, Helm, and cloud platforms (AWS, Azure, or Google Cloud Platform)
  • Expertise in CI/CD using Jenkins, GitHub Actions, GitLab, or Azure DevOps
  • Hands-on experience with Terraform, Ansible, Infrastructure as Code (IaC), and GitOps
  • Proficiency in Python, Bash, and SQL (Go or Java is a plus)
  • Experience with Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, monitoring, and observability
  • Strong knowledge of Incident Management, RCA, automation, and production support
  • Experience supporting AI/ML platforms, model deployment, model monitoring, and MLOps is highly preferred
  • Excellent troubleshooting, communication, and collaboration skills
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 91172085
  • Position Id: 426-43729-1784904516
  • Posted 1 day ago

Company Info

About Octans Group LLC

We enjoy working hard, sometimes long nights, switching between multiple projects, chasing mission impossible deadlines and we don't just keep planning, we do travel all over the world and stay motivated to keep transforming people's lifestyle through technology.

Unlike everybody, Instead of just focusing on technology, our focus is on your business and how you can use “IT” to work for you.

Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

It looks like there aren't any Similar Jobs for this job yet.

Search all similar jobs