Site Reliability Engineer

Hybrid in Schaumburg, IL, US • Posted 8 hours ago • Updated 8 hours ago
Contract W2
12 Months
No Travel Required
Hybrid
Depends on Experience
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • SRE
  • Python
  • ML
  • Pandas
  • Numpy
  • Power BI

Summary

W2 ONLY

 

Position: Site Reliability Engineer

Location: Schaumburg, IL or Secaucus, NJ (Hybrid)

CONTRACT

 

Job Description:

Mandatory Skills : Python/R and ML libraries (scikit-learn, TensorFlow, PyTorch), Data analysis and visualization (Pandas, NumPy, Power BI/Tableau), SQL and database management.

 

Key Responsibilities

  • Monitor, maintain, and improve the reliability, availability, and performance of production systems.
  • Design and implement monitoring, alerting, logging, and observability solutions.
  • Establish and track Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budgets.
  • Automate operational tasks and repetitive processes using scripting and Infrastructure as Code (IaC).
  • Lead incident response activities, troubleshooting, root cause analysis (RCA), and post-incident reviews.
  • Collaborate with development, infrastructure, and platform teams to improve system reliability and resilience.
  • Perform capacity planning, performance tuning, and scalability assessments.
  • Support CI/CD pipelines and deployment automation initiatives.
  • Implement high-availability, disaster recovery, and failover strategies.

 

Required Skills

  • Strong experience with Linux/Unix administration.
  • Proficiency in scripting languages such as Python, Shell, or PowerShell.
  • Hands-on experience with cloud platforms (AWS, Azure, or Google Cloud Platform).
  • Experience with containerization technologies such as Docker and Kubernetes.
  • Knowledge of monitoring and observability tools such as Prometheus, Grafana, ELK, Splunk, Dynatrace, or Datadog.
  • Understanding of CI/CD tools such as Jenkins, GitHub Actions, GitLab CI, or Azure DevOps.
  • Experience with Infrastructure as Code tools such as Terraform, Ansible, or CloudFormation.
  • Strong troubleshooting, debugging, and problem-solving skills.
  • Understanding of networking, security, and distributed systems concepts.

 

Experience

  • 5–10+ years of overall IT experience.
  • 5+ years of hands-on experience in Site Reliability Engineering, Production Support, DevOps, or Cloud Operations roles.
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 90798281
  • Position Id: 9050862
  • Posted 8 hours ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Naperville, Illinois

Today

Easy Apply

Full-time

USD 82500-110000

Remote

Today

Full-time

USD 147,000.00 - 168,000.00 per year

Remote

Today

Full-time

USD 114,000.00 - 148,000.00 per year

No location provided

Today

Full-time

USD 81,100.00 - 187,000.00 per year

Search all similar jobs