Senior DevOps SRE Engineer

Remote • Posted 2 hours ago • Updated 2 hours ago
Contract Independent
Contract Corp To Corp
Contract W2
12 Months
No Travel Required
Remote
Depends on Experience
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Grafana stack
  • Healthcare
  • DevOps
  • SRE
  • Amazon EKS
  • Kubernetes administration.
  • Site Reliability Engineer
  • CI/CD

Summary

Role: Senior DevOps / Site Reliability Engineer (SRE) 

Location: Remote 


Core Focus
We are seeking a Senior DevOps / Site Reliability Engineer with deep expertise in cloud-native observability, Kubernetes, and modern monitoring platforms. This role will design, implement, and optimize observability solutions across our AWS infrastructure, leveraging the Grafana ecosystem, eBPF, and Kubernetes to improve system reliability, performance, and operational visibility.

5-8 years experience
 
Key Responsibilities
Design, implement, and maintain scalable observability solutions across cloud-native environments.
Deploy and optimize the Grafana stack, including Grafana Beyla, Grafana Alloy, and OpenTelemetry-based instrumentation.
Leverage eBPF to enable low-overhead application and infrastructure monitoring, tracing, and performance analysis.
Support and optimize Amazon EKS clusters to ensure high availability, scalability, and operational excellence.
Develop monitoring, alerting, and dashboarding solutions that provide actionable insights into platform health.
Partner with engineering teams to improve reliability, incident response, capacity planning, and performance optimization.
Drive SRE and DevOps best practices through automation, operational standards, and continuous improvement.

Required Qualifications
Strong experience as a DevOps Engineer or Site Reliability Engineer supporting production cloud environments.
Deep expertise with Amazon EKS and Kubernetes administration.
Hands-on experience with the Grafana observability stack, including Grafana, Grafana Beyla, and Grafana Alloy.
Strong knowledge of eBPF for application performance monitoring, tracing, and infrastructure observability.
Experience implementing OpenTelemetry and modern observability frameworks.
Proficiency with AWS, Infrastructure as Code, CI/CD, and cloud-native operational practices.
Strong troubleshooting skills with a focus on reliability, scalability, automation, and performance optimization.
Experience in healthcare or other highly regulated environments is a plus.

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 91100051
  • Position Id: 9086581
  • Posted 2 hours ago
Contact the job poster
AK

Arun Kumar

Recruiter @ FlairTech Solutions
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Remote

6d ago

Easy Apply

Contract, Third Party

$63 - $65

Remote or Pennsylvania

Today

Full-time

USD 155,000.00 - 170,000.00 per year

Remote

Today

Easy Apply

Full-time

$80000 - $120000

Remote

Today

Full-time

USD 170,000.00 - 220,000.00 per year

Search all similar jobs