Site Reliability Engineer (SRE)

Remote • Posted 4 hours ago • Updated 4 hours ago
Contract W2
Contract Corp To Corp
Contract Independent
12 Months
No Travel Required
Remote
$55 - $57/hr
Fitment

Dice Job Match Score™

🔗 Matching skills to job...

Job Details

Skills

  • Splunk
  • Kibana
  • Kubernetes
  • Grafana
  • Prometheus
  • OpenTelemetry
  • Docker
  • Elasticsearch
  • Reliability Engineering
  • Terraform
  • DevOps
  • Amazon Web Services

Summary

 
 
Job Description
 
We are looking for an experienced Lead Site Reliability Engineer (SRE) – Observability to join our team and drive the design, implementation, and support of enterprise-scale observability platforms. The ideal candidate will have strong expertise in Splunk, Elasticsearch (ELK), Grafana, Prometheus, OpenTelemetry, Kafka, Terraform, and Kubernetes, with a solid background in Site Reliability Engineering, DevOps, and cloud technologies.
Required Skills
  • 7+ years of experience in Site Reliability Engineering (SRE), Platform Engineering, or DevOps.
  • Hands-on experience with Splunk Enterprise and/or Splunk Cloud administration.
  • Strong proficiency in Splunk SPL.
  • Experience with Elasticsearch (ELK Stack)KibanaPrometheusGrafanaGrafana Tempo, and OpenTelemetry.
  • Experience implementing distributed tracing, monitoring, logging, and alerting solutions.
  • Strong knowledge of Kafka and observability pipelines.
  • Hands-on experience with Terraform and Infrastructure as Code (IaC).
  • Experience with Kubernetes, Docker, and Linux environments.
  • Strong scripting skills using Python, Go, Ruby, or Bash.
  • Experience with cloud platforms such as AWS, Azure, or Google Cloud Platform.
Responsibilities
  • Design, deploy, and maintain enterprise observability platforms.
  • Administer Splunk infrastructure, including Search Head Clusters, Indexers, Heavy Forwarders, and Deployment Servers.
  • Build and manage Elasticsearch clusters for large-scale log analytics.
  • Develop dashboards, alerts, and monitoring solutions using Splunk, Grafana, and Kibana.
  • Implement distributed tracing using OpenTelemetry and Grafana Tempo.
  • Automate infrastructure deployments using Terraform.
  • Troubleshoot production issues and improve platform reliability, scalability, and performance.
  • Collaborate with development and infrastructure teams to enhance monitoring and operational excellence.
Preferred Qualifications
  • Splunk Certification.
  • Experience with Ansible, Consul, CI/CD pipelines, and Service Mesh technologies.
  • Experience working in FedRAMP or other regulated environments.
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 90769335A
  • Position Id: 9038625
  • Posted 4 hours ago
Contact the job poster
NI

Nancy Infoway

Recruiter @ Info Way Solutions
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Remote

4d ago

Easy Apply

Contract

Depends on Experience

Remote

Today

Full-time

USD 141,800.00 - 195,000.00 per year

Remote

Today

Full-time

USD 132,000.00 - 215,000.00 per year

Remote

Today

Full-time

USD 180,000.00 - 220,000.00 per year

Search all similar jobs