Site Reliability Engineer

Remote • Posted 8 hours ago • Updated 8 hours ago
Full Time
No Travel Required
Remote
$70/hr
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Job Description: Key Responsibilities Support and enhance observability (monitoring
  • logging
  • alerting) across production systems Help maintain SLIs/SLOs for key services Participate in evaluating services for production readiness Collaborate with development teams to identify reliability risks and improve system architecture Contribute to automation of operations
  • including CI/CD pipelines
  • incident response
  • and infrastructure provisioning Participate in incident response and on-call rotations for critical services Contribute to post-incident analysis and drive reliability improvements Partner with security
  • infrastructure
  • and product teams to support performance
  • compliance
  • and operational excellence Must-Haves Willingness to work onsite and participate in a 24/7 on-call rotation as needed 5+ years of experience managing and supporting high-traffic digital platforms Strong experience with CI/CD pipelines and deployment automation Experience with cloud platforms such as AWS and/or GCP Solid scripting skills (e.g.
  • Python
  • Bash
  • Groovy) Hands-on experience with observability and monitoring tools like Datadog
  • New Relic
  • AppDynamics
  • or similar Understanding of web
  • mobile
  • and OTT architectures Experience supporting large scale websites
  • Mobile and OTT applications
  • microservices
  • APIs
  • and distributed systems Experience with infrastructure-as-code tools such as Ansible
  • Terraform
  • or Chef Familiarity with performance testing tools like JMeter or k6 Hands on experience with debugging tools like Charles Proxy or Fiddler Preferred Qualifications Experience working with CDNs (e.g.
  • Akamai) and reverse proxies (e.g.
  • NGINX
  • Varnish) Exposure to video streaming platforms and Familiarity with application/infrastructure security controls and best practices Certifications in SRE
  • DevOps
  • or Performance Engineering are a plus

Summary

Job Description:
 
Key Responsibilities
 
  • Support and enhance observability (monitoring, logging, alerting) across production systems
  • Help maintain SLIs/SLOs for key services
  • Participate in evaluating services for production readiness
  • Collaborate with development teams to identify reliability risks and improve system architecture
  • Contribute to automation of operations, including CI/CD pipelines, incident response, and infrastructure provisioning
  • Participate in incident response and on-call rotations for critical services
  • Contribute to post-incident analysis and drive reliability improvements
  • Partner with security, infrastructure, and product teams to support performance, compliance, and operational excellence 
 
Must-Haves
 
  • Willingness to work onsite and participate in a 24/7 on-call rotation as needed
  • 5+ years of experience managing and supporting high-traffic digital platforms
  • Strong experience with CI/CD pipelines and deployment automation
  • Experience with cloud platforms such as AWS and/or Google Cloud Platform
  • Solid scripting skills (e.g., Python, Bash, Groovy)
  • Hands-on experience with observability and monitoring tools like Datadog, New Relic, AppDynamics, or similar
  • Understanding of web, mobile, and OTT architectures
  • Experience supporting large scale websites, Mobile and OTT applications, microservices, APIs, and distributed systems
  • Experience with infrastructure-as-code tools such as Ansible, Terraform, or Chef
  • Familiarity with performance testing tools like JMeter or k6
  • Hands on experience with debugging tools like Charles Proxy or Fiddler
 
 Preferred Qualifications
 
  • Experience working with CDNs (e.g., Akamai) and reverse proxies (e.g., NGINX, Varnish)
  • Exposure to video streaming platforms and Familiarity with application/infrastructure security controls and best practices
  • Certifications in SRE, DevOps, or Performance Engineering are a plus
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 10283077
  • Position Id: 9071234
  • Posted 8 hours ago
Contact the job poster
NK

Nikesh Kumar

Recruiter @ Bansar Technologies Inc.
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Remote

6d ago

Easy Apply

Full-time

$180,000 - $215,000

Remote

Today

Full-time

Remote

Today

Full-time

USD 141,800.00 - 195,000.00 per year

Remote

Today

Easy Apply

Full-time

Up to $100,000

Search all similar jobs