Senior SRE with AI Automation

Hybrid in San Francisco, CA, US • Posted 2 hours ago • Updated 2 hours ago
Full Time
Hybrid
$80,000 - $120,000/yr
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Dynatrace
  • Splunk
  • Grafana
  • Python
  • Bash
  • PowerShell

Summary

Job Title:Senior SRE with AI Automation

Location: San Francisco, CA

Duration: 12 Months + Extension

Bill Rate: $68/hour

Job Type: W-2 Contract

Client: To Be Discussed Later

Work Authorization: US-Citizen, H-1B, OPT-EAD, GC-EAD

Job Overview

We are looking for a Senior Site Reliability Engineer (SRE) with AI Automation experience to support high-scale digital applications and drive reliability, automation, observability, and operational excellence. The ideal candidate will have strong hands-on experience with Azure, AKS, Kubernetes, Docker, CI/CD, observability tools, scripting, and production support.

Key Responsibilities

  • Design, implement, and maintain highly reliable and scalable production environments.

  • Provide hands-on SRE, DevOps, and production engineering support for high-scale digital applications.

  • Manage and support Azure AKS, Kubernetes, Docker, Service Mesh, and API-driven architectures.

  • Support production environments hosting React front-end applications and Spring Boot microservices.

  • Develop automation using Python, Bash, PowerShell, and YAML to reduce operational toil and improve incident response.

  • Implement and maintain comprehensive observability using Dynatrace, Splunk, Grafana, and Prometheus.

  • Monitor system health, availability, performance, capacity, and reliability.

  • Lead incident management, troubleshooting, root cause analysis (RCA), and permanent corrective actions.

  • Identify recurring operational issues and develop automated solutions to improve reliability.

  • Build and maintain CI/CD pipelines using tools such as Jenkins and GitHub Actions.

  • Apply SRE principles, automation, and reliability engineering practices across application and infrastructure environments.

  • Collaborate closely with development, DevOps, platform, and operations teams to improve developer and operator experience.

  • Identify opportunities to leverage AI/automation to streamline operations, incident response, monitoring, and troubleshooting.

  • Participate in production support and handle critical incidents effectively in a high-pressure environment.

  • Establish and improve operational processes, reliability standards, and best practices.

 

Equal Opportunity Employer:
We are an equal opportunity employer. All aspects of employment including the decision to hire, promote, discipline, or discharge, will be based on merit, competence, performance, and business needs. We do not discriminate on the basis of race, color, religion, marital status, age, national origin, ancestry, physical or mental disability, medical condition, pregnancy, genetic information, gender, sexual orientation, gender identity or expression, national origin, citizenship/ immigration status, veteran status, or any other status protected under federal, state, or local law

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 91099737
  • Position Id: 964SD
  • Posted 2 hours ago
Contact the job poster
KN

Kavya Nair

Recruiter @ QUANTUM TECHNOLOGIES LLC
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Remote or San Francisco, California

Today

Full-time

USD 139,764.00 - 287,749.00 per year

San Francisco, California

Today

Full-time

USD 120,600.00 - 150,900.00 per year

Remote or San Francisco, California

Today

Full-time

USD 152,500.00 - 205,000.00 per year

San Francisco, California

Today

Full-time

Search all similar jobs