Hiring | SRE / Production Support | SFO, CA | Contract

San Francisco, CA, US • Posted 2 days ago • Updated 2 days ago
Contract W2
12 Months
On-site
Depends on Experience
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Amazon Web Services
  • Terraform
  • Splunk
  • Product Development

Summary

Role : SRE / Production Support
Location : SFO, CA
Duration : Long-term Contract
Preferred Skills/Experience :
  • A successful candidate will have at least 5 years of experience in software development or technical support or operations experience.
  • Excellent troubleshooting, debugging and problem-solving skills required.'
  • Expertise in Kubernetes, Docker, Jenkins, and Java stack-based production systems administration.
  • Proficiency in at least two of the technologies -Cassandra, Yugabyte, Kafka, Microservices, Spring Boot, Spark Streaming, Flink desired.
  • Experience with monitoring and logging tools such as Datadog, Kibana, and Splunk.
  • Experience in incident management preferred.
  • Strong experience with load balancing principles (F5, etc.) a plus.
  • Strong experience with infrastructure/cloud technologies (Google Cloud Platform, AWS, Azure) preferred.
  • Experience in systems and multi-tier application and network troubleshooting.
  • Strong experience with configuration management and automation (e.g., Ansible) a plus.
  • Experience programming in core java or python is required.
  • Experience with scripting languages (Shell & Python).
  • Experience with database concepts (SQL or NO SQL).
  • You know Linux. Even Kubernetes still runs on computers! This means you can debug most normal issues with performance, networking, kernel drivers, package management, etc. or have a good idea where to start
  • Demonstrable knowledge of Terraform, Jenkins, Artifactory, a strong plus.
  • A BA/BS/master's degree in the field of CS or related field is preferred but not required.
Required Skills :
  • Create and maintain automation scripts for deployment, scaling, and monitoring.
  • Develop and enhance production monitoring and management capabilities leveraging existing platforms and tools.
  • Work independently and within a team to triage and remediate production system and application incidents.
  • Handle escalations from USBank Consumer Domain partners about critical issues.
  • Work with Domain, Infra and other Support teams inside the USBank in troubleshooting, escalating, and resolving critical site incidents.
  • Identify recurring system and application issues and work with cloud teams, infra teams, product development, vendors, and other stakeholders in investigating and resolving causes.
  • Send communications regarding outages to other USBank teams, partners, and other customers.
  • Maintain accurate documentation of site incidents, including impact details, timelines, steps taken for mitigation/resolution.
  • Develop and maintain technical documentation for all USBank operations infrastructure and practices.
  • Remediate all P1, P2, P3,P4 Vulnerabilities within the approved US Bank SLA.
  • Schedule: TBD
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 91172365
  • Position Id: 9082704
  • Posted 2 days ago
Contact the job poster
SM

Shalman Mohamed Aniba

Recruiter @ Healthcare Triangle Inc
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

San Francisco, California

Today

Full-time

Remote

Today

Easy Apply

Contract

$60 - $70 per hour

Florida

Today

Full-time

USD 144,730.00 - 246,040.00 per year

Remote

4d ago

Easy Apply

Contract

Depends on Experience

Search all similar jobs