Site Reliability Engineer

Austin, TX, US • Posted 10 hours ago • Updated 10 hours ago
Full Time
On-site
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Creative Problem Solving
  • Project Management
  • Preventive Maintenance
  • Performance Management
  • Order Management
  • Finance
  • Conflict Resolution
  • Problem Solving
  • Decision-making
  • Continuous Improvement
  • Collaboration
  • Accountability
  • Knowledge Sharing
  • Leadership
  • Innovation
  • Adaptability
  • Computer Science
  • Information Technology
  • Reliability Engineering
  • Technical Support
  • Java
  • AppDynamics
  • Splunk
  • Grafana
  • BMC Control-M
  • Oracle
  • SQL
  • Database
  • Performance Tuning
  • Linux Administration
  • Red Hat Enterprise Linux
  • Scripting
  • Python
  • Shell Scripting
  • Operational Efficiency
  • Incident Management
  • Management
  • Communication
  • Operational Excellence
  • MEAN Stack
  • ROOT
  • Software Engineering
  • Mentorship
  • Coaching
  • IT Management
  • Production Support
  • Change Management
  • Risk Management
  • Regulatory Compliance

Summary

Your Opportunity

At Schwab, you're empowered to make an impact on your career. Here, innovative thought meets creative problem solving, helping us challenge the status quo and transform the finance industry together. We believe in the importance of in-office collaboration and fully intend for the selected candidate for this role to work on site 4-days per week during night shifts (2pm-10pm CST), and weekends as needed, in the specified location(s).

As a Site Reliability Engineer, you will play a critical role in protecting the stability, performance, and resiliency of Schwab's Order Management System, supporting the technology that enables our clients to navigate their financial futures with confidence. In this highly visible role, you will assess and resolve complex production incidents, drive rapid restoration of critical systems, and collaborate across engineering, infrastructure, databases, and vendor teams to minimize business impact and improve client experiences.

Success in this role requires strong problem-solving capabilities, sound decision-making during high-pressure situations, and the ability to navigate complex distributed environments. You will partner closely with technology teams to strengthen production readiness, improve operational excellence, and identify opportunities to reduce recurring incidents through automation, observability, and continuous improvement initiatives. As a senior member of the team, you will also help shape operational best practices, mentor fellow engineers, and contribute to a culture of collaboration, accountability, and knowledge sharing.

At Schwab, we succeed together as One Schwab. You will have the opportunity to make a meaningful impact while developing your technical expertise, leadership capabilities, and operational excellence skills in a collaborative environment that values innovation, adaptability, and continuous learning.

What you have

Required Qualifications
  • Bachelor's degree in Computer Science, Engineering, Information Technology, or a related field, or equivalent practical experience.
  • 5+ years of experience in Production Support, Site Reliability Engineering (SRE), Software Operations, or a related technology support role.
  • Advanced experience troubleshooting Java-based distributed applications and SQL-backed enterprise systems.
  • Strong experience with application monitoring and observability tools, including AppDynamics, Splunk, Grafana, InfluxDB, and Control-M.
  • Strong Oracle Database knowledge with experience in SQL analysis, database troubleshooting, and performance optimization.
  • Strong Linux administration experience, preferably supporting RHEL 7/8/9 environments.
  • Experience using scripting or automation technologies such as Python or Shell scripting to improve operational efficiency.
  • Experience leading incident response activities and managing high-severity production incidents.
  • Strong written and verbal communication skills with the ability to communicate effectively with both technical and non-technical stakeholders.
  • Availability to support night shifts, weekends, and participation in a rotating on-call support model.

Preferred Qualifications
  • Experience driving operational excellence initiatives that improve availability, reliability, and mean time to resolution (MTTR).
  • Experience developing or enhancing monitoring strategies, operational runbooks, and escalation procedures.
  • Demonstrated ability to identify root causes, implement permanent corrective actions, and reduce incident recurrence.
  • Experience partnering with software engineering teams to improve application supportability and production readiness.
  • Experience leading change implementation planning and supporting medium-to-high risk production releases.
  • Demonstrated mentoring, coaching, or technical leadership experience within production support or SRE teams.
  • Knowledge of change management, risk management, security, and compliance practices in highly regulated environments.

In addition to the salary range, this role is eligible for bonus or incentive opportunities.
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 90989465
  • Position Id: 3fe6b5764906e201349aca6c97b63449
  • Posted 10 hours ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Austin, Texas

Today

Full-time

USD 150,000.00 - 162,000.00 per year

Austin, Texas

10d ago

Full-time

USD 110,700.00 - 171,800.00 per year

Hybrid in Austin, Texas

2d ago

Easy Apply

Contract, Third Party

Depends on Experience

Montana

Today

Full-time

USD 189,000.00 - 232,000.00 per year

Search all similar jobs