Technical Manager / SRE Senior Lead&nbsp

PHOENIX, AZ, US • Posted 3 hours ago • Updated 19 minutes ago
Full Time
Part Time
On-site
Company Branding Image
Fitment

Dice Job Match Score™

🛠️ Calibrating flux capacitors...

Job Details

Skills

  • DOCUMENTATION
  • TREND ANALYSIS
  • ITSM
  • OPERATIONAL EFFICIENCY
  • RELIABILITY ENGINEERING
  • SERVICE LEVEL
  • INCIDENT MANAGEMENT
  • SERVICE IMPROVEMENT
  • DISASTER RECOVERY
  • PREVENTIVE ACTIONS

Summary

Job Title : Technical Manager / SRE Senior Lead

Location : Phoenix,AZ

Client: TCS

Rate: $35/hr on W2


Positions: 2

JD:

Job Title : Technical Manager/SRE Senior Lead

ROLE_DESCRIPTION

"Lead and manage the SRE/Production Support team to ensure client expectations and service level objectives are consistently met.

(Core) Possess a strong understanding of ITSM and Site Reliability Engineering (SRE) processes to effectively manage production support operations.

(Core) Prepare and maintain operational reports and documentation, including bridge updates, RCA tracking, incident trends, service availability, platform health metrics, monthly operational deliverables, and trend analysis.

(Core) Facilitate technical discussions and bridge calls for critical and escalated production issues, ensuring effective communication and stakeholder alignment throughout the incident lifecycle.

Participate in Incident Management bridge calls, driving timely resolution by coordinating with engineering teams and escalating issues to the appropriate Subject Matter Experts (SMEs) as required.

Monitor platform health and identify opportunities to improve system reliability, availability, observability, and operational efficiency.

Drive continuous service improvement initiatives by analyzing recurring incidents, identifying root causes, and recommending preventive actions.

Plan, coordinate, and facilitate Disaster Recovery (DR) exercises, ensuring readiness, documentation, and post-exercise review of outcomes."

SKILLS_REQUIRED

"Lead and manage the SRE/Production Support team to ensure client expectations and service level objectives are consistently met.

(Core) Possess a strong understanding of ITSM and Site Reliability Engineering (SRE) processes to effectively manage production support operations.

(Core) Prepare and maintain operational reports and documentation, including bridge updates, RCA tracking, incident trends, service availability, platform health metrics, monthly operational deliverables, and trend analysis.

(Core) Facilitate technical discussions and bridge calls for critical and escalated production issues, ensuring effective communication and stakeholder alignment throughout the incident lifecycle.

Participate in Incident Management bridge calls, driving timely resolution by coordinating with engineering teams and escalating issues to the appropriate Subject Matter Experts (SMEs) as required.

Monitor platform health and identify opportunities to improve system reliability, availability, observability, and operational efficiency.

Drive continuous service improvement initiatives by analyzing recurring incidents, identifying root causes, and recommending preventive actions.

Plan, coordinate, and facilitate Disaster Recovery (DR) exercises, ensuring readiness, documentation, and post-exercise review of outcomes."

ESSENTIAL_SKILLS

"Lead and manage the SRE/Production Support team to ensure client expectations and service level objectives are consistently met.

(Core) Possess a strong understanding of ITSM and Site Reliability Engineering (SRE) processes to effectively manage production support operations.

(Core) Prepare and maintain operational reports and documentation, including bridge updates, RCA tracking, incident trends, service availability, platform health metrics, monthly operational deliverables, and trend analysis.

(Core) Facilitate technical discussions and bridge calls for critical and escalated production issues, ensuring effective communication and stakeholder alignment throughout the incident lifecycle.

Participate in Incident Management bridge calls, driving timely resolution by coordinating with engineering teams and escalating issues to the appropriate Subject Matter Experts (SMEs) as required.

Monitor platform health and identify opportunities to improve system reliability, availability, observability, and operational efficiency.

Drive continuous service improvement initiatives by analyzing recurring incidents, identifying root causes, and recommending preventive actions.

Plan, coordinate, and facilitate Disaster Recovery (DR) exercises, ensuring readiness, documentation, and post-exercise review of outcomes."

KEYWORDS

Technical Manager/SRE Senior Lead

EXPERIENCE_RANGE_IN_REQUIRED_SKILLS : 10+ Years

Role Descriptions: Lead and manage the SRE/Production Support team to ensure client expectations and service level objectives are consistently met.(Core) Possess a strong understanding of ITSM and Site Reliability Engineering (SRE) processes to effectively manage production support operations.(Core) Prepare and maintain operational reports and documentation| including bridge updates| RCA tracking| incident trends| service availability| platform health metrics| monthly operational deliverables| and trend analysis.(Core) Facilitate technical discussions and bridge calls for critical and escalated production issues| ensuring effective communication and stakeholder alignment throughout the incident lifecycle.Participate in Incident Management bridge calls| driving timely resolution by coordinating with engineering teams and escalating issues to the appropriate Subject Matter Experts (SMEs) as required.Monitor platform health and identify opportunities to improve system reliability| availability| observability| and operational efficiency.Drive continuous service improvement initiatives by analyzing recurring incidents| identifying root causes| and recommending preventive actions.Plan| coordinate| and facilitate Disaster Recovery (DR) exercises| ensuring readiness| documentation| and post-exercise review of outcomes.

Essential Skills: Lead and manage the SRE/Production Support team to ensure client expectations and service level objectives are consistently met.(Core) Possess a strong understanding of ITSM and Site Reliability Engineering (SRE) processes to effectively manage production support operations.(Core) Prepare and maintain operational reports and documentation| including bridge updates| RCA tracking| incident trends| service availability| platform health metrics| monthly operational deliverables| and trend analysis.(Core) Facilitate technical discussions and bridge calls for critical and escalated production issues| ensuring effective communication and stakeholder alignment throughout the incident lifecycle.Participate in Incident Management bridge calls| driving timely resolution by coordinating with engineering teams and escalating issues to the appropriate Subject Matter Experts (SMEs) as required.Monitor platform health and identify opportunities to improve system reliability| availability| observability| and operational efficiency.Drive continuous service improvement initiatives by analyzing recurring incidents| identifying root causes| and recommending preventive actions.Plan| coordinate| and facilitate Disaster Recovery (DR) exercises| ensuring readiness| documentation| and post-exercise review of outcomes.

Desirable Skills:

Keyword:

Skills: Digital : Site Reliability Engineering (SRE)

Experience Required:
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 91171926
  • Position Id: OOJ - 1386-390-1785962058
  • Posted 3 hours ago

Company Info

About StratEdge It consulting INC

We are a specialized IT consulting firm dedicated to providing strategic solutions and robust technical support tailored to meet your organization's unique technology needs. With deep expertise across cloud infrastructure, networking, cybersecurity, and systems management, our team helps businesses optimize their technology landscape, enhance operational efficiency, and drive innovation.

Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

It looks like there aren't any Similar Jobs for this job yet.

Search all similar jobs