Technical Manager / SRE Senior Lead 
Full Time
Part Time
On-site

StratEdge It consulting INC
Fitment
Dice Job Match Score™
🛠️ Calibrating flux capacitors...
Job Details
Skills
- DOCUMENTATION
- TREND ANALYSIS
- ITSM
- OPERATIONAL EFFICIENCY
- RELIABILITY ENGINEERING
- SERVICE LEVEL
- INCIDENT MANAGEMENT
- SERVICE IMPROVEMENT
- DISASTER RECOVERY
- PREVENTIVE ACTIONS
Summary
Job Title : Technical Manager / SRE Senior Lead
Location : Phoenix,AZ
Client: TCS
Rate: $35/hr on W2
Positions: 2
JD:
Job Title : Technical Manager/SRE Senior Lead
ROLE_DESCRIPTION
"Lead and manage the SRE/Production Support team to ensure client expectations and service level objectives are consistently met.
(Core) Possess a strong understanding of ITSM and Site Reliability Engineering (SRE) processes to effectively manage production support operations.
(Core) Prepare and maintain operational reports and documentation, including bridge updates, RCA tracking, incident trends, service availability, platform health metrics, monthly operational deliverables, and trend analysis.
(Core) Facilitate technical discussions and bridge calls for critical and escalated production issues, ensuring effective communication and stakeholder alignment throughout the incident lifecycle.
Participate in Incident Management bridge calls, driving timely resolution by coordinating with engineering teams and escalating issues to the appropriate Subject Matter Experts (SMEs) as required.
Monitor platform health and identify opportunities to improve system reliability, availability, observability, and operational efficiency.
Drive continuous service improvement initiatives by analyzing recurring incidents, identifying root causes, and recommending preventive actions.
Plan, coordinate, and facilitate Disaster Recovery (DR) exercises, ensuring readiness, documentation, and post-exercise review of outcomes."
SKILLS_REQUIRED
"Lead and manage the SRE/Production Support team to ensure client expectations and service level objectives are consistently met.
(Core) Possess a strong understanding of ITSM and Site Reliability Engineering (SRE) processes to effectively manage production support operations.
(Core) Prepare and maintain operational reports and documentation, including bridge updates, RCA tracking, incident trends, service availability, platform health metrics, monthly operational deliverables, and trend analysis.
(Core) Facilitate technical discussions and bridge calls for critical and escalated production issues, ensuring effective communication and stakeholder alignment throughout the incident lifecycle.
Participate in Incident Management bridge calls, driving timely resolution by coordinating with engineering teams and escalating issues to the appropriate Subject Matter Experts (SMEs) as required.
Monitor platform health and identify opportunities to improve system reliability, availability, observability, and operational efficiency.
Drive continuous service improvement initiatives by analyzing recurring incidents, identifying root causes, and recommending preventive actions.
Plan, coordinate, and facilitate Disaster Recovery (DR) exercises, ensuring readiness, documentation, and post-exercise review of outcomes."
ESSENTIAL_SKILLS
"Lead and manage the SRE/Production Support team to ensure client expectations and service level objectives are consistently met.
(Core) Possess a strong understanding of ITSM and Site Reliability Engineering (SRE) processes to effectively manage production support operations.
(Core) Prepare and maintain operational reports and documentation, including bridge updates, RCA tracking, incident trends, service availability, platform health metrics, monthly operational deliverables, and trend analysis.
(Core) Facilitate technical discussions and bridge calls for critical and escalated production issues, ensuring effective communication and stakeholder alignment throughout the incident lifecycle.
Participate in Incident Management bridge calls, driving timely resolution by coordinating with engineering teams and escalating issues to the appropriate Subject Matter Experts (SMEs) as required.
Monitor platform health and identify opportunities to improve system reliability, availability, observability, and operational efficiency.
Drive continuous service improvement initiatives by analyzing recurring incidents, identifying root causes, and recommending preventive actions.
Plan, coordinate, and facilitate Disaster Recovery (DR) exercises, ensuring readiness, documentation, and post-exercise review of outcomes."
KEYWORDS
Technical Manager/SRE Senior Lead
EXPERIENCE_RANGE_IN_REQUIRED_SKILLS : 10+ Years
Role Descriptions: Lead and manage the SRE/Production Support team to ensure client expectations and service level objectives are consistently met.(Core) Possess a strong understanding of ITSM and Site Reliability Engineering (SRE) processes to effectively manage production support operations.(Core) Prepare and maintain operational reports and documentation| including bridge updates| RCA tracking| incident trends| service availability| platform health metrics| monthly operational deliverables| and trend analysis.(Core) Facilitate technical discussions and bridge calls for critical and escalated production issues| ensuring effective communication and stakeholder alignment throughout the incident lifecycle.Participate in Incident Management bridge calls| driving timely resolution by coordinating with engineering teams and escalating issues to the appropriate Subject Matter Experts (SMEs) as required.Monitor platform health and identify opportunities to improve system reliability| availability| observability| and operational efficiency.Drive continuous service improvement initiatives by analyzing recurring incidents| identifying root causes| and recommending preventive actions.Plan| coordinate| and facilitate Disaster Recovery (DR) exercises| ensuring readiness| documentation| and post-exercise review of outcomes.
Essential Skills: Lead and manage the SRE/Production Support team to ensure client expectations and service level objectives are consistently met.(Core) Possess a strong understanding of ITSM and Site Reliability Engineering (SRE) processes to effectively manage production support operations.(Core) Prepare and maintain operational reports and documentation| including bridge updates| RCA tracking| incident trends| service availability| platform health metrics| monthly operational deliverables| and trend analysis.(Core) Facilitate technical discussions and bridge calls for critical and escalated production issues| ensuring effective communication and stakeholder alignment throughout the incident lifecycle.Participate in Incident Management bridge calls| driving timely resolution by coordinating with engineering teams and escalating issues to the appropriate Subject Matter Experts (SMEs) as required.Monitor platform health and identify opportunities to improve system reliability| availability| observability| and operational efficiency.Drive continuous service improvement initiatives by analyzing recurring incidents| identifying root causes| and recommending preventive actions.Plan| coordinate| and facilitate Disaster Recovery (DR) exercises| ensuring readiness| documentation| and post-exercise review of outcomes.
Desirable Skills:
Keyword:
Skills: Digital : Site Reliability Engineering (SRE)
Experience Required:
Location : Phoenix,AZ
Client: TCS
Rate: $35/hr on W2
Positions: 2
JD:
Job Title : Technical Manager/SRE Senior Lead
ROLE_DESCRIPTION
"Lead and manage the SRE/Production Support team to ensure client expectations and service level objectives are consistently met.
(Core) Possess a strong understanding of ITSM and Site Reliability Engineering (SRE) processes to effectively manage production support operations.
(Core) Prepare and maintain operational reports and documentation, including bridge updates, RCA tracking, incident trends, service availability, platform health metrics, monthly operational deliverables, and trend analysis.
(Core) Facilitate technical discussions and bridge calls for critical and escalated production issues, ensuring effective communication and stakeholder alignment throughout the incident lifecycle.
Participate in Incident Management bridge calls, driving timely resolution by coordinating with engineering teams and escalating issues to the appropriate Subject Matter Experts (SMEs) as required.
Monitor platform health and identify opportunities to improve system reliability, availability, observability, and operational efficiency.
Drive continuous service improvement initiatives by analyzing recurring incidents, identifying root causes, and recommending preventive actions.
Plan, coordinate, and facilitate Disaster Recovery (DR) exercises, ensuring readiness, documentation, and post-exercise review of outcomes."
SKILLS_REQUIRED
"Lead and manage the SRE/Production Support team to ensure client expectations and service level objectives are consistently met.
(Core) Possess a strong understanding of ITSM and Site Reliability Engineering (SRE) processes to effectively manage production support operations.
(Core) Prepare and maintain operational reports and documentation, including bridge updates, RCA tracking, incident trends, service availability, platform health metrics, monthly operational deliverables, and trend analysis.
(Core) Facilitate technical discussions and bridge calls for critical and escalated production issues, ensuring effective communication and stakeholder alignment throughout the incident lifecycle.
Participate in Incident Management bridge calls, driving timely resolution by coordinating with engineering teams and escalating issues to the appropriate Subject Matter Experts (SMEs) as required.
Monitor platform health and identify opportunities to improve system reliability, availability, observability, and operational efficiency.
Drive continuous service improvement initiatives by analyzing recurring incidents, identifying root causes, and recommending preventive actions.
Plan, coordinate, and facilitate Disaster Recovery (DR) exercises, ensuring readiness, documentation, and post-exercise review of outcomes."
ESSENTIAL_SKILLS
"Lead and manage the SRE/Production Support team to ensure client expectations and service level objectives are consistently met.
(Core) Possess a strong understanding of ITSM and Site Reliability Engineering (SRE) processes to effectively manage production support operations.
(Core) Prepare and maintain operational reports and documentation, including bridge updates, RCA tracking, incident trends, service availability, platform health metrics, monthly operational deliverables, and trend analysis.
(Core) Facilitate technical discussions and bridge calls for critical and escalated production issues, ensuring effective communication and stakeholder alignment throughout the incident lifecycle.
Participate in Incident Management bridge calls, driving timely resolution by coordinating with engineering teams and escalating issues to the appropriate Subject Matter Experts (SMEs) as required.
Monitor platform health and identify opportunities to improve system reliability, availability, observability, and operational efficiency.
Drive continuous service improvement initiatives by analyzing recurring incidents, identifying root causes, and recommending preventive actions.
Plan, coordinate, and facilitate Disaster Recovery (DR) exercises, ensuring readiness, documentation, and post-exercise review of outcomes."
KEYWORDS
Technical Manager/SRE Senior Lead
EXPERIENCE_RANGE_IN_REQUIRED_SKILLS : 10+ Years
Role Descriptions: Lead and manage the SRE/Production Support team to ensure client expectations and service level objectives are consistently met.(Core) Possess a strong understanding of ITSM and Site Reliability Engineering (SRE) processes to effectively manage production support operations.(Core) Prepare and maintain operational reports and documentation| including bridge updates| RCA tracking| incident trends| service availability| platform health metrics| monthly operational deliverables| and trend analysis.(Core) Facilitate technical discussions and bridge calls for critical and escalated production issues| ensuring effective communication and stakeholder alignment throughout the incident lifecycle.Participate in Incident Management bridge calls| driving timely resolution by coordinating with engineering teams and escalating issues to the appropriate Subject Matter Experts (SMEs) as required.Monitor platform health and identify opportunities to improve system reliability| availability| observability| and operational efficiency.Drive continuous service improvement initiatives by analyzing recurring incidents| identifying root causes| and recommending preventive actions.Plan| coordinate| and facilitate Disaster Recovery (DR) exercises| ensuring readiness| documentation| and post-exercise review of outcomes.
Essential Skills: Lead and manage the SRE/Production Support team to ensure client expectations and service level objectives are consistently met.(Core) Possess a strong understanding of ITSM and Site Reliability Engineering (SRE) processes to effectively manage production support operations.(Core) Prepare and maintain operational reports and documentation| including bridge updates| RCA tracking| incident trends| service availability| platform health metrics| monthly operational deliverables| and trend analysis.(Core) Facilitate technical discussions and bridge calls for critical and escalated production issues| ensuring effective communication and stakeholder alignment throughout the incident lifecycle.Participate in Incident Management bridge calls| driving timely resolution by coordinating with engineering teams and escalating issues to the appropriate Subject Matter Experts (SMEs) as required.Monitor platform health and identify opportunities to improve system reliability| availability| observability| and operational efficiency.Drive continuous service improvement initiatives by analyzing recurring incidents| identifying root causes| and recommending preventive actions.Plan| coordinate| and facilitate Disaster Recovery (DR) exercises| ensuring readiness| documentation| and post-exercise review of outcomes.
Desirable Skills:
Keyword:
Skills: Digital : Site Reliability Engineering (SRE)
Experience Required:
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
- Dice Id: 91171926
- Position Id: OOJ - 1386-390-1785962058
- Posted 3 hours ago
Company Info
About StratEdge It consulting INC
We are a specialized IT consulting firm dedicated to providing strategic solutions and robust technical support tailored to meet your organization's unique technology needs. With deep expertise across cloud infrastructure, networking, cybersecurity, and systems management, our team helps businesses optimize their technology landscape, enhance operational efficiency, and drive innovation.
Create job alert
Similar Jobs
It looks like there aren't any Similar Jobs for this job yet.
Search all similar jobs