Senior IT Triage Engineer

Raleigh, NC, US • Posted 3 hours ago • Updated 3 hours ago
Full Time
On-site
depends on experience
Company Branding Image
Fitment

Dice Job Match Score™

🫥 Flibbertigibetting...

Job Details

Skills

  • FOCUS
  • Incident Management
  • Systems Analysis
  • Cloud Computing
  • Network
  • Bridging
  • ROOT
  • Corrective And Preventive Action
  • Operational Excellence
  • Process Optimization
  • MEAN Stack
  • Process Improvement
  • Recovery
  • IT Management
  • Mentorship
  • Operational Risk
  • Application Development
  • Server Administration
  • Information Security
  • IT Operations
  • Technical Support
  • Management
  • Problem Management
  • Root Cause Analysis
  • IT Service Management
  • Software Architecture
  • Microsoft Windows
  • Linux Administration
  • Virtualization
  • VMware
  • Hyper-V
  • Storage
  • Backup
  • Performance Monitoring
  • Software Performance Management
  • Dashboard
  • Event Management
  • Windows PowerShell
  • Python
  • Bash
  • Scripting
  • Workflow
  • API
  • Orchestration
  • Analytical Skill
  • Accountability
  • Collaboration
  • Stakeholder Management
  • Communication
  • Decision-making
  • Problem Solving
  • Conflict Resolution
  • Continuous Improvement

Summary

Overview

The Senior IT Triage Engineer serves as a critical technical leader responsible for the rapid diagnosis, restoration, and resolution of complex technology incidents across enterprise infrastructure, cloud platforms, applications, and end-user services. This role combines deep technical troubleshooting expertise with a proactive focus on eliminating recurring issues through structured problem management, root cause analysis, and continuous service improvement.

The ideal candidate excels in high-pressure operational environments, drives technical resolution efforts across multiple teams, and leverages automation, observability, and reliability practices to improve service stability and reduce operational risk.

Responsibilities

Key Responsibilities
  • Incident Management, Service Restoration, and System Analysis.
  • Lead technical triage efforts for high-priority incidents and service disruptions.
  • Coordinate cross-functional teams to restore critical business services as quickly as possible.
  • Analyze alerts, logs, monitoring data, and telemetry to identify the source of issues.
  • Serve as a senior escalation point for complex infrastructure, cloud, network, application, and platform incidents.
  • Drive incident bridges, facilitate technical discussions, and maintain clear communication with stakeholders.
  • Problem Management & Root Cause Elimination
  • Lead root cause investigations for recurring or significant incidents.
  • Develop and drive corrective and preventive action plans across technology teams.
  • Track and manage problem records through resolution.
  • Identify systemic issues and technical debt that impact service stability.
  • Attend post-incident reviews and ensure lessons learned are translated into operational improvements.

Reliability & Operational Excellence

  • Continuously improve service availability, resiliency, and operational performance.
  • Partner with engineering teams to improve monitoring, alerting, telemetry, and observability capabilities.
  • Reduce alert noise through event correlation, automation, and process optimization.
  • Drive efforts to improve mean time to detect (MTTD) and mean time to restore service (MTTR).

Automation & Process Improvement
  • Identify opportunities to automate operational workflows and repetitive support activities.
  • Develop runbooks, playbooks, and operational procedures.
  • Collaborate with engineering teams to implement self-healing and automated recovery mechanisms.

Technical Leadership

  • Provide mentorship and guidance to engineers and operational support teams.
  • Influence technical decision-making related to operational readiness and supportability.
  • Act as a trusted advisor for service reliability, supportability, and operational risk management.
  • Participate in change reviews to ensure production readiness and minimize operational impact.

Qualifications

Bachelor's Degree and 8 years of experience in Technical work in Application Development, Server Administration, Information Security, or Engineering OR High School Diploma or GED and 12 years of experience in Technical work in Application Development, Server Administration, Information Security, or Engineering
  • 7+ years of experience in enterprise IT operations, infrastructure engineering, platform operations, or technical support environments.
  • Proven experience managing and resolving critical production incidents across multiple system architectures and infrastructures.
  • Strong background in problem management and root cause analysis methodologies.
  • Experience supporting large-scale enterprise environments.
  • Strong understanding of ITIL service management practices.

Technical Skills
  • Modern application architecture patterns and operations
  • Infrastructure & Platforms
  • Windows and Linux administration
  • Virtualization platforms (VMware, Hyper-V, etc.)
  • Storage and backup technologies
  • Containers and orchestration platforms (OpenShift)

Monitoring & Observability

  • Enterprise monitoring platforms
  • Log aggregation and analysis tools
  • Application performance monitoring (APM)
  • Operational dashboards and telemetry systems
  • Event management solutions

Automation
  • PowerShell, Python, Bash, or equivalent scripting
  • Workflow automation platforms
  • API integration and orchestration
  • Runbook development

Key Competencies
  • Exceptional troubleshooting and analytical skills.
  • Strong sense of ownership and accountability.
  • Ability to perform effectively during major incidents and high-pressure situations.
  • Excellent collaboration and stakeholder management skills.
  • Strong written and verbal communication.
  • Data-driven decision making.
  • Systems thinking and problem-solving mindset.
  • Continuous improvement orientation.

Benefits are an integral part of total rewards and First Citizens Bank is committed to providing a competitive, thoughtfully designed and quality benefits program to meet the needs of our associates. More information can be found at

$descr2

$descr3
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 10122789
  • Position Id: 34784
  • Posted 3 hours ago

Company Info

About First-Citizens Bank & Trust Company

Why First Citizens?

As America’s largest family-controlled bank, we have a unique legacy of strength, stability and vision.In recent years, we’ve grown to become a top 20 U.S. bank and a member of the Fortune 500.

With growth comes opportunities to learn and advance.

Technology is mission-critical here – from innovations that make banking easier to ensuring funds and data are secure, we constantly seek to create efficiencies for our associates and customers.

Our technology team makes a lasting difference by delivering solutions that empower our customers to take charge of their financial future.

Discover what it’s like to make better happen every day at a bank that’s committed to offering you a supportive, vibrant and flexible work environment.

Equal Opportunity Employer.

Member FDIC.

 

Market Position
Top 20 U.S. Bank

Total Assets
$221B

About_Company_One
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Raleigh, North Carolina

Today

Full-time

depends on experience

Remote or North Carolina

Today

Full-time

depends on experience

Raleigh, North Carolina

Today

Full-time

depends on experience

Remote or Raleigh, North Carolina

Today

Full-time

depends on experience

Search all similar jobs