Role Senior DevOps Engineer (DevOps / Site Reliability Engineering)
Project Video Engineering / Production Reliability & Observability - Senior DevOps Backfill
Office Location - Greenwood Village, CO
Onsite Requirements - 4 days a week onsite
Work Status Client is looking for either holder
The client is seeking a Senior DevOps Engineer to support its Video Engineering
environment. This is a senior-level position focused on production reliability, observability,
troubleshooting, and operational support for video platforms and services. The ideal candidate will bring
strong DevOps/SRE experience, hands-on Splunk expertise, and prior experience supporting video
technology or video delivery environments.
Development operations (DevOps) engineers are responsible for the production and ongoing
maintenance of a website platform. They also manage cloud infrastructure and system administration and
work with teams to identify and repair issues on an as-needed basis, so strong communication skills are
important in this position. They are generally expected to work well under pressure with tight deadlines for
certain tasks, and a proactive demeanor and friendly disposition are also helpful. DevOps engineers may
work with junior and senior engineers, project managers, and executives, as well as administrative
assistants, executive assistants, and a receptionist. Hours can be flexible, though they typically work
during regular weekly business hours, and they are not usually responsible for customer/client interaction
or supervising junior employees.
Key Qualifications
Video Experience:
Prior experience supporting video platforms, video engineering, streaming, video delivery, or a
comparable video technology environment is required.
Splunk:
Strong hands-on Splunk experience for production monitoring, log analysis, troubleshooting, alerting, and
root-cause investigation is required.
Cloud / DevOps:
Strong experience with AWS/cloud environments, Kubernetes, containers, microservices, and Linux-
based systems.
Observability / SRE:
Experience supporting highly available production systems using observability and monitoring tools such
as Splunk, Datadog, OpenSearch, or comparable platforms.
Production Troubleshooting:
Demonstrated ability to investigate production issues, analyze telemetry and logs, identify root causes,
and partner with engineering and operations teams through resolution.
Automation / Scripting:
Experience with scripting and automation in support of production operations, monitoring, deployments,
and/or incident response.
Responsibilities
Support the reliability, availability, and performance of production video platforms and services.
Monitor application and infrastructure health and proactively identify production-impacting issues.
Use Splunk and other observability tools to analyze logs, troubleshoot incidents, and perform root-cause
analysis.
Support containerized and Kubernetes-based workloads in AWS/cloud environments.
Partner with application developers, video engineering, operations, and other technical teams to resolve
complex production issues.
Improve monitoring, alerting, automation, and operational processes to increase platform reliability.
Participate in incident response and support customer-impacting escalations as needed.
Additional Skills
Strong analytical and problem-solving skills, with the ability to work independently in a senior-level
capacity.
Excellent communication and cross-functional collaboration skills.
Ability to operate effectively in a fast-paced production environment and prioritize issues based on
customer and business impact.