Dev Ops Engineers - Intermediate

Tampa, FL, US • Posted 2 hours ago • Updated 2 hours ago
Contract Independent
On-site
Company Branding Image
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Operational Excellence
  • Production Support
  • High Availability
  • Storage
  • Incident Management
  • Root Cause Analysis
  • Workflow
  • Performance Monitoring
  • Collaboration
  • Operational Efficiency
  • Scripting
  • Python
  • Bash
  • Durable Skills
  • DevOps
  • Production Engineering
  • Reliability Engineering
  • Scalability
  • Orchestration
  • Computer Networking
  • Storage Management
  • Docker
  • Management
  • Kubernetes
  • Load Balancing
  • Continuous Integration
  • Continuous Delivery
  • Jenkins
  • Bitbucket
  • Infrastructure Architecture
  • Terraform
  • Provisioning
  • Linux
  • Server Administration
  • Cloud Computing
  • Amazon Web Services
  • Dynatrace
  • Splunk
  • Grafana
  • Dashboard
  • Performance Metrics
  • Privacy
  • Marketing

Summary

Location: Tampa, FL Description:
Job Title: Site Reliability Engineer (Kubernetes & Observability)
Location: Tampa, FL
Working Model: Fully onsite
Duration: 6 months Right to Hire


Job Description:


We are seeking a highly skilled Site Reliability Engineer (SRE) with deep expertise in Kubernetes infrastructure, observability platforms, and automation to support and improve large-scale production systems. This role requires a hands-on engineer who can build, scale, and maintain Kubernetes environments while driving reliability, monitoring, and operational excellence. You will work closely with development, infrastructure, production support, and platform engineering teams to design scalable solutions, automate operational processes, and ensure high availability across mission-critical applications.

Key Responsibilities
  • Build and administer Kubernetes clusters across multiple environments
  • Design and implement containerized solutions using Kubernetes and Docker
  • Troubleshoot production issues involving networking, storage, cluster performance, and application reliability
  • Support incident management, root cause analysis, and proactive reliability improvements
  • Develop Infrastructure as Code (IaC) solutions using Terraform
  • Build and maintain CI/CD pipelines using Jenkins, Bitbucket, and GitOps methodologies
  • Automate infrastructure provisioning, deployments, and operational workflows
  • Implement observability and monitoring solutions using Dynatrace, Splunk, Prometheus, Grafana, and OpenTelemetry
  • Create dashboards, alerts, and telemetry solutions to improve application visibility and performance monitoring
  • Collaborate with engineering teams to improve system scalability, availability, and operational efficiency
  • Develop automation tools and scripts using Python, Go, or Bash
  • Support platform modernization initiatives and help establish best practices for SRE and Cloud-Native operations


Required Skills & Experience

Core Skills
  • 6+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Production Engineering
  • Strong hands-on Kubernetes experience with cluster administration and platform build-out
  • Extensive experience supporting production environments and resolving system-level issues
  • Proven expertise in Infrastructure as Code (IaC) using Terraform
  • Strong understanding of reliability engineering principles, system uptime, scalability, and availability


Kubernetes & Container Platforms
  • Hands-on experience building Kubernetes environments from the ground up
  • Expertise with Kubernetes cluster administration, workload orchestration, networking, and storage management
  • Strong understanding of containerization technologies including Docker
  • Experience managing Kubernetes deployments in enterprise environments
  • Knowledge of service discovery, ingress controllers, load balancing, and autoscaling


CI/CD & Tools
  • Experience building and maintaining CI/CD pipelines using Jenkins and Bitbucket
  • Familiarity with GitOps deployment methodologies
  • Experience automating infrastructure provisioning and deployments
  • Ability to create self-service automation solutions that reduce manual intervention


Infrastructure & Cloud
  • Strong understanding of infrastructure architecture and distributed systems
  • Expertise with Terraform and Infrastructure as Code practices
  • Experience provisioning compute instances and automating infrastructure deployments
  • Understanding of Linux environments and server administration
  • Cloud platform experience is beneficial, but AWS expertise is not required


Observability & Monitoring
  • Hands-on experience with Dynatrace, Splunk, Prometheus, Grafana, or similar monitoring platforms
  • Experience building dashboards, alerts, telemetry, and application monitoring solutions
  • Understanding of OpenTelemetry concepts and distributed tracing
Experience troubleshooting applications using performance metrics, logs, and telemetry data
By providing your phone number, you consent to: (1) receive automated text messages and calls from the Judge Group, Inc. and its affiliates (collectively "Judge") to such phone number regarding job opportunities, your job application, and for other related purposes. Message & data rates apply and message frequency may vary. Consistent with Judge's Privacy Policy, information obtained from your consent will not be shared with third parties for marketing/promotional purposes. Reply STOP to opt out of receiving telephone calls and text messages from Judge and HELP for help.
Contact:
This job and many more are available through The Judge Group. Please apply with us today!
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: cxjudgpa
  • Position Id: 1148524
  • Posted 2 hours ago

Company Info

About Judge Group, Inc.

The Judge Group, is a leading professional services firm specializing in talent, technology, and learning solutions. We consult, staff, train, and solve. Through our work we make people and organizations better.

Our services are successfully delivered through a network of more than 30 offices across the United States, Canada, and India. The Judge Group is proud to partner with the best and brightest companies in business today, including over 60 of the Fortune 100. We serve organizations in financial services, healthcare, life sciences, insurance, government (including aerospace and defense), manufacturing, and technology and telecommunications.

About_Company_OneAbout_Company_Two
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Charlotte, North Carolina

Today

Contract

USD 53.00 - 57.00 per hour

Charlotte, North Carolina

Today

Contract

USD 64.00 - 69.00 per hour

Columbus, Ohio

Today

Contract

USD 64.00 - 69.00 per hour

Irving, Texas

Today

Contract

USD 64.00 - 69.00 per hour

Search all similar jobs