Platform Engineer

St.Louis, MO, US • Posted 17 hours ago • Updated 1 hour ago
Contract Independent
Contract W2
Contract Corp To Corp
On-site
$DOE
Fitment

Dice Job Match Score™

🎯 Assessing qualifications...

Job Details

Skills

  • Linux
  • Kubernetes
  • Terraform
  • AI/ML
  • NVIDIA
  • GPUs and NVIDIA
  • GPUs

Summary

Role: Platform Engineer

Duration: 1 month

Location: St. Louis, MO

Role Overview
Highly skilled AI Infrastructure Engineer to design, build, and operate scalable GPU-enabled Kubernetes platforms for AI/ML workloads.

Must have skills: Kubernetes, Linux, Terraform, GPUs and NVIDIA stack

Required Qualifications

  • 3 8+ years' experience
  • Strong Kubernetes knowledge
  • Experience with GPUs and NVIDIA stack
  • Linux (Ubuntu) expertise
  • Experience with Terraform

Preferred Qualifications

  • Longhorn or Ceph experience
  • Canonical ecosystem (MAAS, Juju)
  • AI/ML tools like Kubeflow
  • Certifications (CKA, NVIDIA)

Soft Skills

  • Strong problem-solving and troubleshooting mindset
  • Ability to collaborate with cross-functional teams (ML engineers, data scientists)
  • Clear communication and documentation skills
  • Passion for automation and platform scalability

Key Responsibilities

  • Design and manage Kubernetes clusters
  • Build GPU-enabled infrastructure
  • Deploy Longhorn storage
  • Automate infrastructure using Terraform
  • Monitor systems using Prometheus and Grafana
  • Knowledge Transfer & Client Enablement
  • Provide structured knowledge transfer (KT) sessions to client teams on all core platform components, including:
    • Kubernetes architecture, operations, and troubleshooting
    • GPU infrastructure (NVIDIA stack, scheduling, resource optimization)
    • Longhorn storage management and performance tuning
    • Canonical ecosystem tools (MAAS, Juju, Charmed Kubernetes)
  • Develop and deliver technical documentation, runbooks, and training materials to support ongoing operations
  • Conduct hands-on workshops and guided sessions to enable client teams to independently manage and scale the platform
  • Act as a technical advisor, helping client stakeholders understand best practices in:
  • Cloud-native infrastructure
    o AI/ML platform operations
    o Reliability, performance, and cost optimization
  • Ensure smooth handoff of production systems with full operational readiness and support knowledge
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 91088813
  • Position Id: 2026-1918
  • Posted 17 hours ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

St. Louis, Missouri

5d ago

Easy Apply

Contract

70 - 80

Texas

Today

Third Party, Contract

Remote

10d ago

Easy Apply

Contract

50 - 55

Remote

Today

Easy Apply

Contract

$55 - $65 per hour

Search all similar jobs