
1 results (1 new)




NVIDIA AI Infrastructure & Kubernetes Platform Engineer (DGX Systems) Department: Infrastructure Engineering
Location / Remote Policy: Remote Role Type: Contract 6-month initial engagement
About Our Client
Our client is a technology and professional services firm founded in 2015 on the strength of its founders' 30 years of industry experience. They set out to bridge a gap in professional services to be a true partner rather than just a vendor delivering expert guidance, innovative solutions, and personalized service at a cost-effective rate. Their mission is to empower businesses to succeed in the digital era, harnessing technology to drive transformation, innovation, and growth. Guided by a "make a customer, not a sale" philosophy, they lead with a customer-first approach and a team of senior-level engineers sourced from the world's leading OEMs, including AWS, Palo Alto Networks, Cisco, and Microsoft.
Job Description
Our client is seeking a highly skilled AI Infrastructure & Kubernetes Platform Engineer with a proven track record deploying and managing NVIDIA DGX-based AI clusters, orchestrating containerized AI workloads on Kubernetes, and ensuring secure, high-throughput operations across InfiniBand-powered networks. You'll bring a strong certification foundation across both Kubernetes (CKA, CKAD, CKS) and NVIDIA's AI infrastructure stack, paired with hands-on experience across DGX, BlueField, and high-speed networking.
This role is central to supporting AI/ML infrastructure at scale enabling efficient training and inference for complex models and integrating NVIDIA's compute, storage, and fabric solutions with modern DevOps practices. Day to day, you'll own DGX cluster operations, architect GPU-accelerated Kubernetes platforms, tune InfiniBand fabric for throughput, and harden the environment through DPU-enhanced security.
You'll work at the intersection of infrastructure, DevOps, and AI/ML, keeping the platform reliable and cost-efficient for the teams that depend on it. The ideal candidate is deeply hands-on, obsessed with performance and security, and energized by operating some of the most advanced AI compute available.
Duties and Responsibilities
AI Infrastructure Operations
Kubernetes Platform Engineering
High-Performance Networking & DPUs
Security & Compliance
Monitoring, Telemetry & Optimization
Required Experience/Skills
Certifications
Hands-On Expertise
Technical Skills
Nice-to-Haves
Education
Bachelor's degree in Computer Science, Engineering, or a related field or equivalent hands-on experience.
One IT Corp was founded with the vision of providing advanced IT solutions to our growing customer base. The company adopts high commercial values of transparency and integrity in the exercise of its activities. One IT Corp provides exemplary services through innovation, technical expertise, and fair business practices.
Over the years, One IT Corp has defined, designed and developed business solutions based on technology and processes that help its customers differentiate themselves from others. Focusing on one of the objectives and encouragement for Business Intelligence (BI) tools, application development, systems integration, software development, testing, recruitment, and training in the company, who have created milestones throughout the process.
One IT Corp is proud to build long-term relationships with its customers. We proudly emphasize that our only motto is the delight of the customer, achieved through exemplary service and respect for values and standards.


🔢 Crunching numbers...
It looks like there aren't any Similar Jobs for this job yet.
Search all similar jobs