
593 results (19 new)


Stryde Consulting Services LLC
Santa Clara, California • Today
Easy Apply
Contract
Depends on Experience

Prudent Technologies and Consulting
Santa Clara, California • Today
Easy Apply
Contract
Depends on Experience
Aziro Technologies LLC
Hybrid in San Jose, California • 30+d ago
Easy Apply
Contract
Depends on Experience
Aziro Technologies LLC
Hybrid in San Jose, California • 16d ago
Easy Apply
Contract, Third Party
Depends on Experience










InfiCare
Hybrid in Sunnyvale, California • 6d ago
Easy Apply
Contract, Third Party
Depends on Experience



UnitedHealth Group
Remote or Eden Prairie, Minnesota • Today
Full-time
USD 134,600.00 - 230,800.00 per year





Robert Half
Remote or Bayport, Minnesota • Today
Easy Apply
Full-time
USD 120,000.00 - 140,000.00 per year





WHAT THIS CANDIDATE WILL BE DOING
· Administer large-scale Linux environments supporting AI training, inference, and HPC workloads.
· Own deep troubleshooting of OS, kernel, boot, package, firmware, driver, filesystem, service, and resource-consumption issues across bare-metal server fleets.
· Diagnose failures across BIOS, BMC, PXE, DHCP, DNS, NFS, local disk, RAID, NVMe, systemd, and GPU driver stacks.
· Build and maintain golden images, provisioning pipelines, configuration baselines, and post-deployment validation procedures.
· Partner with network, platform, storage, and validation teams to isolate cross-domain failures affecting cluster readiness or job execution.
· Investigate performance anomalies involving CPU, memory, NUMA, I/O, interrupts, process scheduling, and kernel tuning.
· Automate repeatable administration and remediation tasks with Bash and Python.
· Produce clear runbooks, failure signatures, and escalation criteria for recurring operational issues.
WHAT WE NEED TO SEE
· 7+ years delivering Linux administration in data center, cloud, AI, or HPC environments.
· Deep expertise with RHEL, Ubuntu, Rocky, or similar enterprise Linux distributions.
· Strong troubleshooting skill across boot flow, system logs, networking stack, authentication, service lifecycle, and hardware-software interaction.
· Experience with GPU servers, out-of-band management, firmware coordination, and cluster node bring-up.
· Hands-on knowledge of Ansible, PXE/iPXE, Kickstart, cloud-init, image lifecycle management, and configuration enforcement.
· Strong shell scripting and Python-based automation capability.
· Working knowledge of storage and network dependencies affecting Linux host health.
· Ability to operate independently in ambiguous, high-severity production situations.
Laiba Technology is one of the premier US based IT company. Our corporate office is in USA ,Dubai ,India. We serve government and commercial clients . We provide Software Development,Revenue Cycle Management, Staff Augmentation ,Software Support ,Corporate training etc. We have staffed thousands of contract and full time positions across multiple industries and skill sets. We have steadily grown through word of mouth referrals.
Laiba Technology is one of the reliable and fastest growing Software company serving client globally across the world. The demand for SEO/SMO/PPC, website design and development services and Software solutions worldwide has helped fuel the rapid expansion of We are in international market, where there is great requirement for businesses to increase their online publicity to spur financial growth.
We offer several innovative learning methods and delivery models to cater the unique requirements of a global customer base.
We also provide corporate training's on various cutting edge technologies. We have a team of Certified Trainers with minimum 10+ years of Industry background. Our Training courses are for individuals as well as for corporate. We also undertake customization of the courses as per client requirement.
🔢 Crunching numbers...
It looks like there aren't any Similar Jobs for this job yet.
Search all similar jobs