Cloud Operations Engineer Virtual Data Center W2 role

Remote • Posted 1 hour ago • Updated 1 hour ago
Contract W2
12 Months
Remote
$40,000 - $60,000/yr
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Virtual Private Cloud

Summary

xperience : 8 + yrs

Remote in US ( preferably PST / EST zone )

Cloud Operations Engineer Virtual Data Center

AWS Virtual Data Center Operations

  • Operate and maintain AWS-based Virtual Data Center infrastructure supporting enterprise engineering workloads.
  • Monitor platform health, availability, performance, and capacity utilization.
  • Execute operational activities including provisioning, maintenance, patching, upgrades, and performance tuning.
  • Support incident, problem, and change management processes.

LSF Compute Administration

  • Manage and support IBM Spectrum LSF / batch compute environments.
  • Monitor job execution, queue performance, scheduler health, and resource utilization.
  • Troubleshoot failed, hung, or long-running jobs and coordinate with application teams for resolution.
  • Optimize compute resource allocation and workload scheduling.

NFS Storage Management

  • Administer and maintain high-performance NFS storage environments.
  • Monitor storage performance, capacity, and availability.
  • Troubleshoot file system, mount, and storage-related issues.
  • Support backup, recovery, and storage lifecycle activities.

Data Synchronization & Platform Reliability

  • Manage inter-site and cloud-based data synchronization processes.
  • Monitor data transfer jobs and investigate synchronization failures.
  • Support data integrity validation and operational troubleshooting.
  • Collaborate with infrastructure and application teams to resolve complex data flow issues.

Linux Operations & Troubleshooting

  • Perform Linux system administration and operational support.
  • Analyze system logs, performance metrics, and resource utilization.
  • Troubleshoot compute, storage, networking, and application-related incidents.
  • Develop operational automation using scripting and infrastructure tools.

Monitoring & Incident Management

  • Utilize monitoring and observability tools to proactively identify issues.
  • Participate in incident response, root cause analysis (RCA), and post-incident reviews.
  • Create and maintain runbooks, SOPs, and operational documentation.
  • Support on-call and production support activities as required.

Required Technical Skills

  • AWS Cloud Services (EC2, VPC, EBS, IAM, CloudWatch, S3)
  • Linux Administration (RHEL/CentOS/Ubuntu)
  • IBM Spectrum LSF / Batch Compute Administration
  • NFS Storage Administration
  • Shell Scripting (Bash/Python)
  • Networking fundamentals (TCP/IP, DNS, NFS, VPN)
  • Monitoring and Observability Tools
  • Incident Management and Root Cause Analysis
  • Security and Access Management concepts

Preferred Skills

  • Terraform and Infrastructure as Code (IaC)
  • Ansible Automation
  • Kubernetes and Docker
  • Grafana, ELK, Splunk, or similar monitoring platforms
  • ServiceNow ITSM
  • HPC (High Performance Computing) environments
  • AWS Solutions Architect or SysOps certifications

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 10500016
  • Position Id: 9091999
  • Posted 1 hour ago
Contact the job poster
Rakesh Vangala

Rakesh Vangala

Recruiter @ BURGEON IT SERVICES LLC
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Remote

Today

Easy Apply

Contract

Depends on Experience

Remote or Texas

Today

Easy Apply

Contract

$50.00 - 55.00

Remote or California

Today

Easy Apply

Full-time, Contract

$DOE

Remote

Today

Easy Apply

Contract

$90 - $100

Search all similar jobs