Cloud Infrastructure Site Reliability Engineer

Berkeley Heights, NJ, US • Posted 2 hours ago • Updated 2 hours ago
Contract Independent
On-site
USD $70.00 - 80.00 per hour
Company Branding Image
Fitment

Dice Job Match Score™

📊 Calculating match score...

Job Details

Skills

  • Continuous Improvement
  • IaaS
  • Incident Management
  • Ansible
  • Collaboration
  • Software Engineering
  • Knowledge Sharing
  • Computer Science
  • Software Development
  • Python
  • Java
  • C++
  • Amazon Web Services
  • Google Cloud
  • Google Cloud Platform
  • Microsoft Azure
  • Network Security
  • Storage
  • Data Management
  • Linux
  • Computer Networking
  • File Systems
  • Dashboard
  • Continuous Integration and Development
  • Continuous Integration
  • Continuous Delivery
  • Automated Testing
  • Provisioning
  • Management
  • Root Cause Analysis
  • Service Level
  • Reliability Engineering
  • Terraform
  • Dynatrace
  • Financial Services
  • Cloud Computing
  • Conflict Resolution
  • Problem Solving
  • Communication
  • Mentorship
  • DevOps
  • Privacy
  • Marketing

Summary

Location: Berkeley Heights, NJ Salary: $70.00 USD Hourly - $80.00 USD Hourly Description:
Job Title: Cloud Infrastructure Site Reliability Engineer

Location: Berkeley Heights, NJ / Alpharetta, GA (Onsite 5 Days)

Duration: Contract To Hire

Job Description:

Position Summary:

As a Cloud Infrastructure Site Reliability Engineer (SRE) with expertise in multiple public cloud service provider platforms, you will be responsible for operating infrastructure solutions, following the principles and practices pioneered by Google's SRE model. Your work will ensure our cloud services meet uptime, reliability, and performance targets, and you will drive automation and continuous improvement across our production environments. This role will involve collaborating with cross-functional teams to enhance our cloud reliability posture and streamline processes through automation.

Key Responsibilities:
  • Design, build, and maintain highly available, scalable, and secure cloud infrastructure on platforms such as AWS, Google Cloud Platform, or Azure.
  • Develop and implement automation for provisioning, monitoring, scaling, and incident response using Infrastructure-as-Code tools (e.g., Terraform, CloudFormation, Ansible).
  • Monitor system reliability, capacity, and performance; proactively detect and address issues before they impact users.
  • Respond to production incidents, participate in on-call rotations, and lead post-incident reviews to drive root cause analysis and reliability improvements.
  • Collaborate with software engineering and security teams to ensure new services and features are production-ready and meet reliability standards.
  • Build and maintain tools for deployment, monitoring, and operations; automate manual processes to reduce toil.
  • Document operational processes and system architectures to ensure knowledge sharing and repeatability.
  • Continuously evaluate and implement new technologies to improve system reliability, security, and efficiency.


Qualifications:
  • Bachelor's degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience.
  • 3+ years of experience in software development with proficiency in at least one programming language (e.g., Python, Go, Java, C++).
  • Experience administrating cloud platforms (AWS, Google Cloud Platform, Azure), including networking, security, containerization, storage, data management, and serverless technologies.
  • Solid understanding of Linux systems, networking fundamentals, virtualized, and distributed systems, file systems, system processes and configurations.
  • Deep understanding of observability (monitoring, alerting, and logging) tools in cloud environments. Ability to set up and maintain monitoring dashboards, alerts, and logs.
  • Familiarity with Continuous Integration/Continuous Deployment (CI/CD) tools for automated testing, deployments, provisioning, and observability.
  • Ability to manage and respond to incidents, perform root cause analysis, and implement post-mortem reviews.
  • Understanding of setting, monitoring, and maintaining Service-Level Objectives (SLOs) and Service-Level Agreements (SLAs) for system reliability.


Needs experience with Terraform and Dynatrace
  • Additional Qualifications a Plus: Experience working with enterprise-scale financial services or other regulated industries
  • 5+ years of experience in SRE, DevOps, infrastructure, or cloud engineering roles, preferably supporting large-scale, distributed systems.
  • Excellent problem-solving, troubleshooting, and communication skills.
  • Experience leading technical projects or mentoring junior engineers.
  • Certifications: Certified Engineer, DevOps, SRE, CSREF

By providing your phone number, you consent to: (1) receive automated text messages and calls from the Judge Group, Inc. and its affiliates (collectively "Judge") to such phone number regarding job opportunities, your job application, and for other related purposes. Message & data rates apply and message frequency may vary. Consistent with Judge's Privacy Policy, information obtained from your consent will not be shared with third parties for marketing/promotional purposes. Reply STOP to opt out of receiving telephone calls and text messages from Judge and HELP for help.
Contact:
This job and many more are available through The Judge Group. Please apply with us today!
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: cxjudgpa
  • Position Id: 1146611
  • Posted 2 hours ago

Company Info

About Judge Group, Inc.

The Judge Group, is a leading professional services firm specializing in talent, technology, and learning solutions. We consult, staff, train, and solve. Through our work we make people and organizations better.

Our services are successfully delivered through a network of more than 30 offices across the United States, Canada, and India. The Judge Group is proud to partner with the best and brightest companies in business today, including over 60 of the Fortune 100. We serve organizations in financial services, healthcare, life sciences, insurance, government (including aerospace and defense), manufacturing, and technology and telecommunications.

About_Company_OneAbout_Company_Two
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Berkeley Heights, New Jersey

Today

Contract

USD 70.00 - 80.00 per hour

Berkeley Heights, New Jersey

Today

Contract

USD 75.00 - 78.00 per hour

Burlington, New Jersey

Today

Full-time

USD 130,000.00 - 150,000.00 per year

Philadelphia, Pennsylvania

Today

Contract

USD 50.00 - 55.00 per hour

Search all similar jobs