Platform Engineer / Site Reliability Engineer (SRE)

Remote • Posted 15 hours ago • Updated 2 hours ago
Full Time
On-site
Fitment

Dice Job Match Score™

🫥 Flibbertigibetting...

Job Details

Skills

  • Recruiting
  • Business Operations
  • Industry-specific
  • Innovation
  • Outsourcing
  • Sourcing
  • Health Care
  • IaaS
  • DevOps
  • Scalability
  • Provisioning
  • Operational Efficiency
  • Product QA
  • Continuous Improvement
  • Collaboration
  • Performance Tuning
  • Information Systems
  • Information Technology
  • Computer Science
  • Terraform
  • Ansible
  • Continuous Integration
  • Continuous Delivery
  • Jenkins
  • GitHub
  • Apache Maven
  • JFrog
  • Google Cloud
  • Google Cloud Platform
  • Computer Networking
  • Optimization
  • Cloud Computing
  • Amazon Web Services
  • Microsoft Azure
  • Management
  • Kubernetes
  • Scripting
  • Python
  • Java
  • C++
  • Perl
  • Ruby
  • SQL
  • Workflow
  • Incident Management
  • Root Cause Analysis
  • Reliability Engineering
  • Capacity Management
  • Agile
  • Sprint
  • UPS
  • Law

Summary

At NTT DATA, we know that with the right people on board, anything is possible. The quality, integrity, and commitment of our employees have been key factors in our company's growth and market presence. By hiring the best people and helping them grow both professionally and personally, we ensure a bright future for NTT DATA and for the people who work here.

For more than 25 years, NTT DATA Services have focused on impacting the core of your business operations with industry-leading outsourcing services and automation. With our industry-specific platforms, we deliver continuous value addition, and innovation that will improve your business outcomes. Outsourcing is not just a method of gaining a one-time cost advantage, but an effective strategy for gaining and maintaining competitive advantages when executed as part of an overall sourcing strategy.

NTT Data is seeking a Digital Engineering-Software Architect to join one of our leading organizations.
Project Duration:6 Months

Role Description:
In this role you will play a critical role in designing, building, and operating the infrastructure and automation that support our suite of cloud-native enterprise applications in the rapidly evolving healthcare technology landscape. You will be part of a collaborative engineering team focused on ensuring our platforms are scalable, secure, reliable, and efficient - enabling clinicians and patients to access technology that sustains and improves health outcomes.

Role Responsibilities:
  • Design, implement, and maintain scalable, secure, and highly available cloud infrastructure using Infrastructure-as-Code (IaC) tools and modern DevOps practices.
  • Build and optimize CI/CD pipelines using tools such as Jenkins, GitHub Actions, Maven, and JFrog to ensure fast, reliable, and repeatable software delivery.
  • Develop and manage containerized application environments with Kubernetes, ensuring optimal deployment, scalability, and service reliability.
  • Automate infrastructure provisioning, configuration, and deployment processes to improve operational efficiency and reduce manual interventions.
  • Design and implement comprehensive monitoring, alerting, and observability solutions to ensure system health, performance, and reliability.
  • Collaborate closely with development, product, QA, and security teams to design robust platform solutions aligned with business and technical requirements.
  • Participate actively in Agile ceremonies such as daily stand-ups, sprint planning, demos, and retrospectives, driving continuous improvement and team collaboration.
  • Contribute to operational support processes, including incident response, root cause analysis, capacity planning, and performance optimization for large-scale distributed systems.

Required Experience
  • Bachelor's degree in Information Systems, Information Technology, Computer Science, Engineering, or a related field - or equivalent work experience.
  • Minimum of 5 years of hands-on experience with tools such as Terraform and Ansible for building and managing infrastructure as code, and CI/CD automation tools like Jenkins, GitHub, Maven, and JFrog.
  • Proven hands-on experience deploying and managing infrastructure and services on Google Cloud Platform (Google Cloud Platform)
  • Strong knowledge of Google Cloud Platform networking, IAM, security, cost optimization, and core services (e.g., Cloud Run, Pub/Sub, GKE, Firestore) is essential.
  • Minimum 2 years' experience with AWS or Azure is a plus but not a substitute.
  • Strong experience deploying, scaling, and managing containerized applications using Kubernetes, including service mesh, auto-scaling, and rolling update strategies.
  • Proficiency in one or more programming or scripting languages such as Python, Go, Java, C++, Perl, Ruby, or SQL, with the ability to automate workflows and optimize infrastructure operations.
  • Experience designing and implementing operational support models, including incident response, root cause analysis, monitoring, and alerting strategies.
  • Demonstrated experience with reliability engineering practices such as capacity planning, fault tolerance, and resilience design.
  • Experience working in Agile environments with familiarity in sprints, daily stand-ups, planning sessions, and retrospectives.
Required Experience:
  • Bachelor's degree in Information Systems, Information Technology, Computer Science, Engineering, or a related field - or equivalent work experience.
  • Minimum of 5 years of hands-on experience with tools such as Terraform and Ansible for building and managing infrastructure as code, and CI/CD automation tools like Jenkins, GitHub, Maven, and JFrog.
  • Proven hands-on experience deploying and managing infrastructure and services on Google Cloud Platform (Google Cloud Platform)
  • Strong knowledge of Google Cloud Platform networking, IAM, security, cost optimization, and core services (e.g., Cloud Run, Pub/Sub, GKE, Firestore) is essential.
  • Minimum 2 years' experience with AWS or Azure is a plus but not a substitute.
  • Strong experience deploying, scaling, and managing containerized applications using Kubernetes, including service mesh, auto-scaling, and rolling update strategies.
  • Proficiency in one or more programming or scripting languages such as Python, Go, Java, C++, Perl, Ruby, or SQL, with the ability to automate workflows and optimize infrastructure operations.
  • Experience designing and implementing operational support models, including incident response, root cause analysis, monitoring, and alerting strategies.
  • Demonstrated experience with reliability engineering practices such as capacity planning, fault tolerance, and resilience design.
  • Experience working in Agile environments with familiarity in sprints, daily stand-ups, planning sessions, and retrospectives.
Required:
  • Bachelor's degree in Information Systems, Information Technology, Computer Science, Engineering, or a related field - or equivalent work experience.
  • Minimum of 5 years of hands-on experience with tools such as Terraform and Ansible for building and managing infrastructure as code, and CI/CD automation tools like Jenkins, GitHub, Maven, and JFrog.
  • Proven hands-on experience deploying and managing infrastructure and services on Google Cloud Platform (Google Cloud Platform)
  • Strong knowledge of Google Cloud Platform networking, IAM, security, cost optimization, and core services (e.g., Cloud Run, Pub/Sub, GKE, Firestore) is essential.
  • Minimum 2 years' experience with AWS or Azure is a plus but not a substitute.
  • Strong experience deploying, scaling, and managing containerized applications using Kubernetes, including service mesh, auto-scaling, and rolling update strategies.
  • Proficiency in one or more programming or scripting languages such as Python, Go, Java, C++, Perl, Ruby, or SQL, with the ability to automate workflows and optimize infrastructure operations.
  • Experience designing and implementing operational support models, including incident response, root cause analysis, monitoring, and alerting strategies.
  • Demonstrated experience with reliability engineering practices such as capacity planning, fault tolerance, and resilience design.
  • Experience working in Agile environments with familiarity in sprints, daily stand-ups, planning sessions, and retrospectives.

Where required by law, NTT DATA provides a reasonable range of compensation for specific roles. The starting hourly range for this remote role is $92.00-$97.00/hour. This range reflects the minimum and maximum target compensation for the position across all US locations. Actual compensation will depend on several factors, including the candidate's actual work location, relevant experience, technical skills, and other qualifications.

NTT DATA Services is an equal opportunity employer and considers all applicants without regarding to race, color, religion, citizenship, national origin, ancestry, age, sex, sexual orientation, gender identity, genetic information, physical or mental disability, veteran or marital status, or any other characteristic protected by law. We are committed to creating a diverse and inclusive environment for all employees. If you need assistance or an accommodation due to a disability, please inform your recruiter so that we may connect you with the appropriate team.
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 80166078
  • Position Id: ab30273ec523a5d57b19ccacd9b2c181
  • Posted 15 hours ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Remote or St. Louis, Missouri

Today

Contract

USD95 - USD105

Remote or New York, New York

Today

Full-time

USD 160,000.00 - 200,000.00 per year

Remote

Today

Full-time

Remote

Today

Full-time

Search all similar jobs