Cloud Operations Analyst II

Hybrid in Farmington Hills, MI, US • Posted 2 days ago • Updated 2 days ago
Full Time
No Travel Required
Hybrid
$110,000 - $120,000/yr
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • AWS
  • Azure
  • OCI
  • scripting
  • AWS SysOps
  • Datadog
  • PagerDuty
  • Windows
  • Linux
  • Certified Datadog

Summary

Cloud Operations Analyst II
Hybrid - 3 days a week

Farmington Hills MI (Detroit suburb) 


Summary Statement

The Cloud Operations Analyst II is responsible for operating and enhancing the performance, availability, and reliability of cloud and on-premises infrastructure. This individual consults on more complex observability scenarios, streamlining server and batch operations, and contributing to the efficient management of data center resources.  Consults with cross-functional teams to identify opportunities for process improvements, implement best practices, and support critical business operations. 

 

Essential Tasks/Major Duties

·         Develop, implement, and maintain observability tools to monitor cloud and on-premises systems.  

·         Create dashboards, alerts, and reports to track system health, performance, and availability.  

·         Proactively leverage observability tools and identify opportunities.  

·         Analyze metrics and logs to identify trends, prevent potential issues, and optimize system performance.  

·         Collaborate with FinOps teams to monitor resource utilization and ensure cost-effective operations across cloud environments. 

·         Support the lifecycle of cloud and on-premises servers, including provisioning, patching, configuration, and decommissioning. 

·         Troubleshoot and resolve server-related issues, ensuring minimal downtime. Implement and enforce server security policies and compliance requirements. 

·         Schedule, monitor, and manage batch processes to ensure timely execution of critical tasks. 

·         Identify and resolve batch failures or delays, coordinating with relevant teams to ensure smooth operations.  

·         Optimize batch jobs for improved performance and resource utilization. 

·         Manage on-site and remote data center operations, ensuring proper functioning of hardware, power, cooling, and network infrastructure.  

·         Coordinate with vendors and service providers for hardware maintenance, replacements, and upgrades.  

·         Maintain accurate inventory of data center assets and ensure compliance with organizational standards. 

·         Participate in on-call rotations to address system incidents and outages promptly.  

·         Conduct root cause analysis and implement solutions to prevent recurrence of issues. Document and communicate incident resolution processes to relevant stakeholders. 

·         Work closely with cross-functional teams, including DevOps, Networking, and Application Development, to implement and maintain system integrations.  

·         Maintain and create comprehensive documentation for configurations, processes, and incident resolutions.  

·         Provide training and support to team members and other departments. 

 

Knowledge, Skills & Abilities

·         Bachelor’s degree in computer science, Information Technology, or a related field, or equivalent experience. 

·         3 years of experience working with monitoring and observability tools (e.g., Datadog, PagerDuty). 

·         Certified Datadog Fundamentals or equivalent experience required. 

·         Certified PagerDuty Administrator or equivalent experience required. 

·         3 years of experience in cloud operations or server management roles. 

·         Certified AWS SysOps Administrator or equivalent experience required. 

·         3 years of progressive server administration experience (Windows, Linux). 

·         3 years of experience in designing, implementing, and managing IT workload automation solutions to optimize scheduling, orchestration, and execution of enterprise workflows across on-prem and cloud environments.

·         Experience leveraging artificial intelligence to drive innovation and solve complex problems. Demonstrated ability to utilize AI-driven solutions that optimize processes, enhance decision-making, or create transformative business outcomes.

·         3 years working with cloud platforms (AWS, Azure, OCI).

·         Certified AWS Cloud Practitioner or equivalent experience required.

·         Strong experience with data center infrastructure and knowledge of best practices.  

·         Proficiency in scripting and automation tools (Python, Bash, PowerShell). 

·         Strong understanding of networking and identity management in cloud environments. 

·         Working knowledge of security best practices and compliance standards.

·         Working knowledge of agile methodologies. 

·         Excellent troubleshooting, problem-solving, and communication skills. 

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 91163673
  • Position Id: 9097584
  • Posted 2 days ago
Contact the job poster
PK

Prudhvi Krishna

Team Lead @ Trinite Consulting Group LLC
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Southfield, Michigan

•

9d ago

Easy Apply

Full-time

USD 110,000.00 - 145,000.00 per year

Livonia, Michigan

•

Today

Full-time

Remote

•

Today

Full-time

No location provided

•

Today

Full-time

Search all similar jobs