Machine Learning Engineer ONLY W2

Remote • Posted 1 hour ago • Updated 1 hour ago
Contract W2
Remote
Depends on Experience
Company Branding Image
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Machine Learning (ML)
  • Kubernetes
  • SRE
  • Site Reliability Engineering
  • Amazon Web Services

Summary

Job Summary:

We are seeking a skilled Machine Learning Engineer with a strong Site Reliability Engineering (SRE) mindset to join our team. The ideal candidate will have hands-on experience maintaining applications on both Windows and Linux environments, managing on-premises servers, and working with Kubernetes clusters. This role requires solid Python programming skills, a good understanding of machine learning concepts, and practical knowledge of ML model deployment, monitoring, and debugging.

Key Responsibilities:

  • Maintain and support machine learning applications running on Windows and Linux servers in on-premises environments.
  • Manage and troubleshoot Kubernetes clusters hosting ML workloads.
  • Collaborate with data scientists and engineers to deploy machine learning models reliably and efficiently.
  • Implement and maintain monitoring and alerting solutions using DataDog to ensure system health and performance.
  • Debug and resolve issues in production environments using Python and monitoring tools.
  • Automate operational tasks to improve system reliability and scalability.
  • Ensure best practices in security, performance, and availability for ML applications.
  • Document system architecture, deployment processes, and troubleshooting guides.

Required Qualifications:

  • Proven experience working with Windows and Linux operating systems in production environments.
  • Hands-on experience managing on-premises servers and Kubernetes clusters and Docker containers
  • Strong proficiency in Python programming.
  • Solid understanding of machine learning concepts and workflows.
  • Experience with machine learning model deployment and lifecycle management.
  • Familiarity with monitoring and debugging tools, e.g. DataDog.
  • Ability to troubleshoot complex issues in distributed systems.
  • Experience with CI/CD pipelines for ML applications.
  • Familiarity with AWS cloud platforms
  • Background in Site Reliability Engineering or DevOps practices.
  • Strong problem-solving skills and attention to detail.
  • Excellent communication and collaboration skills

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 10426227
  • Position Id: WAS218
  • Posted 1 hour ago

Company Info

About Aziro Technologies LLC

Aziro (formerly MSys Technologies and pronounced as "Ah-zee-roh") is an AI-native product engineering company driving innovation-led transformation for global enterprises, high-growth ISVs, and AI-first pioneers.

Contact the job poster
WA

Wasim Ahmed

Recruiter @ Aziro Technologies LLC
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Remote

•

Today

Easy Apply

Contract

$60 - $80

Remote

•

Today

Easy Apply

Contract

Depends on Experience

Search all similar jobs