Site Reliability Engineer

Remote • Posted 3 hours ago • Updated 3 hours ago
Contract Corp To Corp
Contract W2
5 Months
Remote
$55 - $58/hr
Fitment

Dice Job Match Score™

🛠️ Calibrating flux capacitors...

Job Details

Skills

  • SITE RELIABILITY ENGINEER
  • SRE
  • PRODUCTION ENGINEER
  • PLATFORM ENGINEER
  • AWS
  • EKS
  • KUBERNETES
  • K8S
  • HELM
  • TERRAFORM
  • GITOPS
  • ARGO CD
  • ARGOCD
  • FLUXCD
  • FLUX CD

Summary

HonorVet Technologies is a Service Disable Veteran-Owned IT staffing firm, ISO 9001 and ISO 27001 certified, working with federal agencies, state governments, and Fortune 500 enterprise clients across the US. What makes us different isn''t a tagline; it''s the way we work. We don''t forward resumes and hope for the best. We take the time to understand where a professional like you is headed and only reach out when we genuinely believe there''s a fit worth exploring

Position Title: Site Reliability Engineer
Location: Remote
Duration: 5 Months

Overview:

We are seeking a highly motivated Site Reliability Engineer (SRE) with a strong operational focus to join our growing team. In this role, you will play a vital role in ensuring the smooth operation and performance of our critical infrastructure and services. You''ll work cross-functionally to create alignment and deliver results alongside builders who have helped to shape the success of companies such as Google, Okta, AWS, Snowflake.

What you will do in this role:
  • Deploy software for Cloud Prem and SAAS customers.
  • Respond to and diagnose system incidents in a timely and efficient manner, minimizing downtime and impact on users.
  • Collaborate with other engineers to establish root causes and implement effective resolutions.
  • Continuously improve incident response processes and documentation for future occurrences.
  • Proactively monitor and maintain the health and performance of our infrastructure and services.
  • Perform routine administrative tasks such as system configuration, user management, and data backups.
  • Identify and implement operational improvements to ensure ongoing system reliability and efficiency.
  • Develop and implement scripts and automated solutions to streamline operational tasks and reduce manual workload.
  • Participate in the on-call rotation to address critical incidents outside of regular business hours.
  • Ensure effective handoff between on-call engineers and document post-incident information for future reference.
  • Document processes for support and create, maintain and execute run-books for identified situations
  • Provide tier 2/3 technical support to customers experiencing platform issues or requiring advanced troubleshooting
  • Work directly with customer technical teams to resolve complex deployment, configuration, and integration challenges
  • Conduct technical onboarding sessions and provide guidance on best practices for customer implementations
  • Collaborate with customer success teams to ensure smooth customer experiences and rapid issue resolution
  • Create and maintain customer-facing technical documentation, troubleshooting guides, and knowledge base articles
  • Escalate customer feedback and feature requests to product and engineering teams
  • Participate in customer calls and technical discussions to provide expert-level platform guidance
  • Track and analyze customer support metrics to identify trends and areas for improvement
What you will need to be successful in this role:

Education:
  • BS degree in Computer Science or related field
Experience:
  • 3+ years of experience in Site Reliability Engineering
  • 2+ years experience working with cloud platform and cloud automation tools especially in AWS
  • Strong experience with Kubernetes, Helm, Linux, AWS networking(VPC) and Terraform
  • Experience with the GitOps model for deployment
  • Familiarity with distributed version control
  • Experience with monitoring and alerting tools (e.g., Prometheus, Grafana).
  • Bazel and CueLang experience a plus
  • Understanding of software configuration best practices
  • Ability to wear multiple hats in a fast-paced environment
  • Hands-on, “can do” attitude and a bias for action
  • Low ego and high intellectual curiosity
  • Comfortable working across time zones to support global customer base
  • Excellent communication skills with ability to explain technical concepts to both technical and non-technical audiences
  • Strong customer service orientation with patience and empathy when working with frustrated customers
 

     
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 90941473
  • Position Id: 26-20786
  • Posted 3 hours ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Remote

Today

Full-time

USD 114,000.00 - 148,000.00 per year

Remote

Today

Full-time

USD 147,000.00 - 168,000.00 per year

Remote

Today

Full-time

USD 125,000.00 - 145,000.00 per year

Remote or San Francisco, California

Today

Full-time

USD 114,297.00 - 235,319.00 per year

Search all similar jobs