Snowflake Site Reliability Engineer
Location: Plano, TX (Dallas-area local candidates only)
Work Schedule: Hybrid – Onsite every Wednesday
Duration: 6+ Months Contract (Extendable)
Interview : Video - Final F2F Meeting
Job Description
We are seeking an experienced Site Reliability Engineer (SRE) with strong Snowflake and AWS experience to support and enhance a highly reliable, scalable, and cost-effective Snowflake data platform.
The ideal candidate will have expertise in Snowflake platform operations, Infrastructure as Code (IaC), cloud infrastructure, automation, reliability engineering, and operational governance.
Key Responsibilities
Design and implement Infrastructure as Code (IaC) solutions to build and manage data pipelines and Snowflake platform infrastructure.
Monitor and maintain Snowflake platform operational health and reliability.
Manage Snowflake accounts, compute resources, warehouses, and capacity planning.
Drive platform automation and operational excellence.
Improve platform scalability, reliability, resilience, and performance.
Implement monitoring, alerting, and reliability best practices.
Support cost optimization and governance across the Snowflake environment.
Develop and maintain operational standards and processes for Snowflake platform management.
Work closely with engineering and operations teams to resolve platform issues and improve service reliability.
Support resilience, availability, and disaster recovery initiatives.
Required Skills
Strong hands-on experience with Snowflake platform administration and operations.
Experience with AWS cloud infrastructure and services.
Strong experience with Infrastructure as Code (IaC) and automation tools, preferably Terraform.
Experience with Snowflake compute and warehouse management.
Knowledge of capacity planning, performance monitoring, and cost optimization.
Experience with Site Reliability Engineering (SRE), DevOps, or platform engineering practices.
Strong understanding of monitoring, alerting, automation, and incident management.
Experience designing and supporting highly available and scalable cloud platforms.
Strong troubleshooting and problem-solving skills.
Preferred Skills
Experience with Snowflake resource governance and cost management.
Experience building automated data platform infrastructure.
Knowledge of resilience and disaster recovery practices.
Experience improving operational processes and platform reliability.