Lead Data Engineer- Snowflake Cortex+ AWS Bedrock

• Posted 2 hours ago • Updated 2 hours ago
Full Time
Part Time
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Snowflake Cortex
  • AWS Bedrock

Summary

Senior/ Lead Data Engineer- Snowflake Cortex+ AWS Bedrock

Location : Remote

We are looking for an experienced Site Reliability Engineer (SRE) to support and operate highly available Snowflake and AWS-based data and AI platforms, with hands-on exposure to Snowflake Cortex and Amazon Bedrock.

The engineer will be responsible for platform reliability, monitoring, incident management, automation, performance optimization, security, and operational support for data and Generative AI workloads.

Key Responsibilities

  • Design, deploy, monitor, and maintain highly available AWS, Snowflake, and GenAI platforms.
  • Support Snowflake Cortex capabilities for AI/ML and Generative AI workloads.
  • Work with Amazon Bedrock to integrate and operate foundation models and GenAI applications.
  • Monitor infrastructure, applications, data pipelines, APIs, and AI workloads.
  • Define and monitor SLIs, SLOs, and SLAs for critical services.
  • Implement observability using CloudWatch, logs, metrics, traces, and alerting.
  • Troubleshoot production incidents and perform Root Cause Analysis (RCA).
  • Participate in on-call / production support and incident management.
  • Automate repetitive operational activities using Python, Shell scripting, Terraform, or AWS services.
  • Build and maintain CI/CD pipelines for application, data, and AI workloads.
  • Optimize Snowflake compute usage, query performance, warehouse utilization, and cost.
  • Monitor and troubleshoot Snowflake data pipelines and workloads.
  • Support AWS services such as S3, Lambda, IAM, CloudWatch, ECS/EKS, API Gateway, and Secrets Manager.
  • Implement security, access controls, encryption, secrets management, and least-privilege IAM.
  • Monitor the reliability and performance of LLM/GenAI applications using Amazon Bedrock.
  • Track AI application metrics such as latency, throughput, errors, token usage, and cost.
  • Establish automated health checks, alerts, recovery mechanisms, and disaster-recovery procedures.
  • Collaborate with Data Engineers, ML Engineers, Developers, Architects, Product Managers, and business stakeholders.
  • Continuously improve platform reliability through automation, observability, capacity planning, and performance engineering.

Thanks, and Regards,

Ajay Pratap Singh
Team Lead - Recruitment

E:

Tekaccel, Inc.

2601 Little Elm Pkwy,

Suite # 1804,

Little Elm, TX 75068

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 91126533
  • Position Id: 2026-18550
  • Posted 2 hours ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Remote

Today

Easy Apply

Contract, Third Party

Depends on Experience

Remote

Today

Easy Apply

Full-time, Part-time, Contract, Third Party

Jersey City, New Jersey

Today

Easy Apply

Full-time

$50 - $60 per hour

Remote

Today

Full-time

Search all similar jobs