Azure Data lead

Hybrid in New York, NY, US • Posted 1 hour ago • Updated 1 hour ago
Contract Independent
Contract W2
12 Months
Travel Required
Hybrid
Depends on Experience
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Azure Data lead - Python
  • Pyspark
  • Databricks
  • ADF
  • Data Lake

Summary

Title: Azure Data Lead

Location: NYC (3 days onsite)

Skills: Azure Data lead - Python, Pyspark, Databricks, ADF, Data Lake

Introduction

The Azure Data Lead will be responsible for leading the modernization and migration of existing Python object-oriented applications into scalable PySpark and Spark SQL data-processing solutions on Azure Databricks. This role requires a strong blend of software engineering, data engineering, cloud architecture, and performance optimization.

Responsibilities

  • Analyze existing Python OOP applications and redesign single-node processing logic for distributed Spark execution.
  • Design, develop, and deploy enterprise-scale data pipelines on Azure Databricks; build reusable PySpark frameworks and utility modules.
  • Implement Delta Lake solutions using the Bronze–Silver–Gold architecture.
  • Build robust ETL/ELT pipelines with Azure Data Factory, ADLS Gen2, and Azure Synapse Analytics.
  • Implement data quality, reconciliation, validation, and monitoring frameworks.
  • Optimize Spark jobs (partitioning, bucketing, caching, broadcast joins, Adaptive Query Execution, Delta optimization) and benchmark converted applications against original Python implementations.

Requirements

Core Skills:

  • Python (expert), OOP, and advanced Python design patterns
  • PySpark, Spark SQL, and SQL
  • Azure Databricks, Azure Data Factory, ADLS Gen2
  • Apache Spark, Delta Lake, Data Lakehouse architecture, distributed computing

Must-have: Python, Azure Databricks, Azure Data Factory (ADF), MS SQL, Oracle PL/SQL.

Good to have: PySpark; certifications in Azure Data Factory, Azure Databricks, SQL, Oracle, or Python.

Experience & Expected Outcome

The ideal candidate for this role will be a senior data engineering leader with proven delivery of large-scale Databricks modernization programs. The expected outcome is to have existing Python applications converted into scalable, cost-efficient, enterprise-grade data solutions on Azure Databricks with proven performance parity.

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 91138575
  • Position Id: 9066027
  • Posted 1 hour ago
Contact the job poster
Sravan Krishna

Sravan Krishna

Recruiter @ StratG Inc
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Hybrid in New York, New York

Today

Easy Apply

Contract

Depends on Experience

Hybrid in New York, New York

Today

Easy Apply

Contract

Depends on Experience

Hybrid in New York, New York

12d ago

Easy Apply

Contract

70 - 80

Hybrid in New York, New York

Today

Easy Apply

Contract

Depends on Experience

Search all similar jobs