Lead Databricks Engineer ( F2F interview | Parsippany, NJ | only W2 )

Hybrid in Parsippany, NJ, US • Posted 1 hour ago • Updated 1 hour ago
Contract W2
12 Months
Hybrid
Depends on Experience
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Delta Lake
  • PySpark
  • SQL
  • Data Architecture

Summary

Position Overview

We are seeking an experienced Lead Data Engineer with strong expertise in Databricks, PySpark, Python, SQL, and modern cloud-based data platforms. The ideal candidate will have a combination of approximately 70% data engineering and 30% analytical responsibilities.

This role involves designing and maintaining scalable data pipelines, optimizing large-scale data processing environments, and supporting financial analytics and data science initiatives.

The successful candidate will bring hands-on experience with financial or payroll data, lakehouse architecture, data governance, and analytical workflows. The individual will collaborate with data scientists, economists, and business stakeholders to deliver reliable data solutions supporting financial research and business intelligence.

Key Responsibilities

Data Engineering & Platform Development

  • Design, develop, and maintain scalable ETL/ELT pipelines for payroll, financial, and macroeconomic datasets.

  • Build and support Databricks-based data platforms using PySpark and Delta Lake.

  • Develop data models and data marts for analytics, reporting, and machine learning use cases.

  • Optimize data processing performance across billions of records and multi-terabyte environments.

  • Ensure data quality, consistency, lineage, governance, and observability.

  • Translate business requirements into scalable and maintainable technical solutions.

Data Analytics & Research Support

  • Perform exploratory data analysis, data profiling, and anomaly detection.

  • Collaborate with data scientists to implement analytical logic using scalable PySpark solutions.

  • Develop validation dashboards and notebooks to verify data quality and pipeline outputs.

  • Support feature engineering, time-series analysis, and complex data aggregations.

  • Utilize Python, pandas, and NumPy for ad hoc analytical tasks.

  • Investigate unusual data patterns and validate analytical results.

Data Architecture & DevOps

  • Work with Databricks, Delta Lake, Unity Catalog, and lakehouse architecture.

  • Participate in architecture discussions involving medallion architecture, catalog design, and data mesh concepts.

  • Implement CI/CD pipelines using Bitbucket Pipelines, Jenkins, and Databricks Asset Bundles.

  • Manage data governance, access controls, and schema design through Unity Catalog.

  • Maintain security, compliance, and data management standards.

AI-Assisted Development

  • Utilize AI-assisted coding tools such as GitHub Copilot, Amazon Q, Kiro, or equivalent.

  • Review AI-generated code for accuracy, performance, scalability, and maintainability.

  • Incorporate AI-assisted development practices into engineering workflows.

Required Qualifications

  • Bachelor's or Master's degree in Computer Science, Data Engineering, Information Systems, Statistics, Finance, Economics, or a related field.

  • 5+ years of experience in Data Engineering or Data Platform Development.

  • Strong hands-on experience with Databricks, PySpark, Python, and SQL.

  • Experience with financial services or payroll data, including compensation, deductions, and pay-period logic.

  • Expertise in large-scale data processing, including billions of records, multi-terabyte datasets, and time-series data.

  • Strong knowledge of Delta Lake, Unity Catalog, Databricks Workflows, and data optimization.

  • Experience with ETL/ELT development, data modeling, and data quality frameworks.

  • Proficiency in pandas and NumPy for exploratory data analysis.

  • Experience with CI/CD tools such as Jenkins, Bitbucket Pipelines, or equivalent.

  • Understanding of statistical concepts, data distributions, correlations, and feature engineering.

  • Experience participating in technical architecture and design decisions.

  • Familiarity with AI-assisted development tools.

Preferred Qualifications

  • Experience with financial markets, macroeconomic, or capital markets datasets.

  • Knowledge of lakehouse architecture, medallion architecture, and data mesh concepts.

  • Experience with Kafka or Spark Structured Streaming.

  • Exposure to Census data, TIGER datasets, and FIPS codes.

  • Infrastructure-as-code experience with Terraform or CDK.

  • Databricks Associate or Professional certification.

  • Experience migrating legacy data platforms to Databricks.

  • Exposure to machine learning and AI-driven analytics.

  • Experience with Power BI, Tableau, or Databricks Dashboards.

  • Scala programming experience is a plus.

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: nm13509
  • Position Id: 9103065
  • Posted 1 hour ago
Contact the job poster
Nitin Bhojwani

Nitin Bhojwani

IT Recruiter @ COOLSOFT
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Hybrid in Parsippany-Troy Hills, New Jersey

•

14d ago

Easy Apply

Contract

Depends on Experience

Hybrid in Parsippany-Troy Hills, New Jersey

•

4d ago

Easy Apply

Contract

Depends on Experience

New York, New York

•

Today

Contract

New York, New York

•

10d ago

Easy Apply

Contract

90 - 100

Search all similar jobs