Senior Technology Consultant - Data / AI Native Engineer

New York, NY, US • Posted 1 hour ago • Updated 6 minutes ago
Full Time
Part Time
6 Months
On-site
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Marketing Operations
  • Extract
  • Transform
  • Load
  • Workflow
  • Application Development
  • Data Governance
  • Access Control
  • Real-time
  • Development Testing
  • Documentation
  • Use Cases
  • Change Data Capture
  • SCD
  • Data Quality
  • Design Review
  • Collaboration
  • Machine Learning (ML)
  • Software Engineering
  • IT Management
  • Scala
  • GitHub
  • Command-line Interface
  • Data Engineering
  • Productivity
  • Artificial Intelligence
  • Databricks
  • Unity
  • Data Processing
  • PySpark
  • Cloud Computing
  • Amazon Web Services
  • Amazon S3
  • Electronic Health Record (EHR)
  • Amazon Redshift
  • Apache Spark
  • Streaming
  • Apache Kafka
  • Amazon Kinesis

Summary

Hello

Description:

Job Title: Senior Technology Consultant Data / AI Native Engineer
Location: Arlington, VA, New York, NY and St. Louis, MO; Hybrid work model | Local candidates preferred.
Overall Experience: 7+ Years

Role Summary:
We are seeking a hands-on Senior Data / AI Native Engineer with strong expertise in Databricks, AWS, PySpark, and modern Lakehouse architecture. The ideal candidate combines deep Data Engineering expertise with strong Software Engineering practices and has experience leading complex enterprise data initiatives.
The role requires daily hands-on use of GitHub Copilot and/or Claude Code CLI to accelerate software and data pipeline development. The candidate should be comfortable rapidly building POCs and MVPs and applying AI directly within data engineering workflows-not just using AI for application development.

Day to Day Job Duties

  • Design, develop, and optimize high-volume enterprise data pipelines using PySpark, Spark, and Databricks.
  • Build scalable Lakehouse solutions using Delta Lake and Medallion Architecture (Bronze/Silver/Gold).
  • Implement data governance, access controls, and cataloging using Databricks Unity Catalog.
  • Design and develop cloud-native data solutions using AWS S3, EMR, Glue, Lambda, and Redshift.
  • Lead complex Data Engineering initiatives and provide technical guidance to engineering teams.
  • Develop batch and real-time data processing pipelines for large-scale enterprise datasets.
  • Use GitHub Copilot and/or Claude Code CLI daily to accelerate coding, pipeline development, testing, troubleshooting, and documentation.
  • Rapidly develop POCs and MVPs using AI-assisted engineering practices.
  • Apply AI within data pipelines for use cases such as schema inference, automated data-quality rule generation, anomaly detection, and PySpark transformation generation.
  • Build AI-ready data platforms supporting RAG, embeddings, vector stores, and AI/ML applications.
  • Develop streaming pipelines using Structured Streaming, Kafka, Kinesis, or Databricks Auto Loader.
  • Implement modern data engineering patterns including CDC, SCD Type 2, schema evolution, idempotent processing, and data contracts.
  • Implement data quality, validation, monitoring, lineage, and observability across data pipelines.
  • Conduct code/design reviews and establish reusable Data Engineering patterns and standards.
  • Collaborate with Data Architects, AI/ML Engineers, Software Engineers, and business stakeholders to deliver enterprise data products.

Basic Qualifications Must Have:

  • 7+ years of experience in Data Engineering and Software Engineering, building production-grade enterprise data solutions.
  • 4+ years of hands-on experience with Databricks, Delta Lake, Medallion Architecture, and PySpark/Spark.
  • 4+ years of experience with the AWS data ecosystem, including S3, EMR, Glue, Lambda, and/or Redshift.
  • Proven experience leading complex Data Engineering initiatives or providing technical leadership to Data Engineering teams.
  • Strong hands-on experience processing high-volume datasets using PySpark, rather than Scala-only Spark development.
  • Demonstrated daily use of GitHub Copilot and/or Claude Code CLI, with the ability to explain specific examples of how these tools improve Data Engineering productivity.
  • Proven experience rapidly developing POCs and MVPs using AI-assisted development practices.

Technical Skills
Data Platform: Databricks, Delta Lake, Unity Catalog
Data Processing: PySpark, Apache Spark
Cloud: AWS S3, EMR, Glue, Lambda, Redshift
Architecture: Lakehouse, Medallion Architecture, Data Lakes
Streaming: Spark Structured Streaming, Kafka, Kinesis, Auto Loader

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 91022079
  • Position Id: 2026-51416
  • Posted 1 hour ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

New York, New York

Today

Easy Apply

Third Party, Contract

New York, New York

Today

Easy Apply

Full-time, Part-time, Contract, Third Party

USD 65-65

New York, New York

Today

Contract

USD70 - USD75

New York, New York

Today

Contract

USD 80.00 - 88.00 per hour

Search all similar jobs