Lead Data Engineer Databricks & Data Product :: Lawrenceville, NJ

Lawrenceville, NJ, US • Posted 22 hours ago • Updated 1 hour ago
Full Time
12 Months
On-site
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Project Management
  • Performance Management
  • Preventive Maintenance
  • Artificial Intelligence
  • Business Intelligence
  • Acceptance Testing
  • Product Scoping
  • Business Rules
  • Technical Drafting
  • Analytical Skill
  • Clustering
  • Performance Tuning
  • Continuous Integration
  • Continuous Delivery
  • Delegation
  • Design Review
  • Code Review
  • Technical Direction
  • SOW
  • Data Engineering
  • Data Modeling
  • Databricks
  • PySpark
  • Apache Spark
  • Unity
  • Offshoring
  • SQL
  • Python
  • Git
  • Workflow
  • Communication
  • Management
  • Pharmaceutics
  • Life Sciences
  • Pharmacy
  • Veeva
  • Customer Relationship Management (CRM)
  • Marketing
  • Analytics
  • Use Cases
  • Data Governance
  • Data Quality
  • Auditing
  • Privacy
  • Access Control
  • Recruiting
  • EXT
  • Oracle UCM

Summary

Job Title: Lead Data Engineer Databricks & Data Product

Work Location: Lawrenceville, NJ 50% onsite

Duration: 12+ Months
Hours: Mon-Fri 8am-5pm


ROLE SUMMARY
The Company's Commercialization Business Insights & Technology (CBI&T) Dat and AI (D&A) New Build team is responsible for executing on prioritized work leveraging Databricks as its data foundation platform.
Seeking US Data Lead to manage delivery of these data assets and data products end to end. This position is the technical point of contact for this work, translating business requirements from our Commercialization AI Strategy & Analytics (CASA) business partners into designed, built, tested, and released data products on Databricks.
Work with US data team members, other commercialization BI&T stakeholders, company offshore engineers and systems integrator resources.

KEY RESPONSIBILITIES
Data delivery
Lead design, build, UAT, releases
Partner with CASA data owners, domain leads, and business SMEs to translate data product scope: critical data elements, business rules, metric definitions, grain, aggregation logic, and acceptance criteria into technical design and execution.
Produce core design artifacts: source-to-target mappings, analytical requirement documents, Bronze, Silver Gold-layer logical data models.
Technical build and data modeling
Hands on understanding of of Databricks tools, (PySpark, Spark SQL, Delta Lake, etc.)
Apply platform standards for Unity Catalog cataloging, table design and clustering, performance tuning, data quality controls, and CI/CD through Git.
Data Product data modeling
Offshore/Vendor Engagement
Day-to-day work of an offshore engineering pod in Hyderabad, India - backlog grooming, task assignment, design review, and code review.
Provide technical direction to systems integrator and consulting partners working on the same domains, and hold their deliverables to company standards.
Review vendor work products against statement-of-work scope and flag gaps early.

REQUIRED QUALIFICATIONS
6 10+ years in data engineering
Databricks experience
Strong data product and data modeling skills. Proven experience defining and delivering data products with business stakeholders, not just executing handed-down specifications.
Demonstrated hands-on Databricks experience: PySpark, Spark SQL, Delta Lake, and experience with medallion architecture (Bronze/Raw, Silver/Refined, Gold/Published) including what logic belongs in which layer and why.
Experience with Unity Catalog: catalogs and schemas, access control, and data classification.
Experience leading distributed and offshore teams, with demonstrated ability to get quality output from an India-based delivery pod.
Advanced SQL, Python, and a Git-based development workflow.
Excellent written and verbal communication; able work independently with business stakeholders
Self-directed. This team needs people who can take an ambiguous business ask and drive it to released data products

PREFERRED QUALIFICATIONS
Pharmaceutical or life sciences commercial data experience, and familiarity with industry data sources such as IQVIA, Symphony Health, Komodo, Flatiron, specialty pharmacy and distributor data, Veeva CRM, and marketing or digital engagement platforms.
Understanding of commercial analytics use cases: field force performance, customer, patient journey and line of therapy
Familiarity with data governance concepts: data ownership, stewardship, critical data elements, and data quality frameworks.
Experience in a regulated environment with audit, privacy, and access-control requirements

Ankita Singh -Recruitment Manager

Email- | Ext: 138 Contact No :

STELLENT IT A Nationally Recognized Minority Certified Enterprise

"Happiness can be found, even in the darkest of times, if one only remembers to turn on the light."
- JK Rowling

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 91022079
  • Position Id: 2026-51819
  • Posted 22 hours ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Lawrence Township, New Jersey

•

Yesterday

Easy Apply

Contract

80 - 82

New Jersey

•

Today

Easy Apply

Contract

$70 - $80 per hour

Hybrid in Raritan, New Jersey

•

7d ago

Easy Apply

Contract, Third Party

80 - 90

New York, New York

•

24d ago

Full-time

Search all similar jobs