Solution Databricks Architect(20 hours/Week, Part-time)

Remote • Posted 5 hours ago • Updated 5 hours ago
Contract Independent
Contract W2
12 Months
No Travel Required
Remote
Depends on Experience
Company Branding Image
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Medallion Architecture
  • Databricks Architecture
  • Databricks Implementation
  • Python
  • PySpark
  • Apache Spark
  • Spark Internals
  • Delta Lake
  • Delta Tables
  • Unity Catalog
  • Batch Processing
  • Legacy-to-Databricks Migration
  • SQL
  • Data Modeling
  • AWS
  • Azure
  • CI/CD
  • Performance Optimization
  • Data Governance
  • Data Security
  • Architecture Documentation
  • Stakeholder Communication
  • End-to-End Technical Ownership

Summary

Job Description: Databricks Architect

Type: Remote

Work Hours : 20 hours (4 hrs a day)

Role Overview

We are looking for an experienced Databricks Architect to support a large-scale data modernization initiative for Client. The consultant will be responsible for designing scalable data architecture, defining Databricks best practices, and leading migration and implementation activities across the enterprise data platform.

The ideal candidate should have strong hands-on experience with Databricks, Apache Spark, Python/PySpark, Delta Lake, Unity Catalog, cloud data platforms, and enterprise data architecture.

Key Responsibilities

  • Design and implement enterprise-scale Databricks architecture for complex data engineering and analytics workloads.
  • Define end-to-end architecture covering data ingestion, transformation, storage, governance, security, orchestration, and consumption.
  • Lead migration of legacy/on-premise data workloads to a modern cloud + Databricks platform.
  • Design and implement Medallion Architecture – Bronze, Silver, and Gold layers.
  • Build and optimize large-scale data pipelines using Python, PySpark, Spark SQL, and Delta Lake.
  • Define standards for Delta Tables, schema evolution, partitioning, OPTIMIZE, Z-Ordering, caching, and performance tuning.
  • Architect and implement Unity Catalog for centralized governance, access control, lineage, and data security.
  • Establish appropriate RBAC, service principals, secrets management, and data-access policies.
  • Design batch and, where required, streaming data processing solutions.
  • Provide architectural guidance around Spark execution, cluster configuration, memory management, shuffle optimization, and performance troubleshooting.
  • Design integration patterns between Databricks and enterprise systems, APIs, databases, data warehouses, and cloud storage.
  • Establish CI/CD and DevOps standards for Databricks notebooks, workflows, jobs, and infrastructure deployments.
  • Define development standards across Dev, QA, UAT, and Production environments.
  • Conduct architecture reviews and provide technical leadership to Data Engineers and Databricks developers.
  • Work closely with enterprise architects, business stakeholders, security teams, data governance teams, and application teams.
  • Translate business and technical requirements into scalable architecture and implementation plans.
  • Support technical assessments, POCs, design documentation, and solution architecture presentations.
  • Identify performance, scalability, security, and operational risks and recommend appropriate solutions.
  • Provide technical ownership from discovery and architecture through implementation and production deployment.

Required Skills

Databricks

  • 8+ years of overall Data Engineering / Data Platform experience.
  • Strong hands-on experience with Azure Databricks or Databricks on AWS.
  • Multiple enterprise-level end-to-end Databricks implementations.
  • Strong understanding of Databricks platform architecture and administration.
  • Hands-on expertise with:
    • Delta Lake
    • Delta Tables
    • Unity Catalog
    • Databricks Workflows / Jobs
    • Auto Loader
    • Databricks SQL
    • Cluster configuration and optimization
    • Performance tuning

Spark / PySpark

  • Advanced knowledge of Apache Spark and PySpark.
  • Strong understanding of Spark internals including:
    • Driver and Executors
    • DAG and stages
    • Lazy evaluation
    • Shuffle operations
    • Partitioning
    • Repartition vs. Coalesce
    • Broadcast joins
    • Data skew
    • Memory management
    • Spark performance optimization

Data Architecture

  • Strong experience designing enterprise data platforms.
  • Expertise in Medallion Architecture.
  • Strong understanding of:
    • Data lakes
    • Lakehouse architecture
    • Data warehouses
    • Data modeling
    • ETL/ELT
    • Batch processing
    • Data quality
    • Metadata management
    • Data lineage

Programming

  • Advanced Python/PySpark development experience.
  • Strong SQL skills.
  • Ability to write and review production-quality data engineering code.

Cloud

Strong experience with at least one major cloud platform:

  • AWS: S3, IAM, Glue, Lambda, Redshift, Secrets Manager, CloudWatch, etc.

OR

  • Azure: ADLS Gen2, ADF, Key Vault, Azure DevOps, Entra ID, etc.

Exposure to multiple cloud platforms would be preferred.

Security & Governance

  • Strong experience with Unity Catalog.
  • Data access control and RBAC.
  • Data lineage.
  • PII/sensitive-data handling.
  • Secrets and credential management.
  • Enterprise data-governance practices.

DevOps / CI-CD

  • Experience implementing CI/CD for Databricks solutions.
  • Git-based development.
  • Azure DevOps / GitHub / Jenkins or similar tools.
  • Automated deployment across multiple environments.
  • Infrastructure-as-Code exposure is preferred.

Preferred Qualifications

  • Databricks Certified Data Engineer Professional or Databricks Architect-level certification.
  • Experience within Insurance, Financial Services, or highly regulated enterprise environments.
  • Experience modernizing large legacy data ecosystems.
  • Experience working directly with US-based enterprise stakeholders.
  • Knowledge of MLOps, AI/ML, or GenAI workloads on Databricks is an advantage.
  • Experience designing reusable frameworks and accelerators for enterprise data engineering.

Ideal Candidate Profile

The ideal consultant should be able to operate as both a hands-on Databricks expert and Solution Architect. The person should not be limited to notebook development; they should be capable of understanding the complete enterprise landscape, making architectural decisions, defining platform standards, guiding engineering teams, and taking responsibility for successful end-to-end delivery.

Key Evaluation Parameters

Skill

Expected Level

Databricks Architecture

Expert

Databricks Implementation

Expert

Python / PySpark

Expert

Apache Spark

Expert

Spark Internals

Advanced

Delta Lake / Delta Tables

Expert

Unity Catalog

Expert

Medallion Architecture

Expert

Batch Processing

Expert

Legacy-to-Databricks Migration

Advanced

SQL

Advanced

Data Modeling

Advanced

Cloud – AWS/Azure

Advanced

CI/CD

Advanced

Performance Optimization

Expert

Data Governance & Security

Advanced

Architecture Documentation

Advanced

Stakeholder Communication

Excellent

End-to-End Technical Ownership

Mandatory

  • Primary Focus: Databricks architecture, modernization, migration, governance, performance, and enterprise data engineering

 

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 91165977
  • Position Id: 9067525
  • Posted 5 hours ago

Company Info

About Balin Technologies LLC

Balin Technologies, headquartered in Cumming, GA, is one of the leading IT consulting firms founded by industry experts with extensive experience in IT consulting services, Talent Acquisition, and SOW outlining scope, timeline, cost, and other aspects between two parties. Our priority is customer satisfaction, the cornerstone of our success.

We provide end-to-end IT consulting services, from requisition to candidate onboarding, across various industry verticals. Our rigorous screening, interviewing, and recruiting processes ensure the right fit for contract, contract-to-hire, and permanent placements, catering to clients of all sizes.

About_Company_OneAbout_Company_Two
Contact the job poster
Phani Kishore

Phani Kishore

Recruiter @ Balin Technologies LLC
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

It looks like there aren't any Similar Jobs for this job yet.

Search all similar jobs