Data Engineer with Databricks-Only W2


cloudingest inc
Dice Job Match Score™
🔢 Crunching numbers...
Job Details
Skills
- Databricks
- Snowflake
- ETL
- Glue
- DBT
Summary
Role : Data Engineer
Location : Remote
We are seeking a highly skilled Data Engineer with handson experience in Databricks, Snowflake, and modern cloud data pipelines. The ideal candidate has deep expertise in PySpark, SQL, Delta Lake, Snowflake ELT, and distributed data processing. This role focuses on building scalable data pipelines, optimizing performance, ensuring data quality, and supporting analytics, AI/ML, and enterprise reporting workloads
Key Responsibilities
Databricks Engineering
Build and optimize ETL/ELT pipelines using PySpark, Spark SQL, Databricks Workflows, and Delta Lake.
Develop Bronze/Silver/Gold Medallion architecture pipelines.
Implement Delta Live Tables (DLT) for automated ingestion and transformation.
Manage and optimize Databricks clusters, jobs, notebooks, repos, and workflows.
Perform Spark performance tuning (partitioning, caching, AQE, broadcast joins).
Implement Unity Catalog governance (catalogs, schemas, tables, permissions).
Integrate Databricks with AWS S3 / Azure Data Lake / Kafka / APIs.
Snowflake Engineering
Design and develop Snowflake ELT pipelines using Snowflake SQL, Streams, Tasks, and Snowpipe.
Build warehouse models, fact/dimension tables, and curated datasets.
Optimize Snowflake performance (clustering, micro-partitioning, query tuning).
Implement RBAC, masking policies, row-level security, and governance.
Integrate Snowflake with Fivetran, DBT, ADF, Glue, Kafka, or custom ingestion frameworks.
Data Pipeline & Integration
Build scalable ingestion frameworks for structured, semistructured, and unstructured data.
Integrate data from databases, APIs, cloud storage, streaming sources, and enterprise systems.
Implement data quality checks, validation rules, reconciliation, and lineage.
Support ML/AI workloads, feature engineering, and model-ready datasets.
Cloud & DevOps
Work with AWS, Azure, or Google Cloud Platform cloud-native services.
Implement CI/CD using GitHub Actions, Azure DevOps, GitLab, Jenkins.
Containerize workloads using Docker and orchestrate with Kubernetes (nice to have).
Monitor pipelines using CloudWatch, Azure Monitor, Databricks metrics, PrometheGrafana.
Required Skills
Core Technical Skills
Databricks (PySpark, Spark SQL, Delta Lake, Workflows, DLT)
Snowflake (SQL, Streams, Tasks, Snowpipe, RBAC)
Python
Advanced SQL
Cloud platforms: AWS / Azure / Google Cloud Platform
Data modeling (Star schema, dimensional modeling)
ETL/ELT pipeline development
Data quality, governance, lineage
CI/CD pipelines
API integration & REST services
Nice-to-Have Skills
DBT
Kafka / Spark Structured Streaming
MLflow / Feature Store
Terraform
Airflow / ADF / Glue / Dataflow
RAG/LLM data preparation (bonus)
Healthcare, finance, or regulated industry experience
- Dice Id: RTX1bc3ac
- Position Id: 9081600
- Posted 1 hour ago
Company Info
About cloudingest inc
CloudIngest is a full-service tech software firm. We possess extensive practical experience in Management, Business, and Economics. We stay up-to-date on emerging trends in terms of both the evolving cloud-based tech stack and client considerations in terms of Financial Billing and Payments. Our team consists of Client Intake Managers, Project Managers, Software Developers, Quality Assurance, and Integrative Solution and NLP Specialists all of whom are fully-equipped and ready to interface with members of your existing team. We are experienced working with client-side Project Managers as well as Designers and C-Suite business executives.
Similar Jobs
It looks like there aren't any Similar Jobs for this job yet.
Search all similar jobs