HI Professionals,
we have a position for Data Engineer at Plano,TX or Alpharetta, or Middletown, NJ (W2)
Title:Data Engineer on W2
Loc:Plano TX or Alpharetta, or Middletown, NJ
Duration : 12 Months+
Skills:Databricks, snowflake, Azure
Job description
At a high level, this is a migration project of data systems and applications.
I need a "builder" that can build and maintain data pipelines.
Primary stack is databricks, snowflake, Azure.
Job Description:
We are seeking a knowledgeable Data Engineer to design, build, and maintain enterprise-scale data platforms using modern cloud-native technologies.
The ideal candidate will have good expertise in Databricks, Snowflake, and Azure services, with a strong background in building scalable ETL/ELT
pipelines and data warehousing solutions.
Daily Responsibilities:
Design and implement end-to-end data pipelines using Palantir Foundry, Databricks, and Snowflake
Develop scalable ETL/ELT workflows using Python, PySpark, and SQL for processing large-scale datasets
Build and maintain data lake and data warehouse architectures on cloud platforms (Azure (Data Factory, Synapse, ADLS), Databricks, Snowflake)
Implement data governance, quality checks, and metadata management frameworks
Create and optimize Databricks notebooks, Delta Lake tables, and Unity Catalog implementations
Collaborate with data scientists, analysts, and business stakeholders to deliver data products
Optimize query performance and storage strategies in Snowflake (clustering, partitioning, materialized views)
Implement real-time streaming pipelines using Spark Streaming, or similar technologies
Develop CI/CD pipelines for data workflows using Git, Jenkins, or similar tools
Required Qualifications:
3+ years of hands-on data engineering experience
Proficiency in Python and PySpark for data processing
Good SQL skills with experience in complex query optimization
Hands-on experience with Databricks, Snowflake
Experience with cloud platforms (Azure) and their data services
Knowledge of data modeling (dimensional modeling, star schema, snowflake schema)
Experience with orchestration tools (Airflow, dbt, Luigi, or similar)
Good understanding of data governance, security, and compliance requirements
Desired Qualifications:
Experience with IBM WatsonX.data or similar AI data platforms
Knowledge of Delta Lake, Iceberg, or similar lakehouse formats
Experience with GenAI, RAG architectures, or LangChain
Familiarity with MCP (Model Context Protocol) for AI agent integrations
Thanks & Regards
Suman|Technical Recruiter |BrightSol