Job Description: Senior Data Engineer (Code Review & Data Platform Governance)
Role Overview
We are seeking a Senior Data Engineer who will play a lead role in governing code quality, reviewing data pipelines, and ensuring scalable data product engineering. The candidate will work closely with distributed data teams to enforce best practices across Snowflake, Redshift, AWS Athena, and Python-based data pipelines, while driving robust data modeling and architecture standards.
________________________________________
Key Responsibilities
1. Code Review & Engineering Governance (Primary Focus)
• Lead code reviews across data engineering teams (SQL, Python, ETL pipelines) to ensure:
o Performance optimization (query tuning, partitioning, clustering)
o Maintainability and readability of code
o Adherence to enterprise coding standards and design principles
• Establish and enforce coding standards, review frameworks, and best practices
• Identify anti-patterns in data pipelines and recommend improvements
• Drive adoption of CI/CD, version control, and automated quality checks
________________________________________
2. Data Platform Engineering (Snowflake / AWS / Redshift / Athena)
• Design, review, and optimize data pipelines and transformations across:
o Snowflake (EDW / Data Lakehouse)
o Amazon Redshift & AWS Athena
• Ensure efficient handling of large-scale structured and semi-structured datasets
• Optimize data ingestion, transformation, and consumption layers
• Guide teams in building scalable, cost-efficient cloud data solutions leveraging AWS services
________________________________________
3. Data Modeling & Architecture
• Define and review logical and physical data models (dimensional, normalized, data vault, etc.)
• Ensure alignment with data product architecture and consumption patterns
• Drive data consistency, reusability, and governance standards
• Implement best practices for data lineage, metadata management, and documentation
________________________________________
4. Data Quality & Validation
• Establish frameworks for:
o Data validation and reconciliation
o Anomaly detection and monitoring
• Ensure reliability of pipelines through robust testing strategies (unit/integration)
________________________________________
5. Stakeholder Collaboration & Mentorship
• Work with data engineers, architects, and product owners to translate requirements into scalable solutions
• Provide technical leadership and mentoring to junior engineers
• Act as a gatekeeper for production-ready data assets and pipelines
________________________________________
Required Skills & Experience
• 8–12+ years in Data Engineering with strong hands-on coding and review experience
• Strong expertise in:
o Snowflake, Redshift, AWS Athena
o Python (data processing, scripting, frameworks)
o Advanced SQL & query optimization
• Proven experience in:
o Data modeling techniques (Star schema, Snowflake schema, Data Vault)
o ETL/ELT pipeline design and optimization
o Working with AWS data services (S3, Glue, etc.)
• Familiarity with:
o CI/CD pipelines, Git-based workflows
o Orchestration tools (Airflow, etc.)
o Data quality and governance frameworks
___________________________________