Mandatory Skills
At least 4-6 years of Developer experience specifically focused on Data Engineering
Strong Hands-on experience in Data Engineering development using Python and Pyspark as an ETL tool
Hands-on experience in AWS services like Glue, RDS, S3, Step functions, Event Bridge, Lambda, MSK (Kafka), EKS etc.
Hands-on experience in Databases like Postgres, SQL Server, Oracle, Sybase
Hands-on experience with SQL database programming, SQL performance tuning, relational model analysis, queries, stored procedures, views, functions, and triggers
Strong technical experience in Design (Mapping specifications, HLD, LLD), Development (Coding, Unit testing).
Good knowledge in CI/CD DevOps process and tools like Bitbucket, GitHub, Jenkins
Strong foundation and experience with data modeling, data warehousing, data mining, data analysis and data profiling.
Good communication and inter-personal skills
Responsibilities:
Provide scoping, estimating, planning, design, development, and support services to a project.
Work with developers and business areas to design, configure, deploy and maintain custom ETL Infrastructure to support project initiatives.
Design and develop data/batch processing, data manipulation, data mining, and data extraction/transformation/loading (ETL Pipelines) into large data domains.
Design, development, test, and implement application code
Follow proper software development lifecycle processes and standards
Quality Analysis of the products, responsible for the Defect tracking and Classification