Senior Data Software Engineer/ AWS, Databricks, PySpark

Remote • Posted 2 hours ago • Updated 2 hours ago
Full Time
Remote
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Microsoft Exchange
  • Apache Airflow
  • Collaboration
  • Software Engineering
  • Python
  • Testing
  • Amazon S3
  • Amazon DynamoDB
  • Kubernetes
  • Amazon Web Services
  • Databricks
  • PySpark
  • Extract
  • Transform
  • Load
  • Microservices
  • Terraform
  • Data Security
  • Privacy
  • Accountability
  • Communication
  • Management
  • Virtual Team
  • Continuous Improvement
  • English
  • Reasoning
  • Data Modeling

Summary

We are looking for a Senior Data Software Engineer to join a team that builds and maintains data infrastructure and platform tooling for a large engineering organization. The work spans cross-region data exchange solutions, Databricks-based data sharing infrastructure, and backend integrations used daily by engineers across multiple regions. Responsibilities Design and implement solutions for cross-region data exchange in environments with strict regulatory and security requirements Build and maintain scalable backend services and integrations using Python Develop and support data pipelines using PySpark and Apache Airflow (MWAA) Deploy and manage Databricks environments on AWS via reusable Terraform modules Establish and maintain Delta Sharing for secure data sharing between Databricks accounts Create microservices and platform tooling that enable collaboration across distributed engineering teams Contribute to platform initiatives used by engineers globally Collaborate with cross-functional stakeholders spanning multiple geographies and business units Requirements 3+ years of experience in Data Software Engineering Strong backend engineering experience in Python, including async patterns, clean code, and testing with pytest Working knowledge of AWS services such as IAM, S3, and DynamoDB Familiarity with Kubernetes on AWS, CodeArtifact, and MWAA Hands-on proficiency in Databricks and PySpark for data pipeline development and platform management Background in designing and building microservices architecture and distributed systems Skills in Infrastructure as Code using Terraform or equivalent Understanding of data security and privacy principles, including GDPR Ownership - proactive accountability for deliverables end-to-end Team player - respectful communication and shared responsibility for team outcomes Organisational skills - self-directed and able to manage workload in a distributed team Engineering mastery - commitment to quality, best practices, and continuous improvement English level B2+, with the ability to communicate fluently with English-speaking stakeholders and share technical reasoning clearly Nice to have Expertise in Delta tables and Delta Sharing Competency in data modeling and extensible schema design
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 10330481
  • Position Id: 2812605126f8b80f7b38cab02b54269e
  • Posted 2 hours ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Remote

•

Today

Full-time

Remote

•

Today

Easy Apply

Full-time

Depends on Experience

Remote

•

Yesterday

Easy Apply

Contract

Depends on Experience

Remote or Washington, District of Columbia

•

Today

Full-time

Search all similar jobs