Data & Software Engineer

Chantilly, VA, US • Posted 2 days ago • Updated 3 hours ago
Full Time
On-site
Fitment

Dice Job Match Score™

🛠️ Calibrating flux capacitors...

Job Details

Skills

  • Data Flow
  • Java
  • Data Security
  • Privacy
  • Regulatory Compliance
  • Extract
  • Transform
  • Load
  • Workflow
  • Orchestration
  • Cloud Computing
  • MySQL
  • PostgreSQL
  • Performance Tuning
  • Query Optimization
  • Analytical Skill
  • Git
  • Geospatial Analysis
  • Bash
  • Scripting
  • Data Processing
  • Machine Learning (ML)
  • Problem Solving
  • Conflict Resolution
  • Debugging
  • Data Quality
  • Data Migration
  • Data Engineering
  • Documentation
  • Design Patterns
  • Apache Spark
  • PySpark
  • Python
  • Pandas
  • NumPy
  • Docker
  • Amazon Web Services
  • Amazon S3
  • Step-Functions
  • SQL
  • NoSQL
  • Amazon DynamoDB
  • Unity
  • Operations Support Systems
  • Apache HTTP Server
  • Terraform
  • PostGIS
  • Health Care
  • Life Insurance
  • Training And Development

Summary

The Data & Software Engineer works with a small team to build complex data flows for a custom application. Successful candidate will have advanced Python programming skills, familiarity with Java, an understanding of data security, privacy, governance and compliance principles and a demonstrated history of building production data pipelines and ETL workflows at scale. Candidate must have experience:
  • Building end-to-end data pipelines leveraging Python
  • Using orchestration tools to deploy data pipelines, including configuring and updating Spark Jobs
  • Containerizing and deploying applications in cloud environments like AWS.
  • Working with MySQL and PostgreSQL including performance tuning, schema design, and query optimization for complex, analytical workloads.
  • Leveraging industry standard tools for code control (Git, IaaC control, etc.)
  • Working with data catalogs, tracking data lineage and handling a variety of data formats, including Geospatial.
  • Using Bash scripting for automation and data processing tasks
  • Integrating Al/ML services and models

Responsibilities:
  • Work with stakeholders to understand data requirements, assess feasibility, and design appropriate solutions with minimal oversight
  • Leverage strong problem-solving and debugging skills for data quality issues, pipeline failures, and performance bottlenecks
  • Leverage a background in large-scale data migration or platform modernization efforts
  • Contribute to data engineering documentation, best practices, and design patterns.

Requirements

Minimum of 5 years' experience with:
  • Apache Spark & PySpark
  • Advanced Python skills (including Pandas & NumPy)
  • Docker, Podman
  • AWS S3, Lambda & Step functions
  • Apache Iceberg, Airflow, etc.
  • SQL (with Trino)
  • NoSQL, DynamoDB
  • Unity Catalog OSS, Apache Polaris
  • Apache Superset
  • Terraform or CloudFormation
  • OpenLineage
  • H3, PostGIS

Benefits

Eligibility requirements apply.

  • Employer-Paid Health Care Plan (Medical, Dental & Vision)
  • Retirement Plan (401k, IRA) with a generous matching program
  • Life Insurance (Basic, Voluntary & AD&D)
  • Paid Time Off (Vacation, Sick & Public Holidays)
  • Short Term & Long Term Disability
  • Training & Development
  • Employee Assistance Program
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 80183499
  • Position Id: b8e263998c6dc8910be484ae40dc2807
  • Posted 2 days ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Chantilly, Virginia

Today

Full-time

McLean, Virginia

Today

Full-time

McLean, Virginia

Today

Full-time

Herndon, Virginia

Today

Full-time

USD 130,000.00 - 260,000.00 per year

Search all similar jobs