Big Data Engineer

Sunnyvale, CA, US • Posted 5 hours ago • Updated 5 hours ago
Full Time
On-site
Fitment

Dice Job Match Score™

👾 Reticulating splines...

Job Details

Skills

  • Management
  • Big Data
  • Apache Hadoop
  • Apache Hive
  • Cloud Computing
  • Amazon Web Services
  • Microsoft Azure
  • Optimization
  • Workflow
  • Writing
  • SQL
  • Data Modeling
  • Analytical Skill
  • Conflict Resolution
  • Problem Solving
  • Data Integration
  • Effective Communication
  • Collaboration
  • Real-time
  • Streaming
  • FOCUS
  • Scalability
  • SLA
  • Apache HTTP Server
  • Redis
  • Elasticsearch
  • GraphQL
  • Stress Testing
  • Dashboard
  • Tableau
  • Electronic Commerce
  • Google Cloud
  • Google Cloud Platform
  • HDFS
  • Apache Spark
  • Scala
  • Python
  • Automic
  • Apache Kafka
  • API

Summary

Big Data Engineer:

Must Have:
Proficiency in managing and manipulating huge datasets in the order of terabytes (TB) is essential
Expertise in big data technologies like Hadoop, Apache Spark (Scala preferred), Apache Hive, or similar frameworks on the cloud (Google Cloud Platform preferred, AWS, Azure etc.) to build batch data pipelines with strong focus on optimization, SLA adherence and fault tolerance.
Expertise in building idempotent workflows using orchestrators like Automic, Airflow, Luigi etc.
Expertise in writing SQL to analyze, optimize, profile data preferably in BigQuery or SPARK SQL
Strong data modeling skills are necessary for designing a schema that can accommodate the evolution of data sources and facilitate seamless data joins across various datasets
Ability to work directly with stakeholders to understand data requirements and translate that to pipeline development / data solution work
Strong analytical and problem-solving skills are crucial for identifying and resolving issues that may arise during the data integration and schema evolution process
Ability to move at rapid pace with quality and start delivering with minimal ramp up time will be crucial to succeed in this initiative
Effective communication and collaboration skills are necessary for working in a team environment and coordinating efforts between different stakeholders involved in the project

Nice to have:
Experience building complex near real time (NRT) streaming data pipelines using Apache Kafka, Spark streaming, Kafka Connect with a strong focus on stability, scalability and SLA adherence.
Good understanding of REST APIs - working knowledge on Apache Druid, Redis, Elastic search, GraphQL or similar technologies. Understanding of API contracts, building telemetry, stress testing etc.
Exposure in developing reports/dashboards using Looker/Tableau
Experience in eCommerce domain.

Tech stack: Google cloud, HDFS, SPARK, Scala, Python (optional), Automic/Airflow, BigQuery, Kafka, API, Druid
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 80183517
  • Position Id: ca0d8a23562b58e26a506ee1d0387af8
  • Posted 5 hours ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Palo Alto, California

Today

Full-time

USD 140,000.00 - 200,000.00 per year

Sunnyvale, California

Today

Full-time

USD 66.00 per hour

San Jose, California

11d ago

Easy Apply

Full-time

140,000 - 190,000

San Jose, California

Today

Full-time

USD 323,000.00 - 428,000.00 per year

Search all similar jobs