Sr Big data with AWS and Databricks- 10 years needed

Hybrid in Tampa, FL, US • Posted 3 hours ago • Updated 3 hours ago
Full Time
Hybrid
$90,032 - $134,200/yr
Fitment

Dice Job Match Score™

🫥 Flibbertigibetting...

Job Details

Skills

  • Pyspark
  • ETL
  • Big data
  • AWS
  • Databricks
  • Python

Summary

About LTIMindtree

LTIMindtree is a global technology consulting and digital solutions company that enables enterprises across industries to reimagine business models, accelerate innovation, and maximize growth by harnessing digital technologies. As a digital transformation partner to more than 750 clients, LTIMindtree brings extensive domain and technology expertise to help drive superior competitive differentiation, customer experiences, and business outcomes in a converging world. Powered by nearly 90,000 talented and entrepreneurial professionals across more than 30 countries, LTIMindtree a Larsen & Toubro Group company combines the industry-acclaimed strengths of erstwhile Larsen and Toubro Infotech and Mindtree in solving the most complex business challenges and delivering transformation at scale. For more information, please visit

Title: AWS Pyspark Developer

Work Location : Irving, Texas/ Tampa, FL

Job Summary

  • We are seeking a highly skilled and motivated AWS Certified Engineer to design build and optimize scalable data solutions within the Amazon Web Services AWS ecosystem The ideal candidate will have strong expertise in big data processing using PySpark and a deep understanding of data warehousing concepts including Hive and modern table formats like Iceberg This role involves developing deploying and managing robust efficient and secure data pipelines and analytics solutions on AWS leveraging core networking and compute services

Responsibilities

  • AWS Solution Design Implementation Design develop and deploy scalable and costeffective data solutions on AWS leveraging services such as S3 for data lakes EC2 EMR Glue Athena Lambda Redshift and Kinesis
  • Data Pipeline Development Build and maintain robust ETLELT data pipelines using PySpark for data ingestion transformation and loading into various data stores including those utilizing open table formats like Iceberg
  • Big Data Processing Develop and optimize big data processing jobs using PySpark on AWS EMR or AWS Glue handling large datasets efficiently and integrating with Iceberg table formats
  • Data Warehousing Design implement and manage data warehousing solutions including schema design data modeling and query optimization with a focus on Hive and modern data lake table formats like Iceberg for historical data and analytical queries
  • Cloud Infrastructure Networking Implement secure and robust cloud infrastructure components including VPCs subnets routing and security groups to ensure proper connectivity and isolation for data solutions
  • Containerized Workloads Design deploy and manage containerized data processing applications on Amazon Elastic Kubernetes Service EKS
  • Performance Tuning Optimization Optimize AWS resources and big data applications Spark Hive Iceberg for performance cost and efficiency
  • Data Governance Security Implement best practices for data security access control and compliance within AWS including IAM policies S3 bucket policies and encryption
  • Monitoring Troubleshooting Set up monitoring ing and logging for data pipelines and AWS infrastructure troubleshoot and resolve issues promptly
  • Automation Develop and maintain automation scripts using Python and shell scripting for infrastructure provisioning deployment and operational tasks
  • Collaboration Work closely with data scientists analysts and other engineering teams to understand data requirements and deliver reliable data solutions

Required Skills , Qualifications

  • AWS Certification Hold at least one AWS certification eg AWS Certified Solutions Architect Associate AWS Certified Data Analytics Specialty AWS Certified Developer Associate
  • AWS Services Expertise Handson experience with key AWS services for data processing and storage including
  • Storage S3 for data lakes EC2
  • Data Processing EMR Glue Athena Lambda
  • Networking VPC Subnets Routing Security Groups
  • Containerization EKS
  • Big Data Processing Strong proficiency in PySpark for developing complex data transformations and analytics
  • Data Lake Table Formats Practical experience with Apache Iceberg for managing and querying data lakes
  • Data Warehousing Indepth knowledge and practical experience with Apache Hive for data storage querying and schema management

Programming Languages

  • Python Expertlevel proficiency in Python for scripting data manipulation and AWS automation Boto3
  • Shell Scripting Proficient in shell scripting for automation and operational tasks
  • Database SQL Strong SQL skills for data querying and manipulation
  • Data Concepts Solid understanding of ETLELT processes data modeling distributed computing and data governance
  • Good to Have Skills
  • Containerization Orchestration Experience with Kubernetes for deploying and managing containerized applications
  • CICD Experience with CICD tools and practices eg AWS CodePipeline GitHub Actions GitLab CI for automating deployment of data solutions
  • Orchestration Experience with workflow orchestration tools like Apache Airflow
  • Version Control Proficient in using Git for source code management
  • Other Big Data Technologies Exposure to other big data technologies like Apache Kafka Flink or Presto
  • Certifications
  • AWS Certified Solutions Architect AssociateProfessional
  • AWS Certified Data Analytics Specialty
  • AWS Certified Developer Associate

Benefits and Perks:

Comprehensive Medical Plan Covering Medical, Dental, Vision

Short Term and Long-Term Disability Coverage

401(k) Plan with Company match

Life Insurance

Vacation Time, Sick Leave, Paid Holidays

Paid Paternity and Maternity Leave

The range displayed on each job posting reflects the minimum and maximum salary target for the position across all Canada locations. Within the range, individual pay is determined by work location and job level and additional factors including job-related skills, experience, and relevant education or training. Depending on the position offered, other forms of compensation may be provided as part of overall compensation like an annual performance-based bonus, sales incentive pay and other forms of bonus or variable compensation.

LTIMindtree is an equal opportunity employer that is committed to diversity in the workplace. Our employment decisions are made without regard to race, colour, creed, religion, sex (including pregnancy, childbirth or related medical conditions), gender identity or expression, national origin, ancestry, age, family-care status, veteran status, marital status, civil union status, domestic partnership status, military service, handicap or disability or history of handicap or disability, genetic information, atypical hereditary cellular or blood trait, union affiliation, affectional or sexual orientation or preference, or any other characteristic protected by applicable federal, state, or local law, except where such considerations are bona fide occupational qualifications permitted by law.

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 10207105
  • Position Id: 9049922
  • Posted 3 hours ago
Contact the job poster
KN

Kapa Nagender

Recruiter @ LTIMindtree
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Remote

3d ago

Easy Apply

Full-time

Depends on Experience

Remote

22d ago

Easy Apply

Full-time

$110,000 - $120,000

Remote

22d ago

Easy Apply

Full-time

$110,000 - $120,000

Remote

3d ago

Easy Apply

Full-time

125000 - 180000

Search all similar jobs