
GSK Solutions Inc.
Hybrid in Sinking Spring, Pennsylvania • Today
Easy Apply
Third Party, Contract
Depends on Experience
1 results (1 new)

GSK Solutions Inc.
Hybrid in Sinking Spring, Pennsylvania • Today
Easy Apply
Third Party, Contract
Depends on Experience



|
Job Title |
Sr. Data Engineer - Flink (Hybrid Onsite) |
|
Location |
Reading, PA |
|
Duration |
4 Months |
|
Pay Rate |
$70/hr on C2C / 1099 all inclusive OR $65/hr on W2 |
|
Interview Type |
Virtual / In-Person |
|
Note |
Local would be great, remote is fine if travel is not an issue. At least couple of times a month. So close by is great if not local. |
|
Description |
Job Summary We are seeking a highly skilled Hands-On Data Engineering Lead with deep expertise in Apache Flink, AWS Managed Service for Apache Flink, Kafka, and AWS. This role requires a technical leader who actively architects, develops, debugs, optimizes, and supports production-grade real-time streaming platforms. The ideal candidate combines hands-on engineering depth with experience leading teams that deliver scalable, resilient, high-throughput, and low-latency data solutions in 24x7 environments.
Required Qualifications • Bachelor's or Master's degree in Computer Science, Engineering, Information Systems, or a related field, or equivalent professional experience. • 10+ years of software engineering, data engineering, or distributed-systems experience, including 7+ years of recent hands-on Apache Flink engineering. • Deep production experience with AWS Managed Service for Apache Flink, including deployments, runtime behavior, scaling, service limits, snapshots, monitoring, troubleshooting, and recovery. • Strong command of the Flink DataStream API and Flink SQL, including stateful processing, event time, watermarks, windows, joins, timers, checkpoints, savepoints, and exactly-once concepts. • Demonstrated experience diagnosing backpressure, checkpoint behavior, state growth, skew, idle partitions, memory issues, restarts, Kafka lag, and performance degradation. • Strong programming skills in Java, Python/PyFlink, and SQL, with experience developing and supporting production-grade streaming applications. • Hands-on experience with Kafka or Confluent Kafka and sharded distributed systems, including topic and shard-key design, partitioning, routing, consumer groups, rebalancing, state migration, cross-shard processing, failure isolation, and recovery. • Working knowledge of AWS services and controls supporting streaming platforms, including CloudWatch, S3, IAM, networking, service quotas, CI/CD, and infrastructure as code. • Experience building and operating highly available, high-throughput, low-latency platforms in 24x7 production environments. • Experience with monitoring, alerting, observability, incident response, root-cause analysis, and recovery planning. • Ability to communicate technical findings clearly and distinguish observed facts, estimates, hypotheses, proposals, and approved decisions. • Demonstrated ability to lead technical teams while remaining hands-on in engineering and troubleshooting activities.
Preferred Qualifications • Experience supporting global, mission-critical streaming platforms and mentoring engineering teams. • AWS certification or demonstrated equivalent AWS platform expertise. • Experience processing very large event volumes using multi-shard architectures and managing multi-terabyte state, workload skew, state migration, or disaster-recovery design. Key Responsibilities Hands-On Technical Leadership • Lead and mentor data engineers and architects while remaining directly involved in solution design, implementation, and troubleshooting. • Serve as the technical authority for Apache Flink, AWS Managed Service for Apache Flink, Kafka-based streaming, and associated AWS services. • Lead architecture, design, code, configuration, deployment, and production-readiness reviews. • Establish engineering standards, coding practices, test discipline, and production-support procedures. • Coordinate technical decisions across application, data, cloud-platform, performance, and operations teams. Real-Time Streaming Architecture & Engineering • Architect, design, and develop production-scale streaming solutions using Apache Flink, AWS Managed Service for Apache Flink, Java or PyFlink, Flink SQL, and Kafka. • Apply stateful and event-time processing patterns, including keyed and broadcast state, state TTL, timers, windows, joins, watermarks, and delivery-semantics controls. • Design resilient checkpointing, savepoint, restart, recovery, and state-migration strategies. • Design Kafka topics, partitions, shard keys, routing, consumer groups, offsets, transactions, schemas, and source/sink integrations. • Design and operate sharded streaming architectures, including workload decomposition, shard-key selection, state distribution, rebalancing, cross-shard processing, parallelism, failure isolation, and recovery. • Build fault-tolerant streaming applications that meet defined throughput, latency, availability, and recovery objectives. Performance Engineering & Optimization • Define measurable entry, exit, and acceptance criteria for load, peak, burst, replay, recovery, and long-running soak tests. • Verify that test inputs, replay behavior, duration, measurement windows, metric units, and outputs are representative and comparable. • Analyze throughput, latency, backpressure, state growth, checkpoints, recovery, resource utilization, data skew, and operating headroom. • Identify operator and stage-level bottlenecks and implement validated code, configuration, partitioning, or scaling improvements. • Document test conditions, findings, qualifications, risks, and recommendations using reproducible evidence. Production Engineering & Troubleshooting • Act as a senior escalation point for complex production issues involving Apache Flink, AWS Managed Service for Apache Flink, and Kafka. • Diagnose checkpoint failures, savepoint recovery issues, backpressure, state growth, idle partitions, data skew, memory pressure, garbage collection, restarts, and throughput or latency degradation. • Analyze JobManager and TaskManager events, runtime configuration, logs, metrics, deployment history, and service behavior. • Lead evidence-based root-cause analysis and define corrective and preventive actions for material incidents. • Use controlled experiments to confirm or reject technical hypotheses and validate remediation effectiveness. • Engage AWS Support and service specialists when deeper platform analysis or service-limit clarification is required. Observability & Operational Readiness • Define and implement monitoring for throughput, lag, backpressure, state size, checkpoints, restarts, failures, service events, and recovery. • Standardize metric definitions, units, aggregation windows, data sources, thresholds, alerting, and escalation paths. • Develop and review deployment, incident-triage, rollback, snapshot/savepoint recovery, scaling, and change-control procedures. • Prepare technical documentation, operational runbooks, and knowledge-transfer materials for engineering and support teams. Software Engineering Excellence • Write, debug, test, optimize, and deploy production-grade Java, Python/PyFlink, and SQL code. • Perform code reviews and promote automated testing, code quality, infrastructure-as-code, and DevOps practices. • Build and improve CI/CD pipelines and deployment automation for streaming applications.
|
|
Top Skills |
Custom Skill Requirements:
|
|
Recruiter |
Contact: Lokesh - - Eight three two - Nine nine zero - Two four two siX |
Our values are integrity, leading change, excellence, and respect for the individual, learning and sharing. Our success is based on our ability to be flexible while adhering to a strict project management methodology.
We inspire personal and professional growth in our people through innovation and creativity. We reward excellence. We earn the trust of our customers and the respect of our employees through exceptional teamwork, good business ethics and a high level of commitment
To see how well you match this job, please log in or create an account.
Once logged in, be sure to complete your profile to get the most accurate match score.
Hybrid in Lansing, Michigan
•
Today
Job Title Salesforce Architect (Hybrid Onsite) Location Lansing, MI Duration 12 Months Interview Type In Person Only Note AN IN-PERSON INTERVIEW IS REQUIRED FOR THIS POSITION AND THE CANDIDATE MUST BE ABLE TO WORK THE SPECIFIC 2 DAYS THE TEAM IS ON-SITE (SEE DETAILS BELOW) - Interview Process: Virtual and in-person. 60-minute Virtual Interview via MS Teams (video required). Candidates should join from a laptop and be prepared to share their screen if requested. A screenshot photo of c
Easy Apply
Third Party, Contract
Depends on Experience
Hybrid in Atlanta, Georgia
•
Today
Job Title Senior | Lead] Software Engineer AI Agents (Google Cloud Platform) (Hybrid Onsite) Location Atlanta, GA Duration 6 Months Pay Rate $70/hr on C2C / 1099 all inclusive (OR) $65/hr on W2 Interview Type Virtual / In-Person Note Hybrid Onsite - 3 days Description Build and deploy AI agents and chatbotson Google Cloud Platform, using ADKand modern agent frameworks. Own delivery from rapid POC production. Responsibilities Develop AI agents/chatbotswith tool integration
Easy Apply
Contract, Third Party
Depends on Experience

.png%3Fformat%3Dwebp&w=1080&q=75)