Senior Performance Engineer

Atlanta, GA, US • Posted 1 hour ago • Updated 1 hour ago
Full Time
On-site
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • ICE
  • Recruiting
  • Regression Analysis
  • Quality Assurance
  • Accountability
  • Derivatives
  • Order Entry
  • Test Suites
  • CPU
  • Thread
  • Database
  • Corrective And Preventive Action
  • Capacity Management
  • Non-functional Testing
  • Failover
  • Scalability
  • Expect
  • PASS
  • Auditing
  • UPS
  • ROOT
  • Computer Science
  • Performance Engineering
  • Performance Testing
  • Apache JMeter
  • HP LoadRunner
  • IT Management
  • Communication
  • Distribution
  • Linux
  • Scheduling
  • JIT
  • Algorithms
  • TCP/IP
  • UDP
  • Multicast
  • Network
  • Java
  • C++
  • Python
  • Artificial Intelligence
  • Workflow
  • Test Scripts
  • Root Cause Analysis
  • Technical Drafting
  • Microsoft Certified Professional
  • Servers
  • Financial Services
  • Trading
  • Market Analysis
  • Computer Networking
  • Remote Direct Memory Access
  • FPGA
  • BIOS
  • Firmware
  • Management
  • PTP
  • PPS
  • Computer Hardware
  • Software Performance Management
  • Dynatrace
  • AppDynamics
  • CHAOS
  • Testing
  • Disaster Recovery
  • Microsoft Exchange

Summary

Overview

Job Purpose

Intercontinental Exchange (ICE) is hiring a Senior Performance Engineer to help own performance for our derivatives trading platform - order entry, matching engine, and market data distribution. These are latency-sensitive, business-critical systems where a regression measured in microseconds is a business event.

You will join a small, established performance engineering team. We are hiring a senior engineer to take direct ownership of a substantial share of the platform's performance work. This is not a support seat - you will run campaigns and own conclusions from day one, with an experienced team around you.

The role spans the full performance lifecycle: designing and running test campaigns, profiling pre-production code to surface bottlenecks before release, and diagnosing degradation in production. It is hands-on and deeply technical, suited to an engineer who is equally comfortable profiling a JVM under load, reading a flame graph, and explaining a latency regression to a development team in terms they can act on.

You will use AI tooling to accelerate test design, result analysis, and incident diagnosis - while remaining personally accountable for measurement quality and for every technical conclusion you present.

Responsibilities
  • Own end-to-end performance validation for assigned areas of the derivatives trading platform: order entry, matching engine, and market data distribution paths.
  • Design and execute load, stress, and capacity test suites against pre-production and production-like environments.
  • Profile pre-production services to identify bottlenecks across CPU, memory, I/O, threading, garbage collection, network, and database layers, and partner with development teams to resolve findings before release.
  • Establish and maintain production performance baselines; proactively identify degradation and emerging risk.
  • Investigate latency and capacity events in production, producing formal root cause analyses and corrective action plans.
  • Contribute to capacity planning and infrastructure sizing recommendations grounded in measured behavior and growth projections.
  • Plan and execute non-functional testing - resiliency, failover, scalability, and capacity validation. Expect this to be roughly 20% of your time.
  • Document methodology, results, and findings to a standard suitable for internal engineering audiences and, where applicable, audit and regulatory review.
  • Contribute to the team's shared tooling, harnesses, and measurement practice.

AI-Assisted Performance Engineering
  • AI tooling is part of the standard workflow here, not an experiment. You will use it to draft workload models and test scaffolding from production metrics, SLAs, and prior incidents; to accelerate analysis of profiler output, logs, traces, and packet captures; and to shorten first-pass RCA and produce engineering- and audit-ready write-ups.
  • The bar is that you validate what comes back. You own the statistical soundness of the workload model, the identification of the bottleneck, the stated root cause, and the corrective actions - regardless of what produced the first draft.

Knowledge and Experience
  • Bachelor's degree in Computer Science, Engineering, or a related technical discipline, or equivalent practical experience.
  • 5+ years in performance engineering and testing.
  • Demonstrated experience designing workload models and executing performance test campaigns against distributed, multi-tier systems, using tools such as JMeter, Gatling, k6, or LoadRunner.
  • Demonstrated experience diagnosing performance problems in latency-sensitive systems using profilers and system tooling - for example perf, eBPF/bpftrace, async-profiler, JFR, flame graphs, or Intel VTune.
  • A rigorous statistical approach to measurement: you understand why averages mislead, how coordinated omission distorts results, and how to build tests that produce defensible numbers.
  • Excellent written and verbal communication, including the ability to explain a latency distribution to a room that contains both kernel engineers and business stakeholders.
  • Deep working knowledge in at least two of the following, with working familiarity across the rest:
    • Linux internals as they relate to performance - scheduling, memory management, interrupt affinity, kernel and network tuning.
    • JVM internals - the Java memory model, JIT behavior, escape analysis, algorithms and their pause characteristics (G1, Z Shenandoah), off-heap memory management.
    • Networking fundamentals - TCP/IP, UDP multicast, packet capture and analysis, and measurement of network-induced latency and loss.

Preferred Knowledge and Experience
  • Proficiency in Java, C++, or Python.
  • Applied experience using AI coding agents in a routine performance or development workflow - test-script authoring, log and metric analysis, RCA drafting. Hands-on experience with Claude Code or a comparable agent, including agent skills or MCP servers connecting agents to APM platforms, logs, or test harnesses, is a plus.
  • Experience with low-latency or high-throughput systems, ideally in financial services, trading, or market data.
  • Kernel-bypass and accelerated networking: Solarflare/Onload, DPDK, RDMA, or FPGA-assisted paths.
  • Hardware and colocation tuning: BIOS/firmware configuration, C-state and frequency management, PTP/PPS time synchronization, hardware timestamping.
  • APM platforms such as Dynatrace, AppDynamics, or OpenTelemetry.
  • Chaos engineering, resiliency testing, or disaster recovery validation.

Schedule
  • Onsite in our Atlanta office five days per week.
  • No on-call rotation.
  • Occasional weekend work to support major platform releases. This is infrequent and scheduled in advance.

#LI-JW1

-

Intercontinental Exchange, Inc. is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to legally protected characteristics.
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 10123334
  • Position Id: 2ca0eff5a2f6d8227cb58ef00465332e
  • Posted 1 hour ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Atlanta, Georgia

Today

Full-time

Remote

Today

Full-time

USD 96,569.00 - 130,651.00 per year

Remote

Today

Easy Apply

Full-time

$32 - $42 per hour

Remote

12d ago

Easy Apply

Full-time

Depends on Experience

Search all similar jobs