HPC/Linux Software Engineer

Remote • Posted 13 hours ago • Updated 13 hours ago
Contract W2
12 Months
No Travel Required
Remote
Depends on Experience
Company Branding Image
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Linux
  • HPC

Summary

Job Description:

We are seeking a Staff Software Engineer to lead the design, development, and evolution of networking software for our massively parallel processing (MPP) platform, the foundation of our database and AI solutions.
You will influence technical direction, mentor engineers, and leverage AI-assisted development tools to accelerate innovation and execution at scale.

Key Responsibilities

  • Architect, design, and evolve scalable, reliable, and fault-tolerant networking software for high-speed, low-latency interconnects, delivering predictable performance across large-scale MPP systems.
  • Evaluate and drive adoption of emerging technologies across operating systems, high-performance networking, adapters, DPUs, accelerators, and interconnect fabrics.
  • Lead complex debugging and root-cause analysis of system-level customer and field issues, including SLES OS crash dump analysis, spanning hardware, firmware, OS, and networking layers.
  • Define and execute targeted research initiatives and proof-of-concepts to validate new technologies, quantify performance, and guide platform decisions.
  • Partner with product, hardware, and systems engineering teams to scope, prototype, benchmark, and productionize platform enhancements.
  • Establish performance benchmarks, validation methodologies, and success metrics for networking and interconnect innovations.
  • Influence platform roadmaps through deep understanding of industry trends, academic research, and partner technologies.
  • Mentor and technically guide other engineers through design reviews, code reviews, and architectural discussions.
  • Leverage AI-assisted coding, analysis, and testing tools to accelerate development cycles and improve code quality and reliability.

Who You'll Work With

In this role, you will operate across the full lifecycle from research and architecture through production deployment, working closely with platform, hardware, and product engineering teams.

What Makes You a Qualified Candidate

Required Technical Skills:

  • Strong background in HPC or large-scale distributed systems development.
  • Proven experience with Linux kernel and driver development in C, including production support.
  • Deep familiarity with bare-metal and virtualized environments, including performance tradeoffs.
  • Expertise in InfiniBand and Ethernet networking, leveraging RDMA and RoCE for low-latency, high-throughput communication.
  • Solid understanding of TCP/IP and UDP networking, along with Linux networking, tuning, and diagnostic tools”.
  • Packet-level analysis and Linux kernel debugging using tools such as tcpdump, kgdb, and crash.
  • Experience designing and optimizing high-throughput, low-latency data transport protocols.
  • Strong knowledge of the Linux kernel, including DKMS, driver lifecycle management, and compatibility across kernel versions.
  • Proficiency in C, Bash, and Python for systems programming, automation, and diagnostics.
  • Experience with massively parallel processing (MPP) using message-passing interfaces.
  • Effective use of modern AI-assisted development tools to accelerate design, coding, and debugging.

Nice to Have:

  • Experience with DPUs, SmartNICs, or hardware offload technologies.
  • Hands-on work with kernel-bypass networking (e.g., RDMA verbs, DPDK, XDP, eBPF).
  • Experience with high-speed Ethernet (100G/200G/400G/800G) and modern interconnect fabrics.
  • Experience tuning systems for NUMA, CPU affinity, cache locality, and memory bandwidth.
  • Exposure to distributed storage or database platforms in production environments.
  • Experience working with hardware vendors (NICs, switches, accelerators) on performance or integration issues.
  • Contributions to open-source networking, kernel, or systems software projects.

Education & Experience

  • Bachelor's degree in Computer Science (distributed systems focus preferred), Computer Engineering, or Electrical Engineering, or equivalent practical experience.
  • 7+ years of experience in high-performance Linux systems or networking software development, with demonstrated technical leadership.

What You'll Bring

  • Confidence and resilience, with the ability to navigate technical disagreement, challenge assumptions, and incorporate feedback constructively.
  • Proven ability to lead and coordinate real-time troubleshooting of critical (P1) customer issues, rapidly diagnosing system-level failures and driving resolution under pressure.
  • Strong influencing skills, capable of aligning cross-functional teams and driving outcomes without direct authority or ownership of resources.
  • Excellent communication skills, with the ability to clearly articulate complex technical findings, remediation plans, and customer impact to both technical and business stakeholders.
  • A collaborative, self-directed mindset paired with strong intellectual curiosity and continuous learning.
  • The ability to thrive in ambiguous, fast-paced environments while bringing clarity, structure, and forward momentum.

Why We Think You'll Love Teradata

We prioritize a people-first culture because we know our people are at the very heart of our success.
We embrace a flexible work model because we trust our people to make decisions about how, when, and where they work.
We focus on well-being because we care about our people and their ability to thrive both personally and professionally.
We are committed to actively working to foster an inclusive environment that celebrates people for all of who they are.

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 10122208
  • Position Id: 9090386
  • Posted 13 hours ago

Company Info

About Abacus Service Corporation

Abacus Service Corporation is a full service employment solutions firm designed around the ability to provide agile contingent workforce solutions.



Formed in 2004 by industry veterans, Abacus Service Corporation implemented guiding principles with best in industry processes and innovative technologies, to form an influential force in employment solutions. Abacus Service Corporation was founded in Farmington Hills, Michigan and has grown to become a nationwide presence with offices in 18 locations and two international offices. Through our locations, Abacus has been able to offer our clients cost effective and quality employment solutions regardless of the geographic coverage based on our successful strategies. Abacus is a privately held company with employees in 27 US states and four Canadian Provinces. Abacus is MBE and WBE certified nationally and upholds our commitment to diversity by adhering to a philosophy of recruiting employees from diverse backgrounds. Our extensive experience, passion to deliver the best in class solutions, and dedication to customer service has allowed Abacus to become the workforce ally of our clientele.
About_Company_OneAbout_Company_Two
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

It looks like there aren't any Similar Jobs for this job yet.

Search all similar jobs