Principal Software Engineer - HW/SW in Fleet Infrastructure

Redmond, WA, US • Posted 1 day ago • Updated 1 hour ago
Full Time
On-site
USD $142,800.00 - 274,800.00 per year
Fitment

Dice Job Match Score™

🧠 Analyzing your skills...

Job Details

Skills

  • Reliability Engineering
  • Operational Excellence
  • IT Management
  • Clarity
  • Leadership
  • Collaboration
  • Accountability
  • Debugging
  • Hardware Troubleshooting
  • ROOT
  • SAN
  • Microsoft Office
  • Microsoft Azure
  • Object Data Manager
  • Oracle Data Mining
  • Scalability
  • Artificial Intelligence
  • Storage
  • Technical Direction
  • Mentorship
  • Screening
  • PASS
  • Cloud Computing
  • Computer Science
  • C
  • C++
  • C#
  • Java
  • JavaScript
  • Python
  • IT Strategy
  • Data Storage
  • Computer Networking
  • New Product Introduction
  • Lifecycle Management
  • Failure Analysis
  • Computer Hardware
  • Firmware
  • Communication
  • IaaS
  • Software Engineering
  • IC
  • Internal Communications
  • Integrated Circuit
  • SAP BASIS
  • Microsoft
  • Immigration
  • Military

Summary

Overview

The Fleet & Capacity organization within Microsoft 365 Core Platform is responsible for ensuring the right infrastructure is available in the right places at the right time to support Microsoft 365 services at hyperscale. The team drives capacity strategy, hardware platform readiness, fleet health, reliability engineering, lifecycle management, and operational excellence across Microsoft's global cloud infrastructure.

We are looking for a Principal Software Engineer - HW/SW in Fleet Infrastructure to provide technical leadership for the evolution of Substrate's hardware and firmware platforms. In this role, you will shape New Product Introduction strategy, hardware and firmware health architecture, live-site reliability, and the future direction of AI-optimized infrastructure.

You will work across Microsoft 365, Azure infrastructure, hardware engineering teams, silicon providers, OEM(original equipment manufacturer)/ODM (original design manufacturer) partners, and operations organizations to define architecture, influence platform strategy, and solve complex infrastructure problems at global scale.

This role requires strong technical depth, architectural judgment, and the ability to create clarity across highly ambiguous, cross-organizational problems. Internal guidance for Principal engineers emphasizes broad scope, cross-team architecture leadership, strategic influence, and creating conditions for successful collaboration across teams.

This position is based in Redmond, Washington and requires the employee be in the office a minimum of 3 days per week.

Microsoft's mission is to empower every person and every organization on the planet to achieve more. As employees, we come together with a growth mindset, innovate to empower others, and collaborate to realize our shared goals. Each day we build on our values of respect, integrity, and accountability to create a culture of inclusion where everyone can thrive at work and beyond.

Responsibilities

  • Partner with broad teams to design, develop, validate, debug, and optimize infrastructure software across OS, firmware, fleet health, hardware repair, telemetry, and monitoring areas.
  • Define the technical strategy and architecture for New Product Introduction across Microsoft 365 fleet infrastructure, including platform bring-up, validation, deployment readiness, and lifecycle management.
  • Lead hardware and firmware reliability architecture, including health telemetry, diagnostics, failure detection, predictive insights, and remediation automation.
  • Drive failure analysis and root cause investigation for critical live-site incidents involving hardware, firmware, OS, storage, networking, or infrastructure platforms, and drive durable improvements that reduce recurrence.
  • Partner across Microsoft 365, Azure Hardware Systems, silicon vendors, OEM/ODM partners, firmware teams, and operations organizations to improve platform reliability, scalability, and serviceability.
  • Establish engineering standards, observability patterns, and operational mechanisms that improve fleet availability, reduce deployment risk, and strengthen hyperscale infrastructure operations.
  • Evolve M365 substrate infrastructure strategy for AI and agentic workloads across compute, memory, storage, and networking domains.
  • Influence cross-organizational technical direction, communicate complex tradeoffs clearly, and help teams make durable architecture decisions.
  • Mentor early in career engineers.
  • Embody Microsoft's culture and values.

Qualifications

Required Qualifications:

  • Bachelor's Degree in Computer Science or related technical field AND 6+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or Python
    • OR equivalent experience.

Other Requirements:

Ability to meet Microsoft, customer and/or government security screening requirements are required for this role. These requirements include but are not limited to the following specialized security screenings:
  • Microsoft Cloud Background Check: This position will be required to pass the Microsoft Cloud background check upon hire/transfer and every two years thereafter.

Preferred Qualifications:
  • Master's Degree in Computer Science or related technical field AND 8+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or Python
    • OR Bachelor's Degree in Computer Science or related technical field AND 12+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or Python
    • OR equivalent experience.
  • Experience leading architecture and technical strategy for large-scale distributed systems, cloud infrastructure, or hyperscale service platforms.
  • Deep experience with hardware platforms, firmware, OS, drivers, datacenter infrastructure software, storage systems, networking, or platform software.
  • Experience with New Product Introduction, platform validation, deployment readiness, lifecycle management, or fleet-scale hardware operations.
  • Experience on hyperscale fleet live site, using telemetry, observability, failure analysis, predictive diagnostics, or automation to improve infrastructure reliability.
  • Experience working with silicon providers, hardware vendors, OEMs, ODMs, firmware teams, or platform engineering organizations.
  • Experience influencing cross-organizational engineering strategy and driving complex technical programs across multiple teams.
  • Communication skills with the ability to explain technical tradeoffs to senior engineering and business leaders.
  • Microsoft Secure environment/US Gov clouds.

#M365Core #FleetAndCapacity #AIInfrastructure #HardwareReliability #CloudInfrastructure

Software Engineering IC5 - The typical base pay range for this role across the U.S. is USD $142,800 - $274,800 per year. There is a different range applicable to specific work locations, within the San Francisco Bay area and New York City metropolitan area, and the base pay range for this role in those locations is USD $188,000 - $304,200 per year.

Certain roles may be eligible for benefits and other compensation. Find additional benefits and pay information here:
;br>
This position will be open for a minimum of 5 days, with applications accepted on an ongoing basis until the position is filled.

Microsoft is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to age, ancestry, citizenship, color, family or medical care leave, gender identity or expression, genetic information, immigration status, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran or military status, race, ethnicity, religion, sex (including pregnancy), sexual orientation, or any other characteristic protected by applicable local laws, regulations and ordinances. If you need assistance with religious accommodations and/or a reasonable accommodation due to a disability during the application process, read more about requesting accommodations.
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 10494596
  • Position Id: 2dd08c3e85be2f71c49ad93e44bc40d0
  • Posted 1 day ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Redmond, Washington

Today

Full-time

USD 119,800.00 - 234,700.00 per year

Redmond, Washington

Today

Full-time

USD 142,800.00 - 274,800.00 per year

Redmond, Washington

Today

Full-time

USD 119,800.00 - 234,700.00 per year

Redmond, Washington

Today

Full-time

USD 142,800.00 - 274,800.00 per year

Search all similar jobs