Position: Lead Software Engineer – Full Stack Cloud Platform
Location: Minnetonka, MN (Onsite)
Position Summary
We are seeking a highly technical and hands-on Lead Software Engineer to drive the design, development, modernization, and operational excellence of large-scale cloud-native platforms. This role will provide technical leadership across application architecture, microservices, cloud infrastructure, integrations, DevOps, observability, resiliency, and security.
The ideal candidate will lead a team of engineers while actively contributing to architecture decisions, software development, cloud deployments, production support, and platform modernization initiatives.
Key Responsibilities
Architecture & Technical Leadership
• Define and evolve enterprise-scale cloud and application architectures.
• Establish architecture standards, design patterns, security controls, and development best practices.
• Drive high-availability, disaster recovery, scalability, and resiliency strategies.
• Review solution designs and perform architecture governance.
Full Stack Application Development
• Design, develop, and maintain highly scalable web applications.
• Build modern responsive front-end applications.
• Develop RESTful and event-driven APIs.
• Ensure performance optimization, caching, security, and application reliability.
Cloud & Platform Engineering
• Design and implement cloud-native solutions in Microsoft Azure.
• Lead containerization initiatives using Docker and Kubernetes (AKS).
• Implement cloud infrastructure automation using Terraform and Infrastructure as Code (IaC).
• Manage cloud networking, traffic routing, CDN integration, API gateways, and security controls.
Microservices Engineering
• Design and develop distributed microservices.
• Implement reactive programming patterns.
• Build resilient services using Spring Boot, Spring WebFlux, Node.js, circuit breakers, caching, asynchronous processing, and API Gateway patterns.
DevOps & Automation
• Build and maintain CI/CD pipelines.
• Automate testing, deployments, monitoring, and operational processes.
• Implement blue-green and zero-downtime deployment strategies.
• Drive engineering productivity and release automation.
Reliability Engineering & Operations
• Lead production support and incident response.
• Implement observability solutions including logging, monitoring, health checks, alerting, and distributed tracing.
• Perform root cause analysis and drive permanent fixes.
Data & Integration Engineering
• Build scalable integrations with internal and external platforms.
• Design solutions utilizing PostgreSQL, NoSQL databases, MySQL, Redis, event streaming, and REST APIs.
• Drive API governance and integration security.
Team Leadership
• Lead and mentor software engineers.
• Conduct code reviews and design reviews.
• Establish engineering standards and quality metrics.
• Collaborate with product, architecture, security, operations, and business stakeholders.
Required Qualifications
Experience
• 10+ years of software engineering experience.
• 5+ years leading engineering teams or major platform initiatives.
• Experience developing and operating mission-critical customer-facing applications.
• Experience supporting large-scale digital platforms with high transaction volumes.
Programming Languages
• Java
• Node.js
• JavaScript / TypeScript
• SQL
Front-End Technologies
• Angular
• React
• HTML5
• CSS3
• Bootstrap
• SPA Architecture
Backend Technologies
• Spring Boot
• Spring MVC
• Spring WebFlux
• REST Services
• Microservices
• Event-Driven Architecture
• Reactive Programming
Cloud & Infrastructure
• Microsoft Azure
• Azure Kubernetes Service (AKS)
• Azure App Services
• Azure API Management
• Azure Traffic Manager
• Azure CDN
• Azure Redis
• Azure AI Services
Database Technologies
• Azure PostgreSQL
• NoSQL DB
• MySQL
• Redis Cache
DevOps
• Git
• Azure DevOps / GitHub Actions
• CI/CD Pipelines
• Docker
• Kubernetes
• Terraform
• Automated Testing
Observability
• Application Monitoring
• Distributed Tracing
• Log Analytics
• Performance Tuning
Site Reliability Engineering (SRE)
Strong understanding of Site Reliability Engineering practices, production operations, resiliency, observability, incident management, performance, and continuous improvement.
Optum | Minnesota (MN)