Software Engineer

Des Moines, IA, US • Posted 2 days ago • Updated 9 hours ago
Contract Independent
On-site
USD $68.00 - 73.00 per hour
Company Branding Image
Fitment

Dice Job Match Score™

🤯 Applying directly to the forehead...

Job Details

Skills

  • Information Architecture
  • Information Assurance
  • Impact Analysis
  • Customer Facing
  • Mortgage
  • Operational Risk
  • Data Link Layer
  • Network Layer
  • Problem Management
  • Dashboard
  • Continuous Improvement
  • Capacity Management
  • Testing
  • Service Level
  • Operational Efficiency
  • Middleware
  • Root Cause Analysis
  • Corrective And Preventive Action
  • Systems Engineering
  • Cloud Computing
  • Command-line Interface
  • Continuous Integration
  • Continuous Delivery
  • Database
  • Banking
  • Financial Services
  • Health Care
  • Insurance
  • Large Language Models (LLMs)
  • Use Cases
  • Budget
  • Analytics
  • Offshoring
  • Accountability
  • Operational Excellence
  • Customer Experience
  • Mentorship
  • ITIL
  • Grafana
  • Splunk
  • AppDynamics
  • Jenkins
  • Terraform
  • Ansible
  • SQL
  • MongoDB
  • Oracle
  • Java
  • .NET
  • Unix
  • Linux
  • Incident Management
  • Change Management
  • Machine Learning (ML)
  • Production Support
  • Artificial Intelligence
  • Reliability Engineering
  • Privacy
  • Marketing

Summary

Location: Des Moines, IA Salary: $68.00 USD Hourly - $73.00 USD Hourly Description:
Site Reliability Engineer (SRE) / Production Support Engineer

We are not accepting C2C or 1099 arrangements.

Location: Des Moines, IA (Preferred) | Minneapolis, MN | Irving (Las Colinas), TX
Duration: 18-Month Contract (Potential Extension to 24 Months)
Work Model: Hybrid (3 days onsite, 2 days remote)
About the Role

This role is ideal for an experienced Site Reliability Engineer (SRE) or Production Support Engineer who thrives in large-scale enterprise environments and enjoys improving reliability, observability, automation, and operational excellence.

You will support critical customer-facing mortgage and home lending platforms, ensuring system stability, performance, and operational resilience. The role combines production support, incident management, observability engineering, automation, and emerging AI-enabled operational practices.

You will partner with engineering, platform, infrastructure, and business teams to reduce operational risk, improve service reliability, and drive self-healing capabilities across complex technology ecosystems.
What You'll Do
  • Lead L2/L3 production support activities for mission-critical applications and platforms.
  • Serve as a primary responder and coordinator for incident management, problem management, and change management activities using ITIL best practices.
  • Monitor the health and performance of applications, infrastructure, and services across hybrid cloud and on-premises environments.
  • Build, enhance, and maintain business observability dashboards using Grafana and related monitoring technologies.
  • Analyze logs, metrics, traces, and events using platforms such as Splunk, AppDynamics, BigPanda, and Application Insights.
  • Drive continuous improvement initiatives to reduce incidents, improve service availability, and optimize system performance.
  • Support reliability engineering efforts, including capacity planning, resiliency testing, operational readiness, and service-level objectives (SLOs).
  • Develop and maintain automation solutions using Ansible and related technologies to improve operational efficiency.
  • Support CI/CD pipelines and deployment processes using tools such as Jenkins, Artifactory, UDeploy, and Terraform.
  • Troubleshoot and resolve complex production issues involving Java, .NET, database, middleware, and infrastructure components.
  • Partner with development teams to implement resilient, scalable, and observable system designs.
  • Leverage knowledge of AI/ML, LLM, and Agentic AI technologies to identify operational efficiencies and innovative support solutions.
  • Participate in major incident reviews and contribute to root cause analysis and preventive action planning.
  • Support Oracle, MSSQL, and other enterprise database technologies through analysis and troubleshooting.
Minimum Qualifications
  • 8+ years of experience in Site Reliability Engineering, Production Support, Systems Engineering, or related technical roles.
  • Experience leading production support operations within large-scale enterprise environments.
  • Strong experience with ITIL-based incident, problem, and change management practices.
  • Hands-on experience with observability and monitoring technologies, including:
    • Grafana
    • Splunk
    • AppDynamics
    • BigPanda
    • Application Insights
  • Experience supporting distributed applications across on-premises, hybrid, and cloud platforms.
  • Strong troubleshooting skills using Unix/Linux command-line tools.
  • Experience supporting large-scale Java and/or .NET applications.
  • Experience with Oracle, MSSQL, MongoDB, or similar database technologies.
  • Experience working with CI/CD tools including Jenkins, Artifactory, UDeploy, and Terraform.
  • Experience implementing automation solutions using Ansible.
  • Strong SQL skills with the ability to analyze and troubleshoot application and database issues.
Preferred Qualifications
  • Site Reliability Engineering (SRE) experience in a highly regulated industry such as banking, financial services, healthcare, or insurance.
  • Experience supporting AI/ML platforms and Large Language Model (LLM)-based systems.
  • Understanding of Agentic AI concepts, use cases, operational impacts, and efficiency improvements.
  • Experience building self-healing and autonomous operational solutions.
  • Strong knowledge of reliability engineering principles, including SLIs, SLOs, and error budgets.
  • Experience designing highly available and resilient systems.
  • Familiarity with modern observability practices, telemetry collection, and operational analytics.
  • Experience collaborating with offshore support teams.
Preferred Attributes

Google values candidates who:
  • Solve ambiguous and complex technical problems with a data-driven approach.
  • Demonstrate ownership and accountability for production systems.
  • Communicate effectively across technical and non-technical stakeholders.
  • Continuously improve processes through automation and operational excellence.
  • Have a passion for reliability, customer experience, and scalable engineering practices.
  • Mentor teams and drive adoption of modern SRE methodologies.
Key Skills
  • Site Reliability Engineering (SRE)
  • Production Support
  • ITIL
  • Observability Engineering
  • Grafana
  • Splunk
  • AppDynamics
  • BigPanda
  • Application Insights
  • Jenkins
  • Terraform
  • Artifactory
  • UDeploy
  • Ansible
  • SQL
  • MongoDB
  • Oracle
  • MSSQL
  • Java
  • .NET
  • Unix/Linux
  • Incident Management
  • Change Management
  • Agentic AI
  • AI/ML Operations
  • LLM Support
  • Automation & Self-Healing Systems

Target Candidate: A senior SRE-minded engineer with strong production support expertise, hands-on observability experience, and a solid understanding of AI-enabled operational practices who can help elevate the team's reliability engineering capabilities.

By providing your phone number, you consent to: (1) receive automated text messages and calls from the Judge Group, Inc. and its affiliates (collectively "Judge") to such phone number regarding job opportunities, your job application, and for other related purposes. Message & data rates apply and message frequency may vary. Consistent with Judge's Privacy Policy, information obtained from your consent will not be shared with third parties for marketing/promotional purposes. Reply STOP to opt out of receiving telephone calls and text messages from Judge and HELP for help.
Contact:
This job and many more are available through The Judge Group. Please apply with us today!
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: cxjudgpa
  • Position Id: 1144684
  • Posted 2 days ago

Company Info

About Judge Group, Inc.

The Judge Group, is a leading professional services firm specializing in talent, technology, and learning solutions. We consult, staff, train, and solve. Through our work we make people and organizations better.

Our services are successfully delivered through a network of more than 30 offices across the United States, Canada, and India. The Judge Group is proud to partner with the best and brightest companies in business today, including over 60 of the Fortune 100. We serve organizations in financial services, healthcare, life sciences, insurance, government (including aerospace and defense), manufacturing, and technology and telecommunications.

About_Company_OneAbout_Company_Two
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Des Moines, Iowa

Today

Contract

USD 120,000.00 - 150,000.00 per year

Des Moines, Iowa

Today

Contract

Columbus, Ohio

Today

Contract

USD 68.00 - 73.00 per hour

Ann Arbor, Michigan

Today

Contract

USD 50.00 - 55.00 per hour

Search all similar jobs