1

Ai Reliability Engineer Jobs in Minnesota (NOW HIRING)

The role offers hands-on experience with application design, software development, automated testing, and Site Reliability Engineering (SRE) practices, leveraging AI and automation to drive ...

AI-driven SRE & DevOps mindset -- familiarity with AIOps (anomaly detection, intelligent alerting, predictive scaling, automated remediation) and comfort applying AI/LLMs to simplify how we operate ...

New

AI Engineer

Minnetonka, MN · On-site

$124K - $164K/yr

... reliability, measurable outcomes, security, maintainability, and appropriate use of AI in health ... software engineering experience, including ownership of production systems. * Strong software ...

New

DevOps Engineer

Minneapolis, MN · Hybrid

$55 - $75.50/hr

Today it is the platform that drives AI adoption across the enterprise. Enterprise work stays ... Participate in evaluating and integrating new technologies to enhance the scalability, reliability ...

... reliability engineering standards embedded into development standards - Embraces emerging ... Developed and deployed AI/GenAI-powered applications utilizing Large Language Models (LLMs ...

Showing results 41-60

Ai Reliability Engineer information

What is an AI reliability engineer?

AI Reliability Engineers are professionals responsible for ensuring that artificial intelligence systems function reliably, safely, and effectively over time. They work on monitoring AI models in production, identifying and mitigating potential failures, and improving the robustness of AI systems. Their tasks often include testing, validation, performance monitoring, and implementing best practices for maintaining AI infrastructure. By focusing on reliability, they help organizations deploy AI solutions that are dependable and trustworthy in real-world environments.

What are some common challenges AI reliability engineers face when ensuring model robustness in production environments?

Ai Reliability Engineers often encounter challenges such as monitoring AI model performance for drift or unexpected behavior, managing data quality issues, and implementing automated alerting systems for anomalies. In production, it's crucial to ensure that AI models operate consistently and remain reliable under varying conditions and data inputs. Collaborating closely with data scientists, software engineers, and DevOps teams is essential to address these challenges and to continuously improve model reliability and uptime.

What are the key skills and qualifications needed to thrive as an AI reliability engineer, and why are they important?

To thrive as an AI Reliability Engineer, you need a solid background in computer science or engineering, expertise in AI/ML concepts, and experience with software testing and reliability methodologies. Familiarity with tools like TensorFlow, PyTorch, CI/CD pipelines, and reliability testing frameworks, along with certifications in cloud platforms (e.g., AWS Certified Machine Learning), is highly valuable. Analytical thinking, problem-solving abilities, and strong collaboration skills set top performers apart in this role. These skills ensure robust, dependable AI systems that meet performance standards and maintain trust in critical applications.

What is the difference between Ai Reliability Engineer vs Data Scientist?

AspectAi Reliability EngineerData Scientist
Required CredentialsBachelor's or master's in CS, engineering, or related; certifications in AI/MLBachelor's or master's in CS, statistics, or related; certifications in data analysis or ML
Work EnvironmentTech companies, AI-focused teams, engineering departmentsResearch labs, tech firms, analytics teams
Employer & Industry UsageAI product development, machine learning systems, reliability testingData analysis, predictive modeling, business insights

While both roles involve AI and ML, Ai Reliability Engineers focus on ensuring AI system robustness and uptime, whereas Data Scientists analyze data to generate insights and models. The roles often collaborate but serve different primary functions within AI projects.

What cities in Minnesota are hiring for Ai Reliability Engineer jobs?

Cities in Minnesota with the most Ai Reliability Engineer job openings:

Sr DevOps Engineer-Remote-(If in MN then Hybrid)

UnitedHealth Group

Eden Prairie, MN • Remote

$132K - $170K/yr

Full-time

Retirement

Posted 8 days ago


UnitedHealth Group rating

7.6

Company rating: 7.6 out of 10

Based on 146 frontline employees who took The Breakroom Quiz

192nd of 898 rated healthcare providers


Job description

Our Company

Optum is a global organization that delivers care, aided by technology to help millions of people live healthier lives. The work you do with our team will directly improve health outcomes by connecting people with the care, pharmacy benefits, and data they need to feel their best. Here, you will find a culture guided by diversity and inclusion, talented peers, comprehensive benefits, and career development opportunities. Come make an impact on the communities we serve as you help us advance health equity on a global scale. Join us to start Caring. Connecting. Growing together.

You will enjoy the flexibility to telecommute* from anywhere within the U.S. as you take on some tough challenges.  

Position Summary

As a Senior DevOps Engineer within Optum Technology, you will lead the end-to-end release lifecycle and build enterprise-grade automation that elevates software delivery performance across cloud platforms. In this role, you will design and optimize robust CI/CD pipelines, manage infrastructure as code using Terraform on Azure, and build production-ready AI agents and AI-driven automation to streamline release engineering, deployment support, and operational workflows. You will partner cross-functionally with development, SRE, security, and platform teams to establish high standards of governance, quality, and operational reliability.

Primary Responsibilities

  • Lead the end-to-end release lifecycle for Java-based applications deployed on Azure, enforcing release governance, change standards, deployment readiness criteria, and environment parity practices
  • Design, implement, and optimize CI/CD pipelines using GitHub Actions and Azure DevOps to ensure secure and efficient software delivery across the SDLC
  • Design, develop, deploy, and operationalize production-ready AI agents and agent-based automation to improve release engineering, deployment support, incident triage, and engineering productivity with an emphasis on responsible AI use
  • Build, provision, and manage Azure infrastructure using Terraform across services including AKS, App Services, networking, storage, identity, and monitoring
  • Implement and support modern deployment strategies including blue/green, canary, and rolling releases to enhance release quality, speed, resilience, and observability
  • Identify and implement enterprise automation solutions using approved tools to reduce manual effort, accelerate root cause analysis, and strengthen release reliability
  • Troubleshoot and resolve complex issues across application, pipeline, infrastructure, deployment, and automation layers
  • Partner closely with development, SRE, security, and platform teams to build reusable automation frameworks, release standards, and operational runbooks
  • Evaluate emerging technologies and automation trends to inform solution design, strategic innovation, and continuous process improvement
  • Design, develop, and deploy AI-powered solutions to address complex business challenges with emphasis on responsible use of AI 

You'll be rewarded and recognized for your performance in an environment that will challenge you and give you clear direction on what it takes to succeed in your role as well as provide development for other roles you may be interested in.

Required Qualifications

  • 3 years of experience in Release Engineering, DevOps, Site Reliability Engineering, or Software Engineering

  • 3 years of hands-on experience supporting Java applications, including Spring Boot, Maven (Must have) and/or Gradle, Microservices

  • 3 years of experience with Azure services (such as AKS, App Services, networking, identity, monitoring, and storage) and managing infrastructure with Terraform across the full lifecycle

  • 3 years of CI/CD experience using GitHub Actions, Azure DevOps, or similar pipeline platforms

  • 2 years of Hands-on experience designing, building, and implementing AI agents or agent-based automation for engineering, release, platform, or operational workflows

  • 2 years of experience with Docker, Kubernetes, Git workflows, artifact repositories, deployment strategies, and pipeline quality gates

Preferred Qualifications

  • Bachelor's degree or 4 years of equivalent software engineering or DevOps experience in lieu of a degree
  • Strong understanding of AI-assisted automation, prompt-driven tooling, agent orchestration, and practical enterprise use cases in software delivery and operations
  • Experience defining release processes, deployment controls, and operational best practices in enterprise environments
  • Strong troubleshooting skills across application, build, deployment, infrastructure, and automation domains
  • Ability to work cross-functionally and communicate effectively with engineering, operations, and security stakeholders

*All employees working remotely will be required to adhere to UnitedHealth Group's Telecommuter Policy

Pay is based on several factors including but not limited to local labor markets, education, work experience, certifications, etc. In addition to your salary, we offer benefits such as, a comprehensive benefits package, incentive and recognition programs, equity stock purchase and 401k contribution (all benefits are subject to eligibility requirements). No matter where or when you begin a career with us, you'll find a far-reaching choice of benefits and incentives. The salary for this role will range from $91,700 to $163,700 annually based on full-time employment. We comply with all minimum wage laws as applicable.  

Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

Application Deadline: This will be posted for a minimum of 2 business days or until a sufficient candidate pool has been collected. Job posting may come down early due to volume of applicants.

At UnitedHealth Group, our mission is to help people live healthier lives and make the health system work better for everyone. We believe everyone-of every race, gender, sexuality, age, location, and income-deserves the opportunity to live their healthiest life. Today, however, there are still far too many barriers to good health which are disproportionately experienced by people of color, historically marginalized groups, and those with lower incomes. We are committed to mitigating our impact on the environment and enabling and delivering equitable care that addresses health disparities and improves health outcomes - an enterprise priority reflected in our mission.

Diversity creates a healthier atmosphere: UnitedHealth Group is an Equal Employment Opportunity/Affirmative Action employer and all qualified applicants will receive consideration for employment without regard to race, color, religion, sex, age, national origin, protected veteran status, disability status, sexual orientation, gender identity or expression, marital status, genetic information, or any other characteristic protected by law.

UnitedHealth Group is a drug-free workplace. Candidates are required to pass a drug test before beginning employment.

#RPO, #GREEN


What UnitedHealth Group employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom