1

Ai Reliability Engineer Jobs in Allen, TX (NOW HIRING)

Site Reliability Engineer III

Plano, TX · On-site

$53.25 - $70.75/hr

As a Site Reliability Engineer III at JPMorgan Chase within the IP, you will solve complex and ... Uses enterprise-authorized AI capabilities within the work environment to accelerate incident ...

Site Reliability Engineer III

Plano, TX · On-site

$53.25 - $70.75/hr

As a Site Reliability Engineer III at JPMorgan Chase within the Employee Platforms, you will solve ... Uses enterprise-authorized AI capabilities within the work environment to accelerate incident ...

Site Reliability Engineer III

Plano, TX · On-site

$53.25 - $70.75/hr

As a Site Reliability Engineer III at JPMorgan Chase within the IP, you will solve complex and ... Uses enterprise-authorized AI capabilities within the work environment to accelerate incident ...

Site Reliability Engineer III

Plano, TX · On-site

$53.25 - $70.75/hr

As a Site Reliability Engineer III at JPMorgan Chase within the Employee Platforms, you will solve ... Uses enterprise-authorized AI capabilities within the work environment to accelerate incident ...

Lead Site Reliability Engineer

Plano, TX · On-site

$54.50 - $72.50/hr

As a Lead Site Reliability Engineer at JPMorgan Chase within the Infrastructure Platforms, Web ... Uses enterprise-authorized AI capabilities within the work environment to accelerate major-incident ...

Senior AI Quality & Reliability Engineer

Irving, TX · On-site

$84K - $114K/yr

You will help advance Vizient's Quality Engineering capabilities beyond traditional software testing toward AI-native validation, AI-assisted testing, runtime observability, reliability engineering ...

Site Reliability Engineer III

Plano, TX

$53.25 - $70.75/hr

As a Site Reliability Engineer III at JPMorgan Chase within the Employee Platforms, you will solve ... Uses enterprise-authorized AI capabilities within the work environment to accelerate incident ...

Site Reliability Engineer III

Plano, TX · On-site

$53.25 - $70.75/hr

As a Site Reliability Engineer III at JPMorgan Chase within the IP, you will solve complex and ... Uses enterprise-authorized AI capabilities within the work environment to accelerate incident ...

Applied AI SRE III - PxE GPS

Dallas, TX · On-site

$20.25 - $27.75/hr

Share Applied AI SRE III - PxE GPS with Facebook Share Applied AI SRE III - PxE GPS with LinkedIn Share Applied AI SRE III - PxE GPS with Twitter Caution against fraudulent job offers. Learn more.

Lead Site Reliability Engineer

Plano, TX · On-site

$54.50 - $72.50/hr

As a Lead Site Reliability Engineer at JPMorgan Chase within the Corporate Technology team , you ... Uses enterprise-authorized AI capabilities within the work environment to accelerate major-incident ...

Applied AI SRE III - PxE GPS

Dallas, TX · On-site

$56.50 - $75/hr

Applied AI Site Reliability Engineer III Role Overview: As an Applied AI Site Reliability Engineer III , you will actively engage in your engineering craft, taking a hands-on approach to the ...

Lead Site Reliability Engineer

Plano, TX · On-site

$53.25 - $70.75/hr

As a Lead Site Reliability Engineer at JPMorgan Chase within the Corporate Technology team, you ... Uses enterprise-authorized AI capabilities within the work environment to accelerate major-incident ...

Lead Site Reliability Engineer

Plano, TX · On-site

$53.25 - $70.75/hr

As a Lead Site Reliability Engineer at JPMorgan Chase within the Corporate Technology team , you ... Uses enterprise-authorized AI capabilities within the work environment to accelerate major-incident ...

Applied AI SRE III - PxE GPS

Dallas, TX · On-site

$56.50 - $75/hr

Applied AI Site Reliability Engineer III Role Overview: As an Applied AI Site Reliability Engineer III , you will actively engage in your engineering craft, taking a hands-on approach to the ...

Lead Site Reliability Engineer

Plano, TX · On-site

$53.25 - $70.75/hr

As Lead Site Reliability Engineer at JPMorgan Chase within the Corporate Technology team, you hold ... Uses enterprise-authorized AI capabilities within the work environment to accelerate major-incident ...

Showing results 21-40

Ai Reliability Engineer information

See Allen, TX salary details

$56.7K

$109.7K

$131.2K

How much do ai reliability engineer jobs pay per year?

As of Sep 2, 2026, the average yearly pay for ai reliability engineer in Allen, TX is $109,735.00, according to ZipRecruiter salary data. Most workers in this role earn between $95,300.00 and $120,000.00 per year, depending on experience, location, and employer.

What is an AI reliability engineer?

AI Reliability Engineers are professionals responsible for ensuring that artificial intelligence systems function reliably, safely, and effectively over time. They work on monitoring AI models in production, identifying and mitigating potential failures, and improving the robustness of AI systems. Their tasks often include testing, validation, performance monitoring, and implementing best practices for maintaining AI infrastructure. By focusing on reliability, they help organizations deploy AI solutions that are dependable and trustworthy in real-world environments.

What are some common challenges AI reliability engineers face when ensuring model robustness in production environments?

Ai Reliability Engineers often encounter challenges such as monitoring AI model performance for drift or unexpected behavior, managing data quality issues, and implementing automated alerting systems for anomalies. In production, it's crucial to ensure that AI models operate consistently and remain reliable under varying conditions and data inputs. Collaborating closely with data scientists, software engineers, and DevOps teams is essential to address these challenges and to continuously improve model reliability and uptime.

What are the key skills and qualifications needed to thrive as an AI reliability engineer, and why are they important?

To thrive as an AI Reliability Engineer, you need a solid background in computer science or engineering, expertise in AI/ML concepts, and experience with software testing and reliability methodologies. Familiarity with tools like TensorFlow, PyTorch, CI/CD pipelines, and reliability testing frameworks, along with certifications in cloud platforms (e.g., AWS Certified Machine Learning), is highly valuable. Analytical thinking, problem-solving abilities, and strong collaboration skills set top performers apart in this role. These skills ensure robust, dependable AI systems that meet performance standards and maintain trust in critical applications.

What is the difference between Ai Reliability Engineer vs Data Scientist?

AspectAi Reliability EngineerData Scientist
Required CredentialsBachelor's or master's in CS, engineering, or related; certifications in AI/MLBachelor's or master's in CS, statistics, or related; certifications in data analysis or ML
Work EnvironmentTech companies, AI-focused teams, engineering departmentsResearch labs, tech firms, analytics teams
Employer & Industry UsageAI product development, machine learning systems, reliability testingData analysis, predictive modeling, business insights

While both roles involve AI and ML, Ai Reliability Engineers focus on ensuring AI system robustness and uptime, whereas Data Scientists analyze data to generate insights and models. The roles often collaborate but serve different primary functions within AI projects.

What are popular job titles related to Ai Reliability Engineer jobs in Allen, TX?

For Ai Reliability Engineer jobs in Allen, TX, the most frequently searched job titles are:

What job categories do people searching Ai Reliability Engineer jobs in Allen, TX look for?

The top searched job categories for Ai Reliability Engineer jobs in Allen, TX are:

What cities near Allen, TX are hiring for Ai Reliability Engineer jobs?

Cities near Allen, TX with the most Ai Reliability Engineer job openings:

Site Reliability Engineer III

JPMorgan Chase & Co

Plano, TX • On-site

$53.25 - $70.75/hr

Full-time

Medical, Retirement

Posted 28 days ago


JPMorgan Chase & Co. rating

8.0

Company rating: 8.0 out of 10

Based on 499 frontline employees who took The Breakroom Quiz

72nd of 174 rated banks


Job description

There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. 
As a Site Reliability Engineer III at JPMorgan Chase within the IP, you will solve complex and broad business problems with simple and straightforward solutions. Through code and cloud infrastructure, you will configure, maintain, monitor, and optimize applications and their associated infrastructure to independently decompose and iteratively improve on existing solutions. You are a significant contributor to your team by sharing your knowledge of end-to-end operations, availability, reliability, and scalability of your application or platform. 
Job responsibilities
  • Guides and assists others in the areas of building appropriate level designs and gaining consensus from peers where appropriate, supporting adoption of site reliability engineering best practices within your team
  • Collaborates with other software engineers and teams to design, develop, test, and implement deployment and reliability approaches using automated continuous integration and continuous delivery pipelines
  • Uses enterprise-authorized AI capabilities within the work environment to accelerate incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements.
  • Implements infrastructure, configuration, and network as code for the applications and platforms in your remit
  • Collaborates with technical experts, key stakeholders, and team members to resolve complex problems and proactively address issues using service level indicators and objectives before they impact customers
  • Applies enterprise-authorized AI capabilities within the work environment to identify patterns in operational signals that indicate reliability risk or recurring toil, prioritizing reuse-first improvements tied to SLO outcomes.
  • Familiar with availability, reliability, scalability, and solutions in their applications and works with partners to improve these outcomes iteratively
  • Proactively recognizes road blocks and identifies improvements to solve business problems, including exploring new technologies where appropriate
Required qualifications, capabilities, and skills
  • Formal training or certification on site reliability engineering concepts and 3+ years applied experience
  • Exposure to or hands-on experience in supporting SRE practices for Data management/migration platforms and products, with familiarity in on-prem/public cloud infrastructure components such as Compute(Linux/windows), Storage, Networks and Database.
  • Understanding of how to apply SRE fundamentals - including monitoring, incident response, capacity awareness, and toil identification - with the ability to define and track relevant SLOs/SLIs.
  • Proficient in site reliability culture and principles and familiarity with how to implement site reliability within an application or platform
  • Proficient in at least one programming language or configuration/resource management tools such as Python, Ansible and Terraform.
  • Experience in observability including white and black box monitoring, service level objective alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, etc.
  • Experience with platforms and applications hosted on public/private/hybrid cloud environments, including container orchestration technologies such as Kubernetes, ECS, and Docker.
  • Experience with continuous integration and continuous delivery tools such as Jenkins, GitLab, or Terraform.
  • Familiarity with troubleshooting common networking technologies and issues.
  • Experience with continuous integration and continuous delivery tooling
  • Familiarity with container and container orchestration and troubleshooting common networking technologies and issues
Preferred qualifications, capabilities, and skills
  • Exposure to or hands-on experience in supporting SRE practices for Data management/migration platforms and products, with familiarity in on-prem/public cloud infrastructure components such as Compute(Linux/windows), Storage, Networks and Database.
  • Understanding of how to apply SRE fundamentals - including monitoring, incident response, capacity awareness, and toil identification - with the ability to define and track relevant SLOs/SLIs.
  • Proficient in site reliability culture and principles and familiarity with how to implement site reliability within an application or platform
  • Proficient in at least one programming language or configuration/resource management tools such as Python, Ansible and Terraform.
  • Experience in observability including white and black box monitoring, service level objective alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, etc.
  • Experience with platforms and applications hosted on public/private/hybrid cloud environments, including container orchestration technologies such as Kubernetes, ECS, and Docker.
 
JPMorganChase, one of the oldest financial institutions, offers innovative financial solutions to millions of consumers, small businesses and many of the world's most prominent corporate, institutional and government clients under the J.P. Morgan and Chase brands. Our history spans over 200 years and today we are a leader in investment banking, consumer and small business banking, commercial banking, financial transaction processing and asset management.

We offer a competitive total rewards package including base salary determined based on the role, experience, skill set and location. Those in eligible roles may receive commission-based pay and/or discretionary incentive compensation, paid in the form of cash and/or forfeitable equity, awarded in recognition of individual achievements and contributions. We also offer a range of benefits and programs to meet employee needs, based on eligibility. These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare, tuition reimbursement, mental health support, financial coaching and more. Additional details about total compensation and benefits will be provided during the hiring process. 

We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants' and employees' religious practices and beliefs, as well as mental health or physical disability needs. Visit our FAQs for more information about requesting an accommodation.

JPMorgan Chase & Co. is an Equal Opportunity Employer, including Disability/Veterans

Our professionals in our Corporate Functions cover a diverse range of areas from finance and risk to human resources and marketing. Our corporate teams are an essential part of our company, ensuring that we're setting our businesses, clients, customers and employees up for success.

What JPMorgan Chase & Co. employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom