1

Ai Reliability Engineer Jobs in Oregon (NOW HIRING)

Site Reliability Engineer TELCOR Inc, a leading innovator in laboratory software, is looking for a Site Reliability Engineer to join our TELCOR AI Systems team! Do you have strong experience in cloud ...

Sr. Site Reliability Engineer

OR · On-site +1

$57 - $75.75/hr

FreedomPay is seeking an experienced Senior Site Reliability Engineer to help ensure the highest ... We expect AI and automation to be a force multiplier in everything you do -- from accelerating root ...

$57 - $75.75/hr

As a DevSecOps/SRE Engineer at LMI, you will be at the forefront of modernizing and maintaining ... ready AI to federal agencies at commercial speed. Leveraging our mission-ready technology and ...

OR

$102K - $128K/yr

Waabi, founded by AI visionary Raquel Urtasun, is the leader in Physical AI. With a world-class ... reliability. - Participate in both pre-test and post-test evaluations, ensuring vehicles are test ...

Senior Site Reliability Engineer, Government

OR · Remote

$57 - $75.75/hr

As a Senior Site Reliability Engineer, you will join our Government SRE team and own both the ... AI is redefining how the world operates and rewriting the rules of security in real time, and ...

OR · On-site

$548K - $899K/yr

Demonstrated experience improving SRE AI fluency. Generally, our compensation structure consists solely of an annual salary; we do not have bonuses. You choose each year how much of your compensation ...

OR

$102K - $128K/yr

The Customer Reliability Engineering team is the deep technical escalation tier for ... the AI era - and beyond. We've been innovating fearlessly for 40 years to create solutions that ...

... AI-powered Intelligent Agreement Management (IAM) platform, our infrastructure must be as innovative as our products. We are seeking a Senior Manager of SRE who operates with a Customer-First mindset ...

Site Reliability Engineer, Linux (Remote)

OR · On-site +1

$57 - $75.75/hr

S. soil. MEET THE TEAM The Linux SRE team at Cisco serves as the foundational backbone of our ... the AI era - and beyond. We've been innovating fearlessly for 40 years to create solutions that ...

$57 - $75.75/hr

... ready AI to federal agencies at commercial speed. Leveraging our mission-ready technology and ... s Engineer with a strong focus on observability and reliability to join our health project team.

Staff Senior Reliability Engineer

OR · Remote

$125K - $171K/yr

What we need You'll be the Senior Reliability Engineer who owns complex production investigations ... Kubernetes, VMware, Linux, Grafana, Prometheus, Zabbix, GitLab, and AI-assisted troubleshooting and ...

Senior Infrastructure Engineer/SRE

OR · On-site +1

$108K - $147K/yr

Building machine learning infrastructure that enables AI teams to train, test, and deploy on large-scale datasets. What we are looking for: * 5+ years experience in DevOps, Site Reliability ...

OR · On-site

Integrate AI review signals with our SRE observability stack (logs, metrics, traces) to correlate code changes with incidents and anomalies, closing the loop from PR to production behavior. * Encode ...

This role emphasizes reliability, control, observability, data quality, and governance for agentic ... DevOps and Reliability for AI Agent Systems Define and track SLIs/SLOs for task completion ...

OR

$179K - $231K/yr

As VP of Engineering, AI Innovations, you will lead the our team of talented software developers ... Deep knowledge of SRE principles: SLIs/SLOs, incident management, error budgets. * Experience ...

... AI agents, and serves the Financial Intelligence Graph to Fortune 500 customers who expect ... Strong SRE instincts: you think in terms of SLOs, capacity planning, incident response, and ...

next page

Showing results 1-20

Ai Reliability Engineer information

What are the key skills and qualifications needed to thrive as an AI Reliability Engineer, and why are they important?

To thrive as an AI Reliability Engineer, you need a solid background in computer science or engineering, expertise in AI/ML concepts, and experience with software testing and reliability methodologies. Familiarity with tools like TensorFlow, PyTorch, CI/CD pipelines, and reliability testing frameworks, along with certifications in cloud platforms (e.g., AWS Certified Machine Learning), is highly valuable. Analytical thinking, problem-solving abilities, and strong collaboration skills set top performers apart in this role. These skills ensure robust, dependable AI systems that meet performance standards and maintain trust in critical applications.

What is the difference between Ai Reliability Engineer vs Data Scientist?

AspectAi Reliability EngineerData Scientist
Required CredentialsBachelor's or master's in CS, engineering, or related; certifications in AI/MLBachelor's or master's in CS, statistics, or related; certifications in data analysis or ML
Work EnvironmentTech companies, AI-focused teams, engineering departmentsResearch labs, tech firms, analytics teams
Employer & Industry UsageAI product development, machine learning systems, reliability testingData analysis, predictive modeling, business insights

While both roles involve AI and ML, Ai Reliability Engineers focus on ensuring AI system robustness and uptime, whereas Data Scientists analyze data to generate insights and models. The roles often collaborate but serve different primary functions within AI projects.

What are AI Reliability Engineers?

AI Reliability Engineers are professionals responsible for ensuring that artificial intelligence systems function reliably, safely, and effectively over time. They work on monitoring AI models in production, identifying and mitigating potential failures, and improving the robustness of AI systems. Their tasks often include testing, validation, performance monitoring, and implementing best practices for maintaining AI infrastructure. By focusing on reliability, they help organizations deploy AI solutions that are dependable and trustworthy in real-world environments.

What are some common challenges Ai Reliability Engineers face when ensuring model robustness in production environments?

Ai Reliability Engineers often encounter challenges such as monitoring AI model performance for drift or unexpected behavior, managing data quality issues, and implementing automated alerting systems for anomalies. In production, it's crucial to ensure that AI models operate consistently and remain reliable under varying conditions and data inputs. Collaborating closely with data scientists, software engineers, and DevOps teams is essential to address these challenges and to continuously improve model reliability and uptime.
What are popular job titles related to Ai Reliability Engineer jobs in Oregon? For Ai Reliability Engineer jobs in Oregon, the most frequently searched job titles are:
What job categories do people searching Ai Reliability Engineer jobs in Oregon look for? The top searched job categories for Ai Reliability Engineer jobs in Oregon are:
What cities in Oregon are hiring for Ai Reliability Engineer jobs? Cities in Oregon with the most Ai Reliability Engineer job openings:
Site Reliability Engineer

Site Reliability Engineer

TELCOR Inc

On-site, Remote

$125K - $165K/yr

Full-time

Posted 3 days ago


Job description

Site Reliability Engineer
TELCOR Inc, a leading innovator in laboratory software, is looking for a Site Reliability Engineer to join our TELCOR AI Systems team! Do you have strong experience in cloud infrastructure, distributed systems and production operations? Do you focus on uptime, resilience, observability and incident response? This may be the role for you! Along with the challenging and rewarding aspects of the position, TELCOR offers a culture where an individual is valued, encouraged to grow, and is a contributor to the future of the Company. TELCOR's headquarters is in Lincoln, NE and we also have an office in San Francisco, CA; this position can also be remote.
Copy and paste the following link into your browser to learn more about TELCOR and what it means for TELCOR to be a certified Great Place To Work®: https://www.greatplacetowork.com/certified-company/7054288​
The Site Reliability Engineer will help ensure the reliability, scalability, and performance of the systems that power our AI products. This role will also design and operate resilient systems across cloud and containerized environments, as well as manage production infrastructure and deployment workflows across environments.
Pay Range: $125,000 - $165,000 annually
Requirements
  • Experience with distributed systems
  • Experience with Redis
  • Experience with queuing systems
  • 2+ years of experience with Kubernetes
  • Experience with AWS
  • 2+ years of experience with Terraform
  • Experience with observability

TELCOR is an equal opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability or protected veteran status, or any other characteristic protected by law. Qualified applicants will be considered based on their individual merit alone.
TELCOR complies with the San Francisco Fair Chance Ordinance, the Los Angeles Fair Chance Initiative for Hiring, and the California Fair Chance Act by considering qualified applicants with arrest and conviction records in accordance with these regulations.
All trademarks, service marks, trade names, trade dress, product names and logos appearing herein are the property of their respective owners.
Microsoft is either registered trademark or trademark of Microsoft Corporation in the United States and/or other countries.