1

Sr Reliability Engineer Jobs in Calgary, AB (NOW HIRING)

We're all in on AWS, combining deep UX capability with senior engineering talent to get AI into ... Drive platform reliability, performance SLAs, and cost optimization across production systems

Senior Developer, Enterprise AI

Calgary, AB · Remote

CA$157K - CA$212K/yr

The Senior Developer, Enterprise AI is a hands-on technical leader responsible for building ... Solid understanding of system design, reliability, observability, and operational best practices.

Partner with senior product engineers to route AI-generated defect fixes through the verification ... Performance and reliability awareness.Knowledge of performance tuning, capacity planning, and ...

Showing results 41-60

Sr Reliability Engineer information

What does a Sr Reliability Engineer do?

A Sr Reliability Engineer is responsible for ensuring that products, systems, or processes operate reliably and efficiently over their expected lifecycle. They analyze failure data, develop reliability test plans, and implement strategies to predict and prevent failures. Their role often involves collaborating with design, manufacturing, and maintenance teams to improve product quality and reduce downtime. Additionally, they may use reliability modeling tools and statistical techniques to assess risk and recommend improvements. This position typically requires advanced engineering knowledge and experience in reliability engineering principles.

What are the key skills and qualifications needed to thrive as a Sr Reliability Engineer?

To thrive as a Sr Reliability Engineer, you need expertise in reliability engineering principles, root cause analysis, and a relevant engineering degree, often with several years of industry experience. Familiarity with reliability software (such as ReliaSoft or Minitab), maintenance management systems, and certifications like Certified Reliability Engineer (CRE) are commonly required. Strong problem-solving abilities, proactive communication, and leadership skills help you drive reliability initiatives and collaborate across teams. These competencies are vital to ensure equipment uptime, reduce failures, and improve operational efficiency in complex industrial environments.

How does a Sr Reliability Engineer typically collaborate with cross-functional teams to improve system reliability?

As a Sr Reliability Engineer, you will regularly work alongside operations, development, and QA teams to identify potential reliability risks and implement solutions. This often involves facilitating root cause analyses after incidents, sharing best practices, and leading reliability-focused design reviews. Effective communication and the ability to translate complex technical findings into actionable recommendations are key, as your insights directly influence infrastructure and product decisions. This collaborative approach helps foster a culture of reliability across the organization.

What is the difference between Sr Reliability Engineer vs Reliability Engineer?

AspectSr Reliability EngineerReliability Engineer
CredentialsBachelor's or higher in engineering, certifications like CRE or Six Sigma often preferredBachelor's degree in engineering or related field, similar certifications
Work EnvironmentTypically in manufacturing, energy, or tech industries focusing on system reliability and failure analysisSimilar industries, focusing on designing, testing, and improving product or system reliability
Employer UsageUsed in companies seeking experienced engineers to lead reliability projectsUsed for entry to mid-level roles focused on reliability assessments

The main difference is experience level and responsibility. Sr Reliability Engineers often lead projects and have more advanced certifications, while Reliability Engineers focus on supporting reliability tasks. Both roles require similar credentials and work in comparable environments, but the senior role involves more leadership and strategic planning.

Are senior reliability engineers in demand?

Senior reliability engineers are in high demand across industries such as manufacturing, energy, and technology due to their expertise in system maintenance, failure analysis, and predictive modeling. Companies seek experienced professionals with skills in data analysis, reliability tools, and certifications like Six Sigma or RCM to improve equipment uptime and reduce costs.

How much do Sr Reliability Engineers get paid?

Senior Reliability Engineers typically earn between $90,000 and $130,000 annually, depending on experience, industry, and location. They often have expertise in systems analysis, failure modes, and reliability tools like FMEA and RCM, which can influence compensation levels.

What cities near Calgary, AB are hiring for Sr Reliability Engineer jobs?

Cities near Calgary, AB with the most Sr Reliability Engineer job openings:

Infographic showing various Sr Reliability Engineer job openings in Calgary, AB as of August 2026, with employment types broken down into 91% Full Time, 5% Part Time, and 4% Contract. Highlights an 86% Physical, 6% Hybrid, and 8% Remote job distribution.

Staff Platform Engineer

Robots and Pencils

Calgary, AB • On-site, Remote

Contractor

Posted 7 days ago


Job description

Staff Platform Engineer 

Location: This position can be located in the following area(s): remote in Canada

This is a 4 month contract assignment with potential to extend

Company Overview 

Robots & Pencils is an applied AI engineering firm building the next frontier of business architecture. We design and ship AI co-workers that integrate into enterprise operations and deliver measurable results for our clients. We're all in on AWS, combining deep UX capability with senior engineering talent to get AI into production fast and keep it there. 
We've earned the trust of leaders across Consumer Products and Retail, Education, Energy, Financial Services, Healthcare, and Manufacturing and more, and earned a reputation as the nimble alternative to traditional global systems integrators. Founded in 2009, with delivery centers in Canada, the United States, Eastern Europe, and Latin America, we are smaller, faster, and more senior by design. Our teams average 15+ years of experience. We move fast, sweat the details, and build things that actually ship. 

Position Overview 

We're looking for a Staff Platform Engineer to define and lead platform engineering strategy across complex, multi-environment cloud systems. This role is ideal for an experienced engineer who can own infrastructure architecture end-to-end, drive DevSecOps and compliance practices, and serve as a technical leader on the engagements they support. 

In this role, you will work as a key technical contributor on a cross-functional team, defining standards and owning platform reliability, performance, and security at scale. You'll mentor engineers, partner with leadership on infrastructure direction, and lead complex migrations and modernization initiatives. 

Why This Role Matters 

At Robots & Pencils, we design AI systems for a human world. Our name says it all. Robots and pencils means engineering paired with creativity, because every agent we ship has to work for real people in real workflows. That balance is baked into how we operate. 
Every role here contributes directly to that mission. Here, you shape how AI systems integrate into enterprise operations, how teams move at real velocity, and how products create measurable impact for clients and the people they serve. We ship production-ready AI in 30 to 45 days. That pace demands people who take ownership, lead with craft, and care deeply about what they put their name on. 

What You'll Do 

Craft & Delivery 

  • Define DevOps strategy and lead infrastructure architecture across multi-environment, multi-region cloud systems
  • Architect and own scalable Kubernetes platforms and containerized infrastructure at scale
  • Own infrastructure as code strategy and standards across environments
  • Lead DevSecOps implementation including secrets management, compliance, auditing, IAM, and zero-trust networking
  • Drive platform reliability, performance SLAs, and cost optimization across production systems
  • Lead complex cloud migrations and platform modernization initiatives
  • Own observability strategy and production reliability practices
  • Lead the design and operation of AI/ML platform infrastructure, including model serving and deployment, GPU workload orchestration, LLM gateway and observability, vector store infrastructure, and CI/CD for AI/ML systems
  • Bring an AI-forward mindset to your daily work, using tools like Claude, Cursor, and other modern AI assistants to ship higher-quality work at pace

Collaboration & Communication 

  • Partner with engineering, product, and leadership to align platform strategy with business and delivery goals
  • Communicate complex infrastructure decisions and tradeoffs clearly to technical and non-technical stakeholders
  • Lead design reviews, architecture discussions, and release readiness assessments

Leadership & Influence 

  • Establish platform engineering standards and best practices on the engagements you support
  • Mentor junior and mid-level engineers, helping them grow their craft, confidence, and impact
  • Act as a technical escalation point on complex infrastructure and platform challenges
  • Evaluate emerging tools and technologies, recommending patterns that improve platform reliability and developer experience

What You'll Bring 

  • 7+ years of professional DevOps or platform engineering experience, with experience leading complex platform initiatives
  • Expert scripting and programming skills (e.g., Python, Go, Java, Bash)
  • Deep cloud expertise across at least one major platform
  • Expert Kubernetes and container orchestration skills
  • Expert IaC skills across multiple tools
  • Strong CI/CD architecture experience at scale
  • Strong DevSecOps experience including secrets management, compliance, and auditing
  • Experience with networking, IAM, security architecture, and zero-trust principles in cloud environments
  • Experience with service mesh, distributed systems, and microservices architecture
  • Strong experience with AI/ML platform infrastructure, including model serving and deployment, GPU workload orchestration, LLM gateway and observability, vector store infrastructure, and CI/CD for AI/ML systems
  • Demonstrated leadership and technical mentoring experience across a team or organization
  • Strong stakeholder communication skills, with the ability to translate technical depth across audiences
  • Demonstrable, day-to-day usage and expert knowledge of AI-forward tools such as Claude and Cursor
  • Excellent problem-solving skills and the ability to navigate highly ambiguous technical and business challenges with sound judgment
  • Cloud certifications (e.g., AWS DevOps Engineer Professional, CKA, Azure DevOps Engineer) or FinOps experience is a plus

Helpful Extras and Unique Skills

  • Designing and provision HPC cluster infrastructure using CI/CD pipeline across AWS, CoreWeave, GCP, and OCI
  • Experience with HPC job schedulers and workload managers such as Slurm or equivalent for job submission and queue management

You'll Do Well Here if You Are 

  • A doer. You see something broken and fix it. You'd rather move on clarity than wait for certainty.
  • A fast learner who knows you don't know everything. The AI landscape changes weekly. You're senior enough to know better and curious enough to keep learning anyway.
  • Direct in a way that makes the work better. You give honest feedback. You'd rather have the hard conversation than blow smoke.
  • Obsessed with craft. You know genius is in the details. You ship exceptional, not perfect, and you don't put your name on work you wouldn't stand behind.
  • Built for ownership. You honor commitments, admit mistakes fast, and back your teammates when a decision costs something. No handoffs, no finger-pointing.
  • All in. You treat clients' businesses like your own. You take the work seriously without taking yourself seriously.
  • Resourceful when the budget, timeline, or team is tight. Constraints don't slow you down. They sharpen you.
  • Glad to be in the room with people who care as much as you do. Our teams average fifteen-plus years of experience. We hire people who push each other to do better work.