1

Site Reliability Engineer Jobs in Alberta (NOW HIRING)

You Have: * 5+ years of DevOps/SRE experience * Strong understanding of security best practices * Experience with: * CI/CD * Kubernetes/Docker * Cloud infrastructure (AWS preferably) * Infrastructure ...

Design and own the SRE function for Level 1 data ingestion across all GCP deployments: alert policy design, SLO definition, incident management, and on-call operations * Eliminate manual data quality ...

Design and own the SRE function for Level 1 data ingestion across all GCP deployments: alert policy design, SLO definition, incident management, and on-call operations * Eliminate manual data quality ...

Document infrastructure and share knowledge to improve long-term maintainability What you bring to the table * 5+ years of experience in DevOps, SRE, or related roles * Strong experience with AWS ...

Drive best practices in CI/CD, observability, security, and SRE principles * Guide teams in leveraging AWS, Azure, or GCP ecosystems effectively Enable Production-Ready Systems * Build platforms that ...

Platform Engineer, Databases

Calgary, AB ยท Remote

CA$137K - CA$157K/yr

Readiness to work with and develop the MySQL knowledge to the rest of the SRE team * A portfolio of successful projects (as well as a collection of lessons learned from failed projects) * A keen ...

AB ยท On-site

... Manager, Reliability & Discipline Engineering. The Senior Rotating Equipment Engineer will be an ... Key Activities and Responsibilities Ensure safe and reliable operation of all site rotating ...

next page

Showing results 1-20

Site Reliability Engineer information

Will SRE be replaced by ai?

Site Reliability Engineers (SREs) focus on maintaining system reliability, automation, and incident response. While AI tools can assist with monitoring and automating routine tasks, SREs' expertise in system design, troubleshooting, and decision-making remains essential, making complete replacement unlikely in the near future.

What Is a Site Reliability Engineer?

A site reliability engineer specializes in site reliability engineering, or SRE, a specific branch of operations first pioneered by Google. You are responsible for ensuring that when a website decides to scale a particular feature for various users to access, it does not break the underlying software or website functions. This means you need to use analytical problem-solving skills to determine how to make specific features on a new software release work on top of existing source code.

What are the key skills and qualifications needed to thrive as a Site Reliability Engineer, and why are they important?

To thrive as a Site Reliability Engineer, you need a strong background in computer science, systems administration, and software engineering, often supported by a degree in a technical field. Familiarity with cloud platforms (like AWS or GCP), container orchestration (such as Kubernetes), infrastructure as code (Terraform or Ansible), and monitoring tools (Prometheus, Grafana) is typically expected. Strong problem-solving skills, effective communication, and a proactive mindset help SREs excel at incident management and cross-functional collaboration. These skills are crucial for maintaining system reliability, minimizing downtime, and driving continuous improvement in complex technical environments.

What engineers make $500,000 a year?

Senior-level engineers in high-demand fields such as software engineering, especially those specializing in cloud infrastructure, distributed systems, or machine learning, can earn $500,000 or more annually. These roles often require extensive experience, advanced skills, and sometimes stock options or bonuses in addition to base salary.

Is SRE a stressful job?

Site Reliability Engineers (SREs) often work in high-pressure environments to ensure system stability and uptime, which can lead to stressful situations during outages or incidents. The role requires strong problem-solving skills, familiarity with monitoring tools, and sometimes on-call responsibilities, but organizations often implement practices to manage workload and reduce stress.

What are some of the most common challenges Site Reliability Engineers face when balancing system reliability with rapid software delivery?

Site Reliability Engineers (SREs) often navigate the challenge of maintaining highly reliable systems while supporting fast-paced software releases. This involves managing incidents, automating processes to reduce manual toil, and working closely with development teams to embed reliability into the software development lifecycle. SREs must carefully prioritize their efforts between proactive improvements and urgent, reactive fire-fighting. Effective communication and collaboration with both operations and development teams are crucial to ensuring service uptime without slowing down innovation.

What does a site reliability engineer do?

A site reliability engineer (SRE) is responsible for maintaining and improving the reliability, availability, and performance of software systems. They use automation, monitoring tools, and scripting to prevent outages, troubleshoot issues, and ensure systems run smoothly at scale. SREs often collaborate with development teams and may hold certifications in cloud platforms or scripting languages.

What is the difference between Site Reliability Engineer vs DevOps Engineer?

AspectSite Reliability EngineerDevOps Engineer
CredentialsTypically requires a computer science degree, certifications like AWS, Google Cloud, or KubernetesSimilar credentials, often with cloud certifications and scripting skills
Work EnvironmentFocuses on maintaining and improving system reliability, often in large-scale production environmentsWorks on automation, CI/CD pipelines, and deployment processes across development and operations teams
Industry UsageCommon in tech, cloud services, and large-scale enterprise companiesWidely used in software development, cloud, and IT organizations

Both roles require strong technical skills and cloud knowledge, but SREs focus more on system reliability and uptime, while DevOps engineers emphasize automation and deployment processes. They often collaborate but have distinct primary responsibilities.

What is a Site Reliability Engineer?

A Site Reliability Engineer (SRE) is a professional who applies software engineering principles to infrastructure and operations problems. Their primary goal is to create scalable and highly reliable software systems, often bridging the gap between development and IT operations. SREs automate tasks, monitor system health, respond to incidents, and work to improve system reliability and performance. They also help define service level objectives (SLOs) and ensure systems meet customer expectations for uptime and availability.
What are the most commonly searched types of Site Reliability Engineer jobs in Alberta? The most popular types of Site Reliability Engineer jobs in Alberta are:
What are popular job titles related to Site Reliability Engineer jobs in Alberta? For Site Reliability Engineer jobs in Alberta, the most frequently searched job titles are:
What job categories do people searching Site Reliability Engineer jobs in Alberta look for? The top searched job categories for Site Reliability Engineer jobs in Alberta are:
What are popular job titles related to Site Reliability Engineer jobs in AB? For Site Reliability Engineer jobs in AB, the most frequently searched job titles are:
Senior DevOps Engineer

Senior DevOps Engineer

Targeted Talent

Calgary, AB โ€ข On-site, Remote

Full-time

Medical, Dental, Vision, Retirement, PTO

Re-posted 26 days ago


Job description

We are looking for an experienced DevOps Engineer role for our client. This is a permanent position, that can either be remote or in-office at Toronto! Our client is a large technology firm with a product that you've likely used.
You Have:
  • You have hands-on experience with enterprise-grade infrastructures, operations and / or systems administration.
  • Familiar with DevOps engineering practices
  • Experience designing and implementing automations, including creation of CI/CD pipelines
  • Operational experience of cloud infrastructure (AWS or similar).
  • Hands-on experience implementing and managing Kubernetes/Docker solutions on Public/Private Cloud Platforms
  • Experience implementing Infrastructure as Code and Configuration as Code automations (GitLab, ARM templates, Terraform, Ansible, Puppet or Chef)
  • High proficiency in coding/scripting using Python/Go/Ruby/Shell/PowerShell languages
  • Experience with operational aspects of software systems such as monitoring, centralized logging and alerting
  • Looking to advance and grow your career in Fintechโ€ฆ
You Have:
  • 5+ years of DevOps/SRE experience
  • Strong understanding of security best practices
  • Experience with:
    • CI/CD
    • Kubernetes/Docker
    • Cloud infrastructure (AWS preferably)
    • Infrastructure as Code (GitLab, ARM templates, Terraform, Ansible)
    • Coding/Scripting in Python/Go/Ruby/Shell/PowerShell languages
    • Monitoring, centralized logging and alerting
Perks:
  • Competitive Salary
  • Vacation & Sick Days
  • Health Dental, Vision and Retirement Benefits
  • Opportunity for Remote Work
  • Flexible work environment
If this role fits your career path, please apply to this posting!

Targeted Talent logo

About Targeted Talent

Sourced by ZipRecruiter

Your single source for HR professional services, we offer job seekers specialized employment services, spanning contract, permanent positions, and project solutions for highly specialized and managerial level talent needs. Our team of specialized recruiters and consultants abilities extend far beyond resume or career counseling. With hundreds of collaborators strategically located throughout the country, our organization possess the local market knowledge and industry relationships that make successful geography-specific reach possible.

Industry

Recruiting and staffing services

Company size

11 - 50 Employees

Headquarters location

Vancouver, BC, CA