ABOUT YOU We are looking for a Site Reliability Engineer (Monetization) who is pragmatic, product ... DevOps, or platform engineering experience: on-call or incident response duty, SLO/monitoring ...
ABOUT YOU We are looking for a Site Reliability Engineer (Monetization) who is pragmatic, product ... DevOps, or platform engineering experience: on-call or incident response duty, SLO/monitoring ...
Director, Infrastructure & SRE
Montreal, QC · On-site
About the Role The Director of Infrastructure & SRE owns the function end-to-end: reliability ... operational communications * Own all TailorCare domain and email infrastructure Developer ...
Director, Infrastructure & SRE
Montreal, QC · On-site
About the Role The Director of Infrastructure & SRE owns the function end-to-end: reliability ... operational communications * Own all TailorCare domain and email infrastructure Developer ...
DevOps Engineer
Montreal, QC · On-site +1
We are looking for an experienced DevOps/SRE Engineer for our client. This is a permanent position, that can either be remote or in-office at Toronto! Our client is a large fintech firm with a ...
Quick apply
DevOps Engineer
Montreal, QC · On-site +1
We are looking for an experienced DevOps/SRE Engineer for our client. This is a permanent position, that can either be remote or in-office at Toronto! Our client is a large fintech firm with a ...
DevOps Engineer
Montreal, QC · On-site +1
We are looking for an experienced DevOps/SRE Engineer for our client. This is a permanent position, that can either be remote or in-office at Toronto! Our client is a large fintech firm with a ...
Quick apply
DevOps Engineer
Montreal, QC · On-site +1
We are looking for an experienced DevOps/SRE Engineer for our client. This is a permanent position, that can either be remote or in-office at Toronto! Our client is a large fintech firm with a ...
Site Reliability Engineer
Montreal, QC · On-site +1
Strong relevant experience in Site Reliability, Cloud, or DevOps Engineering in SaaS or large-scale production environments * Strong experience with AWS (multi-account, VPC, EC2, EKS) and Kubernetes ...
Site Reliability Engineer
Montreal, QC · On-site +1
Strong relevant experience in Site Reliability, Cloud, or DevOps Engineering in SaaS or large-scale production environments * Strong experience with AWS (multi-account, VPC, EC2, EKS) and Kubernetes ...
Site Reliability Engineer
Montreal, QC · On-site
Strong relevant experience in Site Reliability, Cloud, or DevOps Engineering in SaaS or large-scale production environments * Strong experience with AWS (multi-account, VPC, EC2, EKS) and Kubernetes ...
Quick apply
Site Reliability Engineer
Montreal, QC · On-site
Strong relevant experience in Site Reliability, Cloud, or DevOps Engineering in SaaS or large-scale production environments * Strong experience with AWS (multi-account, VPC, EC2, EKS) and Kubernetes ...
Le Site Reliability Engineering (SRE) est une discipline orientee production, axee sur ... Science des donnees et DevOps. Notre programme Expert offre aux professionnels experimentes l'acces ...
Le Site Reliability Engineering (SRE) est une discipline orientee production, axee sur ... Science des donnees et DevOps. Notre programme Expert offre aux professionnels experimentes l'acces ...
Le Site Reliability Engineering (SRE) est une discipline orientee production, axee sur ... Science des donnees et DevOps. Notre programme Expert offre aux professionnels experimentes l'acces ...
Le Site Reliability Engineering (SRE) est une discipline orientee production, axee sur ... Science des donnees et DevOps. Notre programme Expert offre aux professionnels experimentes l'acces ...
Le Site Reliability Engineering (SRE) est une discipline orientee production, axee sur ... Science des donnees et DevOps. Notre programme Expert offre aux professionnels experimentes l'acces ...
Le Site Reliability Engineering (SRE) est une discipline orientee production, axee sur ... Science des donnees et DevOps. Notre programme Expert offre aux professionnels experimentes l'acces ...
DevOps / SRE Engineer (Remote)
Montreal, QC · On-site +1
To do that we are eager to add a highly skilled DevOps / SRE Engineer Engineer to our incredible team. This is a senior role working alongside the current backend and frontend software engineering ...
Quick apply
DevOps / SRE Engineer (Remote)
Montreal, QC · On-site +1
To do that we are eager to add a highly skilled DevOps / SRE Engineer Engineer to our incredible team. This is a senior role working alongside the current backend and frontend software engineering ...
Site Reliability Engineer
Montreal, QC · On-site
... DevOps, les donnees et l'ingenierie logicielle de bout en bout, au service d'une multitude ... Nous sommes a la recherche d'un(e) ingenieur(e) en fiabilite des sites (Site Reliability Engineer ...
Site Reliability Engineer
Montreal, QC · On-site
... DevOps, les donnees et l'ingenierie logicielle de bout en bout, au service d'une multitude ... Nous sommes a la recherche d'un(e) ingenieur(e) en fiabilite des sites (Site Reliability Engineer ...
SRE specialist
Montreal, QC · Hybrid
About the role We are seeking a hands-on Site Reliability Engineer within the Intelligent Operations Department's SRE & Resiliency team. This role operates across Azure, AWS, GCP, and onprem ...
SRE specialist
Montreal, QC · Hybrid
About the role We are seeking a hands-on Site Reliability Engineer within the Intelligent Operations Department's SRE & Resiliency team. This role operates across Azure, AWS, GCP, and onprem ...
SRE specialist
Montreal, QC · Hybrid
About the role We are seeking a hands-on Site Reliability Engineer within the Intelligent Operations Department's SRE & Resiliency team. This role operates across Azure, AWS, GCP, and onprem ...
SRE specialist
Montreal, QC · Hybrid
About the role We are seeking a hands-on Site Reliability Engineer within the Intelligent Operations Department's SRE & Resiliency team. This role operates across Azure, AWS, GCP, and onprem ...
Senior Site Reliability Engineer
Montreal, QC · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Quick apply
Senior Site Reliability Engineer
Montreal, QC · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Senior Site Reliability Engineer
Montreal, QC · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Quick apply
Senior Site Reliability Engineer
Montreal, QC · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Intermediate DevOPS/SRE
Montreal, QC · On-site
The COO/GTE/EPL/SRE team has members in Paris, Bangalore, and Montreal and is responsible for the ... To foster agility, software craftsmanship, and DevOps practices within the bank, the DevOps Service ...
Intermediate DevOPS/SRE
Montreal, QC · On-site
The COO/GTE/EPL/SRE team has members in Paris, Bangalore, and Montreal and is responsible for the ... To foster agility, software craftsmanship, and DevOps practices within the bank, the DevOps Service ...
We are looking for an experienced Site Reliability Engineer or Platform Operations Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or ...
Quick apply
We are looking for an experienced Site Reliability Engineer or Platform Operations Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or ...
We are looking for an experienced Site Reliability Engineer or Platform Operations Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or ...
Quick apply
We are looking for an experienced Site Reliability Engineer or Platform Operations Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or ...
DevOps Engineer
Montreal, QC · On-site +1
You Have: * 3+ years of DevOps/SRE experience * Strong understanding of security best practices * Experience with: * CI/CD * Kubernetes/Docker * Cloud infrastructure (AWS preferably) * Infrastructure ...
Quick apply
DevOps Engineer
Montreal, QC · On-site +1
You Have: * 3+ years of DevOps/SRE experience * Strong understanding of security best practices * Experience with: * CI/CD * Kubernetes/Docker * Cloud infrastructure (AWS preferably) * Infrastructure ...
DevOps Engineer
Montreal, QC · On-site +1
You Have: * 3+ years of DevOps/SRE experience * Strong understanding of security best practices * Experience with: * CI/CD * Kubernetes/Docker * Cloud infrastructure (AWS preferably) * Infrastructure ...
Quick apply
DevOps Engineer
Montreal, QC · On-site +1
You Have: * 3+ years of DevOps/SRE experience * Strong understanding of security best practices * Experience with: * CI/CD * Kubernetes/Docker * Cloud infrastructure (AWS preferably) * Infrastructure ...
Sre Devops information
What is the difference between Sre Devops vs Devops Engineer?
| Aspect | Sre Devops | Devops Engineer |
|---|---|---|
| Primary Focus | Reliability, automation, and system stability | Deployment, CI/CD, and infrastructure automation |
| Credentials | Certifications like AWS, Linux, or Kubernetes | Similar certifications, often including cloud and scripting skills |
| Work Environment | Operations teams, site reliability, and incident management | Development teams, deployment pipelines, and infrastructure setup |
| Industry Usage | Tech companies, cloud providers, large-scale systems | Startups, software companies, and enterprises adopting DevOps practices |
While both roles involve automation and cloud skills, Sre Devops emphasizes system reliability and incident response, whereas Devops Engineer focuses more on deployment pipelines and infrastructure automation. Understanding these differences helps organizations assign the right responsibilities and professionals for their needs.

Full-time
Medical, Dental, Vision, PTO
Posted 25 days ago
Job description
We are looking for a Site Reliability Engineer (Monetization) who is pragmatic, product-minded, and equally comfortable writing code and running production systems to join our Infrastructure department's SRE team. The best candidate will be someone who thrives in a fast-paced, highly collaborative, and exceptionally dynamic setting and is excited to own the application-level infrastructure and reliability of a high-traffic commerce domain end to end - from deploy pipelines and Kubernetes manifests to SLOs, capacity planning, and production readiness.
Strong Kubernetes, observability, and software engineering skills are essential, along with experience in operating production services in a cloud environment (GCP/GKE or comparable) and partnering closely with product development teams. The ability to hold a dual perspective - understanding both how developers ship features and what infrastructure needs to stay reliable - and to bring the reliability lens into design decisions early will be key to your success in this role.
This is a hybrid embedded role: you remain part of the SRE organization (practices, standards, duty rotation) while being functionally embedded into the Monetization product domain. You'll build long-term working relationships with the domain's engineering teams, own a meaningful share of their application infrastructure execution, and co-author the reliability practices used company-wide.
If you're passionate about making complex distributed systems boringly reliable and love building the commerce and monetization backbone that lets game developers around the world get paid, we would love to hear from you!
ABOUT US
Xsolla is a global commerce company with robust tools and services to help developers solve the inherent challenges of the video game industry. From indie to AAA, companies partner with Xsolla to help them fund, distribute, market, and monetize their games. Grounded in the belief in the future of video games, Xsolla is resolute in the mission to bring opportunities together, and continually make new resources available to creators. Headquartered and incorporated in Los Angeles, California, Xsolla operates as the merchant of record and has helped over 1,500+ game developers to reach more players and grow their businesses around the world. With more paths to profits and ways to win, developers have all the things needed to enjoy the game.
For more information, visit xsolla.com.
- Own the application-level infrastructure of the Monetization domain: Helm charts, Terraform configurations, Kubernetes deployments, runtime configuration, and service-level networking and integrations
- Own the domain's observability: design and implement SLOs/SLIs, monitors, alerts, and dashboards for critical services on Datadog and OpenTelemetry-based tooling
- Help to set up and evolve CI/CD pipelines for domain services (GitLab CI, GitHub Actions), including deploy and rollback automation
- Perform capacity planning and performance tuning ahead of expected load - product launches, sales events, and regional rollouts - including load testing and performance regression investigation
- Run Production Readiness Reviews for new services and major changes; define and enforce what "production-ready" means for the domain
- Support domain incident response: assist with deep investigation of complex incidents, contribute to post-mortems, drive follow-up reliability improvements, and maintain runbooks
- Build domain-specific automation that reduces operational toil: runbook automation, deploy helpers, recurring operational scripts
- Maintain and drive a forward-looking reliability roadmap for the domain together with product engineering leads
- Participate in product team planning, refinements, and architecture reviews, bringing the reliability perspective before design decisions become expensive to change
- Co-author company-wide SLO/SLI, capacity, and operational standards together with the broader SRE team; contribute improvements directly to shared SRE-operated subsystems
- Participate in the SRE duty rotation, supporting developers across the company
- 3+ years of proven SRE, DevOps, or platform engineering experience: on-call or incident response duty, SLO/monitoring ownership, deploy pipeline and infrastructure work for production services
- Software development background: you have built and shipped backend services, not only operated them - comfortable reading application code during an investigation and writing production-quality automation in at least one language (e.g., Go, PHP)
- Hands-on Kubernetes experience:Â Helm, manifests, deploy strategies, debugging application-level performance and networking issues (GKE or another managed Kubernetes)
- Solid observability practice: building monitors, dashboards, and SLOs/SLIs on a modern platform (Datadog preferred; Prometheus/Grafana experience also relevant), familiarity with OpenTelemetry
- Infrastructure as Code exposure (Terraform/Terragrunt) for collaboration with platform teams
- GCP experience (IAM, networking, managed services)
- Experience building and maintaining CI/CD pipelines (GitLab CI and/or GitHub Actions)
- Programming/scripting proficiency sufficient to build automation and tooling (e.g., Python, Go, or Bash)
- Practical experience with incident response, post-mortems, and driving reliability improvements from incidents
- Strong collaboration and communication skills - this role works embedded with product development teams daily
- Experience in payments, fintech, e-commerce, or gaming - high-traffic transactional systems
- Kubernetes certifications
- Google Cloud Platform certifications
- HashiCorp certifications
We are passionate about fostering a supportive environment for our team, so we prioritize the physical, mental, and emotional well-being of our employees and their families through a comprehensive Benefits Program. This includes medical, dental, and vision, PTO, and a personalized career roadmap for each employee. By investing in professional development through training and educational opportunities, we ensure that our team thrives both personally and professionally. Together, we're not just building a business; we're cultivating a community that values creativity, collaboration, and the transformative power of play.
Equal Employment Opportunity StatementXsolla is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. We do not discriminate based on race, color, religion, sex, national origin, age, disability, sexual orientation, gender identity, or any other characteristic protected by law. We consider qualified applicants with criminal histories in accordance with the Fair Chance Act.
Criminal History ConsiderationFor the Site Reliability Engineer (Monetization) position, we will conduct a background check that may include the following:
- Criminal history check
- Employment verification
- Education verification
The background check is relevant to this position because of the following role responsibilities:
- Accessing confidential company data
- Handling infrastructure that processes sensitive financial transactions
- Ensuring compliance with regulatory requirements
Applicants are encouraged to inquire about their rights under the Fair Chance Act. If you have questions regarding our hiring practices, please contact [email protected].
By submitting the following job application form, you consent to Xsolla processing your data for career-related inquiries and potential employment opportunities. We process your data in accordance with this Xsolla Privacy Notice for Job Applicants. Please direct any inquiries regarding your data privacy to [email protected].
About Xsolla
Sourced by ZipRecruiter
Industry
Pc games
Company size
201 - 500 Employees
Headquarters location
Los Angeles, CA, US
Year founded
2005