2

Remote Chaos Engineering Jobs (NOW HIRING)

Pre/Post Sales Solutions Architect

$64.50 - $85/hr

As the industry leader in Chaos Engineering and reliability testing, we work with hundreds of the ... But as a remote company, teamwork and collaboration won't happen by accident. We approach every ...

... remote capacity. Awarded as a "best place to work" company, our culture fosters team integrity ... chaos engineering principles and "game days" to proactively test the resilience of Convoso ...

... remote capacity. Awarded as a "best place to work" company, our culture fosters team integrity ... chaos engineering principles and "game days" to proactively test the resilience of Convoso ...

[Remote] Director of DevOps

Los Angeles, CA · On-site +1

$56.75 - $77.75/hr

... remote capacity. Awarded as a "best place to work" company, our culture fosters team integrity ... chaos engineering principles and "game days" to proactively test the resilience of Convoso ...

Champion proactive reliability - chaos engineering, game days, failure-mode analysis, capacity and ... remote on-call teams. * Demonstrated ownership of reliability outcomes for customer-facing SaaS at ...

... remote global workforce. If you are passionate about working on business problems that can be ... chaos engineering, performance engineering, toil reduction, reliability engineering etc

DevOps Engineer

$52.75 - $72.25/hr

Knowledge of chaos engineering and resilience testing * Exposure to AI/ML services (SageMaker ... Remote work environment with emphasis on team collaboration and hands-on problem solving Pay Range ...

DevOps Engineer

$52.75 - $72.25/hr

Knowledge of chaos engineering and resilience testing * Exposure to AI/ML services (SageMaker ... Remote work environment with emphasis on team collaboration and hands-on problem solving Pay Range ...

... remote state, and GitOps-based plan/apply pipelines - no unmanaged resources * Audit the cloud ... one chaos engineering exercise (AWS FIS or equivalent) * Partner with engineering teams to ...

[Remote] Director of DevOps

$54 - $74/hr

... chaos engineering principles and "game days" to proactively test the resilience of Convoso's platform • Establish and maintain a CI/CD center of excellence • Drive automation for CI/CD processes ...

[Remote] Director of DevOps

$54 - $74/hr

... chaos engineering principles and "game days" to proactively test the resilience of Convoso's platform • Establish and maintain a CI/CD center of excellence • Drive automation for CI/CD processes ...

Cloud Engineer

$55.75 - $74.50/hr

Knowledge of chaos engineering and resilience testing * Exposure to AI/ML services (SageMaker ... Remote work environment with emphasis on team collaboration and hands-on problem solving Pay Range ...

New

Cloud Engineer

$57 - $76.25/hr

Knowledge of chaos engineering and resilience testing * Exposure to AI/ML services (SageMaker ... Remote work environment with emphasis on team collaboration and hands-on problem solving Pay Range ...

New

Cloud Engineer

$55.75 - $74.50/hr

Knowledge of chaos engineering and resilience testing * Exposure to AI/ML services (SageMaker ... Remote work environment with emphasis on team collaboration and hands-on problem solving Pay Range ...

New

... remote state, and GitOps-based plan/apply pipelines - no unmanaged resources Audit the cloud estate ... chaos engineering exercise (AWS FIS or equivalent) Partner with engineering teams to instrument ...

... remote state, and GitOps-based plan/apply pipelines - no unmanaged resources Audit the cloud estate ... chaos engineering exercise (AWS FIS or equivalent) Partner with engineering teams to instrument ...

Sr. Production Engineer

$118K - $148K/yr

This role is available as a hybrid opportunity 3 days a week in San Jose, CA or Remote reporting to ... Experience with chaos engineering and disaster recovery planning at scale * Expertise in global ...

Production Engineer

$102K - $128K/yr

This role is available as a hybrid opportunity 3 days a week in San Jose, CA or Remote reporting to ... Experience with chaos engineering and disaster recovery planning at scale * Expertise in global ...

next page

Showing results 1-20

Remote Chaos Engineering information

See salary details

$73K

$194.7K

$254K

How much do remote chaos engineering jobs pay per year?

As of Jul 22, 2026, the average yearly pay for remote chaos engineering in the United States is $194,709.00, according to ZipRecruiter salary data. Most workers in this role earn between $141,500.00 and $253,000.00 per year, depending on experience, location, and employer.

What engineers make 200,000 a year?

Senior engineers in fields like software engineering, cloud engineering, and site reliability engineering often earn $200,000 or more annually, especially with extensive experience, specialized skills, and certifications. Roles involving advanced technical expertise, leadership, or working in high-demand industries tend to have higher compensation levels.

What is Remote Chaos Engineering?

Remote Chaos Engineering is the practice of testing distributed systems' resilience by intentionally introducing failures and disruptions in remote or cloud environments. The goal is to identify weaknesses and improve system reliability by simulating real-world incidents, such as network outages or server crashes, in a controlled manner. This approach helps teams understand how their applications behave under stress and develop strategies to mitigate future incidents. Remote Chaos Engineering is particularly valuable for organizations leveraging cloud infrastructure and remote services, ensuring robust performance even under unexpected conditions.

What are some common challenges faced by professionals working in remote chaos engineering roles?

Professionals in remote chaos engineering often encounter challenges such as coordinating experiments across distributed teams, ensuring clear communication about system vulnerabilities, and managing the complexity of large-scale systems without direct, on-site access. Establishing robust monitoring and rollback procedures is essential to minimize risk during remote testing. Additionally, building trust with development and operations teams is key, as chaos engineering often involves intentionally introducing failures to improve system resilience.

What engineers make 500,000 a year?

Senior engineers in specialized fields such as software engineering, cloud infrastructure, or cybersecurity can earn $500,000 or more annually, especially with experience, advanced skills, and leadership roles. High compensation often includes bonuses, stock options, or profit sharing, particularly in large tech companies or startups with significant growth potential.

What engineers make $300,000 a year?

Senior engineers in specialized fields such as software engineering, cloud engineering, or site reliability engineering can earn $300,000 or more annually, especially with extensive experience, advanced skills in cloud platforms, and relevant certifications. These roles often involve high responsibility, complex problem-solving, and working in high-demand industries like technology and finance.

What are the key skills and qualifications needed to thrive as a Remote Chaos Engineer, and why are they important?

To thrive as a Remote Chaos Engineer, you need a strong background in software engineering, systems architecture, and site reliability, often supported by a degree in computer science or a related field. Familiarity with chaos engineering platforms (such as Gremlin or Chaos Monkey), cloud environments (AWS, Azure, GCP), and automation tools is typically required. Strong problem-solving abilities, clear communication, and a collaborative mindset help you effectively identify weaknesses and drive reliability improvements across distributed teams. These skills are crucial for proactively uncovering system vulnerabilities, ensuring system resilience, and maintaining high availability in complex, remote-first infrastructures.

What is the difference between Remote Chaos Engineering vs Remote Site Reliability Engineer?

AspectRemote Chaos EngineeringRemote Site Reliability Engineer
Primary FocusDesigning and executing chaos experiments to improve system resilienceEnsuring system reliability, availability, and performance through monitoring and automation
Skills & CertificationsKnowledge of chaos engineering tools, scripting, cloud platformsMonitoring tools, scripting, cloud infrastructure, SRE certifications
Work EnvironmentCollaborates with development and operations teams, often in DevOps cultureWorks closely with engineering teams to maintain system health and SLAs

While both roles focus on system stability, Remote Chaos Engineering specializes in testing system resilience through chaos experiments, whereas Remote Site Reliability Engineers focus on maintaining overall system reliability and performance. Both roles require scripting skills and cloud knowledge, but their core objectives differ: one proactively tests, the other maintains system health.

Is chaos engineering still used today?

Chaos engineering is actively used in modern IT environments to improve system resilience by intentionally introducing failures and testing recovery processes. Professionals in roles like remote chaos engineering often utilize tools such as Chaos Monkey and conduct experiments in cloud or distributed systems to identify weaknesses before outages occur.
More about Remote Chaos Engineering jobs
What cities are hiring for Remote Chaos Engineering jobs? Cities with the most Remote Chaos Engineering job openings:
What are the most commonly searched types of Chaos Engineering jobs? The most popular types of Chaos Engineering jobs are:
What states have the most Remote Chaos Engineering jobs? States with the most job openings for Remote Chaos Engineering jobs include:
What job categories do people searching Remote Chaos Engineering jobs look for? The top searched job categories for Remote Chaos Engineering jobs are:
Infographic showing various Remote Chaos Engineering job openings in the United States as of July 2026, with employment types broken down into 9% Locum Tenens, 65% Full Time, 25% Part Time, and 1% Contract. Highlights an 93% Physical, 1% Hybrid, and 6% Remote job distribution, with an average salary of $194,709 per year, or $93.6 per hour.
Pre/Post Sales Solutions Architect

Pre/Post Sales Solutions Architect

Gremlin

Remote

$64.50 - $85/hr

Full-time

Medical, Dental, Vision, Retirement

Re-posted 26 days ago


Job description

Today's complex, fast-paced systems have become a minefield of reliability risks-any of which could cause an outage that costs millions and destroys customer confidence. That's why high-availability teams use Gremlin to find and fix reliability risks before they become incidents. The Gremlin Reliability Platform helps software teams proactively monitor and test their systems for common reliability risks, build and enforce reliability standards, and automate their reliability practices organization-wide. As the industry leader in Chaos Engineering and reliability testing, we work with hundreds of the world's largest organizations where high availability is non-negotiable.
About the Role of Pre/Post Sales Solutions Architect
Gremlin's team is growing, and we're seeking a passionate Solutions Architect to help prove the value of Reliability Management to customers. In this pre- and post-sales role, you will have the opportunity to demonstrate Gremlin Reliability Management and offer guidance on best practices for building reliable architectures. As customers convert to a paid subscription, you will advise on how to design and implement experiments to activate customers for their reliability journey.
In this role, you'll get to:
  • Demonstrate Gremlin in customer calls and webinars
  • Partner with sales team to drive technical wins and grow Gremlin's customer base
  • Participate and lead proof-of-concepts with potential customers
  • Educate potential customers on Reliability and Chaos Engineering
  • Work with existing customers on technical projects and assist in troubleshooting
  • Consult with customers on the resiliency of their applications and architecture, diagnose gaps and recommend solutions
  • Participate in technical workshops and conferences

Collaborate with different functions of the company including Product Marketing, Support, and Engineering
We'll expect you to have:
  • 5+ years of experience as a Solution Architect in a tech company
  • Excellent verbal and written communication skills
  • Strong problem-solving skills
  • Hands on experience with:
    • Kubernetes Platforms, Managed and Unmanaged
      • AKS, EKS, GKE
      • OpenShift, Rancher
      • Certified k8 Administrator is a plus
      • Linux - Shell scripting, Certified Linux Administrator is a plus
      • Container and Container Runtimes
    • Operating Systems concepts (CPU, Memory, and networking)
  • Working knowledge of :
    • Observability solutions - Application Performance Management
    • Load Testing solutions (e.g JMeter, LoadRunner, Grafana K6)
    • CI/CD and Automation Tools (e.g. Jenkins, Ansible)
    • Service Mesh (e.g. Istio), REST APIs and related tools
    • Familiarity with one or more programming Languages - Python, Java, Go
  • Certification and experience with one or more public cloud providers including Amazon Web Services (AWS), Microsoft Azure, or Google Cloud Platform (GCP)

Bonus experience:
  • Experience in a SRE or DevOps role resolving production outages
  • Knowledge of modern DevOps and SRE tools
  • Integration into ITSM Tools

*The role does not offer sponsorship employment benefits.
**If you don't think you meet all of the criteria below but still are interested in the job, please apply. Nobody checks every box-we're looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others.
Gremlin offers a competitive total rewards package, which includes:
  • Base salary
  • Equity
  • Healthcare, dental, and vision benefits
  • 401(k) with employer match.
  • Variable compensation for specific roles.

Compensation is based on the candidate's skills and qualifications.
About Gremlin:
Gremlin is a team of industry veterans and people eager to learn from one another. We set the standard for reliability and equip leading organizations with the mindset and expertise needed to drive reliability improvements that move the world forward. We're backed by top-tier investors Index Ventures, Amplify Partners, and Redpoint Ventures. Our customers love us, and we're thrilled to be a partner in their success.
What Do We Care About:
  • We Care about our People

People are our critical differentiators. The company strives to treat our people with respect, empathy, and dignity. We expect that our people will treat each other similarly. In both cases, we will assume good intent. All are welcome at Gremlin. We know our differences make us stronger and that our best ideas and contributions can come from anyone at any level.
  • We Care about Collaboration

Gremlin is strongest when we come together as one team with shared goals. Be the glue, not the glitter. But as a remote company, teamwork and collaboration won't happen by accident. We approach every challenge as a shared challenge. We rely on each other for diverse perspectives and creative ideas. We celebrate our wins as a team.
  • We Care about Results

Be high productivity, low drama. Results matter. To keep our pace, everyone owns the outcomes of their actions and takes action when needed. We reward speed over perfection. We empower each other to iterate and experiment.
You are welcome at Gremlin for who you are. The more voices and ideas we have represented in our business, the more we will all flourish, contribute, and build a more reliable internet. Gremlin is a place where everyone can grow and is encouraged. However you identify and whatever background you bring with you, please apply if this sounds like a role that would make you excited to come into work everyday. It's in our differences that we will find the power to keep building a more reliable internet by building and designing tools used by the best companies in the world.
Visit our website to learn more - https://www.gremlin.com/about