1

Chaos Engineering Jobs (NOW HIRING)

Stress-test and push our systems to the edge: load testing, chaos engineering, and performance benchmarking. * Own security best practices at the infrastructure layer, from sandboxed compute to ...

Site Reliability Engineer

San Francisco, CA

$67.25 - $89.25/hr

Stress-test and push our systems to the edge: load testing, chaos engineering, and performance benchmarking. * Own security best practices at the infrastructure layer, from sandboxed compute to ...

Coordinate with CHAOS Engineering teams to understand technical constraints and requirements for customer-facing events, and to design exhibits for conferences * Partner with Mission Operations ...

Coordinate with CHAOS Engineering teams to understand technical constraints and requirements for customer-facing events, and to design exhibits for conferences * Partner with Mission Operations ...

Senior Backend Engineer

$220K - $290K/yr

As the industry leader in Chaos Engineering and reliability testing, we work with hundreds of the world's largest organizations where high availability is non-negotiable. About the Role of the Senior ...

Senior Backend Engineer

$220K - $290K/yr

As the industry leader in Chaos Engineering and reliability testing, we work with hundreds of the world's largest organizations where high availability is non-negotiable. About the Role of the Senior ...

Showing results 41-60

Chaos Engineering information

See salary details

$46.5K

$146.9K

$174K

How much do chaos engineering jobs pay per year?

As of Sep 9, 2026, the average yearly pay for chaos engineering in the United States is $146,868.00, according to ZipRecruiter salary data. Most workers in this role earn between $116,500.00 and $173,000.00 per year, depending on experience, location, and employer.

What is chaos engineering?

A Chaos Engineering job involves proactively identifying weaknesses in complex systems by intentionally injecting failures and observing how they respond. Professionals in this role design and execute controlled experiments to improve system resilience, ensuring that services remain reliable under unexpected conditions. They work closely with development, operations, and security teams to enhance fault tolerance and incident response strategies.

What are some typical challenges a chaos engineer faces, and how do they overcome them?

Chaos Engineers often face the challenge of designing effective experiments that simulate real-world failures without disrupting production systems. Balancing the need to discover vulnerabilities with maintaining uptime requires careful planning, communication, and coordination with development and operations teams. They address these challenges by thoroughly testing in controlled environments, documenting procedures, and establishing clear rollback strategies. Continuous learning and cross-functional collaboration are also key to staying ahead of new complexities in evolving systems.

What are the key skills and qualifications needed to thrive in the chaos engineering position, and why are they important?

To thrive in Chaos Engineering, a strong background in software engineering, distributed systems, and reliability testing is essential, often supported by a degree in computer science or a related field. Familiarity with chaos engineering tools like Gremlin or Chaos Monkey and experience with cloud platforms, container orchestration, and monitoring systems are highly valued. Excellent problem-solving abilities, communication skills, and a mindset oriented toward experimentation help engineers collaborate effectively and analyze complex failure modes. These skills are crucial for proactively identifying system weaknesses and ensuring the resilience of large-scale technology infrastructures.

Is chaos engineering still relevant?

Chaos engineering is a valuable practice for proactively identifying system vulnerabilities by intentionally introducing failures. It remains relevant in modern DevOps and cloud environments to improve system resilience and reliability, often utilizing tools like Chaos Monkey and Gremlin. As systems grow more complex, the need for chaos engineering skills continues to increase for engineers focused on fault tolerance and system stability.

What does a chaos engineer do?

A chaos engineer designs and executes experiments to intentionally disrupt systems in order to identify vulnerabilities and improve resilience. They use tools like chaos engineering frameworks to simulate failures and ensure systems can withstand unexpected issues, often working closely with development and operations teams. Strong knowledge of distributed systems, scripting, and monitoring is essential for this role.
More about Chaos Engineering jobs

What cities are hiring for Chaos Engineering jobs?

Cities with the most Chaos Engineering job openings:

What are the most commonly searched types of Chaos Engineering jobs?

The most popular types of Chaos Engineering jobs are:

What states have the most Chaos Engineering jobs?

States with the most job openings for Chaos Engineering jobs include:

Infographic showing various Chaos Engineering job openings in the United States as of September 2026, with employment types broken down into 1% Internship, 91% Full Time, 4% Part Time, 3% Contract, and 1% Nights. Highlights an 83% Physical, 4% Hybrid, and 13% Remote job distribution, with an average salary of $146,868 per year, or $70.6 per hour.

Senior Performance Engineer ID84179

Orange, CT • On-site, Remote

AgileEngine
Software Development • 201 - 500 employees

Full-time

Posted 19 days ago


Job description

AgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards.

WHY JOIN US
If you're looking for a place to grow, make an impact, and work with people who care, we'd love to meet you!

ABOUT THE ROLE
We are looking for a Senior Performance Engineer to own and lead an enterprise-wide Performance Engineering function as a horizontal shared service across a global QSR digital ecosystem including web, mobile, kiosk, POS, loyalty, and payments systems. You will design and execute load, stress, and endurance tests using JMeter, k6, or similar tools, integrate automated performance validation into Azure DevOps CI/CD pipelines, partner with SRE and DevOps teams on chaos engineering experiments, and leverage AI-assisted workflows for performance planning and results triage.

WHAT YOU WILL DO
- Own the enterprise Performance Engineering charter as a horizontal shared service, establishing engagement models and SDLC performance-readiness standards.
- Design, script, execute, and report performance tests (load, stress, endurance, scalability) across high-volume ordering, kiosk, POS, loyalty, and payment flows.
- Identify and resolve bottlenecks across application, database, messaging, and infrastructure layers using observability signals (e.g., Dynatrace).
- Embed automated performance validation into CI/CD pipelines (e.g., Azure DevOps) with baseline and threshold gating.
- Partner with SRE/DevOps teams to define reliability targets and run controlled chaos-engineering experiments (fault/latency injection, dependency degradation).
- Define performance-readiness criteria, provide evidence-based go/no-go release recommendations, and assist with production performance-incident resolution.
- Integrate AI-assisted workflows into performance planning, test-design acceleration, and results triage while maintaining AI-consumable artifacts and safety guardrails.

MUST HAVES
- You must be authorized to work for ANY employer in the US (e.g., Green card holders, TN visa holders, GC EAD, H4 EAD, U4U with EAD), as we are unable to sponsor or take over employment visa sponsorship at this time;
- 4+ years of experience in performance engineering or non-functional testing at enterprise scale.
- Hands-on expertise with performance testing tools (e.g., JMeter, k6, LoadRunner, NeoLoad, or OctoPerf).
- Experience with enterprise observability and cloud monitoring platforms (e.g., Dynatrace).
- Demonstrated experience integrating automated performance validation into CI/CD pipelines (e.g., Azure DevOps).
- Scripting and automation proficiency in Java, JavaScript, or Python.
- Hands-on engineering experience in cloud environments (AWS or Azure).
- Hands-on experience using AI-assisted engineering tools, including reviewing and validating AI-generated output.
- Strong communication and stakeholder leadership skills with the ability to translate technical findings into clear business impact.
- Upper-intermediate English level.

NICE TO HAVES
- Direct experience partnering with SRE/DevOps practices on chaos engineering (fault/latency injection and dependency degradation testing).
- Prior experience supporting digital platforms within quick-service-restaurant (QSR), retail, or high-volume transactional ordering ecosystems.
- Experience maintaining AI-consumable artifacts (NFRs, workload models, baselines, thresholds) and defining AI guardrails.

PERKS AND BENEFITS
- Growth without limits: build your skills through mentorship, internal TechTalks, challenging projects, and a dedicated annual learning budget
- Competitive compensation: get recognition that reflects your skills and impact, with regular performance and compensation reviews
- Flexibility: work 100% remotely with flexible hours that support focus, autonomy, and a healthy work rhythm
- Meaningful, modern projects: build impactful products using modern technologies alongside global teams and leading brands
- Collaborative culture: join a supportive environment with zero micromanagement where ideas are welcomed and contributions are recognized
- Well-being & support: access local well-being programs and people-focused support tailored to your location