2

Director Site Reliability Engineer Remote Jobs in Wisconsin

AgenticOps Engineer

WI ยท On-site +1

... engineering, or SRE/DevOps with an understanding of reliability principles applied to non ... without direct authority. Why Work at Zywave? Zywave empowers insurers and brokers to drive ...

Cloud Infrastructure Engineer

West Bend, WI ยท On-site +1

$104K - $137K/yr

... or SRE roles. * Hands-on experience managing Kubernetes clusters in production, including secure ... PM19 LI#-REMOTE Equal Opportunity Employer This employer is required to notify all applicants of ...

$166K - $191K/yr

We are seeking a talented iOS Engineer to join us in building Poe, an innovative platform that ... You'll have a direct influence on the iOS team's product roadmap and plenty of opportunities to ...

$166K - $191K/yr

We are seeking a talented iOS Engineer to join us in building Poe, an innovative platform that ... You'll have a direct influence on the iOS team's product roadmap and plenty of opportunities to ...

$166K - $191K/yr

We are seeking a talented iOS Engineer to join us in building Poe, an innovative platform that ... You'll have a direct influence on the iOS team's product roadmap and plenty of opportunities to ...

$166K - $191K/yr

We are seeking a talented iOS Engineer to join us in building Poe, an innovative platform that ... You'll have a direct influence on the iOS team's product roadmap and plenty of opportunities to ...

$166K - $191K/yr

We are seeking a talented iOS Engineer to join us in building Poe, an innovative platform that ... You'll have a direct influence on the iOS team's product roadmap and plenty of opportunities to ...

$166K - $191K/yr

We are seeking a talented iOS Engineer to join us in building Poe, an innovative platform that ... You'll have a direct influence on the iOS team's product roadmap and plenty of opportunities to ...

$166K - $191K/yr

We are seeking a talented iOS Engineer to join us in building Poe, an innovative platform that ... You'll have a direct influence on the iOS team's product roadmap and plenty of opportunities to ...

It is also expected that as an engineering resource you will support the team as an expert for structural related topics that come up during the quoting and selling phase, or at site during ...

We are looking for a Cyber Security Engineer to help strengthen our security posture and operations ... direct impact in a healthcare setting. What You'll Do * Detect, investigate, and respond to ...

We augment Qooxdoo to enable dynamic site generation and extend the core components to provide a richer data analytics experience. We are looking for people with both a passion for visual perfection ...

next page

Showing results 1-20

Director Site Reliability Engineer Remote information

What is the difference between Director Site Reliability Engineer Remote vs Site Reliability Engineer?

AspectDirector Site Reliability Engineer RemoteSite Reliability Engineer
CredentialsTypically requires 8+ years of experience, advanced certifications (e.g., AWS, Google Cloud), leadership skillsUsually requires 3-5 years of experience, relevant certifications, strong technical skills
Work EnvironmentRemote leadership role overseeing teams, strategic planning, cross-team collaborationPrimarily technical role, often remote or on-site, focused on system reliability and automation
Employer & Industry UsageUsed in large tech companies, cloud providers, and enterprises with complex infrastructureCommon in tech, cloud, and SaaS companies focusing on system stability

The main difference is that the Director Site Reliability Engineer Remote focuses on leadership, strategy, and overseeing teams, while the Site Reliability Engineer is more hands-on, technical, and focused on system reliability tasks. Both roles may be remote, but their responsibilities and experience levels differ significantly.

What job categories do people searching Director Site Reliability Engineer Remote jobs in Wisconsin look for? The top searched job categories for Director Site Reliability Engineer Remote jobs in Wisconsin are:
What cities in Wisconsin are hiring for Director Site Reliability Engineer Remote jobs? Cities in Wisconsin with the most Director Site Reliability Engineer Remote job openings:

AgenticOps Engineer

Zywave

WI โ€ข On-site, Remote

Full-time

Re-posted 4 days ago


Job description

Description
The Agentic Ops Engineer operates as a cross-team specialist responsible for the health, reliability, and continuous improvement of AI agent output across the organization. They are not embedded in any single team's daily sprint cycle. Instead, they maintain a bird's-eye view of how agents perform across a handful of product engineering teams, identifying systemic patterns, diagnosing recurring failure modes, and tuning the shared prompt and tooling infrastructure that every team depends on.
When a Product Engineer hits a wall with agent output quality-whether it's a spec the agent can't parse, a prompt structure that produces inconsistent results, or a workflow that degrades over time-the Agentic Ops Engineer is who they pull in.
What you will do
Agent Performance Monitoring & Pattern Detection
โ€ข Continuously monitor agent output quality, latency, and failure rates across all three product teams.
โ€ข Identify recurring patterns such as "agents keep struggling with this type of spec" or "this prompt structure consistently produces higher-quality output.
โ€ข Build and maintain dashboards and alerting systems that surface degradation before teams feel it.
โ€ข Conduct periodic reviews of agent interaction logs to flag systemic issues and emerging trends.
Prompt Engineering & System Tuning
โ€ข Own the shared prompt infrastructure: templates, system prompts, few-shot libraries, and chain-of-thought scaffolding used across teams.
โ€ข Iterate on prompt structures based on observed failure modes and A/B performance data.
โ€ข Develop and maintain a prompt playbook documenting what works, what doesn't, and why.
โ€ข Evaluate and integrate new model capabilities, versioning changes, and API updates as they roll out from providers.
Escalation Support & Embedded Problem-Solving
โ€ข Serve as the on-call specialist when a Product Engineer encounters persistent agent output quality issues.
โ€ข Diagnose root causes: Is it the prompt? The spec format? The model's limitations? A context window issue?
โ€ข Pair with Product Engineers to rapidly prototype and test fixes, then roll improvements back into shared systems.
โ€ข Maintain a knowledge base of resolved issues and their solutions to reduce repeat escalations.
Tooling, Evaluation & Infrastructure
โ€ข Build and maintain evaluation harnesses, benchmarks, and regression test suites for agent workflows.
โ€ข Develop internal tooling for prompt version control, output comparison, and automated quality scoring.
โ€ข Collaborate with platform/infra teams to optimize agent execution pipelines (caching, context management, token budgets).
โ€ข Establish and track key metrics: output acceptance rate, revision frequency, time-to-resolution on escalations.
Knowledge Sharing & Team Enablement
โ€ข Run regular cross-team syncs sharing findings, patterns, and updated best practices.
โ€ข Produce internal documentation, guidelines, and training materials on working effectively with agents.
โ€ข Coach Product Engineers on prompt construction, spec formatting, and debugging agent behavior.
โ€ข Serve as the organizational point of contact for agent-related decisions (model selection, provider evaluation, capability assessments).
What you will bring
Required
โ€ข 5+ years of software engineering experience with strong fundamentals in systems thinking and debugging.
โ€ข Hands-on experience building with LLM APIs (prompt design, chain-of-thought, tool use, function calling).
โ€ข Demonstrated ability to diagnose and resolve complex, cross-cutting technical issues.
โ€ข Strong analytical skills: comfortable building dashboards, writing queries, and reasoning about statistical patterns in output quality.
โ€ข Excellent written and verbal communication-this role lives on documentation, cross-team clarity, and knowledge transfer.
Preferred
โ€ข Experience with prompt evaluation frameworks, LLM observability tools (e.g., LangSmith, Braintrust, Humanloop), or building internal evaluation harnesses.
โ€ข Background in developer tooling, platform engineering, or SRE/DevOps with an understanding of reliability principles applied to non-deterministic systems.
โ€ข Familiarity with multiple LLM providers and models; able to reason about trade-offs in capability, cost, and latency.
โ€ข Experience working cross-functionally across multiple product teams without direct authority.
Why Work at Zywave?
Zywave empowers insurers and brokers to drive profitable growth and thrive in today's escalating risk landscape. Only Zywave delivers a powerful Performance Multiplier, bringing together transformative, ecosystem-wide capabilities to amplify impact across data, processes, people, and customer experiences. More than 15,000 insurers, MGAs, agencies, and brokerages trust Zywave to sharpen risk assessment, strengthen client relationships, and enhance operations. Additional information can be found at www.zywave.com.
Equal Opportunity Employer
Zywave is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, veteran status, or any other legally protected status. We are committed to building an inclusive workplace where everyone can do their best work. If you need a reasonable accommodation during the application or interview process, please contact our Talent Acquisition team.
#LI-AK1