2

Site Reliability Engineer Remote Jobs in Wisconsin

AgenticOps Engineer

WI · On-site +1

... or SRE/DevOps with an understanding of reliability principles applied to non-deterministic systems. • Familiarity with multiple LLM providers and models; able to reason about trade-offs in ...

Cloud Infrastructure Engineer

West Bend, WI · On-site +1

$104K - $137K/yr

... or SRE roles. * Hands-on experience managing Kubernetes clusters in production, including secure ... PM19 LI#-REMOTE Equal Opportunity Employer This employer is required to notify all applicants of ...

We augment Qooxdoo to enable dynamic site generation and extend the core components to provide a richer data analytics experience. We are looking for people with both a passion for visual perfection ...

$52 - $71.25/hr

We are looking for an experienced DevOps Engineers interested in building, maintaining, and scaling PlaidCloud on Kubernetes. This position requires a keen eye for detail to ensure consistency in ...

next page

Showing results 1-20

Site Reliability Engineer Remote information

See Wisconsin salary details

$10

$64

$92

How much do site reliability engineer remote jobs pay per hour?

As of Jul 26, 2026, the average hourly pay for site reliability engineer remote in Wisconsin is $64.34, according to ZipRecruiter salary data. Most workers in this role earn between $55.34 and $73.51 per hour, depending on experience, location, and employer.

What are the key skills and qualifications needed to thrive in the Site Reliability Engineer Remote position, and why are they important?

To thrive as a Site Reliability Engineer Remote, you need expertise in systems administration, cloud infrastructure, automation, coding (often in Python or Go), and a solid grasp of networking fundamentals, usually demonstrated with a degree in computer science or equivalent experience. Familiarity with tools such as Docker, Kubernetes, AWS/GCP/Azure, monitoring platforms like Prometheus, and certifications like AWS Certified SysOps Administrator are highly valued. Excellent problem-solving, communication, and collaboration skills are essential, especially when troubleshooting incidents and passing information across distributed teams. These abilities ensure reliable, scalable services and smooth coordination in a remote work environment.

What is a Site Reliability Engineer Remote job?

A Site Reliability Engineer (SRE) in a remote role is responsible for ensuring the reliability, performance, and scalability of software systems while working from a remote location. They bridge the gap between development and operations by implementing automation, monitoring, and incident response strategies. Remote SREs collaborate with distributed teams to improve infrastructure, troubleshoot issues, and optimize system performance. Strong communication skills, proficiency in cloud technologies, and expertise in software development are essential for success in this role.

What are some common challenges faced by Site Reliability Engineers working remotely, and how are they addressed?

Site Reliability Engineers working remotely may encounter challenges like coordinating across multiple time zones, maintaining clear communication during urgent incidents, and managing complex systems without direct on-site access. These are often addressed by leveraging collaborative tools (like Slack, Zoom, and incident management platforms), implementing well-documented processes, and participating in regular team syncs or on-call rotations. Remote SREs also benefit from automation and observability practices that provide in-depth systems insights without needing physical presence. Many organizations support their success through robust onboarding, continuous training, and establishing clear lines of communication for rapid response scenarios. This blend of technical and teamwork strategies helps remote SREs maintain service reliability and stay connected with their colleagues.

What are the most commonly searched types of Site Reliability Engineer jobs in Wisconsin? The most popular types of Site Reliability Engineer jobs in Wisconsin are:
What are popular job titles related to Site Reliability Engineer Remote jobs in Wisconsin? For Site Reliability Engineer Remote jobs in Wisconsin, the most frequently searched job titles are:
What job categories do people searching Site Reliability Engineer Remote jobs in Wisconsin look for? The top searched job categories for Site Reliability Engineer Remote jobs in Wisconsin are:
What cities in Wisconsin are hiring for Site Reliability Engineer Remote jobs? Cities in Wisconsin with the most Site Reliability Engineer Remote job openings:
Infographic showing various Site Reliability Engineer Remote job openings in Wisconsin as of July 2026, with employment types broken down into 91% Full Time, 7% Part Time, and 2% Contract. Highlights an 87% Physical, 5% Hybrid, and 8% Remote job distribution, with an average salary of $133,823 per year, or $64.3 per hour.
AgenticOps Engineer

AgenticOps Engineer

Zywave

WI • On-site, Remote

Full-time

Posted 29 days ago


Job description

Description
The Agentic Ops Engineer operates as a cross-team specialist responsible for the health, reliability, and continuous improvement of AI agent output across the organization. They are not embedded in any single team's daily sprint cycle. Instead, they maintain a bird's-eye view of how agents perform across a handful of product engineering teams, identifying systemic patterns, diagnosing recurring failure modes, and tuning the shared prompt and tooling infrastructure that every team depends on.
When a Product Engineer hits a wall with agent output quality-whether it's a spec the agent can't parse, a prompt structure that produces inconsistent results, or a workflow that degrades over time-the Agentic Ops Engineer is who they pull in.
What you will do
Agent Performance Monitoring & Pattern Detection
• Continuously monitor agent output quality, latency, and failure rates across all three product teams.
• Identify recurring patterns such as "agents keep struggling with this type of spec" or "this prompt structure consistently produces higher-quality output.
• Build and maintain dashboards and alerting systems that surface degradation before teams feel it.
• Conduct periodic reviews of agent interaction logs to flag systemic issues and emerging trends.
Prompt Engineering & System Tuning
• Own the shared prompt infrastructure: templates, system prompts, few-shot libraries, and chain-of-thought scaffolding used across teams.
• Iterate on prompt structures based on observed failure modes and A/B performance data.
• Develop and maintain a prompt playbook documenting what works, what doesn't, and why.
• Evaluate and integrate new model capabilities, versioning changes, and API updates as they roll out from providers.
Escalation Support & Embedded Problem-Solving
• Serve as the on-call specialist when a Product Engineer encounters persistent agent output quality issues.
• Diagnose root causes: Is it the prompt? The spec format? The model's limitations? A context window issue?
• Pair with Product Engineers to rapidly prototype and test fixes, then roll improvements back into shared systems.
• Maintain a knowledge base of resolved issues and their solutions to reduce repeat escalations.
Tooling, Evaluation & Infrastructure
• Build and maintain evaluation harnesses, benchmarks, and regression test suites for agent workflows.
• Develop internal tooling for prompt version control, output comparison, and automated quality scoring.
• Collaborate with platform/infra teams to optimize agent execution pipelines (caching, context management, token budgets).
• Establish and track key metrics: output acceptance rate, revision frequency, time-to-resolution on escalations.
Knowledge Sharing & Team Enablement
• Run regular cross-team syncs sharing findings, patterns, and updated best practices.
• Produce internal documentation, guidelines, and training materials on working effectively with agents.
• Coach Product Engineers on prompt construction, spec formatting, and debugging agent behavior.
• Serve as the organizational point of contact for agent-related decisions (model selection, provider evaluation, capability assessments).
What you will bring
Required
• 5+ years of software engineering experience with strong fundamentals in systems thinking and debugging.
• Hands-on experience building with LLM APIs (prompt design, chain-of-thought, tool use, function calling).
• Demonstrated ability to diagnose and resolve complex, cross-cutting technical issues.
• Strong analytical skills: comfortable building dashboards, writing queries, and reasoning about statistical patterns in output quality.
• Excellent written and verbal communication-this role lives on documentation, cross-team clarity, and knowledge transfer.
Preferred
• Experience with prompt evaluation frameworks, LLM observability tools (e.g., LangSmith, Braintrust, Humanloop), or building internal evaluation harnesses.
• Background in developer tooling, platform engineering, or SRE/DevOps with an understanding of reliability principles applied to non-deterministic systems.
• Familiarity with multiple LLM providers and models; able to reason about trade-offs in capability, cost, and latency.
• Experience working cross-functionally across multiple product teams without direct authority.
Why Work at Zywave?
Zywave empowers insurers and brokers to drive profitable growth and thrive in today's escalating risk landscape. Only Zywave delivers a powerful Performance Multiplier, bringing together transformative, ecosystem-wide capabilities to amplify impact across data, processes, people, and customer experiences. More than 15,000 insurers, MGAs, agencies, and brokerages trust Zywave to sharpen risk assessment, strengthen client relationships, and enhance operations. Additional information can be found at www.zywave.com.
#LI-AK1