1

Observability Site Reliability Engineer Jobs in Oregon

Senior Software Engineer, Site Reliability

OR · On-site +1

$57 - $75.75/hr

The Team Upstart's Site Reliability Engineering (SRE) team owns the reliability, resiliency, and observability of Upstart's production systems. The SRE team builds tooling and automation to monitor ...

Site Reliability Engineer TELCOR Inc, a leading innovator in laboratory software, is looking for a ... Do you focus on uptime, resilience, observability and incident response? This may be the role for ...

Software Engineer, Site Reliability

OR · On-site +1

$57 - $75.75/hr

The Team Upstart's Site Reliability Engineering team enables engineers to operate reliable ... We provide shared observability and reliability capabilities, improve incident response and ...

New

Senior Engineering Manager, Site Reliability

OR · On-site +1

$57 - $75.75/hr

The team advances observability, incident detection and response, service level objectives, operational readiness, and systemic improvements based on incident learnings. SRE partners across product ...

Sr. Site Reliability Engineer

OR · On-site +1

$57 - $75.75/hr

FreedomPay is seeking an experienced Senior Site Reliability Engineer to help ensure the highest ... This full-time salaried position builds on a strong foundation of observability, incident response ...

Site Reliability Engineer II

OR · On-site +1

$57 - $75.75/hr

Overview Are you passionate about technology and ready to dive into the world of Site Reliability Engineering? We are looking for enthusiastic individuals to join our team as Site Reliability ...

Site Reliability Engineer

OR · On-site +1

$57 - $75.75/hr

Additionally, the SRE I role will be an expert in Infrastructure as Code (IaC), AWS Cloud ... Mastery of observability, monitoring, metrics and alerting at scale across regionally and globally ...

Essential Skills & Experience * 3+ years of professional experience in SRE, DevOps, or Cloud Engineering within enterprise software environments. * Azure Mastery: Extensive hands-on experience with ...

The Senior Observability Engineer is responsible for defining, leading, and advancing enterprise ... Collaborate across engineering, site reliability, platform, security, infrastructure, operations ...

Site Reliability Engineer (Tu-Sat, night shift)

OR · On-site +1

$57 - $75.75/hr

Why this role matters: The Site Reliability Engineering role is a blend of infrastructure ... We need your help to take our existing platform to the next level with observability, release ...

Senior Staff Site Reliability Engineer

OR · On-site +1

$170K - $227K/yr

As a Ping Identity Senior Staff Site Reliability Engineer, you will be involved in every facet of ... resiliency, observability and optimized cost. * Communicate proactively and effectively to ...

Senior Site Reliability Engineer- Remote

OR · On-site +1

$57 - $75.75/hr

About the role We are committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building ...

Staff Infrastructure Engineer - Observability

OR · On-site +1

$107K - $140K/yr

Ideal candidates will have * 8+ years experience in Infrastructure Engineering, Site Reliability ... grade observability stacks utilizing Prometheus, Grafana, Thanos (or Mimir/Cortex), and ...

Senior Site Reliability Engineer - Compute Platforms

OR · On-site +1

$57 - $75.75/hr

We are seeking a highly experienced Senior Site Reliability Engineer - Compute Platforms to design, implement, and support Kubernetes on baremetal and hypervisor platforms in a private cloud ...

next page

Showing results 1-20

Observability Site Reliability Engineer information

What is the difference between Observability Site Reliability Engineer vs Monitoring Engineer?

AspectObservability Site Reliability EngineerMonitoring Engineer
FocusEnsuring system reliability through observability, automation, and incident responseImplementing and managing monitoring tools and dashboards
SkillsCloud platforms, scripting, incident management, observability toolsMonitoring tools, alerting systems, data analysis
Work EnvironmentDevOps teams, cloud infrastructure, large-scale systemsOperations teams, infrastructure monitoring

While both roles involve system health, the Observability Site Reliability Engineer focuses on comprehensive system reliability using observability practices, whereas Monitoring Engineers primarily manage monitoring tools and alerts. The SRE role emphasizes automation, incident response, and system resilience, making it broader in scope.

What are popular job titles related to Observability Site Reliability Engineer jobs in Oregon?

For Observability Site Reliability Engineer jobs in Oregon, the most frequently searched job titles are:

What job categories do people searching Observability Site Reliability Engineer jobs in Oregon look for?

The top searched job categories for Observability Site Reliability Engineer jobs in Oregon are:

What cities in Oregon are hiring for Observability Site Reliability Engineer jobs?

Cities in Oregon with the most Observability Site Reliability Engineer job openings:

Senior Software Engineer, Site Reliability

OR • On-site, Remote


Upstart

7.6

Company rating: 7.6 out of 10

Based on 6 frontline employees who took The Breakroom Quiz

Good employer

Respectful managers

Good training


$57 - $75.75/hr

Full-time

Re-posted 8 days ago


Job description

The Team

Upstart's Site Reliability Engineering (SRE) team owns the reliability, resiliency, and observability of Upstart's production systems.  The SRE team builds tooling and automation to monitor the health of our infrastructure and create a fast, reliable, and productive environment for other engineers and a world-class experience for our customers. SRE defines Upstart's strategy for technology operations risk mitigation, which includes disaster planning and on-call procedures. We use data-driven approaches to drive our decisions, and provide reports and insights to the business to improve visibility into the system and customer experience.

As a Senior Software Engineer focused on Site Reliability Tooling, your work will directly impact the success of the SRE team and all of Upstart. Your expertise will inform the team's direction, and your work with other SREs and Upstart engineers will make Upstart's systems as effective as possible for our customers.  SRE at Upstart is ever-changing, and you will be a primary contributor in shaping our future path.

How you'll make an impact:

  • Embody and share SRE principles at Upstart
  • Exercise state-of-the-art SRE practices throughout the company
  • Uphold a culture of visibility, ownership, and responsibility around service reliability
  • Implement standards for monitoring microservices, web apps, mobile apps, databases, Kubernetes clusters, and machine learning platforms, in a fast-paced environment
  • Improve incident response practices, both within SRE and throughout the company
  • Automate away toil that make sense to be automated

What we're looking for: 

  • Minimum requirements:
    • Minimum of 6 years combined experience between Software Engineering, Site Reliability, and/or DevOps Engineering including CI/CD, TDD, internal tooling, observability, and other agile development practices
    • Proficiency coding Python, Go, JavaScript/TypeScript 
    • Proficiency with Infrastructure as Code (Terraform, CDK, Cloudformation, etc.)
    • Software engineering background with experience building internal tooling from scratch, and other agile development techniques
    • Strong software design & architecture skills
    • Fundamentally sound with data structures & algorithms 
    • Experience with on-call and incident management environments
    • Experience with observability, monitoring, and reporting tools (e.g., Datadog, Sumologic, , etc.)
    • Experience supporting SaaS software in a microservice-oriented cloud environment
    • Ability to work with multiple teams for enterprise-wide deliverables
    • Data/metrics-driven mindset
  • Preferred qualifications:
    • Experience with service mesh
    • Full Stack development skills 
    • Experience building tooling for an observability platform
    • Experience leveraging LLM/GenAI to improve SRE efficiency and processes  

Position Location - This role is available in the following locations: Remote, San Mateo, Columbus, Austin, NY

Time Zone Requirements - This team operates across all U.S. time zones.

Travel Requirements - This team has regular on-site collaboration sessions. These occur 3 days per quarter at an Upstart office. If you need to travel to make these meetups, Upstart will cover all travel related expenses.


What Upstart employees say

Pay

Hours and flexibility

Workplace

Get the full story on Breakroom