1

Senior Observability Engineer Jobs in Springfield, MA

Senior Security Engineer

Glastonbury, CT ยท Remote

$95K - $145K/yr

Vancord is seeking a Senior Security Engineer to serve as our SOC Lead. This is primarily a ... A strong understanding of data pipelines, normalization, and security observability. Preferred ...

The Sr. Digital Software Engineer leads the delivery of complex software solutions while guiding ... Familiarity with DevOps practices, CI/CD pipelines, and observability tools. * Effective ...

Sr Cloud Engineer

Hartford, CT ยท Hybrid

$137K - $205K/yr

Sr Cloud Engineer - IE07NE We're determined to make a difference and are proud to be an insurance ... duties, observability, and being operator friendly. * Leverage AI to improve operational ...

Sr Cloud Engineer

Hartford, CT ยท Hybrid

$137K - $205K/yr

Sr Cloud Engineer - IE07NE We're determined to make a difference and are proud to be an insurance ... duties, observability, and being operator friendly. * Leverage AI to improve operational ...

Staff Engineer, Composer Platform

Glastonbury, CT ยท On-site +1

$180K - $250K/yr

Improve platform reliability, scalability, observability, and operational resilience * Establish ... At least 12 years of experience as Staff Engineer or senior-level individual contributor supporting ...

Senior Data Engineer

Hartford, CT ยท On-site

$139K - $230K/yr

As a Senior Data Engineer you will accelerate growth and transformation of our analytics landscape ... backfills, and observability. * Demonstrated experience delivering data solutions in AWS ...

next page

Showing results 1-20

Senior Observability Engineer information

See Springfield, MA salary details

$59.3K

$126.1K

$182.9K

How much do senior observability engineer jobs pay per year?

As of Sep 3, 2026, the average yearly pay for senior observability engineer in Springfield, MA is $126,114.00, according to ZipRecruiter salary data. Most workers in this role earn between $104,100.00 and $143,000.00 per year, depending on experience, location, and employer.

What is a senior observability engineer?

A Senior Observability Engineer is a seasoned IT professional responsible for designing, implementing, and maintaining systems that monitor and provide insights into the performance, health, and reliability of software applications and infrastructure. They utilize tools for logging, monitoring, tracing, and alerting to ensure that systems are observable and any issues can be quickly detected and resolved. In addition to technical expertise, they often collaborate with development and operations teams to establish best practices, improve incident response, and optimize system performance. Their work is crucial for maintaining uptime, enhancing customer experiences, and supporting the scalability of technology platforms.

How does a senior observability engineer typically collaborate with development and operations teams?

A Senior Observability Engineer works closely with both development and operations teams to ensure robust monitoring, logging, and tracing solutions are in place across all applications and infrastructure. They often participate in architecture discussions to advise on best practices for instrumenting code and systems for observability. By analyzing metrics and alerting patterns, they help teams proactively resolve issues and optimize system performance. This role also involves mentoring engineers on observability tools and fostering a culture of transparency and accountability in incident response.

What are the key skills and qualifications needed to thrive as a senior observability engineer, and why are they important?

To thrive as a Senior Observability Engineer, you need expertise in monitoring, logging, and tracing systems, with a solid background in computer science or a related field. Familiarity with tools like Prometheus, Grafana, ELK stack, and cloud platforms, as well as certifications such as AWS Certified DevOps Engineer, are typically required. Strong problem-solving, collaboration, and communication skills are critical for effectively diagnosing and resolving complex infrastructure issues. These skills ensure reliable system performance, rapid incident response, and continuous improvement of the technology environment.

What is the difference between Senior Observability Engineer vs Site Reliability Engineer?

AspectSenior Observability EngineerSite Reliability Engineer
CredentialsExperience with monitoring tools, scripting, cloud platformsSame as Senior Observability Engineer, often with SRE certifications
Work EnvironmentFocus on monitoring, logging, and tracing systemsFocus on system reliability, automation, and incident response
Industry UsageUsed in tech companies emphasizing system observabilityCommon in large-scale tech and cloud services
Search/Comparison IntentOften compared for monitoring rolesCompared for reliability and system stability roles

While both roles require expertise in cloud platforms and scripting, the Senior Observability Engineer primarily focuses on designing and maintaining monitoring, logging, and tracing systems to ensure system visibility. In contrast, a Site Reliability Engineer emphasizes system reliability, automation, and incident management to maintain service uptime. Both roles are vital in tech environments but serve different core functions related to system health and stability.

How much do senior observability engineers make?

Senior observability engineers typically earn between $110,000 and $160,000 annually, depending on experience, location, and company size. They often work with tools like Prometheus, Grafana, and cloud platforms, and may require advanced knowledge of monitoring, logging, and alerting systems.

What does a senior observability engineer do?

A senior observability engineer designs, implements, and maintains systems to monitor the performance and health of software applications and infrastructure. They utilize tools like Prometheus, Grafana, and ELK stack to analyze metrics, logs, and traces, ensuring system reliability and performance. This role often requires strong scripting skills and knowledge of cloud environments and distributed systems.

What are the most commonly searched types of Observability Engineer jobs in Springfield, MA?

The most popular types of Observability Engineer jobs in Springfield, MA are:

What are popular job titles related to Senior Observability Engineer jobs in Springfield, MA?

For Senior Observability Engineer jobs in Springfield, MA, the most frequently searched job titles are:

What job categories do people searching Senior Observability Engineer jobs in Springfield, MA look for?

The top searched job categories for Senior Observability Engineer jobs in Springfield, MA are:

What cities near Springfield, MA are hiring for Senior Observability Engineer jobs?

Cities near Springfield, MA with the most Senior Observability Engineer job openings:

Infographic showing various Senior Observability Engineer job openings in Springfield, MA as of June 2026, with employment types broken down into 41% Full Time, 54% Part Time, and 5% Contract. Highlights an 78% Physical, 6% Hybrid, and 16% Remote job distribution, with an average salary of $126,114 per year, or $60.6 per hour.

Senior Site Reliability Engineer

ISO New England, Inc.

Holyoke, MA โ€ข On-site

$56.25 - $74.75/hr

Full-time

Medical, Dental, Vision, Life, Retirement, PTO

Posted 18 days ago


Job description

ISO New England is the independent system operator responsible for ensuring the safe and reliable flow of electricity in our region and planning for the future of the electric grid. We are at the forefront of New England's ongoing transition to clean energy.
The Senior Site Reliability Engineer (SRE) is a hands-on engineering role responsible for improving the reliability, observability, performance, and operational efficiency of ISO New England's IT services. The SRE works across infrastructure, platform, cyber security, and application teams to reduce operational toil, improve service resilience, and implement scalable automation solutions.
This role has a strong emphasis on observability engineering, automation, Splunk administration, and Infrastructure as Code (IaC). The ideal candidate will possess hands-on experience with Splunk or demonstrate a strong willingness to develop expertise in the platform. Experience with Terraform, automation technologies such as Python and PowerShell, and the ability to leverage AI-assisted development tools to accelerate engineering solutions are key components of the role.
What we offer you:
  • A stable, mission-driven workplace where your impact truly matters
  • A highly engaged work environment that values inclusion, collaboration, and employee safety and wellbeing
  • Competitive compensation with a base salary + performance bonus
  • Robust benefits package, including:
    • Enhanced 401(k) and financial planning support
    • Tuition reimbursement and professional development
    • Wellness programs, including an onsite gym
    • Flexible work hours
    • Employee Business Networks
    • Free coffee at our onsite cafรฉ
  • Hybrid work environment (3 days/week onsite)
  • Distance-based relocation assistance available

How you will make an Impact
  • Build and maintain observability, monitoring, logging, alerting, and telemetry platforms (e.g., Splunk, Dynatrace, PRTG, OpsGenie, StatusPage)
  • Administer, maintain, automate, and continuously improve the Splunk platform, including data onboarding, indexing, search performance, dashboards, access controls, health monitoring, platform scalability, and operational workflows
  • Develop and automate Splunk onboarding, configuration, monitoring, and operational workflows to improve platform reliability and reduce administrative overhead
  • Develop meaningful KPIs and dashboards for business and IT service health
  • Engineer and implement resilience patterns including HA, DR, and automated failover
  • Partner with infrastructure and application teams to plan and execute resilience testing and failover exercises to validate recovery capabilities and observability coverage
  • Conduct performance testing, capacity modeling, forecasting, and right-sizing
  • Participate in major incident response activities, providing technical expertise to accelerate service restoration and identify reliability improvements
  • Identify, prioritize, and eliminate manual operational toil through automation, targeting workflows, runbooks, alerting, platform administration, service management processes, and KPI collection, with a bias toward scalable and repeatable engineering solutions
  • Design, develop, maintain, and support automation solutions, integrations, and operational tooling using Python, PowerShell, Bash, or similar technologies to improve reliability, reduce manual effort, and enhance operational efficiency
  • Design, deploy, and manage infrastructure using Terraform and Infrastructure as Code (IaC) practices, including observability platforms, infrastructure services, and supporting technology stacks, with a focus on consistency, repeatability, and operational sustainability
  • Identify gaps in observability coverage and drive engineering solutions to close them
  • Collaborate with architecture and application teams to ensure production readiness
  • Leverage AI-assisted development tools to accelerate automation initiatives while reviewing, validating, troubleshooting, and refining generated code to ensure reliability, security, maintainability, and operational effectiveness
  • Reduce repeat incidents by engineering permanent fixes and driving continuous improvement

What we are looking for
  • 5+ years of experience in SRE, DevOps, systems engineering, platform engineering, or IT operations
  • Experience with enterprise monitoring and observability platforms. Hands-on experience with Splunk is strongly preferred. Candidates without direct Splunk experience must demonstrate a strong willingness and aptitude to develop expertise in Splunk administration, engineering, and automation.
  • Experience designing, deploying, or managing infrastructure using Terraform and Infrastructure as Code (IaC) practices
  • Strong scripting and automation experience using Python, PowerShell, Bash, or similar technologies, including the development of operational tooling, integrations, and workflow automation in production environments
  • Demonstrated experience designing, developing, and supporting automation solutions that measurably reduced manual operational effort in an enterprise environment
  • Ability to read, understand, review, troubleshoot, and refine code produced by engineering teams or AI-assisted development platforms
  • Knowledge of distributed systems, networking, enterprise infrastructure, and cloud platforms
  • Familiarity with SRE principles including SLOs, error budgets, observability, and toil reduction
  • Ability to analyze and troubleshoot complex technical systems
  • Preferred Qualifications
  • Experience in mission-critical, highly available, or regulated environments
  • Experience utilizing AI-assisted development tools to accelerate automation, operational engineering, or platform management activities
  • Knowledge of ITIL processes and/or SRE best practices
  • Experience with performance testing, capacity planning, resilience testing, or disaster recovery validation

This employer will not sponsor applicants for work visas for this position (ex: H-1B, F-1/CPT/OPT, O-1, E-3, TN, J, etc.).
The expected salary range for this position is $134,000 - $170,000 per year, for a Senior to Lead level candidate. This role is also eligible for an annual performance bonus, comprehensive health insurance (medical, dental and vision), flexible spending and health savings accounts, a 401(k) plan with generous employer contributions and a student debt benefit, life and AD&D insurance, disability insurance, critical illness and hospital indemnity benefits, paid time off, paid leave, a wellness program, an employee assistance program and other great company perks.
#LI-HYBRID
This is a U.S. based role. If the successful candidate resides outside of the U.S., relocation will be required.
Equal Opportunity: We are proud to be an EEO employer. Applicants for employment are considered without regard to race, color, religion, creed, sex (including pregnancy, childbirth, and related medical conditions), gender identity or expression, sexual orientation, citizenship, national origin, age, ancestry, marital status, disability (including learning, mental, intellectual, and physical), service in the uniformed services, genetic information, or any other status protected by applicable law.
Drug Free Environment: We maintain a drug-free workplace and perform pre-employment substance abuse testing.