1

Site Reliability Engineer Jobs in Connecticut (NOW HIRING)

Provide post-market reliability engineering support, including analysis of field performance and failure data * Ensure compliance with applicable regulatory and industry standards, including FDA ...

Provide post-market reliability engineering support, including analysis of field performance and failure data * Ensure compliance with applicable regulatory and industry standards, including FDA ...

Working closely with our SRE team to ensure deployed systems are reliable, resilient, scalable, and perform optimally, using common, modern technologies. * Make best use of our observability systems ...

Working closely with our SRE team to ensure deployed systems are reliable, resilient, scalable, and perform optimally, using common, modern technologies. * Make best use of our observability systems ...

Infrastructure Engineer

Torrington, CT · On-site

$110K - $160K/yr

We are transitioning our Tech Ops culture toward a modern Platform Engineering and SRE model. This role is ideal for a high-velocity learner who is inherently curious, exceptionally organized, and ...

Infrastructure Engineer

Torrington, CT · On-site

$110K - $160K/yr

We are transitioning our Tech Ops culture toward a modern Platform Engineering and SRE model. This role is ideal for a high-velocity learner who is inherently curious, exceptionally organized, and ...

Sr. Software Engineer

Stamford, CT

$130K - $172K/yr

Partner closely with Site Reliability Engineering (SRE) to architect systems that are highly available, resilient, performant, and fault-tolerant across multiple regions. * Leverage observability ...

Sr. Software Engineer

Stamford, CT

$130K - $172K/yr

Partner closely with Site Reliability Engineering (SRE) to architect systems that are highly available, resilient, performant, and fault-tolerant across multiple regions. * Leverage observability ...

DevOps Engineer (USA)

Stamford, CT · On-site

$56.25 - $77/hr

Requirements * 2+ years of experience in a DevOps, Infrastructure, or SRE role within a Linux-centric environment. * Strong hands-on experience with Kubernetes management and deployment strategies.

DevOps Engineer (USA)

Stamford, CT · On-site

$56.25 - $77/hr

Requirements * 2+ years of experience in a DevOps, Infrastructure, or SRE role within a Linux-centric environment. * Strong hands-on experience with Kubernetes management and deployment strategies.

New

Showing results 21-40

Site Reliability Engineer information

See Connecticut salary details

$10

$60

$87

How much do site reliability engineer jobs pay per hour?

As of Aug 21, 2026, the average hourly pay for site reliability engineer in Connecticut is $60.64, according to ZipRecruiter salary data. Most workers in this role earn between $52.12 and $69.28 per hour, depending on experience, location, and employer.

What is a site reliability engineer?

A site reliability engineer specializes in site reliability engineering, or SRE, a specific branch of operations first pioneered by Google. You are responsible for ensuring that when a website decides to scale a particular feature for various users to access, it does not break the underlying software or website functions. This means you need to use analytical problem-solving skills to determine how to make specific features on a new software release work on top of existing source code.

What is a site reliability engineer?

A Site Reliability Engineer (SRE) is a professional who applies software engineering principles to infrastructure and operations problems. Their primary goal is to create scalable and highly reliable software systems, often bridging the gap between development and IT operations. SREs automate tasks, monitor system health, respond to incidents, and work to improve system reliability and performance. They also help define service level objectives (SLOs) and ensure systems meet customer expectations for uptime and availability.

What are the key skills and qualifications needed to thrive as a site reliability engineer?

To thrive as a Site Reliability Engineer, you need a strong background in computer science, systems administration, and software engineering, often supported by a degree in a technical field. Familiarity with cloud platforms (like AWS or GCP), container orchestration (such as Kubernetes), infrastructure as code (Terraform or Ansible), and monitoring tools (Prometheus, Grafana) is typically expected. Strong problem-solving skills, effective communication, and a proactive mindset help SREs excel at incident management and cross-functional collaboration. These skills are crucial for maintaining system reliability, minimizing downtime, and driving continuous improvement in complex technical environments.

What are some of the most common challenges site reliability engineers face when balancing system reliability with rapid software delivery?

Site Reliability Engineers (SREs) often navigate the challenge of maintaining highly reliable systems while supporting fast-paced software releases. This involves managing incidents, automating processes to reduce manual toil, and working closely with development teams to embed reliability into the software development lifecycle. SREs must carefully prioritize their efforts between proactive improvements and urgent, reactive fire-fighting. Effective communication and collaboration with both operations and development teams are crucial to ensuring service uptime without slowing down innovation.

What is the difference between Site Reliability Engineer vs DevOps Engineer?

AspectSite Reliability EngineerDevOps Engineer
CredentialsTypically requires a computer science degree, certifications like AWS, Google Cloud, or KubernetesSimilar credentials, often with cloud certifications and scripting skills
Work EnvironmentFocuses on maintaining and improving system reliability, often in large-scale production environmentsWorks on automation, CI/CD pipelines, and deployment processes across development and operations teams
Industry UsageCommon in tech, cloud services, and large-scale enterprise companiesWidely used in software development, cloud, and IT organizations

Both roles require strong technical skills and cloud knowledge, but SREs focus more on system reliability and uptime, while DevOps engineers emphasize automation and deployment processes. They often collaborate but have distinct primary responsibilities.

Is a site reliability engineer a stressful job?

A site reliability engineer (SRE) role can be stressful due to the responsibility of maintaining system uptime, handling incidents, and ensuring reliability under tight deadlines. The job often requires strong problem-solving skills, familiarity with monitoring tools, and the ability to work in high-pressure situations, but it also offers opportunities for skill development and process improvements.

What are the most commonly searched types of Site Reliability Engineer jobs in Connecticut?

The most popular types of Site Reliability Engineer jobs in Connecticut are:

What are popular job titles related to Site Reliability Engineer jobs in Connecticut?

For Site Reliability Engineer jobs in Connecticut, the most frequently searched job titles are:

What job categories do people searching Site Reliability Engineer jobs in Connecticut look for?

The top searched job categories for Site Reliability Engineer jobs in Connecticut are:

What cities in Connecticut are hiring for Site Reliability Engineer jobs?

Cities in Connecticut with the most Site Reliability Engineer job openings:

What are popular job titles related to Site Reliability Engineer jobs in CT?

For Site Reliability Engineer jobs in CT, the most frequently searched job titles are:

Infographic showing various Site Reliability Engineer job openings in Connecticut as of August 2026, with employment types broken down into 1% As Needed, 80% Full Time, 13% Part Time, 2% Temporary, and 4% Contract. Highlights an 93% Physical, 3% Hybrid, and 4% Remote job distribution, with an average salary of $126,124 per year, or $60.6 per hour.

Principal Reliability Engineer - EDS

The Hartford Financial Services Group, Inc.

Hartford, CT • On-site, Remote

Full-time

Re-posted 20 days ago


The Hartford rating

8.8

Company rating: 8.8 out of 10

Based on 121 frontline employees who took The Breakroom Quiz

56th of 311 rated insurance


Job description

Principal Reliability Engineering - IE06JE
We're determined to make a difference and are proud to be an insurance company that goes well beyond coverages and policies. Working here means having every opportunity to achieve your goals - and to help others accomplish theirs, too. Join our team as we help shape the future.
The Enterprise Data Services (EDS) organization is seeking a Principal Reliability Engineer (Principal RE) to serve as the senior technical authority responsible for the reliability, resilience, availability, and performance of all data platforms, cloud infrastructure, data products, and data pipelines across the enterprise data organization. This role sets the strategic vision for Reliability Engineering within EDS and leads the definition, implementation, and continuous evolution of RE practices, tooling, automation, observability frameworks, and AIOps/AI-driven operations.
As the Principal RE, you will influence architectural direction, lead large-scale, cross-organizational technical initiatives, and drive a culture of engineering excellence, automation-first operations, and proactive reliability improvement. You will partner closely with platform engineering, data engineering, security, architecture, and product teams to embed RE principles into every stage of the data product lifecycle.
This role will have a Hybrid work schedule, with the expectation of working in an office (Columbus, OH, Chicago, IL, Hartford, CT or Charlotte, NC) 3 days a week (Tuesday through Thursday).
Key Responsibilities
Enterprise Reliability Strategy & Leadership
  • Work closely with the AVP, RE & Production Support, EDS defining the Reliability Engineering strategy for data platforms, data cloud environments, and data products.
  • Establish long-term RE roadmaps, target operating models, and architectural patterns that scale with organizational growth.
  • Serve as the highest-level technical escalation point for systemic reliability issues, influencing executive stakeholders and engineering leaders.

Platform & Cloud Reliability (AWS, GCP, Snowflake, EMR, Hadoop, ETL/ELT)
  • Leverage Enterprise provided standards and building blocks to Architect and evolve highly reliable, performant, and cost-efficient cloud-based platforms across AWS and GCP for all EDS services.
  • Influence and work directly with Platform Solution Architecture on new product enablement, hyper automation (end to end blueprint automation).
  • Oversee reliability controls and fail-safe patterns for Snowflake, EMR, Hadoop/Spark clusters, container platforms (e.g., Kubernetes), and mission-critical data systems.
  • Lead the creation and enforcement of SLO/SLI frameworks that span the entire data lifecycle.

AI-Enabled Operations, AIOps & Intelligent Automation
  • Develop and implement AI-driven automation for anomaly detection, alert correlation, autonomous remediation, and predictive capacity management.
  • Leverage LLMs, prompt engineering, and cloud-native AI services (AWS Bedrock, SageMaker, Vertex AI) to build intelligent runbooks, advanced troubleshooting agents, and generative-AI-enabled operational tooling.
  • Champion the adoption of machine learning-based observability and reliability analytics.

End-to-End Observability & Operational Excellence
  • Adopt and architect enterprise-wide data observability frameworks-including logging, metrics, tracing, distributed profiling, and event pipelines-for all data platforms and pipelines.
  • Establish gold-standard incident response patterns, post-incident reviews, and continuous improvement processes.
  • Drive elimination of toil across EDS, focusing on self-healing systems, proactive detection, and autonomous operations.

Data Pipeline & Data Product Reliability
  • Define RE best practices for modern data products, governed data pipelines, real-time/streaming systems, and operational analytics platforms.
  • Ensure data quality, data timeliness, and SLAs for data products through automated checks, lineage-informed alerting, and pipeline reliability tooling.
  • Partner with Data Engineering to embed resilience patterns (idempotency, checkpointing, replayability, disaster recovery) into pipeline architectures.

Engineering Standards, Governance & Cross-Org Influence
  • Set and enforce standards for IaC, CI/CD, platform automation, reliability frameworks, operational readiness, and runbook quality across EDS.
  • Provide technical leadership and mentorship to Staff/Senior Engineers in the RE team and Production Support teams, influencing engineering culture and helping grow RE capabilities across the organization.
  • Represent Reliability Engineering in architectural reviews, enterprise governance forums, and executive-level discussions.

Technical Experience
  • 10+ years in one or more of the following areas: data, cloud, platform engineering, site/reliability engineering, or large-scale distributed systems, with experience in leadership or technology leader roles.
  • Proficiency with data or cloud platforms, including architectural patterns for resilience, networking, security, and distributed data infrastructure.
  • Deep experience supporting or engineering platforms such as Snowflake, EMR, Hadoop/Spark, Data Integration, and cloud-native data ecosystems.
  • Scripting and programming (preferably Python) for large-scale automation, platform tooling, and reliability frameworks.
  • Experience with Infrastructure-as-Code (Terraform, CloudFormation) and enterprise CI/CD.

Preferred Qualifications
  • Experience in regulated or highly complex enterprise environments (financial services, insurance, healthcare).
  • Prior experience as a Senior Staff Engineer, Engineering or Architecture leader with hands on experience, or similar senior technical role.
  • Knowledge of data governance, metadata, lineage systems, and data quality engineering practices.
  • Certifications in AWS, GCP, Kubernetes, or SRE/DevOps frameworks.

AI & AIOps
  • Background applying machine learning to operations-anomaly detection, event correlation, predictive modeling, and automated remediation.
  • Understand of AI-enabled developer/operations tools using LLMs, prompt engineering, or cloud AI services for reliability improvements.

Observability & Platform Operations
  • Expertise with enterprise observability stacks (Prometheus, Grafana, Datadog, Splunk, Dynatrace, OpenTelemetry).
  • Ability to design and enforce advanced SLI/SLO frameworks across complex data ecosystems.

Leadership & Cross-Functional Influence
  • Demonstrated ability to lead technical strategy at scale, influence senior engineering leaders, and set enterprise-wide standards.
  • Strong capability in mentoring engineers, providing architectural guidance, and fostering engineering excellence.
  • Exceptional communication skills for interacting with executives, senior architects, product leaders, and engineering teams.

Candidate must be authorized to work in the US without company sponsorship. The company will not support the STEM OPT I-983 Training Plan endorsement for this position.
Compensation
The listed annualized base pay range is primarily based on analysis of similar positions in the external market. Actual base pay could vary and may be above or below the listed range based on factors including but not limited to performance, proficiency and demonstration of competencies required for the role. The base pay is just one component of The Hartford's total compensation package for employees. Other rewards may include short-term or annual bonuses, long-term incentives, and on-the-spot recognition. The annualized base pay range for this role is:
$152,800 - $229,200
Equal Opportunity Employer/Sex/Race/Color/Veterans/Disability/Sexual Orientation/Gender Identity or Expression/Religion/Age
About Us | Our Culture | What It's Like to Work Here | Perks & Benefits

What The Hartford employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Hartford logo

About Hartford

Sourced by ZipRecruiter

Hartford Financial Services Group, widely recognized as The Hartford, is a renowned company based in Hartford, CT, US. Established in 1810, it has evolved into an industry leader in the insurance and financial services sector, proudly serving more than one million businesses in the US. The Hartford is committed to offering a gamut of insurance products that include homeowners, automobile, and business insurance as well as employee benefits and mutual funds. The company’s core values revolve around customer-focused innovations, diversity and inclusion, and ethical dealings that have earned them a customer-centric reputation. This shapes their mission which revolves around aiding their clients to overcome unforeseen obstacles and enhancing their wealth over time. Among the company's noted accomplishments is being consistently listed among the World's Most Ethical Companies, a testament to their unwavering commitment towards responsible business practices.

Industry

Finance and insurance

Company size

10,000+ Employees

Headquarters location

Hartford, CT, US

Year founded

1810

Social media