1

Customer Reliability Engineer Jobs in Delaware (NOW HIRING)

... (SRE) practices across the software development lifecycle. Works closely with application ... customer experience health.Analyze production telemetry to proactively identify performance ...

New

We aim to help customers make smart financial decisions and stay on track, so they can be money ... The Job As Site Reliability Engineer, you will serve as a reliability subject matter expert who ...

Site Reliability Engineer

Wilmington, DE

$55.25 - $73.50/hr

We aim to help customers make smart financial decisions and stay on track, so they can be money ... The Job As Site Reliability Engineer, you will serve as a reliability subject matter expert who ...

Site Reliability Engineer

Wilmington, DE · On-site

$55.25 - $73.50/hr

We aim to help customers make smart financial decisions and stay on track, so they can be money ... The Job As Site Reliability Engineer, you will serve as a reliability subject matter expert who ...

We aim to help customers make smart financial decisions and stay on track, so they can be money ... The Job As Site Reliability Engineer, you will serve as a reliability subject matter expert who ...

Lead Site Reliability Engineer

Wilmington, DE · On-site

$55.25 - $73.50/hr

As a Lead Site Reliability Engineer at JPMorgan Chase within the Enterprise technology, Corporate ... with customers * Demonstrates a high level of technical expertise within one or more technical ...

Collaborating with SRE and infrastructure teams to enhance monitoring and tooling * Reporting on KPIs, system availability, and customer satisfaction metrics Qualifications What You Bring * 5+ years ...

Collaborating with SRE and infrastructure teams to enhance monitoring and tooling * Reporting on KPIs, system availability, and customer satisfaction metrics What You Bring * 5+ years in a leadership ...

Collaborating with SRE and infrastructure teams to enhance monitoring and tooling * Reporting on KPIs, system availability, and customer satisfaction metrics Qualifications What You Bring * 5+ years ...

next page

Showing results 1-20

Customer Reliability Engineer information

See Delaware salary details

$61.1K

$118.1K

$141.1K

How much do customer reliability engineer jobs pay per year?

As of Aug 8, 2026, the average yearly pay for customer reliability engineer in Delaware is $118,074.00, according to ZipRecruiter salary data. Most workers in this role earn between $102,600.00 and $129,100.00 per year, depending on experience, location, and employer.

How does a customer reliability engineer typically interact with clients and internal engineering teams?

Customer Reliability Engineers serve as a vital bridge between clients and internal technical teams. They regularly communicate with customers to understand their needs, troubleshoot issues, and provide technical guidance. Internally, they collaborate closely with product, support, and development teams to relay customer feedback, help prioritize reliability improvements, and ensure seamless incident resolution. This cross-functional role requires strong communication skills and the ability to translate technical information for different audiences, making every day varied and impactful.

What does a customer reliability engineer do?

A customer reliability engineer (CRE) works with clients to ensure the reliability, performance, and availability of products or services. They analyze system issues, develop solutions, and often collaborate with engineering teams to improve infrastructure and customer experience, typically using monitoring tools and technical expertise. CREs may also provide technical support and guidance to help clients optimize their use of the company's offerings.

What is the difference between Customer Reliability Engineer vs Site Reliability Engineer?

AspectCustomer Reliability EngineerSite Reliability Engineer
CredentialsTypically requires engineering degrees, certifications in cloud platforms (AWS, Azure), and knowledge of customer supportRequires engineering degrees, certifications in cloud and systems management, with a focus on infrastructure
Work EnvironmentCustomer-facing, involves direct interaction with clients to resolve issues and improve reliabilityPrimarily internal, focused on maintaining and improving system reliability and scalability
Employer & Industry UsageUsed by cloud service providers and tech companies with a customer support componentCommon in large tech companies managing large-scale infrastructure and services

The main difference is that Customer Reliability Engineers focus on ensuring customer satisfaction and resolving client-specific issues, while Site Reliability Engineers concentrate on internal system stability and scalability. Both roles require technical expertise and cloud knowledge but serve different operational needs.

What is a customer reliability engineer?

A Customer Reliability Engineer (CRE) is a technical professional who works closely with customers to ensure the reliability, performance, and uptime of software products and services. CREs act as a bridge between customers and engineering teams, helping to identify, troubleshoot, and resolve reliability issues. They often collaborate with multiple departments to implement best practices, monitor systems, and proactively address potential problems, ultimately aiming to improve the overall customer experience.

What skills and qualifications are needed to thrive as a customer reliability engineer?

To thrive as a Customer Reliability Engineer, you need a solid background in systems engineering, incident management, and troubleshooting, often supported by a degree in computer science or related field. Familiarity with cloud platforms (such as AWS or GCP), monitoring tools (like Datadog or Prometheus), and automation scripts is typically required. Exceptional communication, problem-solving abilities, and a customer-centric mindset are vital soft skills for this role. These skills ensure efficient incident resolution, strong client relationships, and reliable system performance under pressure.
What are popular job titles related to Customer Reliability Engineer jobs in Delaware? For Customer Reliability Engineer jobs in Delaware, the most frequently searched job titles are:
What job categories do people searching Customer Reliability Engineer jobs in Delaware look for? The top searched job categories for Customer Reliability Engineer jobs in Delaware are:
What cities in Delaware are hiring for Customer Reliability Engineer jobs? Cities in Delaware with the most Customer Reliability Engineer job openings:
Infographic showing various Customer Reliability Engineer job openings in Delaware as of June 2026, with employment types broken down into 81% Full Time, 10% Part Time, 3% Temporary, and 6% Contract. Highlights an 93% Physical, 2% Hybrid, and 5% Remote job distribution, with an average salary of $118,074 per year, or $56.8 per hour.

Lead Site Reliability Engineer

Luxoft

Wilmington, DE • On-site

$140 - $190/hr

Other

Posted 3 days ago

New


Job description

Project description

Responsible at the expert level for ensuring the reliability, scalability, performance, and operational excellence of critical banking platforms and applications. Serves as a senior individual contributor responsible for designing, implementing, and improving Site Reliability Engineering (SRE) practices across the software development lifecycle. Works closely with application development, infrastructure, platform engineering, and business teams to enhance system resiliency through automation, observability, testing, and proactive operational management while coaching and influencing others.

Responsibilities
  • Design, implement, and support highly available, scalable, and resilient applications and cloud infrastructure following enterprise technology standards and SRE best practices.Lead initiatives to improve system reliability, availability, performance, and operational maturity through automation and engineering excellence.Define, implement, and monitor Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets for critical business services.Develop comprehensive observability strategies leveraging Dynatrace, OpenTelemetry (OTel), distributed tracing, metrics, logging, dashboards, and alerting solutions.Design and maintain end-to-end monitoring solutions that provide actionable insights into application, infrastructure, and customer experience health.Analyze production telemetry to proactively identify performance bottlenecks, reliability risks, and capacity constraints.Lead incident response activities for high-severity production events, coordinating cross-functional teams to restore services and minimize customer impact.Perform and facilitate Root Cause Analysis (RCA) activities, ensuring corrective and preventive actions are identified, prioritized, and implemented.Drive operational excellence through automation of repetitive tasks, operational workflows, deployments, recovery procedures, and reliability controls.Partner with development teams to build reliable and observable services throughout the Software Development Lifecycle (SDLC).Design, develop, and execute automated regression testing strategies to validate application stability, reliability, and performance following deployments and infrastructure changes.Review test coverage and reliability validation approaches to ensure comprehensive testing and risk mitigation.Create, maintain, and improve Infrastructure as Code (IaC) solutions using Terraform for cloud infrastructure provisioning, configuration management, and environment standardization.Support and optimize Microsoft Azure environments, including Azure App Services, resource management, scaling strategies, deployment automation, and application lifecycle management.Utilize Azure-native tools such as Azure Monitor, Application Insights, Log Analytics, and related services to improve platform visibility and reliability.Drive implementation of performance testing, resiliency testing, fault tolerance validation, and disaster recovery preparedness within assigned domains.Establish operational readiness standards and ensure applications meet reliability, scalability, observability, and supportability requirements before production deployment.Review architectural designs and provide recommendations to improve platform resiliency, operational efficiency, and cloud optimization.Lead capacity planning, performance tuning, and workload optimization efforts across production environments.Develop and maintain operational runbooks, incident playbooks, knowledge articles, and standard operating procedures.Serve as a key partner with engineering, infrastructure, cybersecurity, architecture, and support teams to identify and implement continuous process improvements spanning organizational boundaries.Communicate system health, reliability trends, operational risks, and remediation strategies to technical and business stakeholders.Present reliability initiatives, operational metrics, and engineering recommendations at architecture reviews, technical forums, and leadership meetings.Mentor engineers on observability, cloud engineering, automation, SRE principles, and operational best practices.Understand and adhere to the Company's risk and regulatory standards, policies, and controls in accordance with the Company's Risk Appetite.Identify reliability, operational, and technology risks requiring escalation to management.Promote an environment that supports a culture of belonging and reflects the Client brand.Maintain Client internal control standards, including timely implementation of internal and external audit findings and regulatory requirements as applicable.Complete other related duties as assigned.
SKILLS Must have
  • Strong experience in observability and monitoring, including hands-on expertise with:DynatraceOpenTelemetry (OTel)Distributed tracingMetrics collection and analysisCentralized logging and log aggregationAlerting and dashboard developmentProven experience designing and executing automated regression testing frameworks and test suites to ensure application and platform stability following deployments.Strong proficiency in Infrastructure as Code (IaC) using Terraform.Experience with CI/CD pipelines, deployment automation, and operational tooling.Expert knowledge of production systems monitoring, incident management, and operational troubleshooting.Strong understanding of application performance management, distributed systems, and modern cloud-native architectures.Cloud & Platform ExpertiseStrong experience with Microsoft Azure, including:Azure App ServicesResource GroupsAzure networking conceptsScaling and performance optimizationDeployment and release managementApplication lifecycle managementExperience leveraging Azure-native operational tooling such as:Azure MonitorApplication InsightsLog AnalyticsAzure dashboards and alertingExperience supporting cloud-native and hybrid infrastructure environments.Reliability & Engineering PracticesDemonstrated experience implementing and operating SRE practices, including:Service Level Objectives (SLOs)Service Level Indicators (SLIs)Error budgetsIncident managementProblem managementRoot Cause Analysis (RCA)Reliability automationAbility to improve system reliability through:Performance tuningCapacity planningObservability-driven insightsProactive issue detectionReliability engineering initiativesExperience developing automated recovery mechanisms and self-healing solutions.Knowledge of resiliency engineering patterns, disaster recovery planning, and high-availability architectures.
Nice to have

Experience supporting large-scale enterprise applications in regulated environments.Strong analytical and troubleshooting skills related to production systems and distributed architectures.Experience working in Agile and DevOps operating models.Ability to work autonomously and lead complex reliability initiatives.Strong organizational and time management skills.Advanced verbal and written communication skills.Experience driving project milestones and delivery commitments.Proven experience leading major incident response and post-incident improvement efforts.Experience partnering with architecture, infrastructure, cybersecurity, and application development teams.Experience with scripting and automation using PowerShell, Python, Bash, or similar technologies.Industry certifications in Azure, Terraform, Cloud Engineering, or Site Reliability Engineering preferred.

#J-18808-Ljbffr