Wells Fargo is seeking a Principal Engineer to lead observability and reliability engineering for Marketing Technology platforms. This role focuses on enabling end-to-end visibility across a complex ...
Wells Fargo is seeking a Principal Engineer to lead observability and reliability engineering for Marketing Technology platforms. This role focuses on enabling end-to-end visibility across a complex ...
Manager, DevOps & Cloud Platform
Plano, TX · On-site
$161K - $220K/yr
This role combines technical depth with people leadership and drives platform reliability, observability, developer experience, and DevOps maturity across the organization. ' Duties and ...
Manager, DevOps & Cloud Platform
Plano, TX · On-site
$161K - $220K/yr
This role combines technical depth with people leadership and drives platform reliability, observability, developer experience, and DevOps maturity across the organization. ' Duties and ...
Manager, DevOps & Cloud Platform
$161K - $220K/yr
This position combines technical expertise with people leadership to advance platform reliability, observability, developer experience, and DevOps maturity across the organization. Duties and ...
Manager, DevOps & Cloud Platform
$161K - $220K/yr
This position combines technical expertise with people leadership to advance platform reliability, observability, developer experience, and DevOps maturity across the organization. Duties and ...
Manager, DevOps & Cloud Platform
Plano, TX · On-site
$161K - $220K/yr
This role combines technical depth with people leadership and drives platform reliability, observability, developer experience, and DevOps maturity across the organization. ' Duties and ...
Manager, DevOps & Cloud Platform
Plano, TX · On-site
$161K - $220K/yr
This role combines technical depth with people leadership and drives platform reliability, observability, developer experience, and DevOps maturity across the organization. ' Duties and ...
Manager, DevOps & Cloud Platform
Plano, TX · On-site
$161K - $220K/yr
This position combines technical expertise with people leadership to advance platform reliability, observability, developer experience, and DevOps maturity across the organization. Duties and ...
Manager, DevOps & Cloud Platform
Plano, TX · On-site
$161K - $220K/yr
This position combines technical expertise with people leadership to advance platform reliability, observability, developer experience, and DevOps maturity across the organization. Duties and ...
Platform Engineer I
Coppell, TX · On-site
Observability & Reliability Engineering * Design alerts that detect customer-impacting issues early while minimising alert fatigue. * Improve platform visibility through metrics, logs, traces, and ...
Platform Engineer I
Coppell, TX · On-site
Observability & Reliability Engineering * Design alerts that detect customer-impacting issues early while minimising alert fatigue. * Improve platform visibility through metrics, logs, traces, and ...
Platform Engineer I
Coppell, TX · On-site
Observability & Reliability Engineering * Design alerts that detect customer-impacting issues early while minimising alert fatigue. * Improve platform visibility through metrics, logs, traces, and ...
Platform Engineer I
Coppell, TX · On-site
Observability & Reliability Engineering * Design alerts that detect customer-impacting issues early while minimising alert fatigue. * Improve platform visibility through metrics, logs, traces, and ...
Senior Tech Lead, Network Automation Engineer
$102K - $135K/yr
Position Overview Our Observability, Software (SW) Development & Automation Enablement department is in search of a Senior Tech Lead, Network Automation Engineer with excellent time management skills ...
Senior Tech Lead, Network Automation Engineer
$102K - $135K/yr
Position Overview Our Observability, Software (SW) Development & Automation Enablement department is in search of a Senior Tech Lead, Network Automation Engineer with excellent time management skills ...
Your work will directly support enterprise-wide adoption of our observability mandate and ... Chaos engineering tools (e.g., Gremlin, AWS FIS) * AWS and Azure cloud environments
Your work will directly support enterprise-wide adoption of our observability mandate and ... Chaos engineering tools (e.g., Gremlin, AWS FIS) * AWS and Azure cloud environments
Senior SRE (Site Reliability Engineer)
Prosper, TX · On-site
$75/hr
Senior SRE Engineer Location: Washington DC - Hybrid We are seeking a high-caliber Senior SRE ... Lead the design, governance, and rollout of Dynatrace observability for distributed microservices ...
Quick apply
Senior SRE (Site Reliability Engineer)
Prosper, TX · On-site
$75/hr
Senior SRE Engineer Location: Washington DC - Hybrid We are seeking a high-caliber Senior SRE ... Lead the design, governance, and rollout of Dynatrace observability for distributed microservices ...
Principal Resiliency Engineer
Coppell, TX · On-site
Your work will directly support enterprise-wide adoption of our observability mandate and ... Chaos engineering tools (e.g., Gremlin, AWS FIS) * AWS and Azure cloud environments
Principal Resiliency Engineer
Coppell, TX · On-site
Your work will directly support enterprise-wide adoption of our observability mandate and ... Chaos engineering tools (e.g., Gremlin, AWS FIS) * AWS and Azure cloud environments
Senior Tech Lead, Network Automation Engineer
Dallas, TX · On-site
$102K - $141K/yr
Position Overview Our Observability, Software (SW) Development & Automation Enablement department is in search of a Senior Tech Lead, Network Automation Engineer with excellent time management skills ...
Senior Tech Lead, Network Automation Engineer
Dallas, TX · On-site
$102K - $141K/yr
Position Overview Our Observability, Software (SW) Development & Automation Enablement department is in search of a Senior Tech Lead, Network Automation Engineer with excellent time management skills ...
Senior Go Engineer
Plano, TX · On-site
$60 - $79/hr
Minimum of 6 years of programming experience. * A minimum of 2 years of back-end programming and ... Experience with any of the following: continuous integration pipelines, observability, monitoring ...
Quick apply
Senior Go Engineer
Plano, TX · On-site
$60 - $79/hr
Minimum of 6 years of programming experience. * A minimum of 2 years of back-end programming and ... Experience with any of the following: continuous integration pipelines, observability, monitoring ...
Principal Resiliency Engineer
Coppell, TX · On-site
Your work will directly support enterprise-wide adoption of our observability mandate and ... Chaos engineering tools (e.g., Gremlin, AWS FIS) * AWS and Azure cloud environments
Principal Resiliency Engineer
Coppell, TX · On-site
Your work will directly support enterprise-wide adoption of our observability mandate and ... Chaos engineering tools (e.g., Gremlin, AWS FIS) * AWS and Azure cloud environments
Senior Go Engineer
Plano, TX · On-site
$110 - $150/hr
Minimum of 6 years of programming experience. * A minimum of 2 years of back-end programming and ... Experience with any of the following: continuous integration pipelines, observability, monitoring ...
Senior Go Engineer
Plano, TX · On-site
$110 - $150/hr
Minimum of 6 years of programming experience. * A minimum of 2 years of back-end programming and ... Experience with any of the following: continuous integration pipelines, observability, monitoring ...
Devops Engineer
Dallas, TX · Remote
$54 - $74/hr
Position: DevOps Consultant Location: Remote Duration: Long term contract Type: Only W2 About the ... Observability & Operational Readiness * Set up monitoring, alerting, and dashboarding tools for ...
Quick apply
Devops Engineer
Dallas, TX · Remote
$54 - $74/hr
Position: DevOps Consultant Location: Remote Duration: Long term contract Type: Only W2 About the ... Observability & Operational Readiness * Set up monitoring, alerting, and dashboarding tools for ...
Define and enhance enterprise observability standards through metrics, logging, tracing, alerting, and service health monitoring * Partner closely with engineering, infrastructure, application, cloud ...
Define and enhance enterprise observability standards through metrics, logging, tracing, alerting, and service health monitoring * Partner closely with engineering, infrastructure, application, cloud ...
Senior Platform Engineer
Dallas, TX · On-site
$107K - $146K/yr
As a Senior Platform Engineer, you will design and evolve the Internal Developer Platform, enabling ... configure observability without operational bottlenecks. • Design, implement, and maintain ...
Senior Platform Engineer
Dallas, TX · On-site
$107K - $146K/yr
As a Senior Platform Engineer, you will design and evolve the Internal Developer Platform, enabling ... configure observability without operational bottlenecks. • Design, implement, and maintain ...
As a Senior SRE, you should feel exceptionally comfortable bringing architectural design proposals ... Observability is paramount. If we can't measure it, we can't prove it works; if we can't prove it ...
As a Senior SRE, you should feel exceptionally comfortable bringing architectural design proposals ... Observability is paramount. If we can't measure it, we can't prove it works; if we can't prove it ...
Devops Engineer
Irving, TX · On-site
$125K - $140K/yr
... Engineer with deep, hands-on expertise in building and operating enterprise-grade CI/CD platforms ... Observability & Capacity Planning Design and maintain a unified observability stack covering ...
New
Devops Engineer
Irving, TX · On-site
$125K - $140K/yr
... Engineer with deep, hands-on expertise in building and operating enterprise-grade CI/CD platforms ... Observability & Capacity Planning Design and maintain a unified observability stack covering ...
New
Observability Engineer information
See Dallas, TX salary details
$22.06 - $29.65
2% of jobs
$29.65 - $37.24
4% of jobs
$37.24 - $44.83
6% of jobs
$44.83 - $52.42
8% of jobs
$53.72 is the 25th percentile. Wages below this are outliers.
$52.42 - $60.01
23% of jobs
The median wage is $63.81 / hr.
$60.01 - $67.60
12% of jobs
$72.22 is the 75th percentile. Wages above this are outliers.
$67.60 - $75.19
32% of jobs
$75.19 - $82.79
11% of jobs
$82.79 - $90.38
1% of jobs
$90.38 - $97.97
0% of jobs
$97.97 - $105.56
1% of jobs
$22
$63
$105
How much do observability engineer jobs pay per hour?
What does an observability engineer do?
An Observability Engineer is responsible for designing, implementing, and maintaining monitoring, logging, and tracing systems to ensure the health, performance, and reliability of applications and infrastructure. They work with tools like Prometheus, Grafana, OpenTelemetry, and ELK to collect and analyze telemetry data. Their goal is to provide visibility into system behavior, detect and diagnose issues quickly, and improve overall system observability. They collaborate with developers, SREs, and operational teams to create automated and scalable observability solutions.
What are the key skills and qualifications needed to thrive as an observability engineer?
To thrive as an Observability Engineer, you need a solid understanding of monitoring, logging, and tracing systems, as well as expertise in programming, cloud infrastructure, and incident response. Proficiency in tools like Prometheus, Grafana, ELK stack, and familiarity with cloud platforms such as AWS, Azure, or GCP are commonly required, and certifications like AWS Certified DevOps Engineer can be advantageous. Strong analytical thinking, collaborative skills, and effective communication are essential soft skills for diagnosing issues and working across development and operations teams. These competencies are vital for proactively maintaining system reliability, ensuring performance, and resolving complications before they impact business operations.

Full-time
Medical, Life, Retirement, PTO
This job post has expired today. Applications are no longer accepted.
Wells Fargo rating
7.8
Based on 707 frontline employees who took The Breakroom Quiz
90th of 170 rated banks
Job description
About this role:
Wells Fargo is seeking a Principal Engineer to lead observability and reliability engineering for Marketing Technology platforms. This role focuses on enabling end-to-end visibility across a complex data ecosystem spanning SaaS platforms, hybrid cloud, and on-prem systems.
This position is responsible for designing and advancing full-stack observability for ETL pipelines, integrations, and data movement across marketing platforms. The goal is to improve reliability, reduce time to detect and resolve issues, and ensure consistent data flow supporting critical customer engagement channels.
You will join a focused team of engineers and application support professionals driving adoption of observability, automation, and Site Reliability Engineering (SRE) practices across Marketing Technology. The team partners closely with data engineering, platform teams, and vendors to deliver stable, scalable, and highly visible systems.
The role supports all aspects of the platform lifecycle-from instrumentation and monitoring design to incident response and continuous improvement-ensuring strong operational health, proactive alerting, and resilient data pipelines across distributed systems.
The team operates across a global footprint, enabling a follow-the-sun support model and consistent platform reliability.
In this role, you will:
- Define observability strategy
- Establish standards, patterns, and designs for monitoring distributed data systems
- Drive adoption of tools such as Grafana, Splunk, Prometheus, and OpenTelemetry
- Enable end-to-end visibility
- Instrument ETL pipelines, integrations, and APIs with metrics, logs, and traces
- Build dashboards for data flow health, latency, throughput, and error tracking
- Monitor data movement across SaaS, hybrid, and on-prem environments
- Strengthen data reliability
- Implement proactive alerting for pipeline failures, latency spikes, and data integrity risks
- Identify and isolate issues across multi-step data hops and integrations
- Partner with data engineering teams to embed observability into ETL frameworks (Airflow, Informatica)
- Lead incident response and problem resolution
- Drive triage for high-severity data and platform incidents
- Distinguish root causes across data, platform, and vendor layers
- Perform root cause analysis and implement long-term fixes
- Develop runbooks and playbooks for repeatable issue resolution
- Manage change and vendor coordination
- Ensure observability readiness for vendor patches, hotfixes, and configuration changes
- Validate monitoring coverage post-change to maintain visibility and alert accuracy
- Participate in change advisory processes and track vendor release cycles
- Document change impacts and lessons learned for audit and compliance
- Drive automation and continuous improvement
- Automate observability setup for new pipelines and integrations
- Implement anomaly detection and predictive alerting
- Continuously refine dashboards, thresholds, and alert logic based on trends
- Influence architecture and engineering practices
- Partner with Enterprise Architecture and engineering teams to align solutions with standards
- Advocate for scalable, observable system design across Marketing Technology
- Contribute to technology strategy and roadmap decisions
- Support operational excellence
- Share responsibility for production support of critical data flows and applications
- Improve key metrics such as availability, time to detect, and time to recover
- Promote SRE practices including SLIs, SLOs, and error budgets
Required Qualifications:
- 7+ years of Engineering experience, or equivalent demonstrated through one or a combination of the following: work experience, training, military experience, education
- 5+ years designing and implementing observability solutions (e.g., Splunk, Grafana, Prometheus, OpenTelemetry)
- 3+ years of experience with cloud and container platforms (Kubernetes, OpenShift)
Desired Qualifications:
- Experience with databases (Oracle, SQL Server, PostgreSQL, MongoDB) and strong SQL skills
- Familiarity with Java-based applications and performance monitoring (JVM metrics, heap/thread analysis)
- Experience with CI/CD pipelines (Jenkins, GitLab, Harness)
- Strong understanding of APIs, microservices, event-driven architecture, and messaging systems (Kafka, MQ)
- Ability to troubleshoot data movement across SaaS, hybrid, and on-prem systems
- Strong communication skills with the ability to work across technical and business teams
- Strong experience with distributed systems monitoring, logging, and tracing
- Experience supporting complex ETL pipelines and data integration workflows
- Proficiency in scripting and automation (Python, Bash, Ansible)
- Experience with Linux/Unix and Windows server environments
- Experience supporting marketing or SaaS platforms (Adobe, Salesforce, Pega, etc.)
- Knowledge of data governance and compliance frameworks
- Experience troubleshooting OS-level and infrastructure-related issues
- Experience supporting both commercial off-the-shelf and custom-built applications
- Strong analytical mindset with a focus on continuous improvement
- Experience working in Agile or Kanban environments
Job Expectations:
- Ability to work onsite at posted location
- Relocation assistance is not available for this position
- Visa sponsorship is not available for this position
Pay Range
Reflected is the base pay range offered for this position. Pay may vary depending on factors including but not limited to demonstrated examples of prior performance, skills, experience, or work location. Employees may also be eligible for incentive opportunities.
$159,000.00 - $305,000.00Benefits
Wells Fargo provides eligible employees with a comprehensive set of benefits, many of which are listed below. VisitBenefits - Wells Fargo Jobs for an overview of the following benefit plans and programs offered to employees.
- Health benefits
- 401(k) Plan
- Paid time off
- Disability benefits
- Life insurance, critical illness insurance, and accident insurance
- Parental leave
- Critical caregiving leave
- Discounts and savings
- Commuter benefits
- Tuition reimbursement
- Scholarships for dependent children
- Adoption reimbursement
Posting End Date:
2 Aug 2026*Job posting may come down early due to volume of applicants.
We Value Equal Opportunity
Wells Fargo is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, status as a protected veteran, or any other legally protected characteristic.
Employees support our focus on building strong customer relationships balanced with a strong risk mitigating and compliance-driven culture which firmly establishes those disciplines as critical to the success of our customers and company. They are accountable for execution of all applicable risk programs (Credit, Market, Financial Crimes, Operational, Regulatory Compliance), which includes effectively following and adhering to applicable Wells Fargo policies and procedures, appropriately fulfilling risk and compliance obligations, timely and effective escalation and remediation of issues, and making sound risk decisions. There is emphasis on proactive monitoring, governance, risk identification and escalation, as well as making sound risk decisions commensurate with the business unit's risk appetite and all risk and compliance program requirements.
Applicants with Disabilities
To request a medical accommodation during the application or interview process, visitDisability Inclusion at Wells Fargo.
Drug and Alcohol Policy
Wells Fargo maintains a drug free workplace. Please see our Drug and Alcohol Policy to learn more.
Wells Fargo Recruitment and Hiring Requirements:
a. Third-Party recordings are prohibited unless authorized by Wells Fargo.
b. Wells Fargo requires you to directly represent your own experiences during the recruiting and hiring process.
What Wells Fargo employees say
Pay
Benefits
Hours and flexibility
Workplace
Get the full story on Breakroom
About Wells Fargo
Sourced by ZipRecruiter
Wells Fargo & Company (NYSE: WFC) is a leading financial services company that has approximately $1.9 trillion in assets, proudly serves one in three U.S. households and more than 10% of small businesses in the U.S., and is a leading middle market banking provider in the U.S. We provide a diversified set of banking, investment and mortgage products and services, as well as consumer and commercial finance, through our four reportable operating segments: Consumer Banking and Lending, Commercial Banking, Corporate and Investment Banking, and Wealth & Investment Management. Wells Fargo ranked No. 41 on Fortune's 2022 rankings of America's largest corporations. In the communities we serve, the company focuses its social impact on building a sustainable, inclusive future for all by supporting housing affordability, small business growth, financial health and a low-carbon economy.
Industry
Finance and insurance
Company size
10,000+ Employees
Headquarters location
San Francisco, CA, US
Year founded
1852