π½ Job Title: Senior Site Reliability Engineer (SRE)
π Location: Seattle, WA (Highly Preferred Onsite)
πΌ Experience: 10+ Years
Job Description
We are seeking a highly experienced Senior Site Reliability Engineer (SRE) with strong expertise in reliability engineering, production support, monitoring, and cloud infrastructure. The ideal candidate will have hands-on experience with Dynatrace and prior experience in the airline industry. This role focuses on ensuring high availability, performance, scalability, and operational excellence of mission-critical applications.
Key Responsibilities
- Design, implement, and maintain highly reliable and scalable production systems.
- Monitor application and infrastructure performance using Dynatrace.
- Identify, troubleshoot, and resolve production issues with minimal downtime.
- Automate operational processes to improve system reliability and efficiency.
- Collaborate with development, infrastructure, and DevOps teams to improve system performance.
- Perform root cause analysis (RCA) and implement preventive measures.
- Develop monitoring dashboards, alerts, and performance reports.
- Participate in on-call support and incident management activities.
- Drive continuous improvements in system reliability, availability, and observability.
Required Skills
- 10+ years of experience in Site Reliability Engineering, DevOps, or Production Support.
- Strong hands-on experience with Site Reliability Engineering (SRE) principles and best practices.
- Extensive experience with Dynatrace for application performance monitoring and observability.
- Experience supporting large-scale enterprise applications in production.
- Knowledge of cloud platforms such as AWS, Azure, or Google Cloud Platform.
- Experience with Linux/Unix administration and scripting (Python, Shell, or PowerShell).
- Familiarity with CI/CD pipelines, Git, Docker, and Kubernetes.
- Strong incident management, troubleshooting, and root cause analysis skills.
- Excellent communication and collaboration skills.
Preferred Qualifications
- Prior experience in the Airline Industry (highly preferred).
- Experience with Kubernetes, Terraform, and Infrastructure as Code (IaC).
- Experience with monitoring tools such as Prometheus, Grafana, or Splunk is a plus.
- AWS, Azure, Kubernetes, or Dynatrace certifications are a plus.
Employment Type: Contract
Location Preference: Seattle, WA (Onsite highly preferred).