1

Sre Platform Engineer Jobs in Alabama (NOW HIRING)

Site Reliability Engineer

Birmingham, AL · On-site

$53.50 - $71/hr

... platforms fast, secure, and always online You'll work side-by-side with developers, product teams ... in a Site Reliability or similar engineering role. * Deep knowledge of Azure cloud services.

Site Reliability Engineer II

Birmingham, AL · On-site

$53.50 - $71/hr

This role relies on monitoring platforms and is continually taking a holistic view of system health and performance. The SRE will enhance and support cloud-based transformations and is focused on ...

Cloud SRE Intern

Birmingham, AL · On-site

$53.50 - $71/hr

Familiarity with cloud platforms and technologies (Google Cloud preferred) * Familiarity with web technology, DevOps monitoring tools * Familiarity of DevOps practices (e.g. CI/CD pipelines ...

The Site Reliability Center (SRC) is focused on establishing a culture of operational excellence by ... platforms, and applications adhere to SRC onboarding standards that improve reliability, enable ...

New

next page

Showing results 1-20

Sre Platform Engineer information

What cities in Alabama are hiring for Sre Platform Engineer jobs?

Cities in Alabama with the most Sre Platform Engineer job openings:

Site Reliability Engineer III

Birmingham, AL • On-site

GPC - Genuine Parts Company
Retail • 10K+ employees

$53.50 - $71/hr

Full-time

Medical, Retirement, PTO

Re-posted 17 days ago


Key responsibilities

  • Gathers and analyzes metrics from monitoring platforms to assist in performance tuning and fault tolerance.

  • Partners with development teams to improve services through testing and release procedures.

  • Works closely with the incident response team and restoring service to normal operation.


Genuine Parts Company rating

7.3

Company rating: 7.3 out of 10

Based on 63 frontline employees who took The Breakroom Quiz


Job description

Site Reliability Engineer III

SUMMARY:

Under limited supervision, the Site Reliability Engineer III is responsible for improving system reliability and resilience. This role focuses on building automation to reduce manual effort and prevent service-impacting incidents. The SRE combines software and systems engineering to build and support large-scale, distributed, fault-tolerant systems. This role ensures that critical platforms are available, reliable, and able to support a fast rate of improvement. This role relies on monitoring platforms and is continually taking a holistic view of system health and performance. The SRE will enhance and support cloud-based transformations and is focused on pushing capabilities forward, staying ahead of customer needs, and innovating for continuous improvement. The SRE provides operational support and engineering for multiple large-scale distributed software applications.

You must be eligible to work in the US without Visa Sponsorship.

JOB DUTIES

  • Gathers and analyzes metrics from monitoring platforms to assist in performance tuning and fault tolerance.
  • Partners with development teams to improve services through testing and release procedures.
  • Participates in system design, platform management and capacity planning.
  • Balances feature development speed and reliability with service-level objectives.
  • Works closely with the incident response team and restoring service to normal operation.
  • Understands debugging and applying troubleshooting skills.
  • Investigates, blocks and rate-limits unwanted traffic.
  • Utilizes monitoring systems and dashboards for proactive changes and alerting.
  • Establishes continuous process improvement cycles where the process, performance, and supporting technologies are reviewed and enhanced where applicable.
  • Performs other duties as assigned.

EDUCATION & EXPERIENCE

Typically requires a bachelor's degree and five (5) or more years of related experience or an equivalent combination.

KNOWLEDGE, SKILLS, ABILITIES

  • Understanding of Kubernetes, containers, clusters, and elastic scalability.
  • Expertise in SRE principles.
  • Mindset of continually finding ways to drive scalability, stability, and performance.
  • Cloud Services experience with Google Cloud Platform (GCP).
  • Experience with API, service-based or microservice-based architecture.
  • Proficiency in infrastructure, network, database, operating systems, or security troubleshooting and remediation.
  • Architecture-level knowledge of Windows and Linux and Infrastructure systems.
  • Experience with production deployment, monitoring, and operational support for enterprise-class applications (Dynatrace a plus).
  • Experience working with Continuous Integration/ Continuous Deployment tools.
  • Experience in performance diagnostics, capacity planning, performance architecture design, performance tuning, and performance monitoring.
  • A strong mix of software engineering and operational support skills.
  • Knowledge of web technologies - HTTP, proxy, java, etc.
  • Experience with Azure DevOps (ADO), Dynatrace, Prometheus, Terraform and Grafana.

PHYSICAL DEMANDS:

LICENSES & CERTIFICATIONS:

SUPERVISORY RESPONSIBILITY: No Supervisory Responsibility

BUDGET RESPONSIBILITY: No

COMPANY INFORMATION: Motion offers an excellent benefits package which includes options for healthcare coverage, 401(k), tuition reimbursement, vacation, sick, and holiday pay.

DISCLAIMER: This job description illustrates the general nature and level of work performed by employees within this job classification. It is not intended to contain or be interpreted as a comprehensive inventory of all duties, responsibilities and skills required. Management retains the right to add or modify duties at any time.

Not the right fit? Let us know you're interested in a future opportunity by joining our Talent Community on jobs.genpt.com or create an account to set up email alerts as new job postings become available that meet your interest!

GPC conducts its business without regard to sex, race, creed, color, religion, marital status, national origin, citizenship status, age, pregnancy, sexual orientation, gender identity or expression, genetic information, disability, military status, status as a veteran, or any other protected characteristic. GPC's policy is to recruit, hire, train, promote, assign, transfer and terminate employees based on their own ability, achievement, experience and conduct and other legitimate business reasons.


What Genuine Parts Company employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom