1

Software Engineer Site Reliability Engineer Jobs in Hernando, MS

... , site reliability engineering, quality engineering, or software development life cycle (SDLC) delivery roles * 6+ years of experience developing applications, services, or automation using one or ...

The Reliability Engineering Manager role is based in Memphis, TN (onsite) . Be part of a ... Define, implement, and sustain site reliability strategies , including preventive and predictive ...

ON-SITE LOCATION: Memphis, TN HOW YOU WILL MAKE AN IMPACT: Varsity Spirit supports a portfolio of ... Quality and reliability of code delivered (defect rates, rework) * Timely completion of assigned ...

ON-SITE LOCATION: Memphis, TN HOW YOU WILL MAKE AN IMPACT: Varsity Spirit supports a portfolio of ... Quality and reliability of code delivered (defect rates, rework) * Timely completion of assigned ...

The Sr. Software Engineer will be responsible for designing and engineering software solutions for ... Web site appropriate use and privacy policies. • Set and enforce compatibility and ...

Senior Software Engineer

Cordova, TN · On-site

$107K - $141K/yr

The Sr. Software Engineer designs and builds full stack web applications using modern software ... reliability analysis, scalability analysis, performance analysis. * Analyze existing code to ...

Software Engineer Full-time job, 40 hours per week Pay/Salary: $135,304.00 year. Number of Openings: 5 Location: MI Softech Inc, 71 Peyton Parkway, Ste 103, Collierville, TN 38017 Website: Posting ...

Software Engineer Full-time job, 40 hours per week Pay/Salary: $135,304.00 year. Number of Openings: 5 Location: MI Softech Inc, 71 Peyton Parkway, Ste 103, Collierville, TN 38017 Website: Posting ...

Software Engineer Full-time job, 40 hours per week Pay/Salary: $135,304.00 year. Number of Openings: 5 Location: MI Softech Inc, 71 Peyton Parkway, Ste 103, Collierville, TN 38017 Website: Posting ...

... reliability required for frontier AI training and inference. We partner closely with datacenter ... Build and operate the software stacks that make site operations scalable, auditable, and fast ...

Lead Kafka Engineer

Memphis, TN · Hybrid

$99K - $131K/yr

Familiar with SRE concepts which includes evaluating and implementing monitoring and observability tools like Splunk, Data Dog, Dynatrace, CloudWatch and other job, log or dashboard concepts for ...

About Software Engineering Roles at Danaher Are you passionate about building real-world applications, writing clean code, and solving meaningful technical challenges? As a Software Engineering ...

About Software Engineering Roles at Danaher Are you passionate about building real-world applications, writing clean code, and solving meaningful technical challenges? As a Software Engineering ...

Showing results 21-40

Software Engineer Site Reliability Engineer information

See Hernando, MS salary details

$10

$60

$86

How much do software engineer site reliability engineer jobs pay per hour?

As of Sep 7, 2026, the average hourly pay for software engineer site reliability engineer in Hernando, MS is $60.13, according to ZipRecruiter salary data. Most workers in this role earn between $51.68 and $68.70 per hour, depending on experience, location, and employer.

What is the difference between Software Engineer Site Reliability Engineer vs DevOps Engineer?

AspectSoftware Engineer Site Reliability EngineerDevOps Engineer
CredentialsBachelor's in CS or related, sometimes certifications in cloud or SRE practicesBachelor's in CS, IT, or related, with certifications in cloud, automation, or CI/CD tools
Work EnvironmentFocus on reliability, scalability, and automation within software development teamsBridge between development and operations, emphasizing automation, deployment, and infrastructure
Employer & Industry UsageTech companies, cloud providers, large enterprisesStartups, tech firms, organizations adopting DevOps practices

While both roles focus on automation and system stability, Software Engineer Site Reliability Engineers primarily ensure system reliability and performance, whereas DevOps Engineers focus on streamlining development and deployment processes. The roles often overlap but differ in their core focus areas and daily responsibilities.

What are popular job titles related to Software Engineer Site Reliability Engineer jobs in Hernando, MS?

For Software Engineer Site Reliability Engineer jobs in Hernando, MS, the most frequently searched job titles are:

What cities near Hernando, MS are hiring for Software Engineer Site Reliability Engineer jobs?

Cities near Hernando, MS with the most Software Engineer Site Reliability Engineer job openings:

Network Operations Center Specialist

Socket.dev

Southaven, MS • On-site

$60 - $80/hr

Other

Posted yesterday

New


Job description

SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates.

ABOUT THE ROLE:

As a Network Operations Center (NOC) Specialist, you are the eyes and the voice of the campus — never the hands. You watch campus health signals around the clock, detect and verify site-impacting events, assemble the right responders fast, and run incident communications leadership can trust. You make sure no major incident closes without a timeline, a report, and a tracked corrective project. This is a communications-and-judgment role at the center of site operations, not a junior-engineering holding pen. You do not do wrench work, plant operation, deep root-cause analysis, monitoring design, technical SEV command, or tool building — those belong to SiteOps, Facilities, Hardware Failure Analysis, Site SRE, and Software Platforms.

RESPONSIBILITIES:
  • Staff the console per shift schedule and watch the designated signal surface: cluster health, node availability, network health, facility trend panels, storage alarms, and threshold breaches.
  • Acknowledge every page within SLA; classify (actionable / known / noise) and log disposition; feed noise patterns back to SRE so signal quality keeps improving.
  • Detect, verify, and elevate within time budgets; operate the escalation matrix (NOC - on-call SRE - domain owners) and page correctly the first time.
  • Open and run incident bridges; own stakeholder communications (first update within SLA, then fixed cadence); maintain the incident timeline in real time; call out ownership stalls.
  • Produce first-pass RCA framing (what happened, when, what's impacted, who's engaged) and hand it to SRE / Hardware Failure Analysis for depth - the NOC does not publish root cause.
  • Run structured shift handoffs and durable shift logs; maintain cross-site awareness.
  • Write major-incident reports; open corrective projects in Linear and chase them to closure - the NOC is the nag of record.
  • Maintain and continuously improve NOC runbooks, escalation matrices, and communications templates; participate in SRE-run game days.
BASIC QUALIFICATIONS:
  • Experience in a 24/7 operations environment (NOC, SOC, dispatch, mission control, or equivalent).
  • Proven ability to acknowledge, classify, and elevate incidents under SLA in a high-signal environment.
  • Experience opening and running incident bridges, including stakeholder updates on a fixed cadence and live timeline hygiene.
  • Excellent written and verbal communication skills; able to write clear updates while an incident is in progress.
  • Demonstrated pattern recognition across multiple domains (compute, network, storage, and/or facilities signals) and curiosity about how those systems interact.
  • Experience following, maintaining, and improving operational process (runbooks, escalation matrices, handoffs, or similar).
  • Willingness and ability to work a rotating shift schedule, including nights and weekends, as part of continuous campus coverage.
PREFERRED SKILLS AND EXPERIENCE:
  • Prior NOC, data center operations, or campus reliability experience in a high-performance computing, AI/ML infrastructure, or large-scale production environment.
  • Experience writing major-incident reports and driving corrective follow-ups to closed (e.g. tickets, projects, or Linear).
  • Familiarity with Linear or similar work-tracking tools for corrective action programs.
  • Experience partnering with SRE, SiteOps, and Facilities on escalations and post-incident follow-through.
  • Participation in game days, tabletop exercises, or runbook improvement programs.
  • Prior work in a fast-paced startup or tech company like SpaceXAI.

SpaceXAI is an equal opportunity employer. For details on data processing, view our Recruitment Privacy Notice.

#J-18808-Ljbffr