1

Linux Site Reliability Engineer Jobs in Iowa (NOW HIRING)

Reliability Technician

Charles City, IA · On-site

$14.25 - $19.50/hr

Job Overview The Reliability Technician will be responsible for executing the site reliability ... Work with Reliability Engineer to collect data in support of conducting failure analyses.

What we look for 4+ years in Platform, Infrastructure, Cloud, or Site Reliability Engineering. Experience with observability tooling such as AWS Cloudwatch, OTEL, and Prometheus. Deep AWS operational ...

Reliability Technician

Charles City, IA · On-site

$14.25 - $19.50/hr

Job Overview The Reliability Technician will be responsible for executing the site reliability ... Work with Reliability Engineer to collect data in support of conducting failure analyses.

... site for work. ADM is seeking a Principal Reliability Advisor in our corporate Reliability ... Bachelor's degree in engineering (Mechanical or Industrial), or a similar technical discipline * 8+ ...

Showing results 21-40

Linux Site Reliability Engineer information

What is a Linux Site Reliability Engineer?

A Linux Site Reliability Engineer (SRE) is an IT professional responsible for ensuring the reliability, scalability, and performance of systems running on the Linux operating system. They bridge the gap between software development and operations by automating processes, monitoring infrastructure, and managing incidents. Linux SREs focus on system availability, building tools for deployment and monitoring, and improving system robustness through best practices and automation. Their work helps organizations deliver reliable online services and quickly recover from outages or system failures.

What are the key skills and qualifications needed to thrive as a Linux Site Reliability Engineer?

To thrive as a Linux Site Reliability Engineer, you need deep expertise in Linux system administration, scripting (such as Bash or Python), and a solid understanding of networking concepts, usually backed by a computer science degree or equivalent experience. Familiarity with configuration management tools (like Ansible, Puppet, or Chef), containerization (Docker, Kubernetes), and cloud platforms (AWS, GCP, or Azure) is typically required, along with relevant certifications like RHCE or AWS Certified SysOps Administrator. Strong problem-solving skills, effective communication, and the ability to work under pressure are crucial soft skills for this role. These competencies ensure the reliability, scalability, and security of complex infrastructure, minimizing downtime and supporting seamless operations.

What are some common challenges faced by Linux Site Reliability Engineers when scaling infrastructure, and how can they be addressed?

Linux Site Reliability Engineers often encounter challenges related to maintaining system stability and performance as infrastructure scales. Issues such as configuration drift, automation bottlenecks, and monitoring gaps can arise when managing numerous servers or services. Addressing these challenges typically involves implementing robust configuration management tools, investing in automated deployment pipelines, and enhancing observability through comprehensive monitoring and alerting solutions. Collaboration with development and operations teams is essential to ensure that scalability solutions align with business needs and technical requirements.

What is the difference between Linux Site Reliability Engineer vs Linux DevOps Engineer?

AspectLinux Site Reliability EngineerLinux DevOps Engineer
CredentialsLinux certifications, SRE-specific trainingLinux certifications, DevOps tools certifications
Work EnvironmentFocus on system reliability, monitoring, incident responseFocus on automation, CI/CD pipelines, deployment
Employer & IndustryTech companies, cloud providers, large enterprisesStartups, tech firms, software development teams
Search & Comparison IntentUnderstanding reliability roles, incident managementAutomation, deployment, continuous integration

While both roles involve Linux expertise, a Linux Site Reliability Engineer primarily focuses on maintaining system reliability, monitoring, and incident response. In contrast, a Linux DevOps Engineer emphasizes automation, continuous integration, and deployment processes. Both roles require Linux skills and often overlap, but their core responsibilities differ based on organizational needs.

What are popular job titles related to Linux Site Reliability Engineer jobs in Iowa?

For Linux Site Reliability Engineer jobs in Iowa, the most frequently searched job titles are:

What job categories do people searching Linux Site Reliability Engineer jobs in Iowa look for?

The top searched job categories for Linux Site Reliability Engineer jobs in Iowa are:

What cities in Iowa are hiring for Linux Site Reliability Engineer jobs?

Cities in Iowa with the most Linux Site Reliability Engineer job openings:

Infographic showing various Linux Site Reliability Engineer job openings in Iowa as of June 2026, with employment types broken down into 73% Full Time, 22% Part Time, 2% Temporary, and 3% Contract. Highlights an 94% Physical, 3% Hybrid, and 3% Remote job distribution.

Data Center Production Operations Engineer

Cedar Rapids, IA • On-site

Meta
Internet and IT • 10K+ employees

$83K/yr

Full-time

Re-posted 12 days ago


Meta rating

7.8

Company rating: 7.8 out of 10

Based on 45 frontline employees who took The Breakroom Quiz

139th of 247 rated software companies


Job description

Meta is seeking a forward thinking experienced engineer to join the Production Operations team within our Data Centers. These Data Centers are the foundation upon which our rapidly scaling infrastructure efficiently operates and upon which our innovative services are delivered. Meta is at the leading edge of the global data center industry both in terms of how data centers are designed and operated. This person should enjoy working in a fast paced, technical environment where adaptability and flexibility will be key to their success. We seek an IT professional with advanced, hands-on technical skills in server hardware and Linux - ideally in a Data Center environment. Having broad knowledge of server administration and participating in projects in a large-scale distributed data center environment is a core competency of this individual. The candidate should also have working knowledge and experience in a few of the following core areas: Hardware repair, OS management, Tooling and Automation, Networking, or Technical Project Management.
Data Center Production Operations Engineer Responsibilities:
  • Support platform health by successfully resolving and closing tickets, while addressing the overall issue (i.e. addressing root cause) including, but not limited to, remote troubleshooting and physical inspection of services in data halls
  • Participate in deep dives and root cause analysis of highly technical issues within the data center, ranging from automated tooling to hardware failures and network issues
  • Collaborate with cross-functional teams on projects and initiatives related to topics such as process, hardware and automation
  • Point of contact for the introduction of new platforms and hardware to the site, in collaboration with partners and global resources, accelerating the time it takes to bring these products to sustained mass production
  • Use tools and data analysis effectively to identify issues. Take actions to communicate with all stakeholders appropriately and manage or escalate as needed
  • Identify corrective actions of hardware issues, work with internal teams and vendors
  • influence future design changes to ensure ease of serviceability
  • Solve systemic hardware and/or software issues at scale using scripting, automation, and tooling to drive global resolution
  • Continuously evaluate and identify areas for improvement in processes, tools, and systems to optimize efficiency and quality of repairs
  • Use data analytics to drive maximum server up-time and utilization rates, understanding hardware failure rates and service level agreements
  • Support and train team members to evaluate and identify better ways to resolve issues, and define updates to tools and processes
  • Provide engineering support and be a go-to technical resource for the team, leadership, and cross-functional teams in operating and maintaining data center servers
  • Maintain and update documentation i.e. procedures, runbooks and guides
  • Build cross functional relationships and influence policies and procedures that improve global data center operations
  • Participate in 24/7 on-call rotation
  • Ability to travel up to 15% of the time

Minimum Qualifications:
  • BS, BA or BEng in technical field or commensurate experience
  • 5+ years of technical IT experience within an infrastructure environment, in a role such as Systems Administrator, DevOps Engineer, or Site Reliability Engineer
  • Intermediate-level understanding in Linux (or equivalent OS) in a complex IT environment with the ability to triage, debug, and troubleshoot server issues
  • Hands-on experience and knowledge of server hardware and components, including storage
  • Intermediate-level knowledge of the interdependencies of data center functions and technologies including electrical, cooling, structured cabling, security, and network
  • Experience managing technical issues and driving to the root cause
  • Experience participating in technical projects related to areas such as process improvement, technology, and/or automation
  • Ability to communicate effectively, in a clear and concise manner, appropriately tailoring messages to the audience
  • Intermediate-level knowledge of technologies such as HTTP, DNS, RAID, and DHCP
  • Experience in providing technical guidance to external vendors
  • Experience in debugging, modifying and developing commonly used scripting or programming languages in at least one of these languages: Bash, PHP, Python, SQL, Rust, Go or Perl
  • Knowledge of out-of-band/lights-out server communication methods, such as IPMI and serial console
  • Experience using data and metrics to drive decisions

Preferred Qualifications:
  • Experience in fostering growth in others, and driving influence across all organizational levels
  • Experience in a large-scale data center environment
  • Experience with large-scale AI implementations
  • Six Sigma knowledge/certification

About Meta:
Meta builds technologies that help people connect, find communities, and grow businesses. When Facebook launched in 2004, it changed the way people connect. Apps like Messenger, Instagram and WhatsApp further empowered billions around the world. Now, Meta is moving beyond 2D screens toward immersive experiences like augmented and virtual reality to help build the next evolution in social technology. People who choose to build their careers by building with us at Meta help shape a future that will take us beyond what digital connection makes possible today—beyond the constraints of screens, the limits of distance, and even the rules of physics.
Meta is proud to be an Equal Employment Opportunity and Affirmative Action employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender, gender identity, gender expression, transgender status, sexual stereotypes, age, status as a protected veteran, status as an individual with a disability, or other applicable legally protected characteristics. We also consider qualified applicants with criminal histories, consistent with applicable federal, state and local law. Meta participates in the E-Verify program in certain locations, as required by law. Please note that Meta may leverage artificial intelligence and machine learning technologies in connection with applications for employment.
Meta is committed to providing reasonable accommodations for candidates with disabilities in our recruiting process. If you need any assistance or accommodations due to a disability, please let us know at accommodations-ext@meta.com.
$83,990/year to $130,000/year + bonus + equity + benefits
Individual compensation is determined by skills, qualifications, experience, and location. Compensation details listed in this posting reflect the base hourly rate, monthly rate, or annual salary only, and do not include bonus, equity or sales incentives, if applicable. In addition to base compensation, Meta offers benefits. Learn more about benefits at Meta.

What Meta employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom