... error trending, and log correlation to identify hardware issues before they cause customer impact ... internal HWEng teams to ensure new server hardware addresses data path and control path ...
... error trending, and log correlation to identify hardware issues before they cause customer impact ... internal HWEng teams to ensure new server hardware addresses data path and control path ...
... data, error trending, and log correlation to identify hardware issues before they cause customer ... internal HWEng teams to ensure new server hardware addresses data path and control path ...
... data, error trending, and log correlation to identify hardware issues before they cause customer ... internal HWEng teams to ensure new server hardware addresses data path and control path ...
... data, error trending, and log correlation to identify hardware issues before they cause customer ... internal HWEng teams to ensure new server hardware addresses data path and control path ...
... data, error trending, and log correlation to identify hardware issues before they cause customer ... internal HWEng teams to ensure new server hardware addresses data path and control path ...
SAS Grid Administrator
Tarrytown, NY · On-site
... servers and processes, including error messages Execute backups regularly as specified in your ... internal customer infra teams Provide guidance and assistance to SAS developers on operational ...
SAS Grid Administrator
Tarrytown, NY · On-site
... servers and processes, including error messages Execute backups regularly as specified in your ... internal customer infra teams Provide guidance and assistance to SAS developers on operational ...
Junior PSOC Administrator (overnight)
Springfield, VA · On-site
$81K/yr
Actively monitor the approved tools for critical server error notifications * Communicate with on ... Demonstrated ability to effectively communicate and collaborate with diverse internal and external ...
Junior PSOC Administrator (overnight)
Springfield, VA · On-site
$81K/yr
Actively monitor the approved tools for critical server error notifications * Communicate with on ... Demonstrated ability to effectively communicate and collaborate with diverse internal and external ...
Software Engineer, Platform
San Francisco, CA · On-site
$180K - $250K/yr
... servers including provisioning, health monitoring, error detection, and recovery -- and when ... Experience building internal tools or dashboards for infrastructure visibility * Excellent ...
Software Engineer, Platform
San Francisco, CA · On-site
$180K - $250K/yr
... servers including provisioning, health monitoring, error detection, and recovery -- and when ... Experience building internal tools or dashboards for infrastructure visibility * Excellent ...
Sr. System Development Engineer, Edge & High Performance Accelerator Servers for AI/ML
Austin, TX · On-site
$103K - $141K/yr
... error trending, and log correlation to identify hardware issues before they cause customer impact ... internal HWEng teams to ensure new server hardware addresses data path and control path ...
Sr. System Development Engineer, Edge & High Performance Accelerator Servers for AI/ML
Austin, TX · On-site
$103K - $141K/yr
... error trending, and log correlation to identify hardware issues before they cause customer impact ... internal HWEng teams to ensure new server hardware addresses data path and control path ...
Software Engineer, Platform
San Francisco, CA · On-site
$180K - $250K/yr
... servers including provisioning, health monitoring, error detection, and recovery - and when ... Experience building internal tools or dashboards for infrastructure visibility * Excellent ...
Software Engineer, Platform
San Francisco, CA · On-site
$180K - $250K/yr
... servers including provisioning, health monitoring, error detection, and recovery - and when ... Experience building internal tools or dashboards for infrastructure visibility * Excellent ...
Front-End Application & Web Platform Engineer
Bridgeton, MO · On-site
$75K - $85K/yr
KEY RESPONSIBILITIES Internal Applications & Intranet Platform * Application Completion ... State & Error Handling : Implement appropriate loading, success, warning, error, and empty states ...
New
Front-End Application & Web Platform Engineer
Bridgeton, MO · On-site
$75K - $85K/yr
KEY RESPONSIBILITIES Internal Applications & Intranet Platform * Application Completion ... State & Error Handling : Implement appropriate loading, success, warning, error, and empty states ...
New
Software Engineer, Platform
San Francisco, CA · On-site
$180K - $250K/yr
... servers including provisioning, health monitoring, error detection, and recovery -- and when ... Experience building internal tools or dashboards for infrastructure visibility * Excellent ...
Software Engineer, Platform
San Francisco, CA · On-site
$180K - $250K/yr
... servers including provisioning, health monitoring, error detection, and recovery -- and when ... Experience building internal tools or dashboards for infrastructure visibility * Excellent ...
$180K - $250K/yr
... servers including provisioning, health monitoring, error detection, and recovery -- and when ... Experience building internal tools or dashboards for infrastructure visibility * Excellent ...
$180K - $250K/yr
... servers including provisioning, health monitoring, error detection, and recovery -- and when ... Experience building internal tools or dashboards for infrastructure visibility * Excellent ...
Sr. System Development Engineer, Edge & High Performance Accelerator Servers for AI/ML
Austin, TX · On-site
$103K - $141K/yr
... error trending, and log correlation to identify hardware issues before they cause customer impact ... internal HWEng teams to ensure new server hardware addresses data path and control path ...
Sr. System Development Engineer, Edge & High Performance Accelerator Servers for AI/ML
Austin, TX · On-site
$103K - $141K/yr
... error trending, and log correlation to identify hardware issues before they cause customer impact ... internal HWEng teams to ensure new server hardware addresses data path and control path ...
Software Engineer, Infrastructure
San Francisco, CA · On-site
$180K - $250K/yr
... servers including provisioning, health monitoring, error detection, and recovery -- and when ... Experience building internal tools or dashboards for infrastructure visibility * Excellent ...
Software Engineer, Infrastructure
San Francisco, CA · On-site
$180K - $250K/yr
... servers including provisioning, health monitoring, error detection, and recovery -- and when ... Experience building internal tools or dashboards for infrastructure visibility * Excellent ...
KEY RESPONSIBILITIES Internal Applications & Intranet Platform * Application Completion ... State & Error Handling : Implement appropriate loading, success, warning, error, and empty states ...
New
KEY RESPONSIBILITIES Internal Applications & Intranet Platform * Application Completion ... State & Error Handling : Implement appropriate loading, success, warning, error, and empty states ...
New
Sr. System Development Engineer, Edge & High Performance Accelerator Servers for AI/ML
Austin, TX · On-site
$103K - $141K/yr
... error trending, and log correlation to identify hardware issues before they cause customer impact ... internal HWEng teams to ensure new server hardware addresses data path and control path ...
Sr. System Development Engineer, Edge & High Performance Accelerator Servers for AI/ML
Austin, TX · On-site
$103K - $141K/yr
... error trending, and log correlation to identify hardware issues before they cause customer impact ... internal HWEng teams to ensure new server hardware addresses data path and control path ...
Server Patching Engineer
Washington, DC · On-site
$70 - $80/hr
Years 8 Years • 8 yrs break/fix support for Server OS and third-party application issues ... internal resources to focus on strategic issues. AHU Technologies INC was co-founded by visionary ...
Server Patching Engineer
Washington, DC · On-site
$70 - $80/hr
Years 8 Years • 8 yrs break/fix support for Server OS and third-party application issues ... internal resources to focus on strategic issues. AHU Technologies INC was co-founded by visionary ...
Fix what breaks - investigate and resolve defects surfaced through QA and customer support ... In an accounting system, a lost update is not a performance problem; it is a financial error.
Fix what breaks - investigate and resolve defects surfaced through QA and customer support ... In an accounting system, a lost update is not a performance problem; it is a financial error.
DC Technician, Server Break Fix IC3
Ashburn, VA · On-site
$30.45 - $35/hr
... servers, networking equipment, and cabling. Expansion of basic knowledge and experience working ... with internal tools Other tasks as required to support IT service on a local basis What You Will ...
DC Technician, Server Break Fix IC3
Ashburn, VA · On-site
$30.45 - $35/hr
... servers, networking equipment, and cabling. Expansion of basic knowledge and experience working ... with internal tools Other tasks as required to support IT service on a local basis What You Will ...
Fix what breaks - investigate and resolve defects surfaced through QA and customer support ... In an accounting system, a lost update is not a performance problem; it is a financial error.
Fix what breaks - investigate and resolve defects surfaced through QA and customer support ... In an accounting system, a lost update is not a performance problem; it is a financial error.
Fix what breaks -- investigate and resolve defects surfaced through QA and customer support ... In an accounting system, a lost update is not a performance problem; it is a financial error.
Quick apply
Fix what breaks -- investigate and resolve defects surfaced through QA and customer support ... In an accounting system, a lost update is not a performance problem; it is a financial error.
Fix Internal Server Error information
See salary details
$6.49 - $7.65
3% of jobs
$7.65 - $8.81
3% of jobs
$8.81 - $9.97
7% of jobs
$9.97 - $11.12
6% of jobs
$12.17 is the 25th percentile. Wages below this are outliers.
$11.12 - $12.28
5% of jobs
$12.28 - $13.44
5% of jobs
$13.44 - $14.60
12% of jobs
The median wage is $14.87 / hr.
$14.60 - $15.76
32% of jobs
$15.79 is the 75th percentile. Wages above this are outliers.
$15.76 - $16.91
16% of jobs
$16.91 - $18.07
6% of jobs
$18.07 - $19.23
3% of jobs
$6
$14
$19
How much do fix internal server error jobs pay per hour?
What does it mean to fix an internal server error?
What are some common challenges faced when troubleshooting and fixing internal server errors as a web developer?
What are the key skills and qualifications needed to thrive as a systems administrator, and why are they important?
What is the difference between Fix Internal Server Error vs Web Developer?
| Aspect | Fix Internal Server Error | Web Developer |
|---|---|---|
| Required Credentials | Technical troubleshooting skills, knowledge of server-side languages | Programming languages, web development certifications |
| Work Environment | IT support, server management, troubleshooting | Designing, coding, testing websites and applications |
| Industry Usage | IT support teams, server administrators | Web development agencies, tech companies |
| Common Search/Comparison | Technical issue resolution | Web design and development tasks |
The main difference is that Fix Internal Server Error involves troubleshooting server issues to resolve errors, while a Web Developer focuses on creating and maintaining websites. The former is more technical and support-oriented, whereas the latter is development-focused. Both roles require technical skills but serve different purposes within the tech industry.

Systems Development Engineer, AWS Generative AI & ML Servers
Cupertino, CA • On-site
Full-time
Medical, Dental, Vision, Life, Retirement, PTO
Posted 20 days ago
Amazon rating
7.4
Based on 7,162 frontline employees who took The Breakroom Quiz
5th of 39 rated national retailers
Job description
What You Will Do
You will solve complex architectural problems that may not be well-defined in advance. You will own your team's systems, proactively identify deficiencies, and write scalable, robust code to solve issues before they impact customers. You will decompose large, difficult server testability, reliability, and diagnosis problems into straightforward tasks and components - delivering yourself and through others in parallel - using a combination of hardware, software, system design, processor architecture, diagnostics, and operations knowledge.
Key job responsibilities
Fleet Health & Predictive Infrastructure
1. Build and own the automation infrastructure responsible for the health of the accelerator (AI/ML) compute server fleet
2. Design and implement predictive failure detection systems using telemetry, sensor data, error trending, and log correlation to identify hardware issues before they cause customer impact
3. Drive toward zero-touch operations - building automation that detects, diagnoses, triages, and remediates hardware and software faults without human intervention
4. Develop monitoring tools, dashboards, and alerting systems to provide real-time visibility into fleet health across lab and production environments
5. Define and track fleet health metrics (failure rates, mean time to detect, mean time to repair, first-time fix rate, predictive accuracy)
Debugging & Troubleshooting
1. Debug and resolve complex system-level issues across compute, GPU, and networking in production environments
2. Troubleshoot Linux boot and runtime failures across x86 and ARM architectures, including PCIe, power, NIC, NVMe, and GPU subsystems
3. Perform root cause analysis on hardware failures - correlating across firmware, kernel, driver, and physical layer to isolate faults
4. Build diagnostic tooling that automates root cause identification and reduces reliance on manual triage
5. Improve manufacturing throughput and yield through test optimization
Systems Development & Automation
1. Define and develop software, automation, and enabling tools for server hardware programs; track and report progress
2. Design and build scalable system-level software with focus on durability, availability, security, and diagnostics
3. Develop and maintain device drivers for Linux on ARM and x86 architectures
4. Build automation solutions using modern programming languages (Python, Ruby, Java, C/C++, etc.)
5. Work with OS internals and accelerator/GPU software stacks in Linux-based environments
6. Build, manage, and deploy CI/CD pipelines for rapid deployment of code changes to org-owned and customer-owned systems
Cross-Team Collaboration
1. Work across internal HWEng teams to ensure new server hardware addresses data path and control path functionality needed by dependent service teams
2. Work closely with internal customers to identify early any potential problems onboarding new accelerated compute servers into their ecosystem
3. Engage with ODMs and design partners on testability, diagnostic, and automation requirements during hardware design and development (NPI)
4. Contribute to server design to improve robustness, testability, diagnosability, and reliability
5. Partner with datacenter operations teams to close the loop between field failures and design improvements
A day in the life
You will collaborate with a variety of roles (SDEs, SDETs, Mechanical/Electrical/Hardware Engineers, TPMs, Managers, Principals) and organizations through server conception, test validation, qualification, launch, and operations - driving high quality and reliability into current and future designs for AWS accelerated server solutions. From orchestration tooling development to hardware integration to kernel driver debugging, you dive deep into problems across the breadth of AWS.
About the team
The Hardware Engineering AI/ML development team is a group of engineers and technical program managers directly responsible for launching and maintaining server hardware in the fleet - including AI/ML accelerator servers with GPUs. Located in Seattle, Cupertino, and Austin, we work with internal development teams, ODMs, and design partners to deliver servers deployed in datacenters worldwide.
BASIC QUALIFICATIONS
- 2+ years of non-internship professional software development experience
- 1+ years of designing or architecting (design patterns, reliability and scaling) of new and existing systems experience
- Experience programming with at least one modern language such as C++, C#, Java, Python, Golang, PowerShell, Ruby
PREFERRED QUALIFICATIONS
- Familiarity with server hardware architecture, BMC/IPMI, firmware, PCIe topology, and hardware diagnostics
- Experience working with ODMs or hardware design partners
- Exposure to zero-touch or self-healing automation concepts for large-scale infrastructure
- Experience working in large-scale datacenter or cloud environments
- Experience with hardware bring-up, validation, or fleet-wide deployment
- Familiarity with telemetry pipelines, anomaly detection, or operational metrics at scale
Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.
Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.
Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner.
The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits.
USA, CA, Cupertino - 148,700.00 - 201,200.00 USD annually
USA, TX, Austin - 129,200.00 - 174,800.00 USD annually
USA, WA, Seattle - 129,200.00 - 174,800.00 USD annually
About Amazon
Sourced by ZipRecruiter
Amazon.com, Inc., commonly known as Amazon, is an American multinational technology company. It was founded by Jeff Bezos in 1994 and initially started as an online marketplace for books. Since then, Amazon has expanded its operations and become one of the largest e-commerce companies in the world. Amazon's primary business is its online retail platform, where customers can purchase a vast array of products, including electronics, clothing, books, home goods, and much more. The company offers a convenient and user-friendly shopping experience, with features such as fast shipping, customer reviews, and personalized recommendations. In addition to its e-commerce platform, Amazon has diversified its business into various other areas. One of its notable ventures is Amazon Web Services (AWS), a comprehensive cloud computing platform that provides services such as storage, compute power, and database management to individuals and businesses. AWS has become a leader in the cloud computing industry, powering many websites and applications worldwide. Amazon has also developed its own consumer electronics, including the popular Amazon Kindle e-reader, Fire tablets, Fire TV streaming devices, and the Alexa-powered Echo smart speakers. The Alexa voice assistant, integrated into these devices, allows users to interact with their devices using voice commands, perform tasks, and access information. Furthermore, Amazon has expanded into media and entertainment. It operates Prime Video, a streaming service that offers a wide range of movies, TV shows, and original content. Amazon Music provides a platform for streaming and purchasing digital music, while Audible offers audiobooks and other audio content. The company's commitment to customer satisfaction and convenience is demonstrated by its membership program, Amazon Prime. Prime members receive various benefits, including free two-day shipping, access to streaming services, exclusive deals, and more.
Industry
It services, book publishers, retail, real estate, computer and electronic product manufacturing and software development
Company size
10,000+ Employees
Headquarters location
Seattle, WA, US