1

Supercomputer Engineer Jobs (NOW HIRING)

Head of Supercomputing

San Jose, CA ยท On-site

$2.0K/mo

Backed by hundreds of millions from top-tier investors and staffed by leading engineers, Etched is ... We are seeking a Head of Supercomputing to define and lead the architecture, software stack, and ...

Head of Supercomputing

San Jose, CA ยท On-site

$2.0K/mo

Backed by hundreds of millions from top-tier investors and staffed by leading engineers, Etched is ... We are seeking a Head of Supercomputing to define and lead the architecture, software stack, and ...

Showing results 21-40

Supercomputer Engineer information

What does a supercomputer engineer do?

A Supercomputer Engineer is responsible for designing, building, maintaining, and optimizing high-performance computing systems known as supercomputers. Their work involves both hardware and software aspects, ensuring these powerful machines can process vast amounts of data at exceptionally high speeds. Supercomputer Engineers collaborate with scientists, researchers, and IT professionals to support advanced simulations, scientific research, and complex data analysis. They also troubleshoot technical issues and implement upgrades to improve system performance and efficiency.

What are the key skills and qualifications needed to thrive as a supercomputer engineer, and why are they important?

To thrive as a Supercomputer Engineer, you need a strong background in computer engineering, parallel computing, and high-performance hardware design, often supported by a degree in computer science, electrical engineering, or a related field. Familiarity with programming languages like C/C++, MPI, OpenMP, and experience with cluster management systems and advanced networking technologies are typically required. Excellent problem-solving skills, teamwork, and adaptability help you address complex technical challenges and collaborate on large-scale projects. These skills are vital for optimizing system performance, ensuring reliability, and driving innovation in computational research and enterprise applications.

What are some common challenges supercomputer engineers face when optimizing high-performance computing systems?

Supercomputer Engineers often encounter challenges related to balancing computational speed with power efficiency, managing complex cooling systems, and ensuring scalability for rapidly evolving workloads. They must also troubleshoot network bottlenecks and memory hierarchies to maximize system throughput. Collaboration with researchers, software developers, and hardware vendors is essential to identify performance issues and implement effective solutions in a multidisciplinary environment.

What is the difference between Supercomputer Engineer vs High-Performance Computing (HPC) Engineer?

AspectSupercomputer EngineerHPC Engineer
Required CredentialsBachelor's or Master's in Computer Engineering, Computer Science, or related fields; certifications in parallel computing or system administrationSimilar credentials; often includes certifications in HPC systems or network administration
Work EnvironmentDesigning, developing, and maintaining supercomputers and large-scale computing systemsOptimizing and managing high-performance computing clusters and infrastructure
Employer & Industry UsageResearch labs, government agencies, supercomputing centersUniversities, research institutions, tech companies with HPC needs

Supercomputer Engineers focus on building and maintaining the world's most powerful computing systems, while HPC Engineers optimize and manage high-performance computing clusters for research and industry. Both roles require similar skills and credentials but differ in scope and specific responsibilities.

What are popular job titles related to Supercomputer Engineer jobs?

For Supercomputer Engineer jobs, the most frequently searched job titles are:

Infographic showing various Supercomputer Engineer job openings in the United States as of September 2026, with employment types broken down into 100% Full Time. Highlights an 50% In-person, and 50% Remote job distribution.

OT Systems Engineer (Supercomputer Infrastructure)

Southaven, MS โ€ข On-site

Pantera Capital
Finance and Insuranceย โ€ขย 1 - 10 employees

Other

Posted 5 days ago


Job description

SpaceXAIโ€™s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the companyโ€™s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates.

ABOUT THE ROLE:

SpaceXAI is looking for a highly skilled and versatile OT (Operational Technology) Systems Engineer to implement and support the backend infrastructure for next-generation controls and industrial software platforms critical to our hyperscale AI supercomputer campuses and co-located power generation. This role sits with the Supercomputer Physical Infrastructure team and supports dedicated controls, facilities, and power organizations while leveraging adjacent IT expertise, tooling, and technologies. You will balance sustainment of live cooling, power, and facility control systems with modernization and continuous improvement. The ideal candidate thrives in high-stakes, 24/7 environments, brings a strong sense of urgency balanced with operational excellence, and combines deep OT expertise with IT technologies such as virtualization, VDI, GitOps, network-level redundancy, and edge compute to create streamlined and secure IC environments for ultra-dense AI compute.

RESPONSIBILITIES:
  • Design, deploy, and augment next-generation OT environments that support cooling plants, electrical distribution, liquid-cooling loops (CDUs / facility water), on-site generation, BESS, and data hall systems across the Memphis / Southaven campus and future sites.
  • Install, configure, maintain, and support industry-standard controls software platforms (BMS, EPMS, SCADA, PLC) as well as internally developed HMIs.
  • Integrate established and emergent IT technologies to simplify management, improve security, and create scalable, highly available plant and facility environments.
  • Deploy and maintain development, test, and staging environments to enable controlled, systematic change introduction on live critical systems.
  • Provide direct support during cluster bring-up, capacity expansions, commissioning, and production training campaigns.
  • Perform systems/software upgrades and maintenance between critical operations (including evenings and weekends as needed).
  • Proactively monitor services and respond rapidly to incidents to maintain high availability and performance of power, cooling, and environmental control.
  • Leverage automation tools and contribute to infrastructure-as-code and broader DevOps initiatives with a focus on IC / OT environments.
  • Work with controls, mechanical, and electrical engineers to iterate on simulation and emulation environments to enhance test coverage and change control.
  • Write and maintain standards, architectures, best practices, and documentation (system overviews, design drawings, operational procedures), with emphasis on highly reliable and secure industrial environments, especially at IT/OT and data-hall boundaries.
  • Collaborate with cross-functional teams (IT, security, controls, facilities operations, construction, and power) as well as vendors and integrators to design robust OT architectures and resolve technical issues.
  • Ensure OT systems are configured and maintained in compliance with industry and cybersecurity standards (e.g., Purdue Model, IEC 62443).
BASIC QUALIFICATIONS:
  • 3+ years of experience in OT systems engineering or industrial control systems administration.
  • Hands-on experience with multiple industry-standard controls software platforms and tools.
  • Significant experience designing, deploying, supporting, and troubleshooting OT environments in high-reliability settings.
PREFERRED SKILLS AND EXPERIENCE:
  • Experience supporting real-time systems, industrial control networks, or OT environments in data centers, power generation, semiconductor, energy, or similar high-reliability industries.
  • Direct experience with BMS, EPMS, SCADA, and PLC systems serving large cooling plants, medium-voltage distribution, and mission-critical facilities.
  • Experience designing architectures that incorporate hyperconverged, rugged industrial edge, and distributed compute technologies.
  • Working knowledge of industrial protocols (BACnet, Modbus, OPC UA, MQTT, Ethernet/IP, DNP3), controls networks, and OT cybersecurity best practices.
  • Proficiency in scripting (Bash / PowerShell / Python) and automation frameworks (Puppet, Terraform, Ansible, etc.).
  • Experience with configuration management, provisioning, infrastructure as code, and DevOps concepts/tools.
  • Familiarity with Active Directory, multi-platform authentication, and identity environments in OT contexts.
  • System administration experience managing Windows and Linux servers, rudimentary database administration, and storage/backup.
  • Network administration experience and understanding of the OSI model, especially Layer 1/2/3 considerations as they apply to industrial and facility networks (segmentation, VRFs, MDFs/IDFs).
  • Excellent communication skills with the ability to work with internal teams, vendors, and management in both formal and informal settings.
ADDITIONAL REQUIREMENTS:
  • Willingness to participate in an after-hours on-call rotation and work extended hours or weekends as necessary to support live campus operations.
  • Willingness to travel (up to 20%) between Memphis, Southaven, and other sites as the campus expands.
  • Ability to lift 30 lbs.
  • Ability to work at heights and in plant / data hall environments.
  • Ability to drive (active valid driverโ€™s license).
  • Ability to work onsite in the Memphis, TN / Southaven, MS area.

SpaceXAI is an equal opportunity employer. For details on data processing, view our Recruitment Privacy Notice.

#J-18808-Ljbffr