1

Hpc Engineer Jobs in Austin, TX (NOW HIRING)

ClifyX is seeking a Network HPC Engineer to design and deploy HPC clusters using high-performance servers interconnected by advanced networks. The role involves optimizing InfiniBand and RoCE ...

Staff Slurm Cluster & HPC Engineer

Austin, TX · On-site +1

$180K - $260K/yr

We are seeking a Staff Slurm Cluster & HPC Scheduling Engineer to own Slurm as a first-class, productized scheduling layer across that fleet. This person is the single technical owner of Slurm ...

HPC Operations Engineer

Austin, TX · Hybrid

$68K - $93K/yr

As an HPC Operations Engineer at NVIDIA, you will play a pivotal role in ensuring the flawless operation of our high-performance computing (HPC) environment. This opportunity is outstanding as you ...

HPC Operations Engineer

Austin, TX · Hybrid

$68K - $93K/yr

As an HPC Operations Engineer at NVIDIA, you will play a pivotal role in ensuring the flawless operation of our high-performance computing (HPC) environment. This opportunity is outstanding as you ...

Join our customer success team as an HPC Customer Solutions Engineer . We are seeking a highly skilled, customer-centric individual to act as the primary link between our cutting-edge HPC technology ...

Join our customer success team as an HPC Customer Solutions Engineer . We are seeking a highly skilled, customer-centric individual to act as the primary link between our cutting-edge HPC technology ...

We are looking for an experienced Staff Applications Engineer to join the Applications team, with demonstrated ability in moving HPC applications onto new platforms and in evaluating performance on ...

Senior HPC Cluster Engineer

Austin, TX · On-site

$103K - $142K/yr

We are seeking a highly skilled and experienced HPC Cluster Engineer to design, deploy, and operate GPU Compute Clusters for EDA (Electronic Design Automation) and high-performance computing ...

Senior HPC Cluster Engineer

Austin, TX · On-site

$103K - $142K/yr

We are seeking a highly skilled and experienced HPC Cluster Engineer to design, deploy, and operate GPU Compute Clusters for EDA (Electronic Design Automation) and high-performance computing ...

We are looking for an experienced Staff Applications Engineer to join the Applications team, with demonstrated ability in moving HPC applications onto new platforms and in evaluating performance on ...

Hudson River Trading's High Performance Computing (HPC) Network Engineering team designs and engineers the low-latency communications infrastructure that underpins our incredibly large GPU and CPU ...

Senior HPC and LSF Operations Engineer

Austin, TX · Hybrid

$103K - $141K/yr

Experience implementing reliability engineering practices within HPC scheduling environments Deep knowledge of job scheduling systems (LSF, Slurm, etc.) configuration tuning, scheduler internals, and ...

Build software tools that enable developers across a spectrum of markets to optimize their workflows; enable complex computer systems doing ongoing work in High Performance Computing(HPC), Machine ...

Senior HPC and LSF Operations Engineer

Austin, TX · Hybrid

$103K - $141K/yr

Experience implementing reliability engineering practices within HPC scheduling environments * Deep knowledge of job scheduling systems (LSF, Slurm, etc.) configuration tuning, scheduler internals ...

Build software tools that enable developers across a spectrum of markets to optimize their workflows; enable complex computer systems doing ongoing work in High Performance Computing(HPC), Machine ...

next page

Showing results 1-20

Hpc Engineer information

See Austin, TX salary details

$23.8K

$107K

$170.4K

How much do hpc engineer jobs pay per year?

As of Sep 3, 2026, the average yearly pay for hpc engineer in Austin, TX is $106,982.00, according to ZipRecruiter salary data. Most workers in this role earn between $82,700.00 and $132,300.00 per year, depending on experience, location, and employer.

What does an HPC engineer do?

An HPC (High-Performance Computing) Engineer designs, deploys, and optimizes high-performance computing systems used for intensive computational tasks. They work with parallel computing, cluster management, and performance tuning to ensure efficient processing of large-scale simulations, data analysis, and scientific research. Their role often involves configuring hardware, optimizing software, and troubleshooting issues to maximize system performance.

What are the key skills and qualifications needed to thrive as an HPC engineer?

To thrive as an HPC Engineer, you need a solid background in computer science, mathematics, or a related field, with expertise in parallel computing, Linux systems, and high-performance cluster management. Proficiency with job schedulers (like SLURM or PBS), programming languages such as C/C++ or Python, and experience with distributed file systems are highly valuable, and certifications in relevant areas can enhance your qualifications. Strong problem-solving, collaboration, and communication skills help you work efficiently within technical teams and explain complex concepts to non-experts. These skills and qualities are essential for ensuring high performance, reliability, and scalability of computing systems in scientific and enterprise settings.

What are some of the common challenges faced by HPC engineers in their day-to-day work?

HPC Engineers often encounter challenges such as optimizing performance for complex workloads, troubleshooting system failures, and efficiently managing large-scale infrastructures. You may be required to balance the needs of multiple users and projects, quickly address hardware or software issues, and stay ahead of evolving technologies. Collaboration with researchers, IT staff, and software developers is key to designing robust solutions. Overcoming these challenges helps improve computational efficiency and supports critical research and business objectives.

Are HPC engineers in demand?

HPC (High-Performance Computing) engineers are in high demand due to the increasing need for advanced computing in fields like scientific research, data analysis, and artificial intelligence. Employers seek professionals skilled in parallel programming, cluster management, and tools such as Linux and MPI, often requiring relevant certifications and experience with large-scale systems.

How much do HPC engineers make in the US?

HPC (High-Performance Computing) engineers in the US typically earn between $80,000 and $150,000 annually, depending on experience, education, and location. Senior roles or those with specialized skills in parallel computing, cluster management, or specific tools like MPI or CUDA can earn higher salaries.

What are the most commonly searched types of Hpc Engineer jobs in Austin, TX?

The most popular types of Hpc Engineer jobs in Austin, TX are:

What are popular job titles related to Hpc Engineer jobs in Austin, TX?

For Hpc Engineer jobs in Austin, TX, the most frequently searched job titles are:

What job categories do people searching Hpc Engineer jobs in Austin, TX look for?

The top searched job categories for Hpc Engineer jobs in Austin, TX are:

What cities near Austin, TX are hiring for Hpc Engineer jobs?

Cities near Austin, TX with the most Hpc Engineer job openings:

Infographic showing various Hpc Engineer job openings in Austin, TX as of August 2026, with employment types broken down into 91% Full Time, 5% Part Time, and 4% Contract. Highlights an 88% Physical, 4% Hybrid, and 8% Remote job distribution, with an average salary of $106,982 per year, or $51.4 per hour.

Network HPC Engineer

ClifyX

Austin, TX • On-site

Full-time

Re-posted 17 days ago


Job description

Job Summary:
ClifyX is seeking a Network HPC Engineer to design and deploy HPC clusters using high-performance servers interconnected by advanced networks. The role involves optimizing InfiniBand and RoCE networks for low-latency and high-bandwidth communication, ensuring performance, security, and effective vendor management.
Responsibilities:
• Designing and deploying HPC clusters consisting of high-performance servers, interconnected by high-speed networks such as InfiniBand (IB) or Ethernet/RoCE with RDMA capabilities.
• Fabric Design and Configuration: Designing InfiniBand fabrics, including switches, host channel adapters (HCAs), and cables, to ensure optimal performance, scalability, and fault tolerance. Configuring switch ports, virtual lanes (VLs), and routing tables to facilitate efficient data communication within the InfiniBand fabric.
• Topology Optimization: Analyzing workload characteristics and traffic patterns to design InfiniBand topologies (e.g., fat-tree, hypercube) that minimize latency and maximize bandwidth utilization. Implementing routing policies and congestion control mechanisms to optimize traffic flow and prevent network congestion.
• Fabric Monitoring and Management: Monitoring InfiniBand fabric health and performance using management tools such as Subnet Manager (SM) and Performance Monitoring Counters (PMCs). Performing regular maintenance tasks, including firmware updates, port diagnostics, and error detection and correction.
• Quality of Service (QoS): Implementing QoS policies to prioritize traffic based on application requirements and service levels. Configuring traffic classes, service levels, and virtual lanes (VLs) to ensure predictable performance for latency-sensitive applications.
• Security and Access Control: Securing the InfiniBand fabric with features such as subnet partitioning (subnet manager security) and encryption to protect data integrity and confidentiality. Enforcing access controls and authentication mechanisms to restrict unauthorized access to the InfiniBand network.
• Network Design and Configuration: Designing and configuring RoCE networks, including switches, network adapters, and Ethernet fabrics, to provide low-latency, high-bandwidth communication for RDMA traffic. Optimizing network settings such as MTU (Maximum Transmission Unit), buffer sizes, and flow control parameters to maximize RoCE performance.
• Congestion Management: Implementing congestion management mechanisms, such as Priority Flow Control (PFC) and Data Center Bridging (DCB), to prevent congestion and ensure fair allocation of network resources. Monitoring network traffic and congestion levels to dynamically adjust congestion control settings and avoid performance degradation.
• Routing and Switching Optimization: Configuring RoCE-aware switches and routers to support RDMA traffic and enable efficient routing of packets between endpoints. Tuning switch port settings, forwarding tables, and routing protocols to minimize packet loss and maximize throughput for RoCE traffic.
• Performance Monitoring and Tuning: Monitoring RoCE network performance metrics, such as latency, throughput, and packet loss, using tools like Ethernet Performance Monitoring (EPM) and InfiniBand Performance Monitoring (IPM). Analyzing performance data to identify bottlenecks, optimize network configurations, and fine-tune RoCE parameters for optimal performance.
• Security and Authentication: Implementing security measures, such as MACsec (Media Access Control Security) and IPsec (Internet Protocol Security), to encrypt and authenticate RDMA traffic over RoCE networks. Enforcing access controls and certificate-based authentication to ensure secure communication between RoCE endpoints.
• Vendor Management: Coordinating with hardware and software vendors to ensure compatibility and support for products in multi-vendor environments. Developing Billing of Materials. Clearly define technical requirements, including performance, scalability, compatibility, and specific features needed for RoCE. Assess the technical specifications, performance benchmarks, and compatibility with existing infrastructure. Implement a PoC to test the switches in a controlled environment and ensure they meet performance and reliability expectations. Evaluate the vendor's technical support capabilities, including responsiveness, expertise, and available resources. Maintain regular communication with the vendor to stay informed about product updates, potential issues, and upcoming changes. Schedule periodic meetings to review performance/bugs, discuss any concerns, and plan for future needs.
Qualifications:
Required:
• Bachelor's Degree in Computer Science, Information Technology, or related field: A solid educational foundation in computer science or IT is essential for understanding networking principles and protocols.
• In-depth understanding of InfiniBand architecture, protocols (IBTA), and technologies (e.g., Mellanox InfiniBand). Proficiency in RoCE (RDMA over Converged Ethernet) protocols, including RoCEv2 and related standards.
• Experience in designing and configuring high-performance networks, including InfiniBand fabrics and RoCE-enabled Ethernet networks. Knowledge of fabric design principles, topology optimization, and performance tuning techniques.
• Ability to analyze network performance metrics, diagnose bottlenecks, and optimize network configurations for low latency and high throughput. Experience in tuning switch port settings, buffer sizes, and flow control parameters to maximize RoCE performance.
• Familiarity with security measures for InfiniBand and RoCE networks, including subnet partitioning, encryption, and access controls. Knowledge of authentication mechanisms and cryptographic protocols for securing RDMA traffic.
• Proficiency in network monitoring tools and techniques for monitoring InfiniBand and RoCE network health and performance. Ability to troubleshoot network issues, diagnose connectivity problems, and resolve performance-related issues.
• Certification programs offered by vendors such as Mellanox (now NVIDIA Networking) for InfiniBand and RoCE technologies.
• Hands-on experience in deploying, managing, and optimizing high-performance computing (HPC) environments and data center networks. Experience working with RDMA-enabled applications and parallel computing frameworks (e.g., MPI, OpenMP).
• Experience in implementing and troubleshooting complex network configurations, including InfiniBand switches, gateways, and RoCE adapters.
Preferred:
• Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience.
• CCNA, CCIE, or similar
• Ability to work efficiently on multiple projects and under pressure
• Previous experience with network equipment vendor products (e.g., Juniper, Cisco, Arista, OEM).
• Working knowledge of stateful and stateless firewalls
• Comfortable with Linux or other UNIX implementations, with scripting skills.
• Experience with python scripting / ansible for scripting and automation
• Ability to 'read code' as source documentation
• DevOps CI/CD mindset for automation and scale
Company:
ClifyX provides innovative business solutions which satisfy requirements for mission-critical reliability, scalability, interoperations. Founded in 1998, the company is headquartered in South Plainfield, USA, with a team of 501-1000 employees. The company is currently Late Stage.

ClifyX logo

About ClifyX

Sourced by ZipRecruiter

ClifyX is a well-established player in the IT Services sector that specializes in providing result-oriented technological solutions to a wide range of industrial verticals. Based in South Plainfield, New Jersey, ClifyX offers a comprehensive selection of IT services that include project staffing, application development, professional consulting, and other IT-based solutions. While the company's website, clifyx.com, does not divulge the exact founding date, it is clear that ClifyX has grown into a renowned name within their domain, thanks to their unwavering commitment to innovative practices. The company's mission statement revolves around harnessing the power of technology to assist their clientele in steering their respective businesses towards success.

Industry

Recruiting and staffing services

Company size

51 - 200 Employees

Headquarters location

South Plainfield, NJ, US

Year founded

1998