1

Slurm Jobs in Texas (NOW HIRING)

Machine Learning Operations Engineer

Dallas, TX · On-site

$68K - $93K/yr

Familiarity with SLURM clusters or other distributed job schedulers. * Exposure to Kafka, Spark Streaming, or other real-time data processing technologies. * Understanding of ML lifecycle management ...

HPC Systems Engineer

Spring, TX

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

... SLURM) * Experience in HPC technologies such as parallel/distributed files systems (e.g., Lustre, GPFS), high speed interconnect fabrics (e.g., Infiniband, Omni-Path), and HPC batch scheduling ...

AI & HPC Infrastructure Engineer

Austin, TX · On-site

$106K - $139K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

  • PTO

Deploy, configure, and manage XPU-based clusters (GPU, DPU, LPU, CPU) across bare-metal and containerized environments using workload schedulers (Slurm, Run:ai), Kubernetes orchestration, and ...

New

Machine Learning Operations Engineer

Dallas, TX · On-site

$68K - $93K/yr

Familiarity with SLURM clusters or other distributed job schedulers. * Exposure to Kafka, Spark Streaming, or other real-time data processing technologies. * Understanding of ML lifecycle management ...

HPC Systems Engineer

Spring, TX · On-site

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

... SLURM) * Experience in HPC technologies such as parallel/distributed files systems (e.g., Lustre, GPFS), high speed interconnect fabrics (e.g., Infiniband, Omni-Path), and HPC batch scheduling ...

Machine Learning Operations Engineer

Dallas, TX · On-site

$68K - $93K/yr

Familiarity with SLURM clusters or other distributed job schedulers. * Exposure to Kafka, Spark Streaming, or other real-time data processing technologies. * Understanding of ML lifecycle management ...

HPC Systems Engineer

Spring, TX

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

... SLURM) * Experience in HPC technologies such as parallel/distributed files systems (e.g., Lustre, GPFS), high speed interconnect fabrics (e.g., Infiniband, Omni-Path), and HPC batch scheduling ...

Showing results 21-40

Slurm information

What job categories do people searching Slurm jobs in Texas look for?

The top searched job categories for Slurm jobs in Texas are:

What cities in Texas are hiring for Slurm jobs?

Cities in Texas with the most Slurm job openings:

Infographic showing various Slurm job openings in Texas as of August 2026, with employment types broken down into 97% Full Time, 2% Part Time, and 1% Contract. Highlights an 84% Physical, 5% Hybrid, and 11% Remote job distribution.

Machine Learning Operations Engineer

System One Holdings, LLC

Dallas, TX • On-site

$120K/yr

Contractor

Re-posted 9 days ago


Job description

Job Title: Machine Learning Operations Engineer
Location: Dallas, Texas
Type: Contract To Hire
Visa : USC, GC, EAD (Only W2, No Sponsorship)

Responsibilities
  • Optimize and maintain large-scale feature engineering pipelines using PySpark, Pandas, and PyArrow on Hadoop-based infrastructure.
  • Refactor and modularize ML codebases to enhance reusability, maintainability, and performance.
  • Collaborate with platform teams on compute capacity planning, resource allocation, and system upgrades.
  • Integrate with existing model serving frameworks to support testing, deployment, and rollback processes.
  • Monitor and troubleshoot production ML pipelines, ensuring high reliability, low latency, and cost efficiency.
  • Contribute to internal ML platforms by sharing insights, proposing improvements, and documenting best practices.
  • Build near real-time ML pipelines using Kafka and Spark Streaming.
  • Work with AWS and SageMaker MLOps ecosystem.

Requirements
  • 6+ years of experience in software engineering, data engineering, or MLOps roles.
  • Strong programming expertise in Python, with hands-on experience in Pandas, PySpark, and PyArrow.
  • Deep understanding of the Hadoop ecosystem, distributed computing, and performance tuning.
  • Experience with CI/CD pipelines and best practices in ML environments.
  • Hands-on experience with monitoring tools for ML pipeline health and performance.
  • Strong collaboration skills with experience working in cross-functional teams (platform, data science, engineering).
  • Experience contributing to or building internal MLOps frameworks/platforms.
  • Familiarity with SLURM clusters or other distributed job schedulers.
  • Exposure to Kafka, Spark Streaming, or other real-time data processing technologies.
  • Understanding of ML lifecycle management, including versioning, deployment, and drift detection.

#M1
#DI-CB2
#L1 - KB1
Ref: #404-IT Pittsburgh