Experience with NVidia, Cray, managing clusters, Slurm, Linux, CUDA, high-speed networks as well as an understanding of the customer's A&A process. Qualifications: * Active Top Secret/Sensitive ...
Experience with NVidia, Cray, managing clusters, Slurm, Linux, CUDA, high-speed networks as well as an understanding of the customer's A&A process. Qualifications: * Active Top Secret/Sensitive ...
Experience with NVidia, Cray, managing clusters, Slurm, Linux, CUDA, high-speed networks as well as an understanding of the customer's A&A process.
Experience with NVidia, Cray, managing clusters, Slurm, Linux, CUDA, high-speed networks as well as an understanding of the customer's A&A process.
Experience with NVidia, Cray, managing clusters, Slurm, Linux, CUDA, high-speed networks as well as an understanding of the customer's A&A process.
Experience with NVidia, Cray, managing clusters, Slurm, Linux, CUDA, high-speed networks as well as an understanding of the customer's A&A process.
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Hampton, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Hampton, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Richmond, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Richmond, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Richmond, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Richmond, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Norfolk, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Norfolk, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Chesapeake, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Chesapeake, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Hampton, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Hampton, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Lynchburg, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Lynchburg, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Norfolk, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Norfolk, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Lynchburg, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Lynchburg, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Annandale, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Annandale, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Chesapeake, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Machine Learning Engineer
Chesapeake, VA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...
Slurm information
What are popular job titles related to Slurm jobs in Virginia?
For Slurm jobs in Virginia, the most frequently searched job titles are:
What job categories do people searching Slurm jobs in Virginia look for?
The top searched job categories for Slurm jobs in Virginia are:
What cities in Virginia are hiring for Slurm jobs?
Cities in Virginia with the most Slurm job openings:

Job description
iota IT, a subsidiary of VTG, is seeking a High-Performance Computing Engineer in McLean, VA.
What will you do?
The highly skilled High Performance Computing Engineer to design, implement, and maintain advanced computing solutions in support of mission-critical operations.
The ideal candidate will bring deep technical expertise in large-scale HPC environments, cluster management, GPU computing, and high-speed networking.
Do you have what it takes?
- Active Top Secret/Sensitive Compartmented Information (TS/SCI) clearance, with polygraph.
- Bachelor's Degree in Computer Science, Engineering or related field.
- Experience with NVidia, Cray, managing clusters, Slurm, Linux, CUDA, high-speed networks as well as an understanding of the customer's A&A process.
- Active Top Secret/Sensitive Compartmented Information (TS/SCI) clearance, with polygraph.
- Bachelor's Degree in Computer Science, Engineering or related field.
- Experience with NVidia, Cray, managing clusters, Slurm, Linux, CUDA, high-speed networks as well as an understanding of the customer's A&A process.