1

Cloud Machine Learning Engineer Jobs in Illinois

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Sr AI Machine Learning Engineer

Chicago, IL ยท Hybrid

$117K - $175K/yr

The Hartford is seeking Senior AI Machine Learning Engineer to build Machine Learning Operations ... Collaborate with partners Enterprise Data, Applied AI, Business, Cloud Enablement Team, and ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Machine Learning Engineer

Chicago, IL ยท On-site

$62K - $100K/yr

As an AI Engineering team member, you will be instrumental in advancing new features and/or solutions from the Proof of Concept stage to full production readiness. Your role involves refining and ...

As an AI Engineering team member, you will be instrumental in advancing new features and/or solutions from the Proof of Concept stage to full production readiness. Your role involves refining and ...

Machine Learning Engineer

Chicago, IL ยท On-site

$62K - $100K/yr

As an AI Engineering team member, you will be instrumental in advancing new features and/or solutions from the Proof of Concept stage to full production readiness. Your role involves refining and ...

Showing results 41-60

Cloud Machine Learning Engineer information

What cities in Illinois are hiring for Cloud Machine Learning Engineer jobs?

Cities in Illinois with the most Cloud Machine Learning Engineer job openings:

Machine Learning Engineer

Springfield, IL โ€ข On-site

Full-time

Re-posted 24 days ago


Job description

  • Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch

  • Build and maintain the infrastructure around RL training: rollout collection, data curation, reward model serving, and experiment orchestration

  • Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues

  • Build evaluation harnesses and benchmark infrastructure, with held-out sets and contamination controls, so results are trustworthy

  • Read eval signal and training curves to determine whether a change actually helped, and feed findings back to the research and environment teams

  • Integrate RL environments into the training stack, working with environment authors on interfaces, reward plumbing, and agent loop mechanics

  • Implement methods from recent ML papers quickly and turn them into production-grade systems