1

Machine Learning Engineer Two Jobs in Georgia (NOW HIRING)

* Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch * Build and maintain the infrastructure around RL training: rollout collection ...

* Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch * Build and maintain the infrastructure around RL training: rollout collection ...

* Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch * Build and maintain the infrastructure around RL training: rollout collection ...

* Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch * Build and maintain the infrastructure around RL training: rollout collection ...

* Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch * Build and maintain the infrastructure around RL training: rollout collection ...

* Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch * Build and maintain the infrastructure around RL training: rollout collection ...

* Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch * Build and maintain the infrastructure around RL training: rollout collection ...

* Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch * Build and maintain the infrastructure around RL training: rollout collection ...

* Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch * Build and maintain the infrastructure around RL training: rollout collection ...

* Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch * Build and maintain the infrastructure around RL training: rollout collection ...

* Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch * Build and maintain the infrastructure around RL training: rollout collection ...

* Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch * Build and maintain the infrastructure around RL training: rollout collection ...

* Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch * Build and maintain the infrastructure around RL training: rollout collection ...

* Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch * Build and maintain the infrastructure around RL training: rollout collection ...

* Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch * Build and maintain the infrastructure around RL training: rollout collection ...

* Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch * Build and maintain the infrastructure around RL training: rollout collection ...

Showing results 41-60

Machine Learning Engineer Two information

Are machine learning engineers still in demand?

Yes, machine learning engineers are in high demand across various industries due to the increasing adoption of AI and data-driven solutions. They are sought after for their skills in algorithms, programming, and tools like Python and TensorFlow, with job growth expected to continue as AI applications expand.

What cities in Georgia are hiring for Machine Learning Engineer Two jobs?

Cities in Georgia with the most Machine Learning Engineer Two job openings:

Machine Learning Engineer

Kennesaw, GA โ€ข On-site

Full-time

Re-posted 23 days ago


Job description

  • Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch

  • Build and maintain the infrastructure around RL training: rollout collection, data curation, reward model serving, and experiment orchestration

  • Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues

  • Build evaluation harnesses and benchmark infrastructure, with held-out sets and contamination controls, so results are trustworthy

  • Read eval signal and training curves to determine whether a change actually helped, and feed findings back to the research and environment teams

  • Integrate RL environments into the training stack, working with environment authors on interfaces, reward plumbing, and agent loop mechanics

  • Implement methods from recent ML papers quickly and turn them into production-grade systems