... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
San Diego, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
San Diego, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
San Jose, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
San Jose, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
Los Angeles, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
Los Angeles, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
Bakersfield, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
Bakersfield, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
San Luis Obispo, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
San Luis Obispo, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
Fresno, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
Fresno, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
Chico, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
Chico, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
Sacramento, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
Sacramento, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
Merced, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
Merced, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
Northridge, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Machine Learning Engineer
Northridge, CA · On-site
... data curation, reward model serving, and experiment orchestration * Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
Convergence Data information
What are popular job titles related to Convergence Data jobs in California?
For Convergence Data jobs in California, the most frequently searched job titles are:
What job categories do people searching Convergence Data jobs in California look for?
The top searched job categories for Convergence Data jobs in California are:
What cities in California are hiring for Convergence Data jobs?
Cities in California with the most Convergence Data job openings:
Full-time
Re-posted 20 days ago
Job description
-
Own model training and post-training pipelines end to end: SFT, RLHF, PPO, DPO, and reward model training in PyTorch
-
Build and maintain the infrastructure around RL training: rollout collection, data curation, reward model serving, and experiment orchestration
-
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues
-
Build evaluation harnesses and benchmark infrastructure, with held-out sets and contamination controls, so results are trustworthy
-
Read eval signal and training curves to determine whether a change actually helped, and feed findings back to the research and environment teams
-
Integrate RL environments into the training stack, working with environment authors on interfaces, reward plumbing, and agent loop mechanics
-
Implement methods from recent ML papers quickly and turn them into production-grade systems