1

Online Rlhf Jobs (NOW HIRING)

Senior AI Engineer

Los Angeles, CA ยท On-site

$112K - $154K/yr

Own requirements data prep feature engineering classical ML or LLM fine-tuning (LoRA, PEFT, RLHF) offline/online evaluation MLflow registry, with automated drift and quality alerts. * Data & Storage ...

Senior Inference Engineer, AGI

Sunnyvale, CA

$122K - $168K/yr

... RL/RLHF/RLAIF) Ensure train/serve consistency - that the inference path used in RL and evaluation faithfully matches production online behavior (e.g., parity across sampling and logit processing ...

(USA) Principal, Data Scientist

San Mateo, CA ยท On-site

$143K - $286K/yr

Define experimentation strategies, offline evaluation frameworks, and online/offline metric ... or RLHF. * Strong engineering skills in Python and Scala, with experience building large-scale ...

Principal AI Engineer

California, MO ยท On-site

$175.80 - $293/hr

... online evaluators on production traffic, calibrated LLMโ€‘asโ€‘aโ€‘judge graders, and A/B ... Experience with model customization (SFT, RLHF, DPO/GRPO), eval/observability platforms and ...

(USA) Principal, Data Scientist

Sunnyvale, CA ยท On-site

$143K - $286K/yr

Define experimentation strategies, offline evaluation frameworks, and online/offline metric ... or RLHF. * Strong engineering skills in Python and Scala, with experience building large-scale ...

(USA) Principal, Data Scientist

Milpitas, CA ยท On-site

$143K - $286K/yr

Define experimentation strategies, offline evaluation frameworks, and online/offline metric ... or RLHF. * Strong engineering skills in Python and Scala, with experience building large-scale ...

(USA) Principal, Data Scientist

Cupertino, CA ยท On-site

$143K - $286K/yr

Define experimentation strategies, offline evaluation frameworks, and online/offline metric ... or RLHF. * Strong engineering skills in Python and Scala, with experience building large-scale ...

Senior Inference Engineer, AGI

Sunnyvale, CA ยท On-site

$122K - $168K/yr

... RL/RLHF/RLAIF) Ensure train/serve consistency - that the inference path used in RL and evaluation faithfully matches production online behavior (e.g., parity across sampling and logit processing ...

Define experimentation strategies, offline evaluation frameworks, and online/offline metric ... or RLHF. * Strong engineering skills in Python and Scala, with experience building large-scale ...

(USA) Principal, Data Scientist

Hayward, CA ยท On-site

$143K - $286K/yr

Define experimentation strategies, offline evaluation frameworks, and online/offline metric ... or RLHF. * Strong engineering skills in Python and Scala, with experience building large-scale ...

(USA) Principal, Data Scientist

San Jose, CA ยท On-site

$143K - $286K/yr

Define experimentation strategies, offline evaluation frameworks, and online/offline metric ... or RLHF. * Strong engineering skills in Python and Scala, with experience building large-scale ...

(USA) Principal, Data Scientist

Fremont, CA ยท On-site

$143K - $286K/yr

Define experimentation strategies, offline evaluation frameworks, and online/offline metric ... or RLHF. * Strong engineering skills in Python and Scala, with experience building large-scale ...

Showing results 41-60

Online Rlhf information

See salary details

$17.5K

$40.6K

$86K

How much do online rlhf jobs pay per year?

As of Aug 22, 2026, the average yearly pay for online rlhf in the United States is $40,596.00, according to ZipRecruiter salary data. Most workers in this role earn between $25,000.00 and $43,500.00 per year, depending on experience, location, and employer.

What is an online RLHF?

Online RLHF (Reinforcement Learning from Human Feedback) jobs typically involve helping to train AI models by providing human feedback on their outputs. Workers in these roles might review model responses, rate the quality of generated text, or suggest improvements to help the AI learn to produce better results. These jobs are often remote and can be done part-time or as contract work. They play a crucial role in improving the safety, usefulness, and accuracy of AI systems by aligning them more closely with human preferences.

What are some common challenges faced by online RLHF specialists when collaborating with cross-functional teams?

Online RLHF specialists often work closely with machine learning engineers, data annotators, and product managers. A common challenge is ensuring that feedback from human annotators is accurately integrated into model training, which requires clear communication and well-defined annotation guidelines. Additionally, balancing the pace of model updates with the need for high-quality human feedback can be demanding. Effective collaboration and regular syncs are essential to maintain alignment and achieve project goals.

What are the key skills and qualifications needed to thrive as an online RLHF specialist, and why are they important?

To thrive as an Online RLHF Specialist, you need a strong background in machine learning, reinforcement learning, and data analysis, typically supported by a degree in computer science or a related field. Familiarity with technical tools like Python, PyTorch or TensorFlow, and experience with human feedback systems or annotation platforms are highly valuable. Strong problem-solving, attention to detail, and the ability to communicate complex concepts clearly are crucial soft skills. These qualifications ensure the effective training and evaluation of AI models, leading to more accurate and reliable machine learning systems.

What is the difference between Online Rlhf vs Online Rlhf?

AspectOnline RlhfOnline Rlhf
CredentialsTypically requires certification in online health coaching or related fieldsTypically requires certification in online health coaching or related fields
Work EnvironmentRemote, online platform-basedRemote, online platform-based
Industry UsageCommon in health and wellness sectorsCommon in health and wellness sectors
Job FocusProviding health guidance and support onlineProviding health guidance and support online

Online Rlhf and Online Rlhf are the same role, often used interchangeably. Both involve providing health and wellness support remotely, requiring similar certifications and working within the online health industry. The key difference is often in terminology rather than job function.

More about Online Rlhf jobs

What cities are hiring for Online Rlhf jobs?

Cities with the most Online Rlhf job openings:

What are the most commonly searched types of Rlhf jobs?

The most popular types of Rlhf jobs are:

What states have the most Online Rlhf jobs?

States with the most job openings for Online Rlhf jobs include:

Infographic showing various Online Rlhf job openings in the United States as of August 2026, with employment types broken down into 61% Full Time, 36% Part Time, 1% Temporary, and 2% Contract. Highlights an 80% Physical, 1% Hybrid, and 19% Remote job distribution, with an average salary of $40,596 per year, or $19.5 per hour.

Senior AI Engineer

Hop

Los Angeles, CA โ€ข On-site

$112K - $154K/yr

Full-time

Re-posted 29 days ago


Job description

Our client is a family office management company serving investments, foundations, and activities of a prominent family. With a broad mandate, their organization oversees diverse assets and programs, including multiple foundations and institutes. Across their entities, they manage hundreds of employees and oversee significant annual expenditures, ranging from grants and gifts to private investments and operational costs.

They are seeking a highly motivated, innovative, and collaborative Technology staff member to serve as the Senior AI Engineer. The selected candidate will be a member of the Enterprise Technology Data Engineering & AI team, playing a pivotal role in driving innovation across the organization.

Summary

You will architect and develop production-grade LLM agents and RAG pipelines, steer the full ML lifecycle from data prep to GPU-scaled deployment, and weave together modern tools and technologies into a secure, cost-aware platform. If you thrive on turning ambiguous ideas into high-impact GenAI products and mentoring others to do the same, this is your playground.
Responsibilities:
  • Build & Ship Gen AI Apps:Design, prototype, and build GenAI solutions, RAG document pipelines, and task-specific agents to support multiple business functions using tools such as LangChain/LlamaIndex, micro-services, Ray/KubeRay.
  • Agent Workflow Pipelines:Design and orchestrate multi-step agent pipelines, integrating LLM prompts, external APIs, and human-in-the-loop escalations.
  • End-to-End ML Lifecycle:Own requirements data prep feature engineering classical ML or LLM fine-tuning (LoRA, PEFT, RLHF) offline/online evaluation MLflow registry, with automated drift and quality alerts.
  • Data & Storage Architecture:Ingest from BigQuery, object-store lakes (Parquet, Avro); generate embeddings and persist to vector DBs (Qdrant/PgVector); enforce governance via OpenMetadata and column-level ACLs.
  • Scalable Deployment & Ops:Package with Docker, helm-deploy on Kubernetes; implement GPU scheduling, autoscaling, blue-green rollouts, and cost telemetry via Prometheus/Grafana; automate CI/CD in GitHub Actions.
  • Observability & Compliance:Instrument tracking, metrics, and structured logs; run A/B or shadow tests; embed security, privacy, and cost-guardrails in every pipeline.
  • Lead & Mentor:Translate ambiguous business ideas into executable roadmaps, run build-vs-buy analysis, set code standards, and coach peers on agentic patterns and ethical AI.
Requirements:
  • Bachelor's or Master's in Computer Science, Data Science, or equivalent experience.
  • 7+ years designing and shipping ML/AI applications, including 2+ years with LLMs or Generative AI.
  • Demonstrated delivery of RAG or agentic systems in production (e.g. LangChain, LlamaIndex, n8n, or custom).
  • Expert-level Python and SQL; strong Spark, distributed data-processing, and performance-tuning skills.
  • Hands-on fine-tuning of foundation models; comfort with MLflow, Ray/KubeRay, and vector databases.
  • Deep familiarity with cloud warehouses (BigQuery, Redshift), lake formats (Parquet, Avro), and streaming/ingestion tools (e.g. Airbyte, Kafka/Pub-Sub).
  • Production experience with Docker, Kubernetes, Helm, and Git-based CI/CD pipelines.
  • Clear communicator able to gather requirements, set technical direction, and influence cross-functional teams.
Additional Details:
  • Only open to U.S. Citizens or Green Card holders.
  • The role is in-office (LA) - West Hollywood.
  • Compensation includes a strong base + bonus (no equity, as they're private).
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
apply for this job

Hop logo

About Hop

Sourced by ZipRecruiter

Industry

Software development

Company size

11 - 50 Employees

Headquarters location

West Hollywood, CA, US

Year founded

2020