1

Supervised Fine Tuning Jobs (NOW HIRING)

Founding Engineer (AI)

San Francisco, CA · On-site

$150K - $220K/yr

Familiarity and experience with supervised fine-tuning (SFT) and/or reinforcement learning fine tuning (RLFT) on platforms such as fireworks, Together AI, or Baseten * Familiarity with evaluation ...

Senior Data Scientist

Lehi, UT · On-site +1

$133K - $213K/yr

Fine-tune and evaluate foundation models for Entrata-specific use cases using supervised fine-tuning and other post-training methods. * Design and curate high-quality training datasets, including ...

Senior Data Scientist

Lehi, UT · On-site +1

$133K - $213K/yr

Fine-tune and evaluate foundation models for Entrata-specific use cases using supervised fine-tuning and other post-training methods. * Design and curate high-quality training datasets, including ...

Fine-tune and retrain models (LoRA, PEFT, supervised fine-tuning) using data collected from our deployed fleet * Deploy across our inference surfaces: third-party APIs, self-hosted, and on-robot edge

Fine-tune and retrain models (LoRA, PEFT, supervised fine-tuning) using data collected from our deployed fleet * Deploy across our inference surfaces: third-party APIs, self-hosted, and on-robot edge

Fine-tune and retrain models (LoRA, PEFT, supervised fine-tuning) using data collected from our deployed fleet * Deploy across our inference surfaces: third-party APIs, self-hosted, and on-robot edge

Senior Principal Machine Learning Engineer

$128K - $177K/yr

Hands-on supervised fine-tuning of embedding or reranking models with measurable production gains. * Experience with healthcare data (claims, electronic health records, or clinical coding such as ICD ...

Senior Principal Machine Learning Engineer

$128K - $177K/yr

Hands-on supervised fine-tuning of embedding or reranking models with measurable production gains. * Experience with healthcare data (claims, electronic health records, or clinical coding such as ICD ...

Researchers develop and evaluate techniques such as supervised fine-tuning, preference optimization (DPO, RLHF, RLAIF), and continual adaptation to align models with Distyl's enterprise systems. The ...

... supervised fine-tuning. Company : River AI provides an API for fine-tuning and reinforcement learning, enabling users to build and serve personalized AI models. Founded in 2026, the company is ...

Showing results 41-60

Supervised Fine Tuning information

See salary details

$14

$22

$32

How much do supervised fine tuning jobs pay per hour?

As of Sep 14, 2026, the average hourly pay for supervised fine tuning in the United States is $22.57, according to ZipRecruiter salary data. Most workers in this role earn between $18.03 and $30.77 per hour, depending on experience, location, and employer.

What other helpful pages are available for Supervised Fine Tuning?

Other pages related to Supervised Fine Tuning:

Founding Engineer (AI)

San Francisco, CA • On-site

Momentic, Inc.
Advertising and Public Relations Services • 11 - 50 employees

Other

Medical, Dental, Vision, Retirement, PTO

Re-posted 3 days ago


Job description

At Momentic, we’re building the future of quality.

We’re building the all-in-one quality platform powered by state of the art agents to help our customers ensure quality at every stage of the SDLC.

Top engineering teams at companies like Notion, Retool and Webflow use Momentic to ship high quality product. Millions of Momentic tests are executed every single day.

Our product has a very large problem space so there’s a ton of stuff to build and take ownership of - we’d be excited to get your help as we're hiring several extremely talented software engineers across the stack.

The product
  • The AI-native automated testing platform
  • Incredibly sticky. Our customers run us on every pull request and merge and before each deploy as quality gates
  • Leagues ahead of the status quo today (think Selenium/Cypress/Playwright)
  • Check out the demo video on our website.
About us
  • We’re a lean team of 17 (ex Robinhood, Retool, WeWork, Qualtrics, Assembled)
  • Located in-person San Francisco (650 5th St.)
  • Recent Series A raise of $15M led by Standard Capital, with participation from existing investors - Y Combinator, FCVC, and Transpose Platform.
You’re perfect for this role if…
  • You enjoy solving hard problems that have both product and technical ambiguity
  • You have strong engineering fundamentals, code efficiently, and you know what you're great at and what you're less great at
  • You dislike meetings and would much rather focus your time on building, being productive, and shipping code
  • You thrive when you have autonomy, own as many of the details as possible, and project manage your own work
  • You're in SF or you're willing to relocate, you love working in-person, and you're serious about joining us to build a culture we'll all love
Must-have qualifications
  • Experience with LLM performance tuning, including prompt engineering and context management strategies
  • Experience with LLM evaluations and observabilityExperience integrating LLMs into real-world applications using a modern tech stack (Python, TypeScript, etc.)
  • Familiarity and experience with supervised fine-tuning (SFT) and/or reinforcement learning fine tuning (RLFT) on platforms such as fireworks, Together AI, or Baseten
  • Familiarity with evaluation dataset engineering, prompt engineering, LLM as a judge verifiers, and RL environments
  • 3+ years of experience
Nice to have
  • Experience with supervised or reinforced fine-tuning
  • Experience with classical machine learning techniques such as template matching, bounding box detection, and OCR
  • Experience running statistical experiments
Our stack

React, TypeScript, Next.js, Node.js, PostgreSQL, Google Cloud, Kubernetes

Benefits
  • Competitive Medical, Vision, Dental insurance
  • 401K
  • Unlimited PTO
  • Offsites and events (French Laundry last quarter)
  • Fully stocked kitchen (Unlimited supply of sparkling water!)
Sponsorship

We can sponsor TN, E-3, or J-1 Visas. We can't sponsor/transfer H1B Visas or Green Cards (I-140) at this time.

#J-18808-Ljbffr