1

Weekend Llm Training Jobs in California (NOW HIRING)

Utilize approved AI and LLM/machine learning tools to support security analysis, documentation ... Provide information security and compliance-related support and training as required of work ...

Evaluate, pilot, and operationalize AI-powered workflows and tools (including LLM-assisted ... Create and maintain high-quality documentation, runbooks, and knowledge articles; provide training ...

Staff Data Architect

Long Beach, CA · On-site

$155K - $221K/yr

Implement MLOps infrastructure to support model training, deployment, and monitoring * Build real ... Familiarity with vector databases, knowledge graphs, and AI/LLM data architectures * Understanding ...

next page

Showing results 1-20

Weekend Llm Training information

What are the most commonly searched types of Llm Training jobs in California?

The most popular types of Llm Training jobs in California are:

What cities in California are hiring for Weekend Llm Training jobs?

Cities in California with the most Weekend Llm Training job openings:

Research Intern, Agent RL Training

NewsBreak

Mountain View, CA • On-site

$35 - $50/hr

Full-time, Internship

Re-posted 8 days ago


Job description

About the Role

We are looking for a Research Intern to join our Agent RL Training team. You will be paired with a full-time employee as your mentor, working together to explore, from zero to one, how to apply large language models to NewsBreak's core business, including content understanding, recommendation, agentic web browsing, and autonomous multi-step task completion.

This is a hands-on research role. You are expected to independently drive experiments, propose novel ideas, and iterate quickly. We value self-starters with deep intellectual curiosity and the drive to push boundaries in LLM post-training and agent capabilities.

Location: Onsite in Mountain View, CA office

What You'll Work On
  • Collaborate with your full-time mentor to identify high-impact research directions for applying LLMs to NewsBreak's products
  • Independently run end-to-end SFT experiments on LLM-based agents, and assist with RL-related exploration such as reward design and training iteration
  • Curate and build high-quality training datasets: instruction-following, preference pairs, agent trajectories, and synthetic data
  • Contribute to public publications; we encourage and support top-venue submissions during your internship
What We're Looking ForRequirements
  • Highly motivated and committed: willing to put in extra hours when needed to push projects across the finish line
  • Genuine passion for research: you read papers for fun, tinker with models on weekends, and care deeply about advancing the field
  • Independently capable of end-to-end model SFT: with basic understanding of RL-based post-training methods (RLHF, DPO, PPO, GRPO, etc.)
  • Excellent taste in model behavior: able to reason about what "good" looks like across user-facing domains and articulate why
  • Strong Python and PyTorch skills
Preferred Qualifications
  • Publication at a top-tier venue (NeurIPS, ICML, ICLR, ACL, EMNLP, or equivalent)
  • Experience with multi-node distributed training (FSDP, DeepSpeed, Megatron-LM)
  • Proficiency in writing custom GPU kernels with Triton or CUDA
  • Experience building synthetic data pipelines for agent training
  • Familiarity with open-source RL frameworks: TRL, OpenRLHF, veRL/vLLM

Hourly Pay:  $35- $50