1

Large Language Model Llm Jobs in California (NOW HIRING)

The successful candidate will lead the implementation of a Large Language Model (LLM) that will transform the way this team performs its work. The Director also oversees budgeting, talent development ...

Controller

San Francisco, CA · On-site

$125/hr

Test and review outputs generated by AI and Large Language Model (LLM) technologies. What We're Looking For Required Qualifications * 7+ years of progressive accounting or finance experience.

Test and review outputs generated by AI and Large Language Model (LLM) technologies. What We're Looking ForRequired Qualifications * 7+ years of progressive accounting or finance experience.

Hands-on experience with large language model (LLM) inference and/or model training using open-source model frameworks * Strong Python programming skills * Experience working with cloud platforms ...

Showing results 21-40

Large Language Model Llm information

What is a large language model llm?

A Large Language Model (LLM) job typically involves working with advanced AI models designed to understand and generate human-like text. Roles in this field may include research, data engineering, model fine-tuning, prompt engineering, or application development. Professionals in LLM jobs often work with machine learning algorithms, natural language processing (NLP), and large-scale datasets to enhance AI capabilities. These roles are common in AI-driven industries, including tech companies, research institutions, and startups. Strong programming skills, knowledge of deep learning frameworks, and expertise in NLP are often required.

What are some common challenges faced by large language model llm engineers in their day-to-day work?

LLM Engineers often encounter challenges related to scaling models efficiently, optimizing performance on large and complex datasets, and ensuring the responsible use of AI technologies. Balancing the trade-offs between model accuracy, speed, and ethical considerations can be demanding, especially as real-world applications often require rapid iterations and rigorous testing. Additionally, staying updated with the latest research advancements and integrating new methods into production systems is an ongoing responsibility. Many engineers tackle these challenges by working closely with data scientists, researchers, and product teams in collaborative, agile environments.

What are the key skills and qualifications needed to thrive in the large language model llm position?

Excelling in the role of a Large Language Model (LLM) Engineer requires strong expertise in natural language processing, machine learning, and computer programming, often supported by an advanced degree in computer science or a related field. Familiarity with industry-standard frameworks like PyTorch or TensorFlow, as well as experience with cloud computing platforms and large-scale data management, is highly valued. Communication, creativity, and problem-solving are essential soft skills to effectively collaborate with cross-functional teams and innovate solutions. These skills ensure the development, deployment, and refinement of powerful language models that can address diverse business needs and technical challenges.

What jobs can I do with large language model?

Large Language Models (LLMs) are used in roles such as AI research scientist, NLP engineer, data scientist, and machine learning engineer. These jobs involve developing, fine-tuning, and deploying LLMs for applications like chatbots, content generation, and language understanding, often requiring skills in programming, data analysis, and deep learning frameworks.

What are the most commonly searched types of Large Language Model Llm jobs in California?

The most popular types of Large Language Model Llm jobs in California are:

What job categories do people searching Large Language Model Llm jobs in California look for?

The top searched job categories for Large Language Model Llm jobs in California are:

What cities in California are hiring for Large Language Model Llm jobs?

Cities in California with the most Large Language Model Llm job openings:

Infographic showing various Large Language Model Llm job openings in California as of August 2026, with employment types broken down into 1% As Needed, 82% Full Time, 15% Part Time, and 2% Contract. Highlights an 89% Physical, 3% Hybrid, and 8% Remote job distribution.

Engineering Manager, LLM Performance

Nvidia

Santa Clara, CA • Hybrid

Full-time

Re-posted 10 days ago


Nvidia rating

9.6

Company rating: 9.6 out of 10

Based on 18 frontline employees who took The Breakroom Quiz

6th of 247 rated software companies


Job description

At NVIDIA, wearen'tjust powering the AI revolution-we'reaccelerating it.We are accelerating LLM inference across the stackand across allopen sourceLLM frameworks like TensorRT LLM,vLLMandSGLang.With demand for AI exploding, particularly in the realm of large language models (LLMs) and vision language models (VLMs, VLAs), we are significantly expanding our team.

We'reseekinga highly skilled and driven Engineering Manager to take the lead inacceleratingthe next generation of LLM/VLM/VLA inference software technologies that will define the future of AI. This is a high-impact, hands-on leadership role at the intersection of deep technicalexpertiseand world-class management. Youwon'tjust manage;you'llarchitect and guide a brilliant team of engineers who arepushing the performance ofLLM inference. Your work will be highly collaborative, interfacing directly with NVIDIA Researchers, GPU Architects, and other teams across the company to ensure we ship production-grade, lightning-fast software that sets the global standard for AI performance.

WhatYou'llBe Doing:

  • Lead and grow a team responsible forpushing the performance of LLM inference across multiple LLM frameworks, including TensorRT LLM,vLLM,SGLangand Dynamoon our datacenter products.

  • Drive the design,implementationand optimization of features that are key to performance in LLM inference.

  • Continuously improve the performance of LLM inference on current and upcoming NVIDIA datacenter architectures and GPUs.

  • Continuously improvethe performance of LLM inference ofimportant foundation models.

  • Work with inference benchmark teams to helptune performance for key workloads.

  • Integratingcutting-edgetechnologies developed at NVIDIA and offering an intuitive developer experience for LLM deployment.

  • Lead software development execution, with responsibility for project planning, milestone delivery, and cross-functional coordination.

What We Need to See:

  • MS, PhD, or equivalent experience in Computer Science, Computer Engineering, AI, ora relatedtechnical field.

  • 7+ overall years of overall software engineering experience, including 3+ years of technical leadership experience.

  • Proven ability to lead and scale high-performing engineering teams, especially across distributed and cross-functional groups.

  • Strong background in C++ or Python, withexpertisein software design and delivering production-quality software libraries.

  • Demonstratedexpertisein large language models (LLM) and/or vision language models (VLM)and/or inference in general.

Ways to Stand Out from the Crowd:

  • Deep understanding of GPU architecture, CUDA programming, and system-level performance tuning.

  • Background in LLM inference or working with frameworks such as TensorRT-LLM,vLLM, orSGLang.

  • Passion for building scalable, user-friendly APIs and enabling developers in the AI ecosystem.

  • Have a proventrack recordof growing and managing a team that encourages idea sharing, empowers team members, and provides opportunities for professional growth.

#LI-Hybrid

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 224,000 USD - 356,500 USD for Level 3, and 272,000 USD - 431,250 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 9, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

What Nvidia employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Nvidia logo

About Nvidia

Sourced by ZipRecruiter

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology--and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent.

Industry

Computer and electronic product manufacturing

Company size

10,000+ Employees

Headquarters location

Santa Clara, CA, US