Demonstrated expertise in large language models (LLM) and/or vision language models (VLM). Ways to ... Stand Out from the Crowd: * Deep understanding of GPU architecture, CUDA programming, and system ...
Demonstrated expertise in large language models (LLM) and/or vision language models (VLM). Ways to ... Stand Out from the Crowd: * Deep understanding of GPU architecture, CUDA programming, and system ...
Demonstrated expertise in large language models (LLM) and/or vision language models (VLM). Ways to ... Stand Out from the Crowd: * Deep understanding of GPU architecture, CUDA programming, and system ...
Demonstrated expertise in large language models (LLM) and/or vision language models (VLM). Ways to ... Stand Out from the Crowd: * Deep understanding of GPU architecture, CUDA programming, and system ...
Required to productionize Large Language Model (LLM) based solutions. * Capable of taking an open-source Large Language Model (LLM) and fine tuning it to reflect custom data and retrieve data from a ...
Required to productionize Large Language Model (LLM) based solutions. * Capable of taking an open-source Large Language Model (LLM) and fine tuning it to reflect custom data and retrieve data from a ...
LLM Application Engineer, Artificial Intelligence (AI) Required, Work From Home
San Francisco, CA · On-site +1
You will work at the intersection of Large Language Models (LLMs), software engineering, and ... LLM Application Engineer Responsibilities: - Build and ship LLM-powered applications and AI agent ...
LLM Application Engineer, Artificial Intelligence (AI) Required, Work From Home
San Francisco, CA · On-site +1
You will work at the intersection of Large Language Models (LLMs), software engineering, and ... LLM Application Engineer Responsibilities: - Build and ship LLM-powered applications and AI agent ...
Discovery AI Lab is pioneering large language model (LLM) technologies across the full ML lifecycle - pre-training, mid-training, and post-training - to.
Discovery AI Lab is pioneering large language model (LLM) technologies across the full ML lifecycle - pre-training, mid-training, and post-training - to.
LLM Application Engineer, Artificial Intelligence (AI) Required, Work From Home
San Francisco, CA · Remote
You will work at the intersection of Large Language Models (LLMs), software engineering, and ... LLM Application Engineer Responsibilities: - Build and ship LLM-powered applications and AI agent ...
LLM Application Engineer, Artificial Intelligence (AI) Required, Work From Home
San Francisco, CA · Remote
You will work at the intersection of Large Language Models (LLMs), software engineering, and ... LLM Application Engineer Responsibilities: - Build and ship LLM-powered applications and AI agent ...
LLM Application Engineer, Artificial Intelligence (AI) Required, Work From Home
San Francisco, CA · Remote
You will work at the intersection of Large Language Models (LLMs), software engineering, and ... LLM Application Engineer Responsibilities: - Build and ship LLM-powered applications and AI agent ...
LLM Application Engineer, Artificial Intelligence (AI) Required, Work From Home
San Francisco, CA · Remote
You will work at the intersection of Large Language Models (LLMs), software engineering, and ... LLM Application Engineer Responsibilities: - Build and ship LLM-powered applications and AI agent ...
Manager, New Localization Technology
Burbank, CA · On-site
$90 - $150/hr
Conduct research and experimentation to optimize Large Language Model (LLM) prompts for localization tasks * Collaborate with product development and key stakeholders to prioritize the development ...
Manager, New Localization Technology
Burbank, CA · On-site
$90 - $150/hr
Conduct research and experimentation to optimize Large Language Model (LLM) prompts for localization tasks * Collaborate with product development and key stakeholders to prioritize the development ...
Lead Researcher, Large Language Models/LLM, TikTok
San Jose, CA · On-site
$308K - $588K/yr
We are looking for researchers in LLM, VLM and Omni Model domain who are experienced in single ... large language models; - Explore new model architecture and inference-efficient model design for ...
Lead Researcher, Large Language Models/LLM, TikTok
San Jose, CA · On-site
$308K - $588K/yr
We are looking for researchers in LLM, VLM and Omni Model domain who are experienced in single ... large language models; - Explore new model architecture and inference-efficient model design for ...
LLM Inference Engineer
Menlo Park, CA · On-site
About the Role We're seeking an experienced LLM Inference Engineer to optimize our large language model (LLM) serving infrastructure. The ideal candidate has: * Extensive hands-on experience with ...
LLM Inference Engineer
Menlo Park, CA · On-site
About the Role We're seeking an experienced LLM Inference Engineer to optimize our large language model (LLM) serving infrastructure. The ideal candidate has: * Extensive hands-on experience with ...
Machine Learning Researcher, Foundation Models [SWE Org]
Cupertino, CA · On-site
$150K - $277K/yr
You will see your ideas improve the experience of billions of users.","responsibilities":"In this role, you will focus on pretraining, large language model (LLM) architecture, and scientific scaling ...
Machine Learning Researcher, Foundation Models [SWE Org]
Cupertino, CA · On-site
$150K - $277K/yr
You will see your ideas improve the experience of billions of users.","responsibilities":"In this role, you will focus on pretraining, large language model (LLM) architecture, and scientific scaling ...
Test Engineer-AI/LLM
Palo Alto, CA · On-site
... safety of Large Language Models (LLMs) in real-world product scenarios and test end-to-end ... We are also seeking a Contractor based LLM Evaluation & QA Engineer to support the testing and ...
Test Engineer-AI/LLM
Palo Alto, CA · On-site
... safety of Large Language Models (LLMs) in real-world product scenarios and test end-to-end ... We are also seeking a Contractor based LLM Evaluation & QA Engineer to support the testing and ...
Test Engineer-AI/LLM
Palo Alto, CA · On-site
... safety of Large Language Models (LLMs) in real-world product scenarios and test end-to-end ... We are also seeking a Contractor based LLM Evaluation & QA Engineer to support the testing and ...
Test Engineer-AI/LLM
Palo Alto, CA · On-site
... safety of Large Language Models (LLMs) in real-world product scenarios and test end-to-end ... We are also seeking a Contractor based LLM Evaluation & QA Engineer to support the testing and ...
Test Engineer-AI/LLM
Palo Alto, CA · On-site
... safety of Large Language Models (LLMs) in real-world product scenarios and test end-to-end ... We are also seeking a Contractor based LLM Evaluation & QA Engineer to support the testing and ...
Quick apply
Test Engineer-AI/LLM
Palo Alto, CA · On-site
... safety of Large Language Models (LLMs) in real-world product scenarios and test end-to-end ... We are also seeking a Contractor based LLM Evaluation & QA Engineer to support the testing and ...
Principal - AI Architect - Insurance
San Diego, CA · On-site
$154 - $193/hr
Generative AI and Large Language Model (LLM) based solutions * Predictive and prescriptive Machine Learning models * Computer Vision * Natural Language Processing (NLP) and Conversational AI
Principal - AI Architect - Insurance
San Diego, CA · On-site
$154 - $193/hr
Generative AI and Large Language Model (LLM) based solutions * Predictive and prescriptive Machine Learning models * Computer Vision * Natural Language Processing (NLP) and Conversational AI
Research Scientist - NLP
Sunnyvale, CA · On-site
$150K - $450K/yr
Lead the research of technology for improving the efficiency of Large Language Model (LLM) while performing target capabilities or supporting many capabilities, such as novel architectures and ...
Research Scientist - NLP
Sunnyvale, CA · On-site
$150K - $450K/yr
Lead the research of technology for improving the efficiency of Large Language Model (LLM) while performing target capabilities or supporting many capabilities, such as novel architectures and ...
Lead the research of technology for improving the efficiency of Large Language Model (LLM) while performing target capabilities or supporting many capabilities, such as novel architectures and ...
Quick apply
Lead the research of technology for improving the efficiency of Large Language Model (LLM) while performing target capabilities or supporting many capabilities, such as novel architectures and ...
Senior Software Engineer - TensorRT Edge-LLM
Santa Clara, CA · On-site
$152 - $288/hr
JR2012868Are you passionate about pushing the limits of real-time large language model inference? Join NVIDIA's TensorRT Edge-LLM team and help shape the next generation of edge AI for automotive and ...
Senior Software Engineer - TensorRT Edge-LLM
Santa Clara, CA · On-site
$152 - $288/hr
JR2012868Are you passionate about pushing the limits of real-time large language model inference? Join NVIDIA's TensorRT Edge-LLM team and help shape the next generation of edge AI for automotive and ...
Senior Software Engineer - TensorRT Edge-LLM
Santa Clara, CA · Hybrid
$143K - $189K/yr
Are you passionate about pushing the limits of real-time large language model inference? Join NVIDIA's TensorRT Edge-LLM team and help shape the next generation of edge AI for automotive and robotics.
Senior Software Engineer - TensorRT Edge-LLM
Santa Clara, CA · Hybrid
$143K - $189K/yr
Are you passionate about pushing the limits of real-time large language model inference? Join NVIDIA's TensorRT Edge-LLM team and help shape the next generation of edge AI for automotive and robotics.
Senior Software Engineer - TensorRT Edge-LLM
Santa Clara, CA · On-site
$143K - $189K/yr
Are you passionate about pushing the limits of real-time large language model inference? Join NVIDIA's TensorRT Edge-LLM team and help shape the next generation of edge AI for automotive and robotics.
Senior Software Engineer - TensorRT Edge-LLM
Santa Clara, CA · On-site
$143K - $189K/yr
Are you passionate about pushing the limits of real-time large language model inference? Join NVIDIA's TensorRT Edge-LLM team and help shape the next generation of edge AI for automotive and robotics.
Large Language Model Llm information
What is a large language model llm?
A Large Language Model (LLM) job typically involves working with advanced AI models designed to understand and generate human-like text. Roles in this field may include research, data engineering, model fine-tuning, prompt engineering, or application development. Professionals in LLM jobs often work with machine learning algorithms, natural language processing (NLP), and large-scale datasets to enhance AI capabilities. These roles are common in AI-driven industries, including tech companies, research institutions, and startups. Strong programming skills, knowledge of deep learning frameworks, and expertise in NLP are often required.
What are some common challenges faced by large language model llm engineers in their day-to-day work?
LLM Engineers often encounter challenges related to scaling models efficiently, optimizing performance on large and complex datasets, and ensuring the responsible use of AI technologies. Balancing the trade-offs between model accuracy, speed, and ethical considerations can be demanding, especially as real-world applications often require rapid iterations and rigorous testing. Additionally, staying updated with the latest research advancements and integrating new methods into production systems is an ongoing responsibility. Many engineers tackle these challenges by working closely with data scientists, researchers, and product teams in collaborative, agile environments.
What are the key skills and qualifications needed to thrive in the large language model llm position?
Excelling in the role of a Large Language Model (LLM) Engineer requires strong expertise in natural language processing, machine learning, and computer programming, often supported by an advanced degree in computer science or a related field. Familiarity with industry-standard frameworks like PyTorch or TensorFlow, as well as experience with cloud computing platforms and large-scale data management, is highly valued. Communication, creativity, and problem-solving are essential soft skills to effectively collaborate with cross-functional teams and innovate solutions. These skills ensure the development, deployment, and refinement of powerful language models that can address diverse business needs and technical challenges.
What jobs can I do with large language model?
What are the most commonly searched types of Large Language Model Llm jobs in California?
The most popular types of Large Language Model Llm jobs in California are:
What are popular job titles related to Large Language Model Llm jobs in California?
For Large Language Model Llm jobs in California, the most frequently searched job titles are:
- Python Django Developer
- Senior Full Stack Designer
- Senior Blockchain Engineer
- Temporary Internship Full Stack Software Developer
- Internship Full Stack Developer
- Java Python Developer Seasonal
- Senior Python Full Stack Developer
- Python Game Developer From Home
- Senior Blockchain Developer
- Senior Remote Nodejs Developer
What job categories do people searching Large Language Model Llm jobs in California look for?
The top searched job categories for Large Language Model Llm jobs in California are:
What cities in California are hiring for Large Language Model Llm jobs?
Cities in California with the most Large Language Model Llm job openings:

Manager, Large Language Model Inference
Santa Clara, CA • Hybrid
9.6
Based on 18 frontline employees who took The Breakroom Quiz
7th of 246 rated software companies
Great coworkers
People enjoy working here
Good employer
Respectful managers
Learn new skills
Full-time
Re-posted yesterday
Job description
At NVIDIA, we aren't just powering the AI revolution-we're accelerating it. The TensorRT inference platform is the backbone of modern AI, delivering the industry's fastest and most efficient deployment of cutting-edge deep learning models on every NVIDIA GPU. With demand for AI exploding, particularly in the realm of large language models (LLMs) and vision language models (VLMs, VLAs), we are significantly expanding our team. We're seeking a highly skilled and driven Engineering Manager to take the lead in developing the next generation of LLM/VLM/VLA inference software technologies that will define the future of AI. This is a high-impact, hands-on leadership role at the intersection of deep technical expertise and world-class management. You won't just manage; you'll architect and guide a brilliant team of engineers who are building the core LLM inference runtime. Your work will be highly collaborative, interfacing directly with NVIDIA Researchers, GPU Architects, and other teams across the company to ensure we ship production-grade, lightning-fast software that sets the global standard for AI performance.
What You'll Be Doing:
Lead and grow a team responsible for specialized kernel development, runtime optimizations, and frameworks for LLM inference.
Drive the design, development, and delivery of production inference software, targeting NVIDIA's next-generation enterprise and edge hardware platforms.
Integrating cutting-edge technologies developed at NVIDIA and offering an intuitive developer experience for LLM deployment.
Lead software development execution, with responsibility for project planning, milestone delivery, and cross-functional coordination.
What We Need to See:
MS, PhD, or equivalent experience in Computer Science, Computer Engineering, AI, or a related technical field.
7+ overall years of overall software engineering experience, including 3+ years of technical leadership experience.
Proven ability to lead and scale high-performing engineering teams, especially across distributed and cross-functional groups.
Strong background in C++ or Python, with expertise in software design and delivering production-quality software libraries.
Demonstrated expertise in large language models (LLM) and/or vision language models (VLM).
Ways to Stand Out from the Crowd:
Deep understanding of GPU architecture, CUDA programming, and system-level performance tuning.
Background in LLM inference or working with frameworks such as TensorRT-LLM, vLLM, or SGLang.
Passion for building scalable, user-friendly APIs and enabling developers in the AI ecosystem.
Have a proven track record of growing and managing a team that encourages idea sharing, empowers team members, and provides opportunities for professional growth.
We are widely considered to be one of the technology world's most desirable employers, and we have some of the most forward-thinking and hardworking people in the world working with us. Due to outstanding growth, our best-in-class teams are rapidly growing. If you're a creative self-starter with a real passion for technology, then come join us.
#LI-Hybrid
Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 2, and 224,000 USD - 356,500 USD for Level 3.You will also be eligible for equity and benefits.
This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.About Nvidia
Sourced by ZipRecruiter
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology--and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent.
Industry
Computer and electronic product manufacturing
Company size
10,000+ Employees
Headquarters location
Santa Clara, CA, US