OR · On-site
$122K - $161K/yr
Designing and implementing extensible abstractions for LLM serving engines * Building efficient just-in-time domain specific compilers and runtimes * Collaborating closely with other engineers at ...
OR · On-site
$122K - $161K/yr
Designing and implementing extensible abstractions for LLM serving engines * Building efficient just-in-time domain specific compilers and runtimes * Collaborating closely with other engineers at ...
OR · On-site +1
$102K - $134K/yr
They understand commercial LLM platforms, enjoy translating engineering domain knowledge into automated tools, and are comfortable traveling to engineering sites to implement and support new AI ...
New
OR · On-site +1
$102K - $134K/yr
They understand commercial LLM platforms, enjoy translating engineering domain knowledge into automated tools, and are comfortable traveling to engineering sites to implement and support new AI ...
New
Portland, OR · On-site
$108K - $143K/yr
Architect and oversee delivery of LLM-enabled applications including copilots, agentic workflows ... Engineering & Data Foundations * Review and contribute to production-quality code * Guide ...
Portland, OR · On-site
$108K - $143K/yr
Architect and oversee delivery of LLM-enabled applications including copilots, agentic workflows ... Engineering & Data Foundations * Review and contribute to production-quality code * Guide ...
Portland, OR · On-site
$129K - $171K/yr
... Engineer will be responsible for designing and shipping new capabilities for Command ... maintaining the LLM layer, and ensuring the reliability and quality of the product.
Portland, OR · On-site
$129K - $171K/yr
... Engineer will be responsible for designing and shipping new capabilities for Command ... maintaining the LLM layer, and ensuring the reliability and quality of the product.
Portland, OR · On-site
$110K - $152K/yr
Cortex AI, Cortex LLM Functions, Cortex Agents, Arctic Embed * 1+ years of experience leading ... Data engineering experience with Spark, Airflow/dbt, streaming, data modeling or ML/data science ...
Portland, OR · On-site
$110K - $152K/yr
Cortex AI, Cortex LLM Functions, Cortex Agents, Arctic Embed * 1+ years of experience leading ... Data engineering experience with Spark, Airflow/dbt, streaming, data modeling or ML/data science ...
OR · On-site +1
Senior Support Engineer Team: Global Services & Delivery Location: [Remote / Hybrid - specify ... Demonstrated experience building or tuning automation or agentic and LLM-based support tooling, or ...
OR · On-site +1
Senior Support Engineer Team: Global Services & Delivery Location: [Remote / Hybrid - specify ... Demonstrated experience building or tuning automation or agentic and LLM-based support tooling, or ...
$107K - $140K/yr
Experience operating LLM/AI-assisted developer tooling at scale (Claude Code, Copilot, or similar ... inside an enterprise is preferred. * Familiarity with Okta/OIDC and enterprise auth patterns for ...
New
$107K - $140K/yr
Experience operating LLM/AI-assisted developer tooling at scale (Claude Code, Copilot, or similar ... inside an enterprise is preferred. * Familiarity with Okta/OIDC and enterprise auth patterns for ...
New
As a Staff AI-First Information Security Engineer, you own the intersection of AI adoption and ... Conduct threat modelling on agentic and LLM-based systems, accounting for novel attack surfaces ...
As a Staff AI-First Information Security Engineer, you own the intersection of AI adoption and ... Conduct threat modelling on agentic and LLM-based systems, accounting for novel attack surfaces ...
OR · On-site
As a Staff AI-First Information Security Engineer, you own the intersection of AI adoption and ... Conduct threat modelling on agentic and LLM-based systems, accounting for novel attack surfaces ...
The AI Solutions Engineer also ensures solutions are developed in accordance with applicable ... Design, develop, test, deploy, and maintain LLM-powered applications, integrations, and agentic AI ...
The AI Solutions Engineer also ensures solutions are developed in accordance with applicable ... Design, develop, test, deploy, and maintain LLM-powered applications, integrations, and agentic AI ...
Portland, OR · On-site
MLOps Engineer Location: Portland, OR (Complete Onsite) Note: Client Interview Face to Face Key ... Experience with Generative AI, LLM deployment, or RAG-based applications. * Familiarity with Apache ...
New
Portland, OR · On-site
MLOps Engineer Location: Portland, OR (Complete Onsite) Note: Client Interview Face to Face Key ... Experience with Generative AI, LLM deployment, or RAG-based applications. * Familiarity with Apache ...
New
Develop and extend production services in Python (FastAPI) and pydantic.ai for LLM-powered ... Partner with engineers across Enterprise Data, Enterprise Apps, Product Development, and Infosec to ...
Develop and extend production services in Python (FastAPI) and pydantic.ai for LLM-powered ... Partner with engineers across Enterprise Data, Enterprise Apps, Product Development, and Infosec to ...
OR · On-site
$108K - $147K/yr
Experience operating LLM/AI-assisted developer tooling at scale (Claude Code, Copilot, or similar ... inside an enterprise is preferred. * Familiarity with Okta/OIDC and enterprise auth patterns for ...
New
OR · On-site
$108K - $147K/yr
Experience operating LLM/AI-assisted developer tooling at scale (Claude Code, Copilot, or similar ... inside an enterprise is preferred. * Familiarity with Okta/OIDC and enterprise auth patterns for ...
New
Experience integrating LLM solutions with enterprise systems via APIs, microservices, or event ... Solution Engineering * Build AI-enabled solutions, agentic platforms, and workflows across ...
Experience integrating LLM solutions with enterprise systems via APIs, microservices, or event ... Solution Engineering * Build AI-enabled solutions, agentic platforms, and workflows across ...
OR · On-site
$104K - $143K/yr
Develop and extend production services in Python (FastAPI) and pydantic.ai for LLM-powered ... Partner with engineers across Enterprise Data, Enterprise Apps, Product Development, and Infosec to ...
OR · On-site
$104K - $143K/yr
Develop and extend production services in Python (FastAPI) and pydantic.ai for LLM-powered ... Partner with engineers across Enterprise Data, Enterprise Apps, Product Development, and Infosec to ...
OR · Remote
$55.25 - $71.25/hr
The Opportunity Grafana Labs is seeking a Senior Engineer (AI & Automation) to own the AI agent ... You'll build multi-agent architectures, LLM integrations, and backend services that connect AI ...
OR · Remote
$55.25 - $71.25/hr
The Opportunity Grafana Labs is seeking a Senior Engineer (AI & Automation) to own the AI agent ... You'll build multi-agent architectures, LLM integrations, and backend services that connect AI ...
$139K - $168K/yr
Our small engineering team works on challenging problems every day. We have a culture that's rooted ... Experience of transformer models and LLM applications * Strong knowledge of Python or C++, or the ...
$139K - $168K/yr
Our small engineering team works on challenging problems every day. We have a culture that's rooted ... Experience of transformer models and LLM applications * Strong knowledge of Python or C++, or the ...
$139K - $168K/yr
Our small engineering team works on challenging problems every day. We have a culture that's rooted ... Experience of transformer models and LLM applications * Strong knowledge of Python or C++, or the ...
$139K - $168K/yr
Our small engineering team works on challenging problems every day. We have a culture that's rooted ... Experience of transformer models and LLM applications * Strong knowledge of Python or C++, or the ...
$139K - $168K/yr
Our small engineering team works on challenging problems every day. We have a culture that's rooted ... Experience of transformer models and LLM applications * Strong knowledge of Python or C++, or the ...
$139K - $168K/yr
Our small engineering team works on challenging problems every day. We have a culture that's rooted ... Experience of transformer models and LLM applications * Strong knowledge of Python or C++, or the ...
Eugene, OR · On-site
$139K - $168K/yr
Our small engineering team works on challenging problems every day. We have a culture that's rooted ... Experience of transformer models and LLM applications * Strong knowledge of Python or C++, or the ...
Eugene, OR · On-site
$139K - $168K/yr
Our small engineering team works on challenging problems every day. We have a culture that's rooted ... Experience of transformer models and LLM applications * Strong knowledge of Python or C++, or the ...
$26.94 - $31.86
1% of jobs
$31.86 - $36.78
5% of jobs
$36.78 - $41.70
9% of jobs
$45.95 is the 25th percentile. Wages below this are outliers.
$41.70 - $46.63
12% of jobs
$46.63 - $51.55
10% of jobs
The median wage is $56.12 / hr.
$51.55 - $56.47
15% of jobs
$56.47 - $61.39
15% of jobs
$64.88 is the 75th percentile. Wages above this are outliers.
$61.39 - $66.31
13% of jobs
$66.31 - $71.23
10% of jobs
$71.23 - $76.15
10% of jobs
$76.15 - $81.08
2% of jobs
$26
$56
$81
An LLM Engineer designs, develops, and optimizes applications that leverage large language models (LLMs). They fine-tune models, integrate them into products, and improve performance through prompt engineering and model customization. This role requires expertise in machine learning, natural language processing (NLP), and software development. LLM Engineers work closely with data scientists and developers to create AI-driven solutions for various applications such as chatbots, content generation, and code assistance.
To thrive as an LLM Engineer, you need strong expertise in machine learning, natural language processing, and proficiency with Python, along with a solid understanding of transformer-based models and deep learning frameworks like PyTorch or TensorFlow. Familiarity with cloud platforms, version control systems (e.g., Git), and tools such as Hugging Face Transformers is typically required, and certifications in AI or data science can be advantageous. Excellent problem-solving, collaboration, and communication skills help you work effectively with interdisciplinary teams and present complex findings clearly. These skills enable you to develop, fine-tune, and deploy large language models efficiently in real-world applications.

$122K - $161K/yr
Full-time
Re-posted 27 days ago
9.6
Based on 17 frontline employees who took The Breakroom Quiz
8th of 242 rated software companies
We're looking for outstanding AI systems engineers to develop groundbreaking technologies in the inference systems software stack! We build innovative AI systems software to accelerate for AI inference. As a member of the team, you'll develop libraries, code generators, and GPU kernel technologies for NVIDIA's hardware architecture. This means designing and building things like new abstractions, efficient attention kernel implementations, new LLM inference runtimes components, and kernel code generators to accelerate large language models, agents, and other high-impact AI workloads.
What you'll be doing:
Innovating and developing new AI systems technologies for efficient inference
Designing, implementing, and optimizing kernels for high impact AI workloads
Designing and implementing extensible abstractions for LLM serving engines
Building efficient just-in-time domain specific compilers and runtimes
Collaborating closely with other engineers at NVIDIA across deep learning frameworks, libraries, kernels, and GPU arch teams
Contributing to open source communities like FlashInfer, vLLM, and SGLang
What we need to see:
Masters degree in Computer Science, Electrical Engineering, or related field (or equivalent experience); PhD are preferred
6+ years (academic/ industry) experience with ML/DL systems development preferable
Strong experience in developing or using deep learning frameworks (e.g. PyTorch, JAX, TensorFlow, ONNX, etc) and ideally inference engines and runtimes such as vLLM, SGLang, and MLC.
Strong Python and C/C++ programming skills
Strong experience in GPU kernel development and performance optimizations (especially using CUDA C/C++, cuTile, Triton, or similar) with hands-on experience with Matrix Multiplication
Ways to stand out from the crowd:
Background in domain specific compiler and library solutions for LLM inference and training (e.g. FlashInfer, Flash Attention)
Expertise in inference engines like vLLM and SGLang
Expertise in machine learning compilers (e.g. Apache TVM, MLIR)
Open source project ownership or contributions
You will also be eligible for equity and benefits.
This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.Sourced by ZipRecruiter
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology--and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent.
Computer and electronic product manufacturing
10,000+ Employees
Santa Clara, CA, US
1993