Experience with GPU kernel optimization is a plus. * Excellent Python or C++ programming skills, including debugging, profiling, and performance analysis at scale. * Experience with ML infra at ...
Experience with GPU kernel optimization is a plus. * Excellent Python or C++ programming skills, including debugging, profiling, and performance analysis at scale. * Experience with ML infra at ...
SW Optimization Engineer AI/ML
Cupertino, CA · On-site
$172K/yr
Our team is driving performance enhancements in application and system software and developing novel algorithms to deliver integrated, highly optimized solutions based on Apple Silicon. Description ...
New
SW Optimization Engineer AI/ML
Cupertino, CA · On-site
$172K/yr
Our team is driving performance enhancements in application and system software and developing novel algorithms to deliver integrated, highly optimized solutions based on Apple Silicon. Description ...
New
Senior Software Engineer, C/C++ SDK Performance Optimization
San Jose, CA · On-site
$218K - $387K/yr
Responsibilities - Drive end-to-end experience optimization of the effect pipeline during shooting ... performance on mobile devices - you enjoy digging into profilers and chasing down the last ...
Senior Software Engineer, C/C++ SDK Performance Optimization
San Jose, CA · On-site
$218K - $387K/yr
Responsibilities - Drive end-to-end experience optimization of the effect pipeline during shooting ... performance on mobile devices - you enjoy digging into profilers and chasing down the last ...
Experience with GPU kernel optimization is a plus. * Excellent Python or C++ programming skills, including debugging, profiling, and performance analysis at scale. * Experience with ML infra at ...
New
Experience with GPU kernel optimization is a plus. * Excellent Python or C++ programming skills, including debugging, profiling, and performance analysis at scale. * Experience with ML infra at ...
New
Experience with GPU kernel optimization is a plus. * Excellent Python or C++ programming skills, including debugging, profiling, and performance analysis at scale. * Experience with ML infra at ...
Experience with GPU kernel optimization is a plus. * Excellent Python or C++ programming skills, including debugging, profiling, and performance analysis at scale. * Experience with ML infra at ...
Software Engineer, Kernel Performance & AI Tooling
San Francisco, CA · On-site
$164K/yr
Responsibilities : • Build developer tooling and workflows that make kernel development and performance optimization faster, more scalable, and easier to debug, integrate, and deploy. • Develop ...
Software Engineer, Kernel Performance & AI Tooling
San Francisco, CA · On-site
$164K/yr
Responsibilities : • Build developer tooling and workflows that make kernel development and performance optimization faster, more scalable, and easier to debug, integrate, and deploy. • Develop ...
Senior Principal Machine Learning Engineer - Optimization
Redwood City, CA · On-site +1
$153K - $211K/yr
Experience building large-scale prediction or optimization systems PubMatic is the leading AI ... Partner with performance advertising signal engineers to define model-ready features, labels ...
Senior Principal Machine Learning Engineer - Optimization
Redwood City, CA · On-site +1
$153K - $211K/yr
Experience building large-scale prediction or optimization systems PubMatic is the leading AI ... Partner with performance advertising signal engineers to define model-ready features, labels ...
Senior Principal Machine Learning Engineer - Optimization
Redwood City, CA · On-site +1
$153K - $211K/yr
Experience building large-scale prediction or optimization systems PubMatic is the leading AI ... Partner with performance advertising signal engineers to define model-ready features, labels ...
Senior Principal Machine Learning Engineer - Optimization
Redwood City, CA · On-site +1
$153K - $211K/yr
Experience building large-scale prediction or optimization systems PubMatic is the leading AI ... Partner with performance advertising signal engineers to define model-ready features, labels ...
The Role You'll be our performance optimization expert, using advanced profiling tools to identify bottlenecks and implementing solutions that achieve 5-10x speedups. From writing custom CUDA kernels ...
The Role You'll be our performance optimization expert, using advanced profiling tools to identify bottlenecks and implementing solutions that achieve 5-10x speedups. From writing custom CUDA kernels ...
The Role You'll be our performance optimization expert, using advanced profiling tools to identify bottlenecks and implementing solutions that achieve 5-10x speedups. From writing custom CUDA kernels ...
Quick apply
The Role You'll be our performance optimization expert, using advanced profiling tools to identify bottlenecks and implementing solutions that achieve 5-10x speedups. From writing custom CUDA kernels ...
Software Engineer, Kernel Performance & AI Tooling
San Francisco, CA · On-site
$266K - $445K/yr
... development, performance engineering, and hardware-software co-design capabilities, with a ... This person will work at the intersection of kernel optimization, developer tooling, observability ...
Software Engineer, Kernel Performance & AI Tooling
San Francisco, CA · On-site
$266K - $445K/yr
... development, performance engineering, and hardware-software co-design capabilities, with a ... This person will work at the intersection of kernel optimization, developer tooling, observability ...
Frontier AI Workloads - Performance and Scalability Engineer
San Jose, CA · On-site
$164K/yr
Drive performance optimization end-to-end across the stack on leading models and customer-relevant serving configurations, closing competitive gaps through kernel and systems-level optimizations
Frontier AI Workloads - Performance and Scalability Engineer
San Jose, CA · On-site
$164K/yr
Drive performance optimization end-to-end across the stack on leading models and customer-relevant serving configurations, closing competitive gaps through kernel and systems-level optimizations
The Research Engineer - AI Performance & Kernel Optimization will improve and optimize the performance of large-scale language model training and inference stacks, focusing on kernel development and ...
The Research Engineer - AI Performance & Kernel Optimization will improve and optimize the performance of large-scale language model training and inference stacks, focusing on kernel development and ...
EED spans the capability lifecycle from concept and architecture design to performance assessment ... The selected candidate will tackle complex optimization and mission planning challenges ...
EED spans the capability lifecycle from concept and architecture design to performance assessment ... The selected candidate will tackle complex optimization and mission planning challenges ...
Tire-Vehicle Performance Engineer II
$1.7K - $2.3K/wk
This role will guide the clients Race Engineering team in understanding tire behavior, performance trade‑offs, and optimization strategies. The engineer will lead the development, calibration, and ...
Tire-Vehicle Performance Engineer II
$1.7K - $2.3K/wk
This role will guide the clients Race Engineering team in understanding tire behavior, performance trade‑offs, and optimization strategies. The engineer will lead the development, calibration, and ...
As a Research Engineer - AI Performance & Kernel Optimization , you will improve and optimize the performance of our large-scale language model training and inference stacks. You will work closely ...
Quick apply
As a Research Engineer - AI Performance & Kernel Optimization , you will improve and optimize the performance of our large-scale language model training and inference stacks. You will work closely ...
As a Research Engineer - AI Performance & Kernel Optimization , you will improve and optimize the performance of our large-scale language model training and inference stacks. You will work closely ...
As a Research Engineer - AI Performance & Kernel Optimization , you will improve and optimize the performance of our large-scale language model training and inference stacks. You will work closely ...
Windows Performance Engineer, Staff
San Diego, CA · On-site
$134.80 - $202.20/hr
General Summary Qualcomm is looking for an experienced Windows Power/Performance Developer with passion for analyzing and optimizing software running on Windows on Snapdragon. The senior developer ...
Windows Performance Engineer, Staff
San Diego, CA · On-site
$134.80 - $202.20/hr
General Summary Qualcomm is looking for an experienced Windows Power/Performance Developer with passion for analyzing and optimizing software running on Windows on Snapdragon. The senior developer ...
Tire-Vehicle Performance Engineer II
$1.7K - $2.3K/wk
This role will guide the clients Race Engineering team in understanding tire behavior, performance trade‑offs, and optimization strategies. The engineer will lead the development, calibration, and ...
Tire-Vehicle Performance Engineer II
$1.7K - $2.3K/wk
This role will guide the clients Race Engineering team in understanding tire behavior, performance trade‑offs, and optimization strategies. The engineer will lead the development, calibration, and ...
EED spans the capability lifecycle from concept and architecture design to performance assessment ... The selected candidate will tackle complex optimization and mission planning challenges, requiring ...
New
EED spans the capability lifecycle from concept and architecture design to performance assessment ... The selected candidate will tackle complex optimization and mission planning challenges, requiring ...
New
Performance Optimization information
What are the key skills and qualifications needed to thrive in performance optimization?
To excel in Performance Optimization, you need a strong analytical mindset, expertise in data interpretation, and a relevant degree in fields like engineering, computer science, or business analytics. Familiarity with tools such as SQL, Python, Tableau, and process improvement methodologies like Lean or Six Sigma is often required. Excellent problem-solving, communication, and project management skills set standout candidates apart. These abilities are essential for identifying improvement opportunities and effectively driving process enhancements within an organization.
What are some common challenges faced in a performance optimization role, and how are they typically addressed?
Professionals in Performance Optimization often encounter challenges such as managing complex data sets, aligning diverse teams around process changes, and quantifying the impact of optimizations. Overcoming these hurdles usually involves effective stakeholder communication, strong analytical processes, and the use of advanced data visualization tools to make insights actionable. Teams often collaborate closely with IT, operations, and management to ensure solutions are both technically sound and practical. Regular training and adapting to new industry best practices help maintain high performance and continue driving meaningful improvements.
What is a performance optimization?
A Performance Optimization job focuses on analyzing, improving, and maintaining the efficiency of systems, processes, or applications. Professionals in this role identify bottlenecks, implement solutions, and enhance overall performance using data-driven strategies. They may work with software, business processes, or operational workflows to maximize productivity and resource utilization. This role often requires expertise in analytics, problem-solving, and technical tools specific to the industry.
What are the most commonly searched types of Performance Optimization jobs in California?
The most popular types of Performance Optimization jobs in California are:
What are popular job titles related to Performance Optimization jobs in California?
For Performance Optimization jobs in California, the most frequently searched job titles are:
What job categories do people searching Performance Optimization jobs in California look for?
The top searched job categories for Performance Optimization jobs in California are:

Principal ML Engineer - Large Scale Training Performance Optimization
San Jose, CA • On-site
$210K/yr
Full-time
Re-posted 26 days ago
Advanced Micro Devices rating
8.6
Based on 13 frontline employees who took The Breakroom Quiz
27th of 157 rated electronics manufacturers
Job description
At AMD, our mission is to build great products that accelerate next-generation computing experiences-from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you'll discover the real differentiator is our culture. We push the limits of innovation to solve the world's most important challenges-striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond. Together, we advance your career.
THE ROLE:
We are looking for a Principal Machine Learning Engineer to join our Models and Applications team. If you are excited by the challenge of distributed training of large models on a large number of GPUs, and if you are passionate about improving training efficiency while innovating and generating new ideas, then this role is for you. You will be part of a world class team focused on addressing the challenge of training generative AI at scale.
THE PERSON:
The ideal candidate should have experience with distributed training pipelines, be knowledgeable in distributed training algorithms (Data Parallel, Tensor Parallel, Pipeline Parallel, Expert Parallel ZeRO), and be familiar with training large models at scale.
KEY RESPONSIBILITIES:
- Train large models to convergence on AMD GPUs at scale.
- Improve the end-to-end training pipeline performance.
- Optimize the distributed training pipeline and algorithm to scale out.
- Contribute your changes to open source.
- Stay up-to-date with the latest training algorithms.
- Influence the direction of AMD AI platform.
- Collaborate across teams with various groups and stakeholders.
PREFERRED EXPERIENCE:
- Experience with ML/DL frameworks such as PyTorch, JAX, or TensorFlow.
- Experience with distributed training and distributed training frameworks, such as Megatron-LM, MaxText, TorchTitan.
- Experience with LLMs or computer vision, especially large models, is a plus.
- Experience with GPU kernel optimization is a plus.
- Excellent Python or C++ programming skills, including debugging, profiling, and performance analysis at scale.
- Experience with ML infra at kernel, framework, or system level
- Strong communication and problem-solving skills.
ACADEMIC CREDENTIALS:
- A master's degree or PhD degree in Computer Science, Artificial Intelligence, Machine Learning, or a related field.
LOCATION:
- San Jose, CA or Bellevue, WA preferred. May consider other US markets within proximity of US AMD offices.
#LI-MV1
#HYBRID
Benefits offered are described: AMD benefits at a glance.
AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants' needs under the respective laws throughout all stages of the recruitment and selection process.
AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD's "Responsible AI Policy" is available here.
This posting is for an existing vacancy.
What Advanced Micro Devices employees say
Pay
Benefits
Hours and flexibility
Workplace
Get the full story on Breakroom
About Advanced Micro Devices (AMD)
Sourced by ZipRecruiter
Industry
Computer and electronic product manufacturing and manufacturing
Company size
5,001 - 10,000 Employees
Headquarters location
Sunnyvale, CA, US