They are seeking a Staff Engineer - AI Model Optimization Architect to lead model transformation and optimization efforts for various AI models on Qualcomm's inference accelerators. Responsibilities ...
They are seeking a Staff Engineer - AI Model Optimization Architect to lead model transformation and optimization efforts for various AI models on Qualcomm's inference accelerators. Responsibilities ...
Solutions Architect - AI Model Specialist
San Francisco, CA · On-site
$74.25 - $97.75/hr
About the job FriendliAI is seeking a Solution Architect specializing in open-source AI models, AI inference API integration, and agentic systems. You will work closely with our customers to ...
Solutions Architect - AI Model Specialist
San Francisco, CA · On-site
$74.25 - $97.75/hr
About the job FriendliAI is seeking a Solution Architect specializing in open-source AI models, AI inference API integration, and agentic systems. You will work closely with our customers to ...
We are seeking a Staff Engineer - AI Model Optimization Architect to lead end-to-end model transformation and optimization for LLMs, VLMs, diffusion, and multimodal models on Qualcomm inference ...
We are seeking a Staff Engineer - AI Model Optimization Architect to lead end-to-end model transformation and optimization for LLMs, VLMs, diffusion, and multimodal models on Qualcomm inference ...
As a leader in Robotics AI Model, you will own the critical pipeline that transforms pretrained foundation models into deployable robot policies - turning general-purpose AI into systems that can ...
As a leader in Robotics AI Model, you will own the critical pipeline that transforms pretrained foundation models into deployable robot policies - turning general-purpose AI into systems that can ...
iOS Engineer - AI Model Evaluator
San Francisco, CA · Remote
$85/hr
Use frontier AI coding agents to complete and evaluate complex engineering tasks. * Review model-generated mobile application code for correctness, quality, maintainability, and performance.
Quick apply
iOS Engineer - AI Model Evaluator
San Francisco, CA · Remote
$85/hr
Use frontier AI coding agents to complete and evaluate complex engineering tasks. * Review model-generated mobile application code for correctness, quality, maintainability, and performance.
Senior Director, AI Model Lifecycle
San Francisco, CA · On-site
$301.75 - $355/hr
Experience in Generative AI (Large Language Models, Multimodal). * Hands‑on experience training, fine‑tuning, and aligning LLMs using Reinforcement Learning and Reinforcement Fine‑Tuning (RFT ...
Senior Director, AI Model Lifecycle
San Francisco, CA · On-site
$301.75 - $355/hr
Experience in Generative AI (Large Language Models, Multimodal). * Hands‑on experience training, fine‑tuning, and aligning LLMs using Reinforcement Learning and Reinforcement Fine‑Tuning (RFT ...
As a leader in Robotics AI Model, you will own the critical pipeline that transforms pretrained foundation models into deployable robot policies -- turning general-purpose AI into systems that can ...
As a leader in Robotics AI Model, you will own the critical pipeline that transforms pretrained foundation models into deployable robot policies -- turning general-purpose AI into systems that can ...
Sr. Staff Engineer, Robotics AI Model Optimization
San Jose, CA · On-site
$178K/yr
The AI Models and Applications team at AMD is looking for a specialized Sr. Staff or Principal level engineer who is passionate about enabling innovative and efficient Generative AI training ...
Sr. Staff Engineer, Robotics AI Model Optimization
San Jose, CA · On-site
$178K/yr
The AI Models and Applications team at AMD is looking for a specialized Sr. Staff or Principal level engineer who is passionate about enabling innovative and efficient Generative AI training ...
As a leader in Robotics AI Model, you will own the critical pipeline that transforms pretrained foundation models into deployable robot policies -- turning general-purpose AI into systems that can ...
As a leader in Robotics AI Model, you will own the critical pipeline that transforms pretrained foundation models into deployable robot policies -- turning general-purpose AI into systems that can ...
Sr. Staff Engineer, Robotics AI Model Optimization
San Jose, CA · Hybrid
$124K - $169K/yr
The AI Models and Applications team at AMD is looking for a specialized Sr. Staff or Principal level engineer who is passionate about enabling innovative and efficient Generative AI training ...
Sr. Staff Engineer, Robotics AI Model Optimization
San Jose, CA · Hybrid
$124K - $169K/yr
The AI Models and Applications team at AMD is looking for a specialized Sr. Staff or Principal level engineer who is passionate about enabling innovative and efficient Generative AI training ...
As a Tech Lead in Robotics AI Model, you will own the critical pipeline that transforms pretrained foundation models into deployable robot policies, directly impacting the capabilities and safety of ...
As a Tech Lead in Robotics AI Model, you will own the critical pipeline that transforms pretrained foundation models into deployable robot policies, directly impacting the capabilities and safety of ...
Senior Software Engineer, AI Model Lifecycle
San Francisco, CA · On-site
$144K - $190K/yr
The Senior Software Engineer for the AI Model Lifecycle team will be responsible for building a managed platform for the application development lifecycle, focusing on leveraging Machine Learning ...
Senior Software Engineer, AI Model Lifecycle
San Francisco, CA · On-site
$144K - $190K/yr
The Senior Software Engineer for the AI Model Lifecycle team will be responsible for building a managed platform for the application development lifecycle, focusing on leveraging Machine Learning ...
Distinguished Technologist - AI Model Performance Architect
Palo Alto, CA · On-site
$196K/yr
Distinguished Technologist - AI Model Performance Architect Description - This role is responsible for bridging system memory architecture with AI model behavior ensuring optimal performance through ...
Distinguished Technologist - AI Model Performance Architect
Palo Alto, CA · On-site
$196K/yr
Distinguished Technologist - AI Model Performance Architect Description - This role is responsible for bridging system memory architecture with AI model behavior ensuring optimal performance through ...
Senior Software Engineer, AI Model Lifecycle
San Francisco, CA · On-site
$144K - $190K/yr
The Senior Software Engineer for the AI Model Lifecycle team will play a crucial role in building a comprehensive managed platform for the entire application development lifecycle, with a specific ...
Senior Software Engineer, AI Model Lifecycle
San Francisco, CA · On-site
$144K - $190K/yr
The Senior Software Engineer for the AI Model Lifecycle team will play a crucial role in building a comprehensive managed platform for the entire application development lifecycle, with a specific ...
Senior Software Engineer, AI Model Lifecycle
San Francisco, CA · On-site
$144K - $190K/yr
The Senior Software Engineer for the AI Model Lifecycle team will play a crucial role in building a comprehensive managed platform for the entire application development lifecycle, with a specific ...
Quick apply
Senior Software Engineer, AI Model Lifecycle
San Francisco, CA · On-site
$144K - $190K/yr
The Senior Software Engineer for the AI Model Lifecycle team will play a crucial role in building a comprehensive managed platform for the entire application development lifecycle, with a specific ...
Senior Software Engineer, AI Model Lifecycle
Sunnyvale, CA · On-site
$143K - $189K/yr
The Senior Software Engineer for the AI Model Lifecycle team will play a crucial role in building a comprehensive managed platform for the entire application development lifecycle, with a specific ...
Quick apply
Senior Software Engineer, AI Model Lifecycle
Sunnyvale, CA · On-site
$143K - $189K/yr
The Senior Software Engineer for the AI Model Lifecycle team will play a crucial role in building a comprehensive managed platform for the entire application development lifecycle, with a specific ...
Senior Software Engineer, AI Model Lifecycle
San Francisco, CA · On-site
$166 - $201/hr
AI/ML Expertise * Familiarity with Generative AI (Large Language Models, Multimodal). * Experience with AI infrastructure components (training, inference). Preferred Qualifications * Proficiency in ...
New
Senior Software Engineer, AI Model Lifecycle
San Francisco, CA · On-site
$166 - $201/hr
AI/ML Expertise * Familiarity with Generative AI (Large Language Models, Multimodal). * Experience with AI infrastructure components (training, inference). Preferred Qualifications * Proficiency in ...
New
Sr. Staff Engineer, Robotics AI Model Optimization (San Jose)
San Jose, CA · On-site
$192K - $288K/yr
Sr. Staff Engineer, AI Models and Applications Join to apply for the Sr. Staff Engineer, AI Models and Applications role at AMD Base pay range $192,000.00/yr - $288,000.00/yr WHAT YOU DO AT AMD ...
Sr. Staff Engineer, Robotics AI Model Optimization (San Jose)
San Jose, CA · On-site
$192K - $288K/yr
Sr. Staff Engineer, AI Models and Applications Join to apply for the Sr. Staff Engineer, AI Models and Applications role at AMD Base pay range $192,000.00/yr - $288,000.00/yr WHAT YOU DO AT AMD ...
DevOps Engineer - AI Model Evaluator
San Francisco, CA · Remote
$85/hr
Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. * Review model-generated implementations involving cloud platforms , Kubernetes , CI/CD systems , and ...
Quick apply
DevOps Engineer - AI Model Evaluator
San Francisco, CA · Remote
$85/hr
Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. * Review model-generated implementations involving cloud platforms , Kubernetes , CI/CD systems , and ...
Staff Software Engineer, AI Model Lifecycle
San Francisco, CA · On-site
$204 - $247/hr
AI/ML Expertise: * Experience with Generative AI (Large Language Models, Multimodal). * Experience with AI infrastructure, including training, inference. * Preferred Qualifications: * Proficiency in ...
New
Staff Software Engineer, AI Model Lifecycle
San Francisco, CA · On-site
$204 - $247/hr
AI/ML Expertise: * Experience with Generative AI (Large Language Models, Multimodal). * Experience with AI infrastructure, including training, inference. * Preferred Qualifications: * Proficiency in ...
New
Ai Model information
What is the difference between Ai Model vs Data Scientist?
| Aspect | Ai Model | Data Scientist |
|---|---|---|
| Required Credentials | Knowledge of machine learning, programming skills, sometimes certifications in AI/ML | Degree in data science, statistics, computer science; certifications beneficial |
| Work Environment | Focus on developing, training, and deploying AI models | Data analysis, interpretation, and visualization; often collaborates with AI teams |
| Industry Usage | Used in AI development, automation, and predictive modeling | Applied across industries for insights, reporting, and decision-making |
While both roles involve working with data and algorithms, an Ai Model primarily focuses on creating and refining AI systems, whereas a Data Scientist analyzes data to generate insights and supports AI development. The roles often overlap but serve distinct functions within the data and AI ecosystem.
What are some common challenges faced by professionals working as AI model developers, and how can they address them?
What is an AI model?
What are the key skills and qualifications needed to thrive as an AI model, and why are they important?
How to get into AI modeling?

Qualcomm rating
8.8
Based on 11 frontline employees who took The Breakroom Quiz
51st of 242 rated software companies
Job description
Qualcomm Technologies, Inc. is at the forefront of Cloud AI, developing platforms for efficient inference of large-scale models. They are seeking a Staff Engineer – AI Model Optimization Architect to lead model transformation and optimization efforts for various AI models on Qualcomm's inference accelerators.
Responsibilities:
• Architect and deliver model optimization strategies that transform PyTorch models for efficient inference on Qualcomm accelerators.
• Drive graph capture and deployment using PyTorch, ONNX, and torch.compile, including model rewrites and graph-level transformations.
• Design and implement fusion kernels using DSL based approaches (e.g., Triton), enabling fused operations and performance critical algorithmic rewrites.
• Partner deeply with compiler, performance, and accuracy teams to co-design lowering strategies, kernel fusion, layout decisions, and runtime integration.
• Profile and optimize LLM/VLM/diffusion inference for throughput and latency across batch sizes, sequence lengths, and serving modes.
• Own transformer specific optimizations including KVcache management, decoding behavior, and long context performance.
• Enable and optimize continuous batching (dynamic/iteration-level scheduling), understanding its impact on memory, scheduling, and tail latency.
• Architect and scale distributed inference strategies (e.g., sharding and parallelism) across multi-core and multi-device systems.
• Establish reusable approaches to scale model optimizations to new hardware architectures, creating robust patterns and tooling.
• Debug complex performance or stability issues to root cause and drive production ready solutions.
Qualifications:
Required:
• Expert level expertise in PyTorch and inference focused model optimization; strong Python engineering skills.
• Hands on experience with torch.compile / TorchDynamo or related graph capture and compilation workflows.
• Deep understanding of transformer architectures, attention mechanisms, MoEs, and performance trade-offs.
• Practical experience with KVcache behavior, serving time optimizations, and memory/performance tradeoffs.
• Strong foundation in computer architecture, ML accelerators, and distributed systems.
• Proven ability to lead cross-functional technical efforts and influence design decisions.
• MS in Computer Science, Machine Learning, Computer Engineering, or Electrical Engineering, or equivalent experience.
• Bachelor's degree in Computer Science, Engineering, Information Systems, or related field and 4+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience.
• Master's degree in Computer Science, Engineering, Information Systems, or related field and 3+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience.
• PhD in Computer Science, Engineering, Information Systems, or related field and 2+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience.
Preferred:
• Experience developing fusion kernels using Triton or similar DSLs, and collaborating with ML compiler teams.
• Familiarity with LLM serving stacks and continuous batching systems.
• Background in numerical methods, performance/accuracy trade-off analysis, or evaluation frameworks.
• PhD in a relevant field.
Company:
Qualcomm designs wireless technologies and semiconductors that power connectivity, communication, and smart devices. Founded in 1985, the company is headquartered in San Diego, USA, with a team of 10001+ employees. The company is currently Late Stage.
About Qualcomm
Sourced by ZipRecruiter
Qualcomm is enabling a world where everyone and everything can be intelligently connected. You interact with products and technologies made possible by Qualcomm every day, including 5G-enabled smartphones that double as pro-level cameras and gaming devices, smarter vehicles and cities, and the technology behind the smart, connected factories that manufactured your latest purchase. Our powerful connectivity solutions keep you connected—even in remote areas. Qualcomm 5G and AI innovations are the power behind the connected intelligent edge. You’ll find our technologies behind and inside the innovations that deliver significant value across multiple industries and to billions of people every day.
Industry
Technology, communication and media
Company size
10,000+ Employees
Headquarters location
San Diego, CA, US
Year founded
1985