Senior AI Performance Engineer
$143K - $189K/yr
You'll work hands-on with some of the most advanced models in the world - such as DeepSeek R1, GPT OSS, and other frontier architectures - to push the limits of throughput, latency, and efficiency.
$143K - $189K/yr
You'll work hands-on with some of the most advanced models in the world - such as DeepSeek R1, GPT OSS, and other frontier architectures - to push the limits of throughput, latency, and efficiency.
$143K - $189K/yr
You'll work hands-on with some of the most advanced models in the world - such as DeepSeek R1, GPT OSS, and other frontier architectures - to push the limits of throughput, latency, and efficiency.
You'll work hands-on with some of the most advanced models in the world - such as DeepSeek R1, GPT OSS, and other frontier architectures - to push the limits of throughput, latency, and efficiency.
You'll work hands-on with some of the most advanced models in the world - such as DeepSeek R1, GPT OSS, and other frontier architectures - to push the limits of throughput, latency, and efficiency.
San Jose, CA · On-site
Enable customer success by deeply optimizing open-sourced models (Llama 3, DeepSeek, Mixtral) and proprietary models for our specific hardware topology, utilizing tools like vLLM and TensorRT-LLM ...
San Jose, CA · On-site
Enable customer success by deeply optimizing open-sourced models (Llama 3, DeepSeek, Mixtral) and proprietary models for our specific hardware topology, utilizing tools like vLLM and TensorRT-LLM ...
San Jose, CA · On-site
$122K - $167K/yr
You'll work hands-on with some of the most advanced models in the world - such as DeepSeek R1, GPT OSS, and other frontier architectures - to push the limits of throughput, latency, and efficiency.
San Jose, CA · On-site
$122K - $167K/yr
You'll work hands-on with some of the most advanced models in the world - such as DeepSeek R1, GPT OSS, and other frontier architectures - to push the limits of throughput, latency, and efficiency.
San Jose, CA · On-site
$140 - $190/hr
You'll work hands-on with some of the most advanced models in the world -- such as DeepSeek R1, GPT OSS, and other frontier architectures -- to push the limits of throughput, latency, and efficiency.
San Jose, CA · On-site
$140 - $190/hr
You'll work hands-on with some of the most advanced models in the world -- such as DeepSeek R1, GPT OSS, and other frontier architectures -- to push the limits of throughput, latency, and efficiency.
You'll work hands-on with some of the most advanced models in the world - such as DeepSeek R1, GPT OSS, and other frontier architectures - to push the limits of throughput, latency, and efficiency.
You'll work hands-on with some of the most advanced models in the world - such as DeepSeek R1, GPT OSS, and other frontier architectures - to push the limits of throughput, latency, and efficiency.
The use of AI tools, including but not limited to ChatGPT, Microsoft Copilot, Gemini, DeepSeek, or any other AI-assisted software, is strictly prohibited during the interview process. This includes ...
The use of AI tools, including but not limited to ChatGPT, Microsoft Copilot, Gemini, DeepSeek, or any other AI-assisted software, is strictly prohibited during the interview process. This includes ...
Sunnyvale, CA · On-site
$140 - $210/hr
Hands-on experience building and shipping AI systems with LLMs, VLMs, or multimodal models, including the ability to run, adapt, and evaluate open-source models such as DeepSeek, Qwen, Llama, or ...
Sunnyvale, CA · On-site
$140 - $210/hr
Hands-on experience building and shipping AI systems with LLMs, VLMs, or multimodal models, including the ability to run, adapt, and evaluate open-source models such as DeepSeek, Qwen, Llama, or ...
Menlo Park, CA · On-site
$120 - $160/hr
Claude, ChatGPT, Kimi, DeepSeek, GLM). The role We are looking for a Forward Deployed Engineer to sit at the intersection of engineering and customer success for Remy. You will work directly with ...
Menlo Park, CA · On-site
$120 - $160/hr
Claude, ChatGPT, Kimi, DeepSeek, GLM). The role We are looking for a Forward Deployed Engineer to sit at the intersection of engineering and customer success for Remy. You will work directly with ...
Carlsbad, CA · On-site
$25 - $40/hr
Run coding LLMs (Qwen, DeepSeek, Gemini) on MaxLinear servers, connected to Bitbucket, Jira and Confluence via an MCP server * Analyze test logs and publish execution summaries and reports to ...
Carlsbad, CA · On-site
$25 - $40/hr
Run coding LLMs (Qwen, DeepSeek, Gemini) on MaxLinear servers, connected to Bitbucket, Jira and Confluence via an MCP server * Analyze test logs and publish execution summaries and reports to ...
San Francisco, CA · On-site
$120 - $180/hr
About Inference.net We combine idle GPU capacity from around the world into a single cohesive plane of compute capable of serving models like DeepSeek and Llama 4. At any moment, 5,000+ GPUs and ...
San Francisco, CA · On-site
$120 - $180/hr
About Inference.net We combine idle GPU capacity from around the world into a single cohesive plane of compute capable of serving models like DeepSeek and Llama 4. At any moment, 5,000+ GPUs and ...
San Francisco, CA · Remote
$191K - $239K/yr
Tune expert gateway router kernels for MoE models like Qwen3-235B, DeepSeek V3, GLM-5 etc * Hardware & Ecosystem Mastery: Act as the subject matter expert on modern GPU families (NVIDIA/AMD) and ...
San Francisco, CA · Remote
$191K - $239K/yr
Tune expert gateway router kernels for MoE models like Qwen3-235B, DeepSeek V3, GLM-5 etc * Hardware & Ecosystem Mastery: Act as the subject matter expert on modern GPU families (NVIDIA/AMD) and ...
San Francisco, CA · On-site
$120K - $180K/yr
About Inference.net We combine idle GPU capacity from around the world into a single cohesive plane of compute capable of serving models like DeepSeek and Llama 4. At any moment, 5,000+ GPUs and ...
San Francisco, CA · On-site
$120K - $180K/yr
About Inference.net We combine idle GPU capacity from around the world into a single cohesive plane of compute capable of serving models like DeepSeek and Llama 4. At any moment, 5,000+ GPUs and ...
$79 - $105.75/hr
If you're passionate about pushing the boundaries of GPU hardware and software performance and understand terms like disaggregated serving, data parallel attention, MoE, Qwen3.5, DeepSeek, GPT-OSS ...
$79 - $105.75/hr
If you're passionate about pushing the boundaries of GPU hardware and software performance and understand terms like disaggregated serving, data parallel attention, MoE, Qwen3.5, DeepSeek, GPT-OSS ...
Santa Clara, CA · On-site
At NVIDIA, our team focuses on improving community models like Nemotron, Llama, Gemma, DeepSeek, and Qwen. As a Product Manager for Open Models, you will work with groundbreaking technology to ...
Santa Clara, CA · On-site
At NVIDIA, our team focuses on improving community models like Nemotron, Llama, Gemma, DeepSeek, and Qwen. As a Product Manager for Open Models, you will work with groundbreaking technology to ...
At NVIDIA, our team focuses on improving community models like Nemotron, Llama, Gemma, DeepSeek, and Qwen. As a Product Manager for Open Models, you will work with groundbreaking technology to ...
At NVIDIA, our team focuses on improving community models like Nemotron, Llama, Gemma, DeepSeek, and Qwen. As a Product Manager for Open Models, you will work with groundbreaking technology to ...
Santa Clara, CA · On-site
$79 - $105.75/hr
If you're passionate about pushing the boundaries of GPU hardware and software performance and understand terms like disaggregated serving, data parallel attention, MoE, Qwen3.5, DeepSeek, GPT-OSS ...
Santa Clara, CA · On-site
$79 - $105.75/hr
If you're passionate about pushing the boundaries of GPU hardware and software performance and understand terms like disaggregated serving, data parallel attention, MoE, Qwen3.5, DeepSeek, GPT-OSS ...
Knowledge of how to leverage OpenAI, DeepSeek, Zapier, SQL, and automation tools to move fast. Minimal engineering skills required, but strong understanding of how to prototype and build within ...
Knowledge of how to leverage OpenAI, DeepSeek, Zapier, SQL, and automation tools to move fast. Minimal engineering skills required, but strong understanding of how to prototype and build within ...
San Francisco, CA · On-site
Knowledge of how to leverage OpenAI, DeepSeek, Zapier, SQL, and automation tools to move fast. Minimal engineering skills required, but strong understanding of how to prototype and build within ...
San Francisco, CA · On-site
Knowledge of how to leverage OpenAI, DeepSeek, Zapier, SQL, and automation tools to move fast. Minimal engineering skills required, but strong understanding of how to prototype and build within ...
San Jose, CA · On-site
$232 - $313/hr
Familiarity with leading open LLM architectures (e.g., Kimi, Llama, DeepSeek, Mixtral, Qwen) and their impact on system-level communication. * Experience solving rack-, node-, and cluster-scale ...
New
San Jose, CA · On-site
$232 - $313/hr
Familiarity with leading open LLM architectures (e.g., Kimi, Llama, DeepSeek, Mixtral, Qwen) and their impact on system-level communication. * Experience solving rack-, node-, and cluster-scale ...
New
| Aspect | Deepseek | Data Analyst |
|---|---|---|
| Required Credentials | Typically requires a background in computer science, data science, or related fields; certifications in data analysis or machine learning are common | Usually requires a degree in statistics, mathematics, or related fields; certifications like Microsoft Excel, Tableau, or SQL are beneficial |
| Work Environment | Primarily technical, involving data processing, algorithm development, and machine learning model training | Primarily analytical, involving data interpretation, reporting, and visualization |
| Employer & Industry Usage | Used in tech companies, AI firms, and research institutions focusing on machine learning and AI solutions | Used across various industries including finance, marketing, healthcare, and consulting for data-driven decision making |
Deepseek focuses on developing AI and machine learning models, requiring technical expertise in algorithms and programming. Data Analysts interpret and visualize data to support business decisions. While both roles work with data, Deepseek is more technical and research-oriented, whereas Data Analysts focus on insights and reporting.
For Deepseek jobs in California, the most frequently searched job titles are:
The top searched job categories for Deepseek jobs in California are:
Cities in California with the most Deepseek job openings:

$143K - $189K/yr
Full-time
Re-posted 5 days ago
We are seeking a talented and driven ML performance engineer to optimize and scale state-of-the-art foundation models on SambaNova's reconfigurable dataflow platform. You'll work hands-on with some of the most advanced models in the world - such as DeepSeek R1, GPT OSS, and other frontier architectures - to push the limits of throughput, latency, and efficiency. In this role, you'll bridge the gap between deep learning and systems performance, collaborating across compiler, runtime, and hardware layers to deliver world-record performance for large-scale AI inference.
Responsibilities