Responsibilities : • Write and optimize compute kernels for a custom AI accelerator -- tensor operations, data movement patterns, memory hierarchy exploitation • Develop and maintain profiling ...
Responsibilities : • Write and optimize compute kernels for a custom AI accelerator -- tensor operations, data movement patterns, memory hierarchy exploitation • Develop and maintain profiling ...
Write and optimize compute kernels for a custom AI accelerator - tensor operations, data movement patterns, memory hierarchy exploitation * Develop and maintain profiling infrastructure to measure ...
Write and optimize compute kernels for a custom AI accelerator - tensor operations, data movement patterns, memory hierarchy exploitation * Develop and maintain profiling infrastructure to measure ...
About the role Own verification of our AI compute core - tensor pipelines, MAC arrays, accumulator logic, and the compute memory interconnect. Work with chip-design and software teams driving ...
About the role Own verification of our AI compute core - tensor pipelines, MAC arrays, accumulator logic, and the compute memory interconnect. Work with chip-design and software teams driving ...
Senior Software Engineer - Edge AI/GenAI & Multimedia
San Diego, CA · On-site
$130K - $171K/yr
Strong ability to handle tensor pre-processing and post-processing and integrate AI models into end-to-end pipelines across inputs such as camera, audio, and text. • Strong background in Linux ...
Senior Software Engineer - Edge AI/GenAI & Multimedia
San Diego, CA · On-site
$130K - $171K/yr
Strong ability to handle tensor pre-processing and post-processing and integrate AI models into end-to-end pipelines across inputs such as camera, audio, and text. • Strong background in Linux ...
Inference Infrastructure Engineer, Serving
Palo Alto, CA · On-site
$180 - $240/hr
Implement multi-GPU and multi-node model parallelism for serving (tensor or pipeline parallel) * Build autoscaling and load balancing for production ML services * Establish standards for reliability ...
Inference Infrastructure Engineer, Serving
Palo Alto, CA · On-site
$180 - $240/hr
Implement multi-GPU and multi-node model parallelism for serving (tensor or pipeline parallel) * Build autoscaling and load balancing for production ML services * Establish standards for reliability ...
Kernel Engineer (Compute / Accelerator)
Mountain View, CA · On-site
$260K - $320K/yr
Write and optimize compute kernels for a custom AI accelerator -- tensor operations, data movement patterns, memory hierarchy exploitation * Develop and maintain profiling infrastructure to measure ...
Quick apply
Kernel Engineer (Compute / Accelerator)
Mountain View, CA · On-site
$260K - $320K/yr
Write and optimize compute kernels for a custom AI accelerator -- tensor operations, data movement patterns, memory hierarchy exploitation * Develop and maintain profiling infrastructure to measure ...
Kernel Engineer (Compute / Accelerator)
Mountain View, CA · On-site
$260K - $320K/yr
Write and optimize compute kernels for a custom AI accelerator - tensor operations, data movement patterns, memory hierarchy exploitation * Develop and maintain profiling infrastructure to measure ...
Kernel Engineer (Compute / Accelerator)
Mountain View, CA · On-site
$260K - $320K/yr
Write and optimize compute kernels for a custom AI accelerator - tensor operations, data movement patterns, memory hierarchy exploitation * Develop and maintain profiling infrastructure to measure ...
Senior Software Engineer -- cuEquivariance
Santa Clara, CA · On-site
$143K - $189K/yr
Responsibilities : • Build, implement, and optimize CUDA kernels for equivariant neural network primitives -- tensor products, segmented polynomials, and triangle-based operations -- targeting peak ...
Senior Software Engineer -- cuEquivariance
Santa Clara, CA · On-site
$143K - $189K/yr
Responsibilities : • Build, implement, and optimize CUDA kernels for equivariant neural network primitives -- tensor products, segmented polynomials, and triangle-based operations -- targeting peak ...
Lead the definition of mechanisms for efficient movement of tensor activations, weights, and outputs through on-chip and off-chip memory pathways and high-throughput DMA architecture. * Partner ...
Lead the definition of mechanisms for efficient movement of tensor activations, weights, and outputs through on-chip and off-chip memory pathways and high-throughput DMA architecture. * Partner ...
Senior Software Engineer, CUTLASS Platform
Santa Clara, CA · On-site
$143K - $189K/yr
If you are passionate about designing abstractions for Tensor Core and related GPU hardware features in MLIR, Python, and C++ that enable writing high performance kernels, apply to join the CUTLASS ...
Senior Software Engineer, CUTLASS Platform
Santa Clara, CA · On-site
$143K - $189K/yr
If you are passionate about designing abstractions for Tensor Core and related GPU hardware features in MLIR, Python, and C++ that enable writing high performance kernels, apply to join the CUTLASS ...
Senior Software Engineer, CUTLASS Kernels
Santa Clara, CA · On-site
$143K - $189K/yr
Write Tensor Core-based deep learning kernels such as grouped-GEMM, attention, and convolution using CUTLASS CUDA C++ and Python DSL for Blackwell, Rubin, and future architectures. * Optimize kernels ...
Senior Software Engineer, CUTLASS Kernels
Santa Clara, CA · On-site
$143K - $189K/yr
Write Tensor Core-based deep learning kernels such as grouped-GEMM, attention, and convolution using CUTLASS CUDA C++ and Python DSL for Blackwell, Rubin, and future architectures. * Optimize kernels ...
Staff Embedded Software Engineer - Edge AI/GenAI & Multimedia
San Diego, CA · On-site
$139K - $183K/yr
Strong ability to handle tensor pre-processing and post-processing and integrate AI models into end-to-end pipelines across inputs such as camera, audio, and text. * Strong background in Linux system ...
Staff Embedded Software Engineer - Edge AI/GenAI & Multimedia
San Diego, CA · On-site
$139K - $183K/yr
Strong ability to handle tensor pre-processing and post-processing and integrate AI models into end-to-end pipelines across inputs such as camera, audio, and text. * Strong background in Linux system ...
Senior Software Engineer - Edge AI/GenAI & Multimedia
San Diego, CA · On-site
$130K - $171K/yr
Strong ability to handle tensor pre-processing and post-processing and integrate AI models into end-to-end pipelines across inputs such as camera, audio, and text. * Strong background in Linux system ...
Senior Software Engineer - Edge AI/GenAI & Multimedia
San Diego, CA · On-site
$130K - $171K/yr
Strong ability to handle tensor pre-processing and post-processing and integrate AI models into end-to-end pipelines across inputs such as camera, audio, and text. * Strong background in Linux system ...
We're looking for someone who finds joy in memory hierarchies, tensor cores, and profiler output. While San Francisco and Boston are preferred, we are open to other locations. What We're Looking For ...
We're looking for someone who finds joy in memory hierarchies, tensor cores, and profiler output. While San Francisco and Boston are preferred, we are open to other locations. What We're Looking For ...
Lead the definition of mechanisms for efficient movement of tensor activations, weights, and outputs through on-chip and off-chip memory pathways and high-throughput DMA architecture. * Partner ...
Lead the definition of mechanisms for efficient movement of tensor activations, weights, and outputs through on-chip and off-chip memory pathways and high-throughput DMA architecture. * Partner ...
Senior Software Engineer, CUTLASS Platform
Santa Clara, CA · On-site
$143K - $189K/yr
If you are passionate about designing abstractions for Tensor Core and related GPU hardware features in MLIR, Python, and C++ that enable writing high performance kernels, apply to join the CUTLASS ...
Senior Software Engineer, CUTLASS Platform
Santa Clara, CA · On-site
$143K - $189K/yr
If you are passionate about designing abstractions for Tensor Core and related GPU hardware features in MLIR, Python, and C++ that enable writing high performance kernels, apply to join the CUTLASS ...
AI Core DV Engineer
Mountain View, CA · On-site
$180K - $320K/yr
About the role Own verification of our AI compute core - tensor pipelines, MAC arrays, accumulator logic, and the compute ↔ memory interconnect. Work with chip-design and software teams driving ...
Quick apply
AI Core DV Engineer
Mountain View, CA · On-site
$180K - $320K/yr
About the role Own verification of our AI compute core - tensor pipelines, MAC arrays, accumulator logic, and the compute ↔ memory interconnect. Work with chip-design and software teams driving ...
AI Core DV Engineer
Mountain View, CA · On-site
$180K - $320K/yr
About the role Own verification of our AI compute core - tensor pipelines, MAC arrays, accumulator logic, and the compute ↔ memory interconnect. Work with chip-design and software teams driving ...
AI Core DV Engineer
Mountain View, CA · On-site
$180K - $320K/yr
About the role Own verification of our AI compute core - tensor pipelines, MAC arrays, accumulator logic, and the compute ↔ memory interconnect. Work with chip-design and software teams driving ...
Senior Software Engineer, CUTLASS Kernels
Santa Clara, CA · On-site
$143K - $189K/yr
Write Tensor Core-based deep learning kernels such as grouped-GEMM, attention, and convolution using CUTLASS CUDA C++ and Python DSL for Blackwell, Rubin, and future architectures. * Optimize kernels ...
Senior Software Engineer, CUTLASS Kernels
Santa Clara, CA · On-site
$143K - $189K/yr
Write Tensor Core-based deep learning kernels such as grouped-GEMM, attention, and convolution using CUTLASS CUDA C++ and Python DSL for Blackwell, Rubin, and future architectures. * Optimize kernels ...
Senior Vehicle Manufacturing Engineer - SKD & Factory QA (San Jose)
San Jose, CA · On-site
$75K - $300K/yr
A leading automotive technology firm in San Jose is seeking an experienced professional for a production management role. Responsibilities include supervising manufacturing activities and managing ...
Senior Vehicle Manufacturing Engineer - SKD & Factory QA (San Jose)
San Jose, CA · On-site
$75K - $300K/yr
A leading automotive technology firm in San Jose is seeking an experienced professional for a production management role. Responsibilities include supervising manufacturing activities and managing ...
Tensor information
See California salary details
$45.4K - $63.1K
1% of jobs
$63.1K - $80.8K
2% of jobs
$80.8K - $98.6K
4% of jobs
$98.6K - $116.3K
9% of jobs
$131.3K is the 25th percentile. Wages below this are outliers.
$116.3K - $134K
11% of jobs
$134K - $151.7K
7% of jobs
The median wage is $157.4K / yr.
$151.7K - $169.4K
50% of jobs
$169.4K - $187.2K
2% of jobs
$187.2K - $204.9K
1% of jobs
$204.9K - $222.6K
0% of jobs
$222.6K - $240.3K
13% of jobs
$45.4K
$162.9K
$240.3K
How much do tensor jobs pay per year?
What is a tensor?
A Tensor job typically refers to a role involving tensors, which are mathematical objects used in machine learning, AI, and scientific computing. These jobs often require expertise in deep learning frameworks like TensorFlow or PyTorch, where tensors represent multi-dimensional arrays for data processing. A Tensor job can involve designing and optimizing neural networks, performing large-scale data analysis, or working with high-performance computing.
What are the key skills and qualifications needed to thrive as a tensor?
What are some common challenges faced by TensorFlow developers when working on large-scale machine learning projects?
What is the difference between Tensor vs Data Scientist?
| Aspect | Tensor | Data Scientist |
|---|---|---|
| Required Credentials | Knowledge of machine learning, programming skills, often a degree in computer science or related fields | Degree in statistics, computer science, or related fields; strong analytical skills |
| Work Environment | Tech companies, AI research labs, software development teams | Business, finance, healthcare, and tech industries analyzing data to inform decisions |
| Industry Usage | Primarily in AI, machine learning, and deep learning projects | Across industries for data analysis, predictive modeling, and insights |
While a Tensor is a fundamental data structure used in machine learning frameworks like TensorFlow, a Data Scientist analyzes data to extract insights and build models. Tensors are tools that Data Scientists often work with, but they are not roles themselves. Understanding tensors is essential for Data Scientists involved in AI and machine learning projects.
What are popular job titles related to Tensor jobs in California?
For Tensor jobs in California, the most frequently searched job titles are:
What job categories do people searching Tensor jobs in California look for?
The top searched job categories for Tensor jobs in California are:
What cities in California are hiring for Tensor jobs?
Cities in California with the most Tensor job openings:

Full-time
Re-posted 20 days ago
Job description
DensityAI is a company focused on AI technology, and they are seeking a Kernel Engineer to write and optimize compute kernels for a custom AI accelerator. The role involves collaborating with architecture and compiler teams to ensure high performance of ML workloads on hardware.
Responsibilities:
• Write and optimize compute kernels for a custom AI accelerator — tensor operations, data movement patterns, memory hierarchy exploitation
• Develop and maintain profiling infrastructure to measure kernel performance against architectural targets
• Define and document shuffle patterns for ML kernel primitives across CPU-like control, tensor cores, and CUTLASS-style operations
• Drive kernel DSL design decisions — thread spawn mechanisms, register passing conventions, and memory management strategies
• Enable end-to-end kernel execution on the architectural simulator
• Collaborate with the compiler team on the MLIR dialect — your kernels are the primary validation target
• Create onboarding documentation and kernel writing guides for the broader team
Qualifications:
Required:
• C/C++ — production-grade systems code, not scripted glue. You'll write performance-critical kernels.
• CUDA or equivalent accelerator programming — deep experience writing GPU kernels, understanding warp/wavefront execution, memory coalescing, shared memory optimization. The mental model transfers directly.
• Computer architecture — you need to reason about pipelines, memory hierarchies, data movement costs, and how software maps to hardware.
• Performance profiling and optimization — you live in profilers. Identifying bottlenecks, measuring throughput, and iterating until kernels meet targets is the core loop.
• Tensor operations — practical understanding of GEMM, convolution, attention, reduction, and scatter/gather as they map to hardware.
• Python — for scripting, DSL integration, and profiling automation.
Preferred:
• RISC-V, x86, or ARM64 ISA experience
• MLIR or LLVM compiler infrastructure
• HPC or scientific computing background (large-scale parallel compute intuition)
• FPGA or Verilog/SystemVerilog (ability to read RTL and reason about the hardware you're targeting)
• Familiarity with CUTLASS, Triton, or similar kernel libraries
Company:
DensityAI is an infrastructure for data centers serving automotive, robotics, and industrial applications Founded in 2025, the company is headquartered in Mountain View, USA, with a team of 51-200 employees. The company is currently Early Stage.