Preferred : • RISC-V, x86, or ARM64 ISA experience • MLIR or LLVM compiler infrastructure • HPC or scientific computing background (large-scale parallel compute intuition) • FPGA or Verilog ...
Preferred : • RISC-V, x86, or ARM64 ISA experience • MLIR or LLVM compiler infrastructure • HPC or scientific computing background (large-scale parallel compute intuition) • FPGA or Verilog ...
... MLIR, LLVM, or equivalent) and their use representing complex program structures • Solid foundation in program analysis and optimization techniques (e.g., SSA form, loop optimizations ...
... MLIR, LLVM, or equivalent) and their use representing complex program structures • Solid foundation in program analysis and optimization techniques (e.g., SSA form, loop optimizations ...
... MLIR, LLVM, or equivalent) and their use representing complex program structures • Solid foundation in program analysis and optimization techniques (e.g., SSA form, loop optimizations ...
... MLIR, LLVM, or equivalent) and their use representing complex program structures • Solid foundation in program analysis and optimization techniques (e.g., SSA form, loop optimizations ...
We are looking for a Principal Software Development Engineer to lead technical development in compiler technology, MLIR-based infrastructure, and model-to-hardware optimization, with a strong focus ...
We are looking for a Principal Software Development Engineer to lead technical development in compiler technology, MLIR-based infrastructure, and model-to-hardware optimization, with a strong focus ...
Senior Software Engineer, CUTLASS Platform
Austin, TX · On-site
$121K - $160K/yr
Contribute to the advancement of the MLIR-based backend compiler stack for the CUTLASS Python DSL by designing dialects and associated compiler passes. * Author example kernels utilizing CUTLASS ...
Senior Software Engineer, CUTLASS Platform
Austin, TX · On-site
$121K - $160K/yr
Contribute to the advancement of the MLIR-based backend compiler stack for the CUTLASS Python DSL by designing dialects and associated compiler passes. * Author example kernels utilizing CUTLASS ...
... MLIR, LLVM, or equivalent) and their use representing complex program structures • Solid foundation in program analysis and optimization techniques (e.g., SSA form, loop optimizations ...
... MLIR, LLVM, or equivalent) and their use representing complex program structures • Solid foundation in program analysis and optimization techniques (e.g., SSA form, loop optimizations ...
Software Engineer, AI Compiler
Austin, TX · On-site
$100K - $500K/yr
In this role you will lead development on TT-Forge, our MLIR-based compiler, and manage a team focused on scaling graph transformations, lowering passes, and kernel-level optimizations. You'll help ...
Software Engineer, AI Compiler
Austin, TX · On-site
$100K - $500K/yr
In this role you will lead development on TT-Forge, our MLIR-based compiler, and manage a team focused on scaling graph transformations, lowering passes, and kernel-level optimizations. You'll help ...
Principal Software Development Engineer - Compiler & ML Acceleration
San Jose, CA · On-site
$180 - $240/hr
Design and implement MLIR-based compiler flows to lower high-level ML representations into highly optimized hardware-specific code * Drive model compilation and data movement optimization for ML ...
Principal Software Development Engineer - Compiler & ML Acceleration
San Jose, CA · On-site
$180 - $240/hr
Design and implement MLIR-based compiler flows to lower high-level ML representations into highly optimized hardware-specific code * Drive model compilation and data movement optimization for ML ...
... ONNX, MLIR, TVM, XLA, IREE, PyTorch), with contributions to MLIR or LLVM projects a plus • Experience with optimization methods (LP/MIP, CP, SAT/SMT) using solvers like Gurobi or OR-Tools for ...
... ONNX, MLIR, TVM, XLA, IREE, PyTorch), with contributions to MLIR or LLVM projects a plus • Experience with optimization methods (LP/MIP, CP, SAT/SMT) using solvers like Gurobi or OR-Tools for ...
HLO, MLIR, operator fusion, SPMD partitioning, and sharding strategies. • Strong performance tuning instincts across compiler and runtime layers • Experience with distributed inference systems (e ...
HLO, MLIR, operator fusion, SPMD partitioning, and sharding strategies. • Strong performance tuning instincts across compiler and runtime layers • Experience with distributed inference systems (e ...
Senior Staff Compiler Software Engineer
San Jose, CA · On-site
$178K/yr
You will design compiler technologies using MLIR and LLVM , develop optimized code-generation flows, and work closely with GPU architecture, runtime, and AI framework teams to deliver industry ...
Senior Staff Compiler Software Engineer
San Jose, CA · On-site
$178K/yr
You will design compiler technologies using MLIR and LLVM , develop optimized code-generation flows, and work closely with GPU architecture, runtime, and AI framework teams to deliver industry ...
... ONNX, MLIR, TVM, XLA, IREE, PyTorch), with contributions to MLIR or LLVM projects a plus • Experience with optimization methods (LP/MIP, CP, SAT/SMT) using solvers like Gurobi or OR-Tools for ...
... ONNX, MLIR, TVM, XLA, IREE, PyTorch), with contributions to MLIR or LLVM projects a plus • Experience with optimization methods (LP/MIP, CP, SAT/SMT) using solvers like Gurobi or OR-Tools for ...
Sr. Engineer, Software - AI Compiler
Austin, TX · On-site
$100K - $500K/yr
You'll work on TT-Forge, our MLIR-based compiler that enables developers to run AI on all configurations of Tenstorrent hardware using an open-source, performant, and general-purpose compiler. You ...
Sr. Engineer, Software - AI Compiler
Austin, TX · On-site
$100K - $500K/yr
You'll work on TT-Forge, our MLIR-based compiler that enables developers to run AI on all configurations of Tenstorrent hardware using an open-source, performant, and general-purpose compiler. You ...
... MLIR-based dialects • Define and evolve the interface between external model representations and our internal compiler IR • Ensure correctness and completeness of operator coverage across ...
... MLIR-based dialects • Define and evolve the interface between external model representations and our internal compiler IR • Ensure correctness and completeness of operator coverage across ...
Senior LLVM Compiler Engineer
Redmond, WA · On-site
$117K - $160K/yr
Responsibilities : • Work closely with LLVM, Clang, MLIR, and related open‑source communities to upstream compiler features, refactors, and infrastructure originating from NVIDIA's downstream ...
Senior LLVM Compiler Engineer
Redmond, WA · On-site
$117K - $160K/yr
Responsibilities : • Work closely with LLVM, Clang, MLIR, and related open‑source communities to upstream compiler features, refactors, and infrastructure originating from NVIDIA's downstream ...
Staff Engineer, Compiler
San Jose, CA · On-site
Medical
Dental
Vision
Life
Retirement
PTO
Triton, Helion, MLIR, XLA, TVM, Inductor, IREE, CUTLASS, or a proprietary equivalent (More experienced candidates will also be considered at relevant levels). * Experience designing a kernel DSL or ...
Staff Engineer, Compiler
San Jose, CA · On-site
Medical
Dental
Vision
Life
Retirement
PTO
Triton, Helion, MLIR, XLA, TVM, Inductor, IREE, CUTLASS, or a proprietary equivalent (More experienced candidates will also be considered at relevant levels). * Experience designing a kernel DSL or ...
On-Device ML Compiler Engineer, Model Compilation, Graphics, Games and Machine Learning
Cupertino, CA · On-site
We have an MLIR-based compiler stack, and use it to target the neural engine, GPU, and CPU in order to harness the full capabilities of the system for ML workflows and execution. Minimum ...
On-Device ML Compiler Engineer, Model Compilation, Graphics, Games and Machine Learning
Cupertino, CA · On-site
We have an MLIR-based compiler stack, and use it to target the neural engine, GPU, and CPU in order to harness the full capabilities of the system for ML workflows and execution. Minimum ...
Senior Staff Compiler Software Engineer
San Jose, CA · Hybrid
$143K - $189K/yr
You will design compiler technologies using MLIR and LLVM , develop optimized code-generation flows, and work closely with GPU architecture, runtime, and AI framework teams to deliver industry ...
Senior Staff Compiler Software Engineer
San Jose, CA · Hybrid
$143K - $189K/yr
You will design compiler technologies using MLIR and LLVM , develop optimized code-generation flows, and work closely with GPU architecture, runtime, and AI framework teams to deliver industry ...
Senior LLVM Compiler Engineer
Redmond, WA · On-site
$117K - $160K/yr
In this role, you will work directly with LLVM, Clang, MLIR, and related opensource projects to upstream compiler functionality that currently lives in NVIDIA's downstream repositories. Your ...
Senior LLVM Compiler Engineer
Redmond, WA · On-site
$117K - $160K/yr
In this role, you will work directly with LLVM, Clang, MLIR, and related opensource projects to upstream compiler functionality that currently lives in NVIDIA's downstream repositories. Your ...
... ONNX, MLIR, TVM, XLA, IREE, PyTorch), with contributions to MLIR or LLVM projects a plus • Experience with optimization methods (LP/MIP, CP, SAT/SMT) using solvers like Gurobi or OR-Tools for ...
... ONNX, MLIR, TVM, XLA, IREE, PyTorch), with contributions to MLIR or LLVM projects a plus • Experience with optimization methods (LP/MIP, CP, SAT/SMT) using solvers like Gurobi or OR-Tools for ...
Mlir information
What is MLIR?
What skills and qualifications are needed to work with MLIR?
What is the difference between Mlir vs Machine Learning Engineer?
| Aspect | Mlir | Machine Learning Engineer |
|---|---|---|
| Required Credentials | Technical knowledge of compiler infrastructure, programming skills in C++/Python | Degree in Computer Science, Data Science, or related fields; experience with ML frameworks |
| Work Environment | Research and development in compiler and software infrastructure teams | Developing, testing, and deploying machine learning models in various industries |
| Employer & Industry Usage | Tech companies, AI research labs, compiler development firms | Tech companies, startups, AI-focused organizations |
| Common Search & Comparison Intent | Understanding technical roles in compiler infrastructure | Learning about careers in machine learning and AI |
While Mlir focuses on compiler infrastructure and software development for optimizing machine learning models, Machine Learning Engineers primarily design and implement ML models for practical applications. Both roles require technical expertise, but Mlir is more specialized in compiler technology, whereas Machine Learning Engineers work directly on AI solutions.
How does an engineer working with MLIR collaborate with different teams?

Full-time
Re-posted 11 days ago
Job description
DensityAI is a company focused on AI technology, and they are seeking a Kernel Engineer to write and optimize compute kernels for a custom AI accelerator. The role involves collaborating with architecture and compiler teams to ensure high performance of ML workloads on hardware.
Responsibilities:
• Write and optimize compute kernels for a custom AI accelerator — tensor operations, data movement patterns, memory hierarchy exploitation
• Develop and maintain profiling infrastructure to measure kernel performance against architectural targets
• Define and document shuffle patterns for ML kernel primitives across CPU-like control, tensor cores, and CUTLASS-style operations
• Drive kernel DSL design decisions — thread spawn mechanisms, register passing conventions, and memory management strategies
• Enable end-to-end kernel execution on the architectural simulator
• Collaborate with the compiler team on the MLIR dialect — your kernels are the primary validation target
• Create onboarding documentation and kernel writing guides for the broader team
Qualifications:
Required:
• C/C++ — production-grade systems code, not scripted glue. You'll write performance-critical kernels.
• CUDA or equivalent accelerator programming — deep experience writing GPU kernels, understanding warp/wavefront execution, memory coalescing, shared memory optimization. The mental model transfers directly.
• Computer architecture — you need to reason about pipelines, memory hierarchies, data movement costs, and how software maps to hardware.
• Performance profiling and optimization — you live in profilers. Identifying bottlenecks, measuring throughput, and iterating until kernels meet targets is the core loop.
• Tensor operations — practical understanding of GEMM, convolution, attention, reduction, and scatter/gather as they map to hardware.
• Python — for scripting, DSL integration, and profiling automation.
Preferred:
• RISC-V, x86, or ARM64 ISA experience
• MLIR or LLVM compiler infrastructure
• HPC or scientific computing background (large-scale parallel compute intuition)
• FPGA or Verilog/SystemVerilog (ability to read RTL and reason about the hardware you're targeting)
• Familiarity with CUTLASS, Triton, or similar kernel libraries
Company:
DensityAI is an infrastructure for data centers serving automotive, robotics, and industrial applications Founded in 2025, the company is headquartered in Mountain View, USA, with a team of 51-200 employees. The company is currently Early Stage.