LLM Inference Engineer
Menlo Park, CA · On-site
$165K/yr
By day 90, you will have shipped a measurable improvement to our inference serving stack (reduced ... Experience with CUDA programming and GPU optimization Nice-to-Have: * Contributions to open-source ...