LLM Inference Engineer
OR · On-site +1
San Francisco or Remote About The Role The NEAR AI team is building decentralized and confidential ... CuTe, CUDA, etc. * Proven track record in designing and maintaining end-to-end high-traffic LLM ...
OR · On-site +1
San Francisco or Remote About The Role The NEAR AI team is building decentralized and confidential ... CuTe, CUDA, etc. * Proven track record in designing and maintaining end-to-end high-traffic LLM ...
OR · On-site +1
San Francisco or Remote About The Role The NEAR AI team is building decentralized and confidential ... CuTe, CUDA, etc. * Proven track record in designing and maintaining end-to-end high-traffic LLM ...
OR · On-site +1
Practical experience optimizing ML workflows using CUDA/GPU acceleration. * Background in feature ... Remote-US Time zone requirements The team operates on the East/West coast time zones. Travel ...
OR · On-site +1
Practical experience optimizing ML workflows using CUDA/GPU acceleration. * Background in feature ... Remote-US Time zone requirements The team operates on the East/West coast time zones. Travel ...
Hillsboro, OR · On-site +1
$195K - $275K/yr
Enjoy a safe, flexible, and supportive work environment-remote or onsite-focused on employee ... GPU optimizations (OpenCL, CUDA, SYCL/DPC++, C for Metal or similar) * Parallel programming (OpenMP ...
Hillsboro, OR · On-site +1
$195K - $275K/yr
Enjoy a safe, flexible, and supportive work environment-remote or onsite-focused on employee ... GPU optimizations (OpenCL, CUDA, SYCL/DPC++, C for Metal or similar) * Parallel programming (OpenMP ...
| Aspect | Remote Cuda Developer | Remote Machine Learning Engineer |
|---|---|---|
| Required Credentials | CUDA programming certifications, computer science degree | Machine learning certifications, data science background |
| Work Environment | Software development, GPU optimization | Model development, data analysis |
| Industry Usage | High-performance computing, gaming, AI | AI, data science, predictive modeling |
Remote Cuda Developers focus on GPU programming and optimization using CUDA, primarily in high-performance computing and AI applications. Remote Machine Learning Engineers develop and deploy machine learning models, often utilizing GPU resources but with a broader focus on data and algorithms. While both roles may involve GPU expertise, Cuda Developers specialize in low-level programming, whereas Machine Learning Engineers work on model development and deployment.

Other
Posted 16 days ago
Locations: San Francisco or Remote
About The Role
The NEAR AI team is building decentralized and confidential machine learning infrastructure to enable user-owned AI. Our mission is to build highly scalable and efficient infrastructure for open-source AI at a global scale.
We are specifically seeking an expert in high-performance LLM serving systems and inference optimization. In this role, you will push the boundaries of how large language models are served.
What You'll Be Doing
What We're Looking For
We'd Love If You Have
Please let us know if you require any special requirements for your interview and we'll do our best to accommodate.