Near Ai

60 Near Ai Jobs Hiring Near You

San Francisco or Remote About The Role The NEAR AI team is building decentralized and confidential machine learning infrastructure to enable user-owned AI. Our mission is to build highly scalable and ...

San Francisco or Remote About The Role The NEAR AI team is building decentralized and confidential machine learning infrastructure to enable user-owned AI. Our mission is to build highly scalable and ...

About Traversal Traversal is the AI Site Reliability Engineer (SRE) for the enterprise--already ... Traversal is fully in‑office, 5 days a week, based in New York near Madison Square Park. We have ...

Job Title: - Senior AI Agent Engineer Location: - McLean, VA - 5 days onsite Employment Type: - Contract Need 10+ years of experience resumes and need locals to VA or nearby location. - We are ...

Implement sophisticated retrieval systems that surface contextually relevant information in near ... AI Innovation Playground: Transform cutting-edge AI concepts into magical experiences * Direct ...

SAIGroup's portfolio serves 2,000+ global enterprise customers, generates nearly $800M in annual revenue, and employs 4,000+ people worldwide -- providing JazzX AI with long-term capital, deep ...

Raise the quality of nearby systems through code review, documentation, design discussion, and ... AI-native working habits paired with independent technical judgment. Benefits * Competitive Salary ...

Raise the quality of nearby systems through code review, documentation, design discussion, and ... AI-native working habits paired with independent technical judgment. * Competitive Salary and Stock ...

Raise the quality of nearby systems through code review, documentation, design discussion, and ... AI-native working habits paired with independent technical judgment. Benefits * Competitive Salary ...

Raise the quality of nearby systems through code review, documentation, design discussion, and ... AI-native working habits paired with independent technical judgment. Benefits * Competitive Salary ...

Raise the quality of nearby systems through code review, documentation, design discussion, and ... AI-native working habits paired with independent technical judgment. Benefits * Competitive Salary ...

Raise the quality of nearby systems through code review, documentation, design discussion, and ... AI-native working habits paired with independent technical judgment. Benefits * Competitive Salary ...

Raise the quality of nearby systems through code review, documentation, design discussion, and ... AI-native working habits paired with independent technical judgment. Benefits * Competitive Salary ...

Raise the quality of nearby systems through code review, documentation, design discussion, and ... AI-native working habits paired with independent technical judgment. Benefits * Competitive Salary ...

$112 - $140/hr

Motive serves nearly 100,000 customers - from Fortune 500 enterprises to small businesses - across ... As a GTM AI Engineer on the AI Operations team, you'll be at the forefront of our AI strategy ...

next page

Showing results 1-20

LLM Inference Engineer

Near AI

OR • On-site, Remote

Full-time

Re-posted 23 days ago


Job description

Locations: San Francisco or Remote

About The Role

The NEAR AI team is building decentralized and confidential machine learning infrastructure to enable user-owned AI. Our mission is to build highly scalable and efficient infrastructure for open-source AI at a global scale.

We are specifically seeking an expert in high-performance LLM serving systems and inference optimization. In this role, you will push the boundaries of how large language models are served.

What You'll Be Doing

  • Architect and maintain production high-traffic LLM serving systems.
  • Optimize throughput, latency, and cost for leading open-source LLMs.

What We're Looking For

  • Strong hands-on experience in LLM inference, with expertise debugging and optimizing major inference engines such as SGLang, vLLM, or TensorRT.
  • Deep knowledge of state-of-the-art GPU architectures, and effectively exploit them using PyTorch, Triton, CuTe, CUDA, etc.
  • Proven track record in designing and maintaining end-to-end high-traffic LLM serving systems.
  • Strong problem-solving skills and ability to communicate technical ideas clearly.

We'd Love If You Have

  • Experience with Trusted Execution Environments (TEE).
  • Active contributor to open-source LLM inference engines.

Please let us know if you require any special requirements for your interview and we'll do our best to accommodate.