1

Fastapi Programmer Jobs in California (NOW HIRING)

$127K - $166K/yr

We are looking for a GenAI Engineer with strong expertise in LLM infrastructure, model deployment ... Develop AI platform services and APIs using Python, FastAPI, Microservices, and Kubernetes.

$121K - $159K/yr

We are looking for a GenAI Engineer with strong expertise in LLM infrastructure, model deployment ... Develop AI platform services and APIs using Python, FastAPI, Microservices, and Kubernetes.

Python Backend Engineer

Sunnyvale, CA · On-site

$60 - $65/hr

Build and optimize RESTful APIs using FastAPI, Django, or similar frameworks. Develop and integrate ... Experience with asynchronous programming and task queues. Understanding of software architecture ...

$104K - $137K/yr

We are looking for a GenAI Engineer with strong expertise in LLM infrastructure, model deployment ... Develop AI platform services and APIs using Python, FastAPI, Microservices, and Kubernetes.

Product Engineer

San Francisco, CA · On-site

$115K - $150K/yr

Role We're hiring a Product Engineer to join our team and help us scale reliably as we grow. As we ... You'll work across our FastAPI (Python) / Postgres / React stack, touching everything from schema ...

AI Engineer

San Francisco, CA · On-site

$200K - $250K/yr

Python, LLM APIs, FastAPI As the AI Engineer, you'll be responsible for ensuring LiteLLM unifies the format for calling LLM APIs in the broader OpenAI + Anthropic spec. This involves writing ...

Python Backend Engineer

San Jose, CA · On-site

$125 - $150/hr

Build and optimize RESTful APIs using FastAPI, Django, or similar frameworks. * Develop and ... Experience with asynchronous programming and task queues. * Understanding of software architecture ...

Product Engineer

San Francisco, CA · On-site +1

$150K - $250K/yr

TypeScript, Python, SQL, cloud infrastructure, LLMs (React frontend, Python/FastAPI backend) Requirements * 0 to 5 years of experience as a full-stack or product engineer. * A track record of ...

Showing results 41-60

Fastapi Programmer information

What job categories do people searching Fastapi Programmer jobs in California look for?

The top searched job categories for Fastapi Programmer jobs in California are:

What cities in California are hiring for Fastapi Programmer jobs?

Cities in California with the most Fastapi Programmer job openings:

GenAI Engineer - LLM Infrastructure & Inference Services

On-site

2T Consulting
IT Services • 51 - 200 employees

$127K - $166K/yr

Full-time

Posted 11 days ago


Job description

We are looking for a GenAI Engineer with strong expertise in LLM infrastructure, model deployment, and high-performance inference services. The ideal candidate will build and manage scalable enterprise GenAI platforms across GPU infrastructure and cloud environments.

Key Responsibilities
  • Deploy, host, and manage Large Language Models (LLMs) on GPU infrastructure for production environments.
  • Build scalable, high-performance inference services using vLLM, TensorRT-LLM, Triton Inference Server, and Ray Serve.
  • Optimize model serving for latency, throughput, GPU utilization, and cost efficiency.
  • Develop AI platform services and APIs using Python, FastAPI, Microservices, and Kubernetes.
  • Implement RAG pipelines, vector databases, and agentic AI frameworks such as LangChain and LangGraph.
  • Manage GPU infrastructure, containerization, and cloud deployments across AWS, Azure, or GCP.
  • Establish MLOps/LLMOps practices including CI/CD, model deployment, monitoring, observability, and governance.
  • Perform performance tuning, benchmarking, capacity planning, and production support for enterprise GenAI platforms.
  • Collaborate with architects, data scientists, and product teams to deliver scalable, secure, and reliable AI solutions.
Core Technologies
  • LLM: vLLM, TensorRT-LLM, Triton Inference Server, Ray Serve
  • AI/GenAI: RAG, LangChain, LangGraph, Vector Databases
  • Development: Python, FastAPI, Microservices
  • Infrastructure: Kubernetes, Docker, GPU Infrastructure
  • Cloud: AWS, Azure, GCP
  • MLOps/LLMOps: CI/CD, Monitoring, Observability, Model Deployment, Governance