1

Python Llm Jobs in Oakland, CA (NOW HIRING)

The ideal candidate will have strong experience working with Python, LLM solution patterns and tools (RAG, Vector DB, Agentic workflows, LoRA, etc.) cloud platforms (AWS, Databricks and Azure), and ...

Agentic AI/Python Engineer Location: Concord, CA (3days in hybrid) Duration: 12+ Months Contract ... Develop LLM-powered applications using RAG, vector search, embeddings, and related technologies.

LLM Inference Engineer

Menlo Park, CA · On-site

$165K/yr

Proficiency in Python and C++ * Experience with CUDA programming and GPU optimization Nice-to-Have: * Contributions to open-source inference frameworks such as vLLM, SGLang, or TensorRT-LLM

Lead AI / LLM Engineer

San Mateo, CA · On-site

$200 - $250/hr

As Lead AI/LLM Engineer, this is a hands-on role, you'll build and deploy the AI systems powering ... Snowflake -- SQL and stored procedures (JavaScript, Python) * Event-driven data pipelines ...

Agentic AI Engineer II

San Mateo, CA · On-site

$116K - $157K/yr

Python * LLM-powered systems or backend frameworks Visa sponsorship not available for this role #LI-CH2 Benefits We offer versatile health perks, including flexible spending accounts, HSA, a 401(k) ...

Showing results 21-40

Python Llm information

See Oakland, CA salary details

$15

$67

$99

How much do python llm jobs pay per hour?

As of Sep 7, 2026, the average hourly pay for python llm in Oakland, CA is $67.33, according to ZipRecruiter salary data. Most workers in this role earn between $55.48 and $76.49 per hour, depending on experience, location, and employer.

What is a Python LLM?

A Python LLM job involves working with Large Language Models (LLMs) using Python to develop, fine-tune, and deploy AI models. Responsibilities may include data preprocessing, prompt engineering, model optimization, and integration with applications. Professionals in this role often work with frameworks like TensorFlow, PyTorch, or Hugging Face Transformers. They may also contribute to improving model efficiency, reducing bias, and ensuring ethical AI usage.

What are the key skills and qualifications needed to thrive in the Python LLM position, and why are they important?

To excel as a Python LLM (Large Language Model) Engineer, you need strong skills in Python programming, machine learning, and natural language processing, typically supported by a degree in computer science or a related field. Proficiency with libraries such as TensorFlow, PyTorch, Hugging Face Transformers, and experience with model deployment platforms are often essential, alongside certifications in AI or data science. Effective communication, problem-solving abilities, and collaboration are important soft skills for working in interdisciplinary teams and delivering results in dynamic environments. These skills ensure the development, fine-tuning, and deployment of advanced language models that meet both technical and business objectives.

What are some common challenges faced by Python LLM engineers in their daily work?

Python LLM Engineers often encounter challenges related to optimizing model performance, managing large datasets, and adapting models to specific business needs. Working with large-scale language models requires balancing computational resource limitations with the need for high accuracy and efficiency. Collaboration with data scientists, product managers, and DevOps engineers is routine to ensure seamless model integration and deployment. Staying updated on the latest advancements in NLP and continuously improving models based on user feedback are also important aspects of the role.

What are popular job titles related to Python Llm jobs in Oakland, CA?

For Python Llm jobs in Oakland, CA, the most frequently searched job titles are:

What job categories do people searching Python Llm jobs in Oakland, CA look for?

The top searched job categories for Python Llm jobs in Oakland, CA are:

What cities near Oakland, CA are hiring for Python Llm jobs?

Cities near Oakland, CA with the most Python Llm job openings:

Infographic showing various Python Llm job openings in Oakland, CA as of August 2026, with employment types broken down into 1% Internship, 86% Full Time, 8% Part Time, and 5% Contract. Highlights an 78% Physical, 6% Hybrid, and 16% Remote job distribution, with an average salary of $140,036 per year, or $67.3 per hour.

GenAI Engineer - LLM Infrastructure & Inference Services

2T Consulting

Mountain View, CA • On-site

$126K - $166K/yr

Full-time

Posted 10 days ago


Job description

We are looking for a GenAI Engineer with strong expertise in LLM infrastructure, model deployment, and high-performance inference services. The ideal candidate will build and manage scalable enterprise GenAI platforms across GPU infrastructure and cloud environments.

Key Responsibilities
  • Deploy, host, and manage Large Language Models (LLMs) on GPU infrastructure for production environments.
  • Build scalable, high-performance inference services using vLLM, TensorRT-LLM, Triton Inference Server, and Ray Serve.
  • Optimize model serving for latency, throughput, GPU utilization, and cost efficiency.
  • Develop AI platform services and APIs using Python, FastAPI, Microservices, and Kubernetes.
  • Implement RAG pipelines, vector databases, and agentic AI frameworks such as LangChain and LangGraph.
  • Manage GPU infrastructure, containerization, and cloud deployments across AWS, Azure, or GCP.
  • Establish MLOps/LLMOps practices including CI/CD, model deployment, monitoring, observability, and governance.
  • Perform performance tuning, benchmarking, capacity planning, and production support for enterprise GenAI platforms.
  • Collaborate with architects, data scientists, and product teams to deliver scalable, secure, and reliable AI solutions.
Core Technologies
  • LLM: vLLM, TensorRT-LLM, Triton Inference Server, Ray Serve
  • AI/GenAI: RAG, LangChain, LangGraph, Vector Databases
  • Development: Python, FastAPI, Microservices
  • Infrastructure: Kubernetes, Docker, GPU Infrastructure
  • Cloud: AWS, Azure, GCP
  • MLOps/LLMOps: CI/CD, Monitoring, Observability, Model Deployment, Governance