1

Python Llm Jobs in Mountain View, CA (NOW HIRING)

Python+ React FDE with AI/ML

Sunnyvale, CA ยท On-site

$59.75 - $82.50/hr

Candidate should be able to build production-grade backend services in Python and modern UIs in ... LLM APIs (Claude, GPT-4, Gemini) and agentic AI workflows into customer-facing products โ€ข ...

New

Senior AI Engineer

South San Francisco, CA ยท On-site

$125K - $138K/yr

The ideal candidate will have strong experience working with Python, LLM solution patterns and tools (RAG, Vector DB, Agentic workflows, LoRA, etc.) cloud platforms (AWS, Databricks and Azure), and ...

... LLM GenAI. The role involves writing efficient machine learning workflows, evaluating model ... Python NLP libraries (NLTK, gensim, spacy etc.) โ€ข Good experience with deploying production ...

LLM Inference Engineer

Menlo Park, CA ยท On-site

$165K/yr

Proficiency in Python and C++ * Experience with CUDA programming and GPU optimization Nice-to-Have: * Contributions to open-source inference frameworks such as vLLM, SGLang, or TensorRT-LLM

Lead AI / LLM Engineer

San Mateo, CA ยท On-site

$116K - $153K/yr

As Lead AI/LLM Engineer, this is a hands-on role, you'll build and deploy the AI systems powering ... Snowflake -- SQL and stored procedures (JavaScript, Python) * Event-driven data pipelines ...

Python * LLM-powered systems or backend frameworks Visa sponsorship not available for this role #LI-CH2 Benefits We offer versatile health perks, including flexible spending accounts, HSA, a 401(k) ...

Senior Python Engineer - Agentic AI Location: San Francisco, California (Hybrid) Duration: 3 months ... LLM routing and fallback strategies Fine-tuning evaluation Model evaluation frameworks ...

New

Showing results 41-60

Python Llm information

See Mountain View, CA salary details

$15

$69

$101

How much do python llm jobs pay per hour?

As of Sep 12, 2026, the average hourly pay for python llm in Mountain View, CA is $69.15, according to ZipRecruiter salary data. Most workers in this role earn between $57.02 and $78.56 per hour, depending on experience, location, and employer.

What is a Python LLM?

A Python LLM job involves working with Large Language Models (LLMs) using Python to develop, fine-tune, and deploy AI models. Responsibilities may include data preprocessing, prompt engineering, model optimization, and integration with applications. Professionals in this role often work with frameworks like TensorFlow, PyTorch, or Hugging Face Transformers. They may also contribute to improving model efficiency, reducing bias, and ensuring ethical AI usage.

What are the key skills and qualifications needed to thrive in the Python LLM position, and why are they important?

To excel as a Python LLM (Large Language Model) Engineer, you need strong skills in Python programming, machine learning, and natural language processing, typically supported by a degree in computer science or a related field. Proficiency with libraries such as TensorFlow, PyTorch, Hugging Face Transformers, and experience with model deployment platforms are often essential, alongside certifications in AI or data science. Effective communication, problem-solving abilities, and collaboration are important soft skills for working in interdisciplinary teams and delivering results in dynamic environments. These skills ensure the development, fine-tuning, and deployment of advanced language models that meet both technical and business objectives.

What are some common challenges faced by Python LLM engineers in their daily work?

Python LLM Engineers often encounter challenges related to optimizing model performance, managing large datasets, and adapting models to specific business needs. Working with large-scale language models requires balancing computational resource limitations with the need for high accuracy and efficiency. Collaboration with data scientists, product managers, and DevOps engineers is routine to ensure seamless model integration and deployment. Staying updated on the latest advancements in NLP and continuously improving models based on user feedback are also important aspects of the role.

What are popular job titles related to Python Llm jobs in Mountain View, CA?

For Python Llm jobs in Mountain View, CA, the most frequently searched job titles are:

What job categories do people searching Python Llm jobs in Mountain View, CA look for?

The top searched job categories for Python Llm jobs in Mountain View, CA are:

What cities near Mountain View, CA are hiring for Python Llm jobs?

Cities near Mountain View, CA with the most Python Llm job openings:

GenAI Engineer - LLM Infrastructure & Inference Services

San Jose, CA โ€ข On-site

2T Consulting
IT Servicesย โ€ขย 51 - 200 employees

$126K - $166K/yr

Full-time

Posted 15 days ago


Job description

We are looking for a GenAI Engineer with strong expertise in LLM infrastructure, model deployment, and high-performance inference services. The ideal candidate will build and manage scalable enterprise GenAI platforms across GPU infrastructure and cloud environments.

Key Responsibilities
  • Deploy, host, and manage Large Language Models (LLMs) on GPU infrastructure for production environments.
  • Build scalable, high-performance inference services using vLLM, TensorRT-LLM, Triton Inference Server, and Ray Serve.
  • Optimize model serving for latency, throughput, GPU utilization, and cost efficiency.
  • Develop AI platform services and APIs using Python, FastAPI, Microservices, and Kubernetes.
  • Implement RAG pipelines, vector databases, and agentic AI frameworks such as LangChain and LangGraph.
  • Manage GPU infrastructure, containerization, and cloud deployments across AWS, Azure, or GCP.
  • Establish MLOps/LLMOps practices including CI/CD, model deployment, monitoring, observability, and governance.
  • Perform performance tuning, benchmarking, capacity planning, and production support for enterprise GenAI platforms.
  • Collaborate with architects, data scientists, and product teams to deliver scalable, secure, and reliable AI solutions.
Core Technologies
  • LLM: vLLM, TensorRT-LLM, Triton Inference Server, Ray Serve
  • AI/GenAI: RAG, LangChain, LangGraph, Vector Databases
  • Development: Python, FastAPI, Microservices
  • Infrastructure: Kubernetes, Docker, GPU Infrastructure
  • Cloud: AWS, Azure, GCP
  • MLOps/LLMOps: CI/CD, Monitoring, Observability, Model Deployment, Governance