1

Llm Backend Engineer Jobs in California (NOW HIRING)

The company is hiring a Backend Engineer to build the reliable systems behind its AI products ... Experience building AI, LLM, agent, or machine-learning infrastructure * Experience integrating ...

Our team has built the open-source LLM, NexusRaven-V2, rivaling GPT-4 in function calling with a ... Our Backend Engineers package up our technology in models and last-mile quality tooling. Our ...

Senior Backend Engineer

San Francisco, CA · On-site

$200K - $260K/yr

  • Medical

  • Dental

  • Vision

Senior Backend Engineer LiteLLM is the world's most popular AI Gateway, trusted by top companies ... Build and design CPU-level guardrails to cover common attacks on LLM API's / MCP servers / Agents

We are looking for a senior engineer to join our team in San Francisco. As one of the founding ... Design and own the core backend infrastructure, including our multi-agent LLM architecture that ...

Backend Engineer

San Francisco, CA · On-site

$120K - $200K/yr

  • Medical

  • Dental

  • Retirement

  • PTO

Python, FastAPI, PostgreSQL, AWS, Docker, Kubernetes (nice to have), LLM/AI (bonus) As a Backend Engineer , you'll contribute directly to the core of our product and infrastructure. This role is ...

Backend Engineer

San Francisco, CA · On-site

  • Medical

  • Dental

  • Retirement

  • PTO

Python, FastAPI, PostgreSQL, AWS, Docker, Kubernetes (nice to have), LLM/AI (bonus) As a Backend Engineer , you'll contribute directly to the core of our product and infrastructure. This role is ...

Sr Python Backend Engineer, AIML

Cupertino, CA

$184K - $324K/yr

  • Medical

  • Dental

  • Retirement

Build and integrate LLM-powered workflows, AI agents, and agentic pipelines into production backend services. Collaborate with front-end developers, product managers, and cross-functional ...

Staff Backend Engineer

San Francisco, CA · Remote

$220K - $270K/yr

About the Role We're seeking a Staff Backend Engineer to help scale the backbone of our platform ... AI Development Tools: AI-powered coding assistants (Claude Code, Cursor, etc.), LLM-based ...

About the Role We're seeking a Staff Backend Engineer to help scale the backbone of our platform ... AI Development Tools: AI-powered coding assistants (Claude Code, Cursor, etc.), LLM-based ...

Staff Backend Engineer

San Francisco, CA · On-site

$220K - $270K/yr

About the Role We're seeking a Staff Backend Engineer to help scale the backbone of our platform ... AI Development Tools: AI-powered coding assistants (Claude Code, Cursor, etc.), LLM-based ...

Staff Backend Engineer

Los Angeles, CA · On-site

$200K - $230K/yr

About the Role Join our fast-growing startup as a Staff Backend Engineer and help scale the ... AI Development Tools: AI-powered coding assistants (Claude Code, Cursor, etc.), LLM-based ...

Staff Backend Engineer

Los Angeles, CA · On-site +1

$200K - $230K/yr

About the Role Join our fast-growing startup as a Staff Backend Engineer and help scale the ... AI Development Tools: AI-powered coding assistants (Claude Code, Cursor, etc.), LLM-based ...

About the Role Join our fast-growing startup as a Staff Backend Engineer and help scale the ... AI Development Tools: AI-powered coding assistants (Claude Code, Cursor, etc.), LLM-based ...

... LLM-powered workflows, AI agents, and modern AI-assisted development practices. Minimum ... backend development experience, with solid foundational programming skills (algorithms, data ...

next page

Showing results 1-20

Llm Backend Engineer information

What are some common challenges faced by LLM backend engineers when deploying large language models in production?

LLM Backend Engineers often encounter challenges such as optimizing inference latency, managing high resource consumption, and ensuring scalability for production workloads. Balancing model performance with cost efficiency requires careful selection of hardware, batching strategies, and model quantization techniques. Additionally, they must address security and privacy concerns associated with handling sensitive data processed by the models. Collaboration with data scientists and DevOps teams is essential to streamline model updates and monitor system health.

What is an LLM backend engineer?

LLM Backend Engineers are software engineers who specialize in designing, building, and optimizing the backend infrastructure that supports large language models (LLMs) like GPT-4. They focus on integrating LLMs into products and services, ensuring scalable APIs, managing data pipelines, and optimizing inference performance. Their work often involves deploying models in cloud environments, monitoring system reliability, and collaborating with AI researchers to bring advancements into production. LLM Backend Engineers play a critical role in making AI-powered applications robust, efficient, and accessible to end users.

What are the key skills and qualifications needed to thrive as an LLM backend engineer?

To thrive as an LLM Backend Engineer, you need a solid foundation in software engineering, backend architecture, and experience working with large language models, typically supported by a degree in computer science or a related field. Proficiency with programming languages like Python or Java, cloud platforms (AWS, GCP, Azure), and machine learning frameworks such as TensorFlow or PyTorch is essential, along with familiarity with APIs and containerization tools like Docker or Kubernetes. Strong problem-solving, collaboration, and communication skills distinguish top performers in this role. These skills ensure robust, scalable, and efficient deployment of LLM-powered applications while enabling effective teamwork and innovation.

What cities in California are hiring for Llm Backend Engineer jobs?

Cities in California with the most Llm Backend Engineer job openings:

Infographic showing various Llm Backend Engineer job openings in California as of August 2026, with employment types broken down into 88% Full Time, 5% Part Time, 2% Temporary, and 5% Contract. Highlights an 86% Physical, 5% Hybrid, and 9% Remote job distribution.

Backend Engineer

ProNexus

San Francisco, CA • On-site

Full-time

Re-posted 21 days ago


Job description

We are hiring on behalf of a venture-backed AI startup building intelligent software that automates complex business workflows.

The platform uses large language models, proprietary data, and human-in-the-loop systems to complete work that previously required significant manual effort. Its backend systems power AI agents, customer integrations, data processing, evaluation infrastructure, and production workflows.

The company is hiring a Backend Engineer to build the reliable systems behind its AI products.

About the role

As a Backend Engineer, you will design and build the services, APIs, data models, and infrastructure that power the company’s product.

You will work primarily in Python and PostgreSQL while collaborating closely with product engineers, AI engineers, customers, and company leadership.

This is not a narrow API-development role. You will work on production architecture, data systems, model integrations, security, reliability, and product capabilities.

What you’ll do
  • Design, build, and maintain production backend services using Python
  • Develop APIs that support customer-facing applications and AI workflows
  • Design scalable relational data models using PostgreSQL
  • Build integrations with model providers, customer systems, and third-party platforms
  • Develop systems for asynchronous jobs, event processing, and long-running AI workflows
  • Deploy and operate services using AWS
  • Debug production issues across application, data, model, and infrastructure layers
  • Help shape backend architecture and engineering practices as the company grows
What we’re looking for
  • Approximately 2–9 years of professional software engineering experience
  • Strong professional experience with Python
  • Experience designing and operating production APIs or backend services
  • Strong understanding of relational databases, SQL, and data modeling
  • Experience with PostgreSQL or a comparable production database
  • Familiarity with AWS, GCP, or another major cloud platform
  • Understanding of application security, authentication, and authorization
  • Experience debugging and operating systems in production
  • Ability to independently own technical projects from design through deployment
  • Comfort making pragmatic engineering decisions in an early-stage environment
  • Strong communication and cross-functional collaboration skills
Nice to have (not required)
  • Experience building AI, LLM, agent, or machine-learning infrastructure
  • Experience integrating OpenAI, Anthropic, or other model providers
  • Familiarity with LLM evaluation, retrieval, embeddings, or vector databases
  • Experience with Docker, Kubernetes, Terraform, or CI/CD systems
  • Experience with Redis, Kafka, queues, or event-driven architectures
  • Experience building multi-tenant SaaS products
  • Previous experience at an early-stage startup
  • Experience with data-intensive or distributed systems