1

Llm Backend Engineer Jobs in California (NOW HIRING)

Our team has built the open-source LLM, NexusRaven-V2, rivaling GPT-4 in function calling with a ... Our Backend Engineers package up our technology in models and last-mile quality tooling. Our ...

We are looking for a senior engineer to join our team in San Francisco. As one of the founding ... Design and own the core backend infrastructure, including our multi-agent LLM architecture that ...

Backend Engineer

Los Angeles, CA · On-site

$100 - $140/hr

Our platform gives engineering teams end-to-end visibility into AI agents, LLM pipelines, and ML ... As a Backend Engineer, you will design, build, and operate the core infrastructure that makes this ...

Backend Engineer The Company Neon Redwood is a data services consulting company working on cutting ... LLM integrations. Responsibilities * Implement multi‑step agentic workflows using LLMs

Backend Engineer

San Francisco, CA · On-site

$120K - $200K/yr

Python, FastAPI, PostgreSQL, AWS, Docker, Kubernetes (nice to have), LLM/AI (bonus) As a Backend Engineer , you'll contribute directly to the core of our product and infrastructure. This role is ...

Python, FastAPI, PostgreSQL, AWS, Docker, Kubernetes (nice to have), LLM/AI (bonus) As a Backend Engineer , you'll contribute directly to the core of our product and infrastructure. This role is ...

Build and integrate LLM-powered workflows, AI agents, and agentic pipelines into production backend services. Collaborate with front-end developers, product managers, and cross-functional ...

... LLM-powered workflows, AI agents, and modern AI-assisted development practices. Minimum ... backend development experience, with solid foundational programming skills (algorithms, data ...

About the Role We're looking for a strong, generalist backend engineer to help build the platform ... LLM/AI model integrations and products * Strong product instincts and a builder's mindset ...

... backend engineer excited to build new features and scale out our ingestion and AI systems ... Shipped or otherwise interested in AI/LLM-native applications. * Iterated rapidly with customers or ...

Backend AI Engineer

San Francisco, CA · On-site

$140 - $230/hr

Design, build, and maintain backend services and APIs that power GenAI, LLM, and Computer Vision ... Mentor engineers and contribute to internal backend engineering best practices. Qualifications * 4 ...

We are seeking an experienced Python Backend Engineer to design, develop, and maintain scalable ... Knowledge of AI/LLM integration frameworks such as LangChain or LlamaIndex. Strong experience with ...

Python Backend Engineer

San Jose, CA · On-site

$100 - $150/hr

We are seeking an experienced Python Backend Engineer to design, develop, and maintain scalable ... Knowledge of AI/LLM integration frameworks such as LangChain or LlamaIndex. * Strong experience ...

... LLM responses, batch matching jobs, payment processing, and real-time chat -- while maintaining ... Design and build scalable backend services in Node.js/TypeScript, including RESTful APIs ...

next page

Showing results 1-20

Llm Backend Engineer information

What is an LLM backend engineer?

LLM Backend Engineers are software engineers who specialize in designing, building, and optimizing the backend infrastructure that supports large language models (LLMs) like GPT-4. They focus on integrating LLMs into products and services, ensuring scalable APIs, managing data pipelines, and optimizing inference performance. Their work often involves deploying models in cloud environments, monitoring system reliability, and collaborating with AI researchers to bring advancements into production. LLM Backend Engineers play a critical role in making AI-powered applications robust, efficient, and accessible to end users.

What are the key skills and qualifications needed to thrive as an LLM backend engineer?

To thrive as an LLM Backend Engineer, you need a solid foundation in software engineering, backend architecture, and experience working with large language models, typically supported by a degree in computer science or a related field. Proficiency with programming languages like Python or Java, cloud platforms (AWS, GCP, Azure), and machine learning frameworks such as TensorFlow or PyTorch is essential, along with familiarity with APIs and containerization tools like Docker or Kubernetes. Strong problem-solving, collaboration, and communication skills distinguish top performers in this role. These skills ensure robust, scalable, and efficient deployment of LLM-powered applications while enabling effective teamwork and innovation.

What are some common challenges faced by LLM backend engineers when deploying large language models in production?

LLM Backend Engineers often encounter challenges such as optimizing inference latency, managing high resource consumption, and ensuring scalability for production workloads. Balancing model performance with cost efficiency requires careful selection of hardware, batching strategies, and model quantization techniques. Additionally, they must address security and privacy concerns associated with handling sensitive data processed by the models. Collaboration with data scientists and DevOps teams is essential to streamline model updates and monitor system health.

What cities in California are hiring for Llm Backend Engineer jobs?

Cities in California with the most Llm Backend Engineer job openings:

Infographic showing various Llm Backend Engineer job openings in California as of August 2026, with employment types broken down into 91% Full Time, 6% Part Time, and 3% Contract. Highlights an 86% Physical, 5% Hybrid, and 9% Remote job distribution.

Full-time

Re-posted 28 days ago


Job description

About Nexusflow.ai

Modern enterprise copilots & agents call for last-mile quality, enterprise-grade robustness and scalable operation costs, beyond simplified programming interfaces for generative AI. Nexusflow tackles this challenge, enabling enterprises to own their workflow copilots & agents stacked on top of powerful yet cost-effective, compact LLMs. We train large language models and build last-mile quality dev tooling for copilots & agents on your enterprise workflows. Our team has built the open-source LLM, NexusRaven-V2, rivaling GPT-4 in function calling with a 100X smaller model size. Our team members are also behind the scenes of Starling, the #1 ranked compact 7B chat model based on human evaluation in Chatbot Arena.

Position: Backend Engineer

Nexusflow is currently adding Backend Engineers to our team. Our Backend Engineers package up our technology in models and last-mile quality tooling. Our Backend Engineers will be the driving force to build our products and solutions, in extensive collaboration with our ML Engineers and Front-end Engineers.

Responsibilities
  • API system development for copilot & agent quality tooling

  • API system development for copilot serving and integration with a focus on enterprise-grade requirements in the following areas

    • Integration with on-prem & cloud compute vendors

    • Integration with software tools required in customer oriented solutions

  • Distributed system and optionally GPU performance optimization

  • Wear many hats and collaborate with the whole team for product development, deployment and customer success

Qualification Required
  • Experience in ML model or ML data pipeline deployment (on-prem or on cloud)

  • Experience in building backend for application or platform API systems

Preferred
  • Working experience in fast-pace team environment  
  • Experience in using or contributing to modern compute frameworks for LLMs (e.g. Deepspeed, Huggingface TGI, FSDP)

  • Experience in projects involving LLMs