1

Operations Support Engineer Jobs in California (NOW HIRING)

Operations Support Specialist

Palo Alto, CA · On-site

$60K - $81K/yr

IT Operations Support Specialist/Release Engineer Work as a member of the Production Deployment operations staff that is responsible for Change management of Ariba Cloud based solutions. Collaborate ...

Bachelor's, Electronic Engineering, Electronic Engineering Technology, or equivalent military training * 3-5 years of relevant experience in test engineering, lab support, or technical operations

Ground Support Engineer

Mojave, CA · On-site

$110K - $140K/yr

  • Medical

  • Dental

  • Vision

  • Retirement

Ground Support Engineer Department: Assembly, Integration & Test Employment Type: Full Time ... Working closely with propulsion, controls, manufacturing, and test operations teams, you'll help ...

Support Engineer

San Francisco, CA · Remote

  • Medical

  • Dental

  • Vision

  • PTO

Our innovative platform is engineered from the ground up to boost operations efficiency and enhance support capabilities for property management business across the US and Canada, a ~$200B market.

Client is currently seeking multiple DevOps Support Engineers to join our team in Santa Clara, CA. As a member of the customer success team and reporting to the director of customer success, you will ...

Support Engineer

San Francisco, CA · On-site +1

  • Medical

  • Dental

  • Vision

  • PTO

Our innovative platform is engineered from the ground up to boost operations efficiency and enhance support capabilities for property management business across the US and Canada, a ~$200B market.

Product Support Engineer

San Diego, CA

  • Medical

  • Dental

  • Vision

  • Life

  • PTO

Hands-on experience with Azure DevOps, Jira, Confluence, or comparable systems * Ability to read ... support benefits. In addition to elective benefit options, benefited employees receive firm-paid ...

Client is currently seeking multiple DevOps Support Engineers to join our team in Santa Clara, CA. As a member of the customer success team and reporting to the director of customer success, you will ...

Product Support Engineer

Santa Monica, CA · On-site

  • Medical

  • Dental

  • Vision

  • Life

  • PTO

Hands-on experience with Azure DevOps, Jira, Confluence, or comparable systems * Ability to read ... support benefits. In addition to elective benefit options, benefited employees receive firm-paid ...

next page

Showing results 1-20

Operations Support Engineer information

See California salary details

$35.5K

$83.9K

$133.2K

How much do operations support engineer jobs pay per year?

As of Aug 13, 2026, the average yearly pay for operations support engineer in California is $83,915.00, according to ZipRecruiter salary data. Most workers in this role earn between $68,600.00 and $92,800.00 per year, depending on experience, location, and employer.

What are the key skills and qualifications needed to thrive as an operations support engineer?

To thrive as an Operations Support Engineer, you need a solid understanding of systems administration, troubleshooting, and incident management, typically backed by a degree in computer science or a related field. Familiarity with monitoring tools like Nagios or Splunk, scripting languages (such as Python or Bash), and ITIL or relevant technical certifications is highly valued. Strong problem-solving abilities, effective communication, and the capacity to remain calm under pressure help you excel in this role. These skills and qualities are crucial for minimizing downtime, ensuring system reliability, and supporting business continuity.

What does an operations support engineer do?

An Operations Support Engineer is responsible for maintaining and improving the day-to-day operations of IT systems and infrastructure. They monitor system performance, troubleshoot technical issues, and provide support to internal teams to ensure business processes run smoothly. Their duties often include incident response, deploying updates, and collaborating with other IT professionals to optimize system reliability and efficiency. This role is critical in minimizing downtime and ensuring that technology services meet organizational needs.

How does an operations support engineer typically collaborate with other departments to resolve technical issues?

Operations Support Engineers frequently work cross-functionally, partnering with teams such as development, quality assurance, and customer support to identify, troubleshoot, and resolve technical problems. They often act as a bridge between technical and non-technical staff, translating issues and coordinating solutions efficiently. Regular communication, participation in incident response meetings, and documentation of solutions are essential aspects of this collaboration. This team-oriented approach not only resolves issues faster but also helps prevent future incidents by sharing knowledge across departments.
What are popular job titles related to Operations Support Engineer jobs in California? For Operations Support Engineer jobs in California, the most frequently searched job titles are:
What job categories do people searching Operations Support Engineer jobs in California look for? The top searched job categories for Operations Support Engineer jobs in California are:
Infographic showing various Operations Support Engineer job openings in California as of August 2026, with employment types broken down into 1% As Needed, 71% Full Time, 23% Part Time, and 5% Contract. Highlights an 91% Physical, 2% Hybrid, and 7% Remote job distribution, with an average salary of $83,915 per year, or $40.3 per hour.

Systems Operations Support Engineer - Linux

Vast.ai Inc

Los Angeles, CA • On-site

$90K - $150K/yr

Full-time

Medical, Dental, Vision, Life, Retirement

Posted 12 days ago


Job description

About Us
Vast.ai's cloud powers AI projects and businesses all over the world. We are democratizing and decentralizing AI computing - reshaping our future for the benefit of humanity. Our mission is to organize, optimize, and orient the world's computation.
We value elegance, ownership, integrity, and continuous learning. You'll have the opportunity to dive into state-of-the-art AI systems while collaborating with a globally distributed team.
About the Role
This is a systems operations support role focused on deep-diving into escalated infrastructure issues that go beyond frontline triage. You'll be the engineering resource our L1 support team relies on when tickets become complex, investigating and resolving issues across the full infrastructure stack-including hardware, BIOS and firmware, networking, Ubuntu, Docker, NVIDIA CUDA and GPUs, and KVM-based virtual machines.
You'll own complex escalations end-to-end: gathering evidence, reproducing issues, identifying the root cause, proposing solutions, and working with the appropriate teams to bring each issue to resolution. The best engineers in this role don't just resolve individual tickets-they identify recurring patterns, improve operational tooling, and build runbooks that prevent future incidents. You'll collaborate directly with the engineering and host support teams on systemic infrastructure issues.
Strong Linux systems knowledge, technical depth, and support experience are the primary requirements. You should be comfortable working autonomously in Ubuntu environments, troubleshooting hardware, networking, containers, virtual machines, and GPU workloads, and clearly communicating your findings and proposed solutions to both technical and non-technical audiences.
Vast.ai users or hosts strongly preferred.
Location and Schedule
This is a full-time position based in our Westwood, Los Angeles office.
Available schedules:
  • Monday-Friday: Fully on-site
  • Sunday-Thursday: Four days on-site and one day working from home

Key Responsibilities
  • Handle escalated support tickets involving GPU workload failures, container issues, networking problems, account infrastructure, and host-side configuration
  • Provide managed support for supplier onboarding and ongoing machine management, acting as a technical resource through installation, configuration, and post-setup troubleshooting
  • Assist clients and infrastructure suppliers working with TensorFlow, PyTorch, and other GPU-accelerated workloads
  • Provide coverage for L1 support overflow during peak periods or incidents
  • Diagnose and resolve issues across Docker, NVIDIA CUDA/GPU drivers, and KVM virtualization environments
  • Troubleshoot network-layer issues, including VLAN, DNS, DHCP, VPN, NAT, firewall rules, and connectivity failures on host machines
  • Investigate performance issues involving GPU utilization, container resource constraints, thermal throttling, driver conflicts, and disk I/O bottlenecks
  • Advise suppliers on installation best practices, including hardware setup, driver configuration, BIOS/firmware settings, and network configuration for optimal performance
  • Write and maintain internal runbooks, escalation guides, and knowledge base articles to reduce repeat escalations
  • Build diagnostic and automation tooling in Python and Bash to reduce manual triage overhead
  • Collaborate with the engineering and support teams to flag and document systemic or recurring platform issues

You Are
  • Experienced with Linux, especially Ubuntu, and comfortable troubleshooting from the command line
  • Someone who enjoys debugging difficult problems and fixing broken systems
  • Methodical and focused on finding root causes, not just temporary fixes
  • Able to manage complex tickets independently
  • A clear written communicator with an interest in AI infrastructure and GPU computing

Must-Haves
  • Strong Linux systems operations experience with Ubuntu, RHEL/CentOS, or Debian, including networking, storage, services, and permissions
  • Proficiency with Docker, including container debugging, Docker Compose, image management, cgroup limits, and Docker storage and filesystem troubleshooting
  • Experience with virtualization platforms such as Proxmox VE, VMware, or similar hypervisors, including VM provisioning and troubleshooting
  • Strong networking fundamentals, including VLANs, DNS, DHCP, NAT, VPNs, firewall rules, and L2/L3 troubleshooting
  • Hands-on experience with NVIDIA GPU drivers, CUDA, and GPU workload troubleshooting
  • Python and Bash scripting skills for automation and diagnostic tooling
  • Strong written English communication that is clear, professional, and technically precise
  • Experience providing technical support in a customer-facing or internal help desk environment
  • Ability to prioritize across a concurrent queue of escalated tickets, triaging by severity and customer impact, balancing reactive resolution against proactive documentation and tooling work, and making clear judgment calls on when to escalate versus own resolution end-to-end

Nice-to-Haves
  • Familiarity with AI/ML frameworks (TensorFlow, PyTorch) and running GPU-accelerated containers
  • Monitoring and observability experience (Prometheus, Grafana)
  • Relevant certifications: RHCSA, CompTIA Linux+, or similar
  • Knowledge of the Vast.ai platform as a client or infrastructure supplier
Interview Process (~1 week)
After you submit your application, our technical team will review your experience and qualifications. Selected candidates will proceed through the following stages:
  • 15 minutes - Initial Screening (Virtual): A brief conversation about your background, availability, and interest in the role
  • 45 minutes - Experience Interview (Virtual): An introduction to Vast.ai and a deeper discussion of your technical and support experience
  • 2 hours - Meet and Greet and Technical Assessment (On-site): Meet the team and complete an LLM-assisted Linux systems operations assessment

Annual Salary Range
$90,000 - $160,000 + equity + benefits
Vast.ai is hiring across all experience levels with compensation commensurate with background, experience and potential.
Benefits
  • Comprehensive health, dental, vision, and life insurance
  • 401(k) with company match
  • Meaningful early-stage equity
  • Onsite meals, snacks, and close collaboration with founders/tech leaders
  • Ambitious, fast-paced startup culture where initiative is rewarded