1

Ceph Storage Jobs in California (NOW HIRING)

Senior HPC Storage Engineer

Santa Clara, CA · On-site

$184 - $287.50/hr

Distributed Storage Expertise - Extensive experience with parallel and distributed filesystems (Ceph, Weka.io, Vast, Lustre, GPFS) and Linux storage kernel development. * GPU & AI Infrastructure ...

S3, NFS) and file systems such as Ceph, DAOS, or similar. * Proficiency in a systems programming ... Familiarity with storage observability tools and telemetry pipelines (e.g., ClickHouse, Prometheus ...

Hands-on experience with distributed storage systems (e.g., Ceph, GlusterFS, OpenEBS, Vast, Lightbits) * Contributions to open-source storage projects or the Linux storage stack * Experience with ...

Hands-on experience with distributed storage systems (e.g., Ceph, GlusterFS, OpenEBS, Vast, Lightbits) * Contributions to open-source storage projects or the Linux storage stack * Experience with ...

Extensive experience with parallel and distributed filesystems (Ceph, Weka.io, Vast, Lustre, GPFS) and Linux storage kernel development. * GPU & AI Infrastructure: Proficient with NVIDIA GPUs, CUDA ...

Contribute to engineering projects on cloud storage features by collaborating with product and ... Ceph, GlusterFS, OpenEBS). • Proven experience in system programming with C, C++, and/or Rust ...

Showing results 21-40

Ceph Storage information

What is Ceph Storage?

Ceph Storage is an open-source, distributed storage system designed to provide excellent performance, reliability, and scalability. It supports object, block, and file storage in a unified platform, making it suitable for cloud infrastructure and enterprise environments. Ceph automatically manages data replication and recovery, helping to ensure high availability and fault tolerance. The system can run on commodity hardware, reducing costs and simplifying expansion as storage needs grow.

What are the key skills and qualifications needed to thrive as a Ceph Storage engineer, and why are they important?

To thrive as a Ceph Storage Engineer, you need a strong background in Linux systems administration, networking, and distributed storage concepts, often supported by a relevant degree or certifications like Red Hat Certified Engineer (RHCE). Familiarity with tools such as Ceph, Ansible, monitoring systems like Prometheus, and scripting languages is typically required. Critical soft skills include problem-solving, attention to detail, and effective communication for troubleshooting and collaborating with IT teams. These abilities are vital to ensuring high availability, scalability, and reliability of storage solutions in enterprise environments.

What are some common challenges faced by professionals working in Ceph Storage administration, and how can they be addressed?

Professionals managing Ceph Storage environments often encounter challenges such as maintaining cluster health, balancing performance and scalability, and troubleshooting hardware or network failures. These issues can be addressed by regularly monitoring cluster metrics, using Ceph's built-in tools for diagnosis, and proactively planning for capacity expansion. Collaborating closely with system administrators and network engineers is also crucial to ensure optimal integration and quick resolution of issues. Staying updated with Ceph community best practices and documentation helps administrators adapt to evolving technologies and maintain robust, high-performing storage solutions.

What is the difference between Ceph Storage vs Storage Engineer?

AspectCeph StorageStorage Engineer
CredentialsKnowledge of distributed storage, Linux, and storage protocolsCertifications like Cisco, CompTIA Storage+, or vendor-specific certifications
Work EnvironmentData centers, cloud environments, large-scale storage deploymentsData centers, enterprise IT teams, cloud providers
Industry UsageOpen-source storage solutions, cloud infrastructureDesign, implementation, and management of storage systems
Search/Comparison IntentTechnical understanding of storage solutionsCareer options, job roles, and skills

Ceph Storage is an open-source, distributed storage platform used to build scalable storage clusters, often in cloud and data center environments. Storage Engineers design, deploy, and maintain storage systems, including Ceph, focusing on performance and reliability. While Ceph Storage refers to the technology itself, Storage Engineers are professionals who work with Ceph and other storage solutions to meet organizational needs.

What job categories do people searching Ceph Storage jobs in California look for?

The top searched job categories for Ceph Storage jobs in California are:

What cities in California are hiring for Ceph Storage jobs?

Cities in California with the most Ceph Storage job openings:

Infographic showing various Ceph Storage job openings in California as of August 2026, with employment types broken down into 76% Full Time, 8% Part Time, and 16% Contract. Highlights an 92% In-person, and 8% Remote job distribution.

$277K/yr

Full-time

Posted 27 days ago


Job description

Special Instructions to Applicants
This position will consider both hybrid and fully remote candidates. On-site presence may be required as needed based on operational, maintenance, or project requirements. Participation in an on-call rotation for emergency hardware and software support is required. Occasional evenings and weekends may be required to support maintenance windows or respond to incidents. Required to accommodate scheduled based on Pacific Standard Time.
Department Summary
Advanced Research Computing melds expert staff and technical infrastructure to amplify and accelerate the impact of UCLA research in the age of networked data and computation. OARC's expertise and resources are available to all UCLA researchers engaged in digital research and scholarship. We work with faculty, student, and postdoctoral researchers; instructors; and staff and administrators.
OARC is a relationship-building organization. We enable digital scholarship through collaborations, partnerships, and networked communities to advance cutting-edge research capabilities at UCLA and beyond. OARC supports and enhances the university mission of education, research, and service through the development and execution of innovative and sustainable technology practices, programs, services, infrastructure, policies, and partnerships.
Position Summary
Join UCLA's Office of Advanced Research Computing (OARC) and help build the research data infrastructure that powers scientific discovery at UCLA and across a national network of high-performance computing centers. We are seeking a Senior HPC Storage Engineer to lead the design, deployment, and operation of large-scale storage systems supporting data-intensive research, AI/ML, interactive computing, and multi-institutional collaborations.
This senior technical role combines architecture with hands-on engineering. You will evaluate emerging storage technologies, lead deployment of petabyte-scale storage platforms, optimize performance, and develop secure, reliable storage services spanning campus, cloud, and federated HPC environments. As part of UCLA's NSF-supported iDLab initiative, you will help build a national-scale research data infrastructure that enables seamless access to data across geographically distributed computing resources.
At OARC, you will work alongside leading researchers, engineers, and partner institutions while influencing the long-term direction of research cyberinfrastructure at UCLA. This is an opportunity to solve challenging technical problems, work with advanced storage technologies, and make a lasting impact on research computing at both the campus and national levels.
Salary & Compensation
*UCLA provides a full pay range. Actual salary offers consider factors, including budget, prior experience, skills, knowledge, abilities, education, licensure and certifications, and other business considerations. Salary offers at the top of the range are not common. Visit UC Benefit package to discover benefits that start on day one, and UC Total Compensation Estimator to calculate the total compensation value with benefits.
Qualifications
  • 7 years or more Experience in research, enterprise, or hyperscale storage environments with responsibility for large-scale production storage services. (Required)
  • Large scale storage Experience with petabyte-scale storage, federated storage, multi-site research infrastructure, or HPC/cloud-integrated storage services. (Preferred)
  • 1. Advanced knowledge of high-performance computing, data science, and cyberinfrastructure environments supporting large-scale research workloads (Required)
  • 2. Demonstrated knowledge of scale-out, parallel, distributed, object, and federated storage architectures, including performance characteristics, consistency models, caching strategies, and operational tradeoffs. (Required)
  • 3. Hands-on experience architecting, deploying, operating, or substantially supporting one or more large-scale storage platforms such as Lustre, VAST Data, GPFS/Spectrum Scale, Ceph, BeeGFS, WekaFS, or MinIO. Experience with multiple platforms and petabyte-scale deployments is preferred. (Required)
  • 4. Demonstrated experience operating large-scale production storage systems, including monitoring, capacity planning, lifecycle management, upgrades, performance tuning, incident response, backup, replication, and disaster recovery. (Required)
  • 5. Advanced Linux systems administration skills, including kernel-level troubleshooting, performance profiling, storage hardware diagnostics, and tuning across InfiniBand, RoCE, and Ethernet fabrics. (Required)
  • 6. Demonstrated experience diagnosing storage performance problems across Linux clients, metadata services, storage servers, object services, network fabrics, protocol layers, and application I/O patterns. (Required)
  • 7. Proficiency in scripting and automation using Bash, Python, or similar languages; familiarity with configuration management tools such as Ansible and version control such as Git. (Required)
  • 8. Experience integrating storage systems with local, campus-wide, and federated identity providers such as Active Directory, LDAP, CILogon, OIDC, SAML, or Globus Auth. (Required)
  • 9. Demonstrated ability to design and execute storage benchmarks, validate vendor claims, characterize representative scientific workloads, and document results for engineering and executive audiences. (Required)
  • 10. Demonstrated ability to design, operate, or support shared research storage services, including allocation models, user-facing service delivery, large-scale data movement, and escalation support for complex research workflows. (Preferred)
  • 11. Demonstrated ability to communicate complex technical information clearly to technical staff, researchers, leadership, vendors, and external research and education audiences. (Required)
  • 12. Demonstrated ability to work independently and collaboratively, manage competing priorities, lead technical working groups, and sustain production service quality while delivering multi-month projects. (Required)

  • Education, Licenses, Certifications & Personal Affiliations
  • Bachelor's Degree Bachelor's degree in Computer Science, Computational Science, Data Science, Engineering, or a related field; or equivalent combination of education and experience. (Required)
  • Master's Degree Master's in Computer Science, Computational Science, Data Science, Engineering, or a related field; or equivalent advanced professional experience. (Preferred) Or
  • PhD PhD in Computer Science, Computational Science, Data Science, Engineering, or a related field; or equivalent advanced professional experience. (Preferred)

  • Special Conditions for Employment
  • Background Check: Continued employment is contingent upon the completion of a satisfactory background investigation.
  • Live Scan Background Check: A Live Scan background check must be completed prior to the start of employment.
  • This position will consider both hybrid and fully remote candidates. On-site presence may be required as needed based on operational, maintenance, or project requirements. Participation in an on-call rotation for emergency hardware and software support is required. Occasional evenings and weekends may be required to support maintenance windows or respond to incidents. (Required)

  • Schedule
    8:00 a.m. to 5:00 p.m
    Union/Policy Covered
    RP-Research and Public Service PR
    Complete Position Description
    https://universityofcalifornia.marketpayjobs.com/ShowJob.aspx?EntityID=38&JDName=Computational%20and%20Data%20Science%20Research%20Specialist%204%20RP%20(TBD_1000126)