1

Ceph Storage Jobs in Virginia (NOW HIRING)

Senior DevOps Engineer

Reston, VA · On-site

$150 - $200/hr

Experience operating storage solutions (CSI, Rook/Ceph). * Cloud-native observability experience -- Prometheus, Grafana, Loki, and Tempo. * Experience with Security Standards -- FIPS, CVE mitigation ...

Senior DevOps Engineer

Reston, VA · On-site

$135K - $173K/yr

Experience operating storage solutions (CSI, Rook/Ceph). * Cloud-native observability experience - Prometheus, Grafana, Loki, and Tempo. * Experience with Security Standards - FIPS, CVE mitigation ...

Showing results 41-44

Ceph Storage information

What is Ceph Storage?

Ceph Storage is an open-source, distributed storage system designed to provide excellent performance, reliability, and scalability. It supports object, block, and file storage in a unified platform, making it suitable for cloud infrastructure and enterprise environments. Ceph automatically manages data replication and recovery, helping to ensure high availability and fault tolerance. The system can run on commodity hardware, reducing costs and simplifying expansion as storage needs grow.

What are the key skills and qualifications needed to thrive as a Ceph Storage engineer, and why are they important?

To thrive as a Ceph Storage Engineer, you need a strong background in Linux systems administration, networking, and distributed storage concepts, often supported by a relevant degree or certifications like Red Hat Certified Engineer (RHCE). Familiarity with tools such as Ceph, Ansible, monitoring systems like Prometheus, and scripting languages is typically required. Critical soft skills include problem-solving, attention to detail, and effective communication for troubleshooting and collaborating with IT teams. These abilities are vital to ensuring high availability, scalability, and reliability of storage solutions in enterprise environments.

What are some common challenges faced by professionals working in Ceph Storage administration, and how can they be addressed?

Professionals managing Ceph Storage environments often encounter challenges such as maintaining cluster health, balancing performance and scalability, and troubleshooting hardware or network failures. These issues can be addressed by regularly monitoring cluster metrics, using Ceph's built-in tools for diagnosis, and proactively planning for capacity expansion. Collaborating closely with system administrators and network engineers is also crucial to ensure optimal integration and quick resolution of issues. Staying updated with Ceph community best practices and documentation helps administrators adapt to evolving technologies and maintain robust, high-performing storage solutions.

What is the difference between Ceph Storage vs Storage Engineer?

AspectCeph StorageStorage Engineer
CredentialsKnowledge of distributed storage, Linux, and storage protocolsCertifications like Cisco, CompTIA Storage+, or vendor-specific certifications
Work EnvironmentData centers, cloud environments, large-scale storage deploymentsData centers, enterprise IT teams, cloud providers
Industry UsageOpen-source storage solutions, cloud infrastructureDesign, implementation, and management of storage systems
Search/Comparison IntentTechnical understanding of storage solutionsCareer options, job roles, and skills

Ceph Storage is an open-source, distributed storage platform used to build scalable storage clusters, often in cloud and data center environments. Storage Engineers design, deploy, and maintain storage systems, including Ceph, focusing on performance and reliability. While Ceph Storage refers to the technology itself, Storage Engineers are professionals who work with Ceph and other storage solutions to meet organizational needs.

What job categories do people searching Ceph Storage jobs in Virginia look for?

The top searched job categories for Ceph Storage jobs in Virginia are:

What cities in Virginia are hiring for Ceph Storage jobs?

Cities in Virginia with the most Ceph Storage job openings:

Infographic showing various Ceph Storage job openings in Virginia as of September 2026, with employment types broken down into 71% Full Time, and 29% Nights. Highlights an 74% In-person, and 26% Remote job distribution.

Senior DevOps Engineer

Reston, VA • On-site

$150 - $200/hr

Other

Medical, Dental, Vision, Life, Retirement, PTO

Posted 4 days ago


Job description

We are looking for talented Platform/DevOps engineer with deep expertise in Kubernetes, Database technologies, Terraform, and Ansible to help scale ••••• ’s AI platform across on-premises, cloud, and SaaS environments. The right candidate will design infrastructure and deploy ••••• software to support our US Government customers in fully realizing the value of AI.

This role demands a strong foundation in Linux, networking (both traditional and Kubernetes), container technologies, and automation. You’ll collaborate closely with internal and external engineering teams, own critical infrastructure, and solve challenging operational and scalability problems in fast-paced, dynamic environments.

From your first day, you will make a valuable — and valued — contribution. We are a fast-growing company where no one is a bystander. We offer you the opportunity to delight millions of consumers around the world while gaining meaningful experience across a variety of disciplines.

We are looking for talented Platform/DevOps engineer with deep expertise in Kubernetes, Database technologies, Terraform, and Ansible to help scale ••••• ’s AI platform across on-premises, cloud, and SaaS environments. The right candidate will design infrastructure and deploy ••••• software to support our US Government customers in fully realizing the value of AI.

This role demands a strong foundation in Linux, networking (both traditional and Kubernetes), container technologies, and automation. You’ll collaborate closely with internal and external engineering teams, own critical infrastructure, and solve challenging operational and scalability problems in fast-paced, dynamic environments.

From your first day, you will make a valuable — and valued — contribution. We are a fast-growing company where no one is a bystander. We offer you the opportunity to delight millions of consumers around the world while gaining meaningful experience across a variety of disciplines.

Duties and Responsibilities:

  • Design, architect, and operate highly available, multi-tenant Kubernetes platforms across cloud and on-premises environments.
  • Own the full networking stack — CNI, service mesh, ingress, DNS, load balancing, and network policy.
  • Own database operations for production Postgres and Opensearch clusters — availability, performance, backup, and recovery.
  • Partner with engineering teams to define platform and delivery standards.
  • Automate infrastructure provisioning, configuration, and lifecycle management.
  • Enforce security best practices across the platform — network policies, RBAC, secrets management, and vulnerability patching.
  • Lead incident response and post-mortems for platform-level failures; implement systemic fixes.
  • Proactively identify and remediate scalability and reliability risks across the platform.

Skills and Qualifications:

  • Active U.S. DoD Top Secret clearance required
  • 5+ years in Platform Engineering, Infrastructure Engineering, or DevOps supporting large-scale distributed systems.
  • 5+ years of Kubernetes experience — cluster architecture, multi-tenancy, RBAC, scheduling, security standards, and autoscaling across cloud and bare-metal. Experience with k3s is a plus.
  • Strong Kubernetes networking knowledge — CNI (Calico, Cilium), service mesh (Istio, Traefik), ingress controllers, and NetworkPolicy.
  • Linux networking fundamentals — TCP/IP, DNS, BGP, and network troubleshooting.
  • Experience designing and operating infrastructure across hybrid environments — on-premises, edge, and multiple cloud providers (AWS, Azure, OCI).
  • Infrastructure as code proficiency — Terraform and Ansible.
  • Working knowledge of Helm/Kustomize for application packaging and deployment.
  • Proficiency in Python, Go, and Bash for automation and tooling.
  • Postgres experience — replication, HA/failover, connection pooling (PgBouncer), query tuning, backup/recovery, and Kubernetes operators (CloudNativePG).
  • Experience operating stateful workloads on Kubernetes including Elasticsearch/Opensearch and Postgres.
  • Experience operating storage solutions (CSI, Rook/Ceph).
  • Cloud-native observability experience — Prometheus, Grafana, Loki, and Tempo.
  • Experience with Security Standards — FIPS, CVE mitigation, FedRAMP, and STIGs.
  • Demonstrated ability to deliver results on time and with high quality.
  • Effective communication skills and the ability to work effectively across multiple business and technical teams.

About the Company:

••••• is a leader in explainable and trustworthy artificial intelligence designed to power mission-critical decisions in enterprises, government, and regulated industries. SeekrFlow™, our end-to-end AI platform, provides secure, auditable AI solutions tailored to sectors where transparency, accuracy, and compliance are paramount. Available across cloud, on-premises, and edge environments, SeekrFlow reduces bias, strengthens data integrity, and simplifies model oversight so organizations can rely on trusted AI decisions in high-stakes settings that impact society’s most sensitive and vital systems. Trusted by leading enterprises and government agencies, we partner with defense, finance, telecom, and critical infrastructure leaders to enable AI solutions that drive real-world results with unmatched transparency and control. We are a team of strategic thinkers and problem-solvers tackling the toughest challenges facing critical infrastructure and global enterprises through best-in-class AI models and customer deployment. Our team operates with unwavering commitment to our core values and mission:

  • We are driven by outcomes—our customers' success is what we strive for every day.
  • We believe trust is earned, which is why we build explainability and transparency into the entire AI lifecycle.
  • We take our responsibility to deliver secure AI seriously.
  • We believe innovation drives progress—we are building the technologies that power the systems our society depends on.

Company Benefits:

  • Meaningful Mission & Impact -Work with a deeply talented, collaborative team solving some of the toughest AI challenges that matter.
  • Equity Ownership– RSUs that let you share directly in ••••• ’s long‑term success and growth.
  • Time Off That Respects Real Life– Unlimited PTO plus 14 paid company holidays to truly recharge.
  • Work Your Way– A flexible hybrid work environment with offices in Reston, VA and Austin, TX.
  • Competitive Total Rewards– A role‑appropriate compensation structure that supports long‑term growth, including base salary, bonuses, or commission plans depending on role.
  • 401(k) with Company Match– Build your future with a retirement plan that includes employer matching.
  • Comprehensive Health & Wellness– Medical, dental, vision, and life insurance coverage starting day one—for you and your family.
  • Parental Leave– Paid parental leave to support employees as they welcome a new child through birth, adoption, or foster placement.

Important Notice:
To conform to U.S. Government international trade regulations, applicant must be a U.S. Citizen, lawful permanent resident of the U.S., protected individual as defined by 8 U.S.C. 1324b(a)(3), or eligible to obtain the required authorizations from the U.S. Department of State or U.S. Department of Commerce.
••••• is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, creed, sex, sexual orientation, gender identity, marital status, national origin, age, veteran status, disability, or any other characteristic protected by applicable law.

#J-18808-Ljbffr