1

Ceph Jobs in Texas (NOW HIRING)

DevOps Engineer

San Antonio, TX · On-site

$47.50 - $65.25/hr

Hadoop, Accumulo, Ceph, Spark, NiFi, Kafka, PostgreSQL, ElasticSearch, Hive, Drill, Impala, Trino, Presto, etc. • Work could possibly require some on-call work. Company : Swift is a privately held ...

Open vSwitch, KVM, QEMU). • Experience with distributed object, block, and file storage (e.g., Ceph). • Experience in end-to-end deployment automation and CI of containerized services. Complete ...

Exposure to IoT device management Experience integrating Red Hat technologies such as IdM, Satellite, Ceph, or RHV. Hands-on experience with AWS cloud services (e.g., EC2, S3, Lambda, RDS, ALB/NLB)

DevOps Engineer - 28337

San Antonio, TX · On-site

$47.50 - $65.25/hr

Hadoop, Accumulo, Ceph, Spark, NiFi, Kafka, PostgreSQL, ElasticSearch, Hive, Drill, Impala, Trino, Presto, etc. • Experience with containers, EKS, Diode, CI/CD, and Terraform are a plus. Company

Experience with Red Hat technologies (e.g., IdM , Satellite , RHV , Ceph , etc.) * Hands-on experience with AWS cloud services (EC2, S3, Lambda, RDS) * Experience in designing, optimizing, and ...

... Ceph, etc. * AWS Cloud experience including core services EC2, S3, ALB/NLB, Lambda, RDS * Experience of designing, optimizing and troubleshooting public cloud platforms associated with large, complex ...

... Ceph, Rook, Longhorn) • Familiarity with Big Data ecosystems such as Cloudera, Hortonworks, or MapR • DoD 8140 / 8570 IAT Level II certification or equivalent preferred Company : Swift is a ...

Showing results 41-60

Ceph information

What is a Ceph?

A Ceph job typically involves managing, deploying, and maintaining Ceph, an open-source distributed storage system. Responsibilities may include configuring storage clusters, monitoring performance, troubleshooting issues, and ensuring data redundancy and scalability. Professionals in this role often work with Linux, networking, and automation tools to optimize storage solutions for enterprises and cloud environments.

What does a Ceph do?

As a Ceph Storage Administrator, your typical day involves monitoring the health and performance of the Ceph cluster, responding to alerts, and proactively addressing issues to maintain data integrity and uptime. You will also be responsible for upgrading and patching the cluster, managing user access, and performing capacity planning. Collaboration with application and infrastructure teams is key to understanding storage needs and ensuring seamless data workflows. Additionally, you may create and maintain technical documentation and participate in on-call rotations as part of a broader IT operations team. These tasks help ensure the stability and scalability of mission-critical storage systems.

What are the key skills and qualifications needed for a Ceph?

To thrive as a Ceph (Ceph Storage Administrator or Engineer), you need a solid background in Linux systems administration, distributed storage concepts, and networking fundamentals, typically supported by a degree in computer science or a related field. Proficiency with Ceph deployment tools, monitoring platforms, and scripting languages like Python or Bash, as well as relevant certifications such as Red Hat Certified Engineer (RHCE), are highly valuable. Strong problem-solving skills, meticulous attention to detail, and the ability to communicate technical concepts clearly with colleagues are important soft skills for this role. These competencies ensure high availability and reliability of storage systems, effective troubleshooting, and smooth collaboration within IT teams.

What are popular job titles related to Ceph jobs in Texas?

For Ceph jobs in Texas, the most frequently searched job titles are:

What cities in Texas are hiring for Ceph jobs?

Cities in Texas with the most Ceph job openings:

Infographic showing various Ceph job openings in Texas as of August 2026, with employment types broken down into 71% Full Time, 24% Part Time, and 5% Contract. Highlights an 82% In-person, and 18% Remote job distribution.

CV/ML Platform Engineer

Allen Control Systems

Austin, TX • On-site

Full-time

Re-posted 29 days ago


Job description

Job Summary:
Allen Control Systems (ACS) is a cutting-edge defense startup focused on developing autonomous technologies. They are seeking an experienced CV/ML Platform Engineer to design, build, and manage the infrastructure for their Computer Vision and Machine Learning team.
Responsibilities:
• Deploy and operate Kubernetes clusters on bare-metal infrastructure hosting 130+ NVIDIA GPUs, with hybrid burst capability to AWS for scalable compute and storage workloads.
• Manage NVIDIA GPU clusters for ML training.
• Own the ACS CV/ML CI/CD pipeline.
• Improve and maintain core ML infrastructure, such as model registration and versioning, experiment tracking, and model and data provenance tracking.
• Improve and maintain ML model testing, performance analysis, and reporting tools.
• Automate repetitive model training and testing tasks to increase developer velocity.
• Work with Software Team Platform Engineers to ensure efficient coordination and minimal duplication between CV/ML infrastructure and wider Software infrastructure.
• Collaborate with the Software Team to automate the optimization of models (TensorRT/quantization) for deployment on NVIDIA Jetson and other edge hardware.
Qualifications:
Required:
• 2+ years of experience in Platform Engineering or DevOps/MLOps.
• Strong programming skills are required for automating ML lifecycles and building custom CLI tools for CV engineers.
• Hands-on experience with NVIDIA GPU infrastructure, including managing CUDA libraries and development environments, GPU Operator, device plugins, and scheduling (MIG, Volcano, or fractional GPU sharing).
• Experience implementing and maintaining MLOps platforms such as Kubeflow, MLflow, Weights & Biases (W&B), or DVC for experiment tracking and model versioning.
• Familiarity with high-performance storage solutions (e.g., MinIO, WEKA, or Ceph) and data orchestration tools capable of handling terabytes of video/image data.
• Proven track record building CI/CD pipelines that include automated model validation, performance benchmarking, and artifact management for both cloud and edge targets.
• Experience with model optimization toolchains, including TensorRT, ONNX, and quantization techniques, specifically for cross-compilation to ARM targets like NVIDIA Jetson.
• Proficiency with observability stacks (ELK, Prometheus/Grafana) adapted for ML, including monitoring GPU health, training throughput, and model inference metrics.
• Strong Linux systems knowledge (Debian/Ubuntu), including networking for high-throughput data, storage, and security hardening for defense-grade production environments.
Company:
Allen Control Systems develops autonomous defense technologies designed to detect, track, and counter unmanned aerial threats. Founded in 2022, the company is headquartered in Austin, USA, with a team of 201-500 employees. The company is currently Growth Stage.