DevOpsEngineer
$49.25 - $67.50/hr
hackajob is collaborating with Leo Technologies to connect them with exceptional professionals for ... Experience operating GPU infrastructure (NVIDIA drivers, CUDA, MIG, GPU operator, scheduling) in ...
$49.25 - $67.50/hr
hackajob is collaborating with Leo Technologies to connect them with exceptional professionals for ... Experience operating GPU infrastructure (NVIDIA drivers, CUDA, MIG, GPU operator, scheduling) in ...
$49.25 - $67.50/hr
hackajob is collaborating with Leo Technologies to connect them with exceptional professionals for ... Experience operating GPU infrastructure (NVIDIA drivers, CUDA, MIG, GPU operator, scheduling) in ...
| Aspect | Professional Cuda | Cuda Developer |
|---|---|---|
| Required Credentials | Typically requires a degree in Computer Science or related field, with certifications in CUDA programming | Often requires similar degrees and certifications, focusing on CUDA expertise |
| Work Environment | Works in research labs, tech companies, or industries utilizing GPU computing | Works in software development teams, research, or hardware optimization projects |
| Industry Usage | Used across high-performance computing, AI, and scientific research sectors | Commonly employed in software development, gaming, and simulation industries |
Both roles involve CUDA programming, but a Professional Cuda typically emphasizes advanced GPU computing skills in research or industry applications, while a Cuda Developer focuses on software development and optimization using CUDA technology. The roles often overlap, but the Professional Cuda may have a broader scope in high-performance computing projects.
hackajob is collaborating with Leo Technologies to connect them with exceptional professionals for this role.
DevOps / Site Reliability Engineer – AI Systems for Corrections & Intelligence
Location: Palm Beach, FL: Full-time Reports to: Chief AI & Data Officer
About the Role
We're hiring a DevOps / SRE to deploy, operate, and harden the AI systems that support corrections operations and intelligence analysis. Our data scientists build LLM-powered agents, RAG pipelines, and ontology-driven analytics — your job is to make sure those systems run reliably, securely, and auditably in environments where uptime, data segregation, and chain-of-custody actually matter. You'll own the path from a trained model or agent prototype to a production system that analysts depend on, in infrastructure that meets CJIS, FedRAMP, or equivalent standards.
What You'll Do
What You Bring
Required
Nice to Have
How We Think About This Work
In corrections and intelligence environments, an outage isn't just a missed SLA — it can mean analysts lose access to tools during a developing situation, or an audit trail gets broken at exactly the wrong time. We expect rigor: changes are reviewed, deploys are reversible, access is least-privilege, and every action affecting sensitive data is logged in a way that survives scrutiny. We also expect honesty about AI system risk. If a model regression slipped through, or a retrieval index is serving stale or wrong data, we want it caught and surfaced — not papered over. People who treat reliability and security as core features rather than overhead will thrive here.
What Success Looks Like
In your first 90 days, you'll have inventoried the current deployment surface, stood up or hardened CI/CD for at least one production AI service, and established baseline observability covering both infrastructure and model-level signals. Within six months, you'll own the AI platform's reliability posture — including SLOs, incident response, and the security controls that let us deploy into the most sensitive environments our customers operate.
Sourced by ZipRecruiter
It services
11 - 50 Employees
Los Angeles, CA, US
2018