1

Prometheus Jobs in Florida (NOW HIRING)

DevOps Engineer

Tampa, FL · On-site

$49.75 - $68.25/hr

Utilizing observability tools (Prometheus, Grafana, ELK stack). Required Skills and Competencies * Expertise in Kubernetes, Docker, Terraform, Ansible, and CI/CD tools (Jenkins, GitLab CI/CD, Azure ...

DevOps Engineer

Miami, FL · On-site

$51 - $70/hr

Cluster monitoring via Prometheus/VictoriaMetrics, logging via Loki/ViactoriaLogs * Logging agents - vector, fluentbit Security * Source code and containers images security scanning - trivy ...

DevOps Engineer

Tampa, FL · On-site

$110 - $160/hr

Utilizing observability tools (Prometheus, Grafana, ELK stack). Required Skills and Competencies * Expertise in Kubernetes, Docker, Terraform, Ansible, and CI/CD tools (Jenkins, GitLab CI/CD, Azure ...

Kubernetes Engineer

Doral, FL · On-site

$54.50 - $72.50/hr

Integrate Kubernetes-native monitoring and logging solutions (e.g., Prometheus, Fluentd, or OpenTelemetry) to support observability and IL-compliant reporting across environments. Job Requirements ...

Implement Prometheus metrics and leverage Spring Boot Actuator for monitoring. * Contribute to REST APIs, GraphQL subgraphs, composite services, and integration gateways deployed on Kubernetes in AWS.

DevOps Engineer

Miami, FL · On-site

$51 - $70/hr

Cluster monitoring via Prometheus/VictoriaMetrics, logging via Loki/ViactoriaLogs * Logging agents - vector, fluentbit Security * Source code and containers images security scanning - trivy ...

DevOps Engineer

Miami, FL · On-site

$51 - $70/hr

Cluster monitoring via Prometheus/VictoriaMetrics, logging via Loki/ViactoriaLogs * Logging agents - vector, fluentbit Security * Source code and containers images security scanning - trivy ...

Java Developer

Tampa, FL · On-site

$48.25 - $62.25/hr

Has experience in alerts and monitoring tools like Grafana, Kibana, Prometheus, Splunk, and Graphite and is able to debug through logs and dashboards. * Has experience on GIT or similar repository ...

Monitor and analyze performance and telemetry using tools such as Splunk, AppDynamics, or Prometheus * Contribute to secure edge architecture decisions and implementation * Participate in ongoing ...

Showing results 21-40

Prometheus information

See Florida salary details

$9

$15

$47

How much do prometheus jobs pay per hour?

As of Aug 22, 2026, the average hourly pay for prometheus in Florida is $15.73, according to ZipRecruiter salary data. Most workers in this role earn between $10.77 and $12.93 per hour, depending on experience, location, and employer.

What is a Prometheus engineer?

Prometheus engineers are IT professionals who specialize in deploying, configuring, and maintaining Prometheus, an open-source systems monitoring and alerting toolkit. They design and implement monitoring solutions for infrastructure, applications, and services using Prometheus and its ecosystem (like Grafana for visualization). Their responsibilities include setting up metrics collection, writing queries, tuning alerts, and ensuring high availability and scalability of monitoring setups. Prometheus engineers often work closely with DevOps teams to provide observability and actionable insights for system reliability.

What skills and qualifications are needed to thrive as a Prometheus engineer?

To thrive as a Prometheus Monitoring Engineer, you need a solid understanding of systems administration, networking, and monitoring concepts, often supported by experience in DevOps or Site Reliability Engineering roles. Proficiency in Prometheus, Grafana, alerting tools, and infrastructure-as-code platforms, as well as familiarity with cloud services and container orchestration systems like Kubernetes, is essential. Strong problem-solving skills, attention to detail, and effective communication help you proactively address issues and collaborate with diverse technical teams. These skills ensure the reliability, scalability, and optimal performance of critical IT infrastructure.

What are common challenges faced by Prometheus engineers, and how can they be addressed?

Prometheus monitoring engineers often face challenges related to scaling the system to handle large volumes of metrics data and optimizing query performance. Dealing with high cardinality metrics and managing storage efficiently can also be complex. These challenges can be addressed by carefully designing label schemas, employing federation or sharding, and using long-term storage integrations. Collaboration with development and operations teams is key to ensure effective instrumentation and alerting, which helps maintain system reliability and performance.

What is the difference between Prometheus vs Grafana Developer?

AspectPrometheusGrafana Developer
Primary RoleMonitoring and alerting system configurationDashboard and visualization development
Required SkillsPromQL, server setup, metrics collectionGrafana, data sources, visualization design
Work EnvironmentDevOps, monitoring infrastructureData visualization, front-end development
CertificationsNone specific, familiarity with monitoring toolsGrafana certifications beneficial

Prometheus and Grafana Developer roles often overlap in monitoring and visualization. Prometheus focuses on metrics collection and alerting, while Grafana Developers specialize in creating dashboards for data visualization. Both roles are essential in DevOps environments but serve different functions within the monitoring ecosystem.

What are popular job titles related to Prometheus jobs in Florida?

For Prometheus jobs in Florida, the most frequently searched job titles are:

What job categories do people searching Prometheus jobs in Florida look for?

The top searched job categories for Prometheus jobs in Florida are:

Infographic showing various Prometheus job openings in Florida as of August 2026, with employment types broken down into 82% Full Time, and 18% Contract. Highlights an 100% In-person job distribution, with an average salary of $32,714 per year, or $15.7 per hour.

Information Technology_USA - USA_Engineer

Real Soft, Inc.

Jacksonville, FL • On-site

$52.75 - $70.25/hr

Contractor

Re-posted 21 days ago


Job description

ALL CAPS, NO SPACES B/T UNDERSCORES PTN_US_GBAMSREQID_
Candidate BeelineID i.e. PTN_US_9999999_SKIPJOHNSON0413
MSP Owner: Thomas Hodges
Targeted - -
REQUIREMENT_CITY - Alpharetta, GA - Need to work from office 3 Days a week - Will be Face2Face Round of Interview
REQUIREMENT_ID-10780881
Role Name - AI SRE
ROLE_DESCRIPTION -
Skill Set - Expertise in UNIX + LINUX Administration + AWS/ AZURE Cloud monitoring + Terraform/ Ansible + Prometheus/ Grafana observability experience).
Work Location - Alpharetta
Experience required for role - 6+ years
• Production experience in SRE / Infrastructure / ops for large-scale systems
• Strong programming/scripting skills (Python, Go, Java, or equivalent)
• Deep experience with containerization (Docker), orchestration (Kubernetes, etc.)
• Infrastructure-as-code (Terraform, Helm, CloudFormation, Ansible, etc.)
• Familiarity with GPU / AI compute clusters, high-performance data storage, and distributed architectures
• Experience with monitoring / observability / logging / alerting tools (Prometheus, Grafana, ELK / EFK, Datadog, etc.)
• Networking & systems engineering knowledge (TCP/IP, DNS, routing, load balancing, distributed storage)
• Solid experience in capacity planning, performance tuning, scaling, and incident response
• Demonstrated ability to lead RCAs, deploy fixes, and drive reliability improvements
• Experience in regulated environments (financial services, compliance, audit, security) is a strong plus
• Excellent communication, documentation, and cross-team collaboration skills
• Proven track record of reducing operational toil via automation
Experience: 6+ years of experience as a Site Reliability Engineer or in a similar role, with hands-on experience in supporting IaaS platforms with networking and system engineering knowledge.
• Operate, monitor, and maintain the infrastructure supporting GenAI applications (training, inference, feature store, data ingestion, model serving)
• Design and build automation for core platform capabilities, reducing manual toil
• Develop and maintain infrastructure-as-code (IaC) for provisioning and managing compute, storage, network, GPU clusters, Kubernetes / container orchestration, etc.
• Establish, monitor, and enforce SLOs/SLIs/SLAs, error budgets, alerting, and dashboards
• Lead incident response, root cause analysis (RCA), postmortems, and systemic remediation
• Perform capacity planning, scaling strategies, workload scheduling, and resource forecasting
• Optimize cost vs. performance tradeoffs in large-scale compute environments
• Harden systems for security, compliance, auditability, and data governance
• Collaborate across teams (cloud engineers, data engineers, infrastructure, security) to ensure safe deployment, rollout, rollback, and integration of new systems
• Define disaster recovery (DR) strategies, backup/restore practices, fault tolerance mechanisms
• Maintain runbooks, operational playbooks, documentation, and training materials
• Participate in on-call rotations and respond to production incidents 24/7 as needed
• Continuously evaluate and integrate new tools, frameworks, or technologies to enhance platform reliability
Skills: Digital : Python~Digital : Docker~Digital : Kubernetes~Digital : Site Reliability Engineering (SRE)
Experience Required: 6-8, Project Code :