1

Grafana Loki Jobs (NOW HIRING)

Site Reliability Engineer

Washington, DC · On-site

$64.25 - $85.50/hr

OpenTelemetry, Prometheus, Grafana, Loki, and Tempo. • Experience with FinOps and optimizing cloud resource consumption. • Experience supporting high-scale distributed systems in a secure ...

OR · Hybrid

$129K - $166K/yr

Enabling interoperability (e.g., Magnum, storage backends, network overlays) and shared observability stacks (Prometheus, Grafana, Loki/EFK). * SRE Principles : Applying uniform practices across both ...

Grafana Engineer

Morristown, NJ · On-site

$58.75 - $78/hr

Prometheus|Loki|Tempo|OpenTelemetry|Elasticsearch/OpenSearch|Splunk (preferred)Understanding of ... Keywords: Grafana| Prometheus| Loki| Tempo| OpenTelemetry| Kubernetes| Cloud Monitoring ...

Grafana Engineer

Morristown, NJ · On-site

$58.75 - $78/hr

Prometheus|Loki|Tempo|OpenTelemetry|Elasticsearch/OpenSearch|Splunk (preferred)Understanding of ... Grafana Enterprise deployment experience.OpenTelemetry implementation experience.Experience with ...

Implement and optimize observability operations using OpenTelemetry, Prometheus, Grafana, Loki, or Tempo. * Oversee capacity planning, performance optimization, and FinOps practices. * Define and ...

You'll work across the full observability stack: from distributed tracing adoption (OpenTelemetry, Jaeger) to log infrastructure (Loki, Alloy) to metrics (Cortex, Prometheus, Grafana). You'll partner ...

Senior Site Reliability Engineer I

Seattle, WA

$64.75 - $86.25/hr

You'll work across the full observability stack: from distributed tracing adoption (OpenTelemetry, Jaeger) to log infrastructure (Loki, Alloy) to metrics (Cortex, Prometheus, Grafana). You'll partner ...

next page

Showing results 1-20

Grafana Loki information

See salary details

$11K

$88.5K

$99K

How much do grafana loki jobs pay per year?

As of Jul 27, 2026, the average yearly pay for grafana loki in the United States is $88,479.00, according to ZipRecruiter salary data. Most workers in this role earn between $82,000.00 and $94,000.00 per year, depending on experience, location, and employer.

What are the key skills and qualifications needed to thrive as a Grafana Loki Engineer, and why are they important?

To excel as a Grafana Loki Engineer, you need a solid background in system administration, log management, cloud infrastructure, and experience with observability tools, preferably with a degree in computer science or related field. Familiarity with Grafana Loki, Prometheus, Kubernetes, and scripting languages like Python or Bash, as well as relevant certifications (e.g., CNCF) are highly valuable. Strong analytical thinking, problem-solving abilities, and effective communication skills are crucial for collaborating with diverse technical teams and resolving issues swiftly. These competencies ensure efficient log aggregation, monitoring, and troubleshooting, which are vital for maintaining system reliability and performance.

What is the difference between Grafana Loki vs Prometheus?

AspectGrafana LokiPrometheus
Primary FunctionLog aggregation and queryingMetrics collection and alerting
Data TypeLogsMetrics
Work EnvironmentCloud-native, containerized environmentsServer monitoring, cloud and on-premises
Common UsageMonitoring logs in Kubernetes and microservicesMonitoring system and application metrics

Grafana Loki and Prometheus are both open-source tools used in monitoring, but they serve different purposes. Loki focuses on log aggregation, making it ideal for analyzing logs from cloud-native environments. Prometheus specializes in collecting and alerting on metrics data, suitable for performance monitoring. Many organizations use both together for comprehensive observability.

What are some common challenges faced by engineers working with Grafana Loki in a production environment?

Engineers working with Grafana Loki often encounter challenges related to scaling and managing log ingestion rates, especially in high-traffic environments. Ensuring efficient log retention and querying performance requires careful configuration and monitoring of storage backends. Additionally, integrating Loki seamlessly with other observability tools and maintaining consistent log labeling can be complex, necessitating strong collaboration with DevOps and SRE teams. Proactively addressing these areas helps ensure reliable log management and smooth operations.

What is Grafana Loki?

Grafana Loki is an open-source log aggregation system designed to store and query logs from various sources. Unlike traditional log management tools, Loki is optimized for cost-efficiency and scalability by indexing only metadata, not the full log content. It works seamlessly with Grafana for visualization, making it easy to correlate logs with metrics and traces. Loki is often used in cloud-native environments and integrates well with Prometheus, Kubernetes, and other monitoring tools.
More about Grafana Loki jobs
What cities are hiring for Grafana Loki jobs? Cities with the most Grafana Loki job openings:
What states have the most Grafana Loki jobs? States with the most job openings for Grafana Loki jobs include:
Infographic showing various Grafana Loki job openings in the United States as of July 2026, with employment types broken down into 94% Full Time, 1% Part Time, and 5% Contract. Highlights an 81% Physical, 5% Hybrid, and 14% Remote job distribution, with an average salary of $88,479 per year, or $42.5 per hour.
Site Reliability Engineer

Site Reliability Engineer

MANTECH

Washington, DC • On-site

$64.25 - $85.50/hr

Full-time

Posted 5 days ago


ManTech rating

9.0

Company rating: 9.0 out of 10

Based on 14 frontline employees who took The Breakroom Quiz

33rd of 245 rated software companies


Job description

Job Summary:
MANTECH seeks a motivated Site Reliability Engineer (SRE) for a new initiative that supports the rapid design and operation of enterprise-scale AI and data capabilities. The role focuses on ensuring operational reliability and optimizing system performance for enterprise AI systems.
Responsibilities:
• Apply core reliability engineering principles to ensure high availability and stability of production systems.
• Manage incident response, root cause analysis, and post-mortem processes for the AI platform.
• Implement and optimize observability operations using OpenTelemetry, Prometheus, Grafana, Loki, or Tempo.
• Oversee capacity planning, performance optimization, and FinOps practices.
• Define and continuously monitor Service Level Objectives (SLOs) and Service Level Agreements (SLAs).
Qualifications:
Required:
• Bachelor’s degree in Computer Science, Engineering, or a related technical discipline.
• 5 or more years of experience in Site Reliability Engineering (SRE), DevOps, or production operations.
• Extensive experience with cloud-native infrastructure, particularly Kubernetes.
• Deep knowledge of monitoring, alerting, and logging systems.
• Proven ability to automate operational tasks and reduce toil.
• For onsite work, a TS/SCI clearance with Poly will be required.
• The person in this position must be able to remain in a stationary position 50% of the time.
• Frequently communicates with co-workers, management, and customers, which may involve delivering presentations.
• Constantly operates a computer and other office productivity machinery.
Preferred:
• Hands-on experience with the full observability stack: OpenTelemetry, Prometheus, Grafana, Loki, and Tempo.
• Experience with FinOps and optimizing cloud resource consumption.
• Experience supporting high-scale distributed systems in a secure environment.
Company:
ManTech is a technology company that offers cyber, IT, and data analytics technologies and solutions for security programs. Founded in 1968, the company is headquartered in Herndon, USA, with a team of 5001-10000 employees. The company is currently Late Stage.

What ManTech employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom