ClickHouse is looking for an experienced engineer to join our Observability organization. We are ... Experience with TypeScript. #LI-Remote
ClickHouse is looking for an experienced engineer to join our Observability organization. We are ... Experience with TypeScript. #LI-Remote
... remote opportunity. We are considering candidates from US and Canada only (EST time zone). The Opportunity: The Grafana Cloud Observability department is focused on enabling developers to understand ...
New
... remote opportunity. We are considering candidates from US and Canada only (EST time zone). The Opportunity: The Grafana Cloud Observability department is focused on enabling developers to understand ...
New
As an Observability Architect, you will be the technical owner of long-term customer success. You ... You'll build strategic relationships with engineers, SREs, and architects-ensuring Grafana ...
As an Observability Architect, you will be the technical owner of long-term customer success. You ... You'll build strategic relationships with engineers, SREs, and architects-ensuring Grafana ...
Senior Product Manager, Infrastructure Observability | US | Remote
OR ยท Remote
$162K - $194K/yr
This is a fully remote position and we're considering candidates in the US EST. About Product ... Product Managers are the bridge that connect Engineering, Design, and UX to GTM. You are a ...
Senior Product Manager, Infrastructure Observability | US | Remote
OR ยท Remote
$162K - $194K/yr
This is a fully remote position and we're considering candidates in the US EST. About Product ... Product Managers are the bridge that connect Engineering, Design, and UX to GTM. You are a ...
Grafana Enterprise is our composable observability platform focused on large-scale operators with ... engineering and operational practices across Grafana Labs * As we are remote-first, we provide ...
Grafana Enterprise is our composable observability platform focused on large-scale operators with ... engineering and operational practices across Grafana Labs * As we are remote-first, we provide ...
Software Engineer (US-Remote)
OR ยท Remote
Software Engineer (US-Remote) ID: 1191 Location: US-Remote or Marlton, NJ area Description A ... Monitor and troubleshoot applications in distributed environments using logging, observability, and ...
Software Engineer (US-Remote)
OR ยท Remote
Software Engineer (US-Remote) ID: 1191 Location: US-Remote or Marlton, NJ area Description A ... Monitor and troubleshoot applications in distributed environments using logging, observability, and ...
At Grafana, we build observability tools that help users understand, respond to, and improve their systems - regardless of scale, complexity, or tech stack. The Grafana AI teams play a key role in ...
At Grafana, we build observability tools that help users understand, respond to, and improve their systems - regardless of scale, complexity, or tech stack. The Grafana AI teams play a key role in ...
Senior Site Reliability Engineer (REMOTE)
Beaverton, OR ยท Remote
$140K/yr
Observability (Datadog, Sentry) * Agentic AI (Claude Code) * Scripting (Shell, Python) * Track record of collaboration and mentorship * Excellent written communication and documentation skills
Senior Site Reliability Engineer (REMOTE)
Beaverton, OR ยท Remote
$140K/yr
Observability (Datadog, Sentry) * Agentic AI (Claude Code) * Scripting (Shell, Python) * Track record of collaboration and mentorship * Excellent written communication and documentation skills
Staff Backend Engineer - Adaptive Telemetry, Databases This is a remote position. We are looking ... Grafana Cloud is our composable observability platform that integrates metrics, logs, traces, and ...
Staff Backend Engineer - Adaptive Telemetry, Databases This is a remote position. We are looking ... Grafana Cloud is our composable observability platform that integrates metrics, logs, traces, and ...
Senior Infrastructure Engineer - Postgres
OR ยท Remote
$108K - $147K/yr
Own observability and monitoring , ensuring robust alerting, metrics, and tracing across ... remote
Senior Infrastructure Engineer - Postgres
OR ยท Remote
$108K - $147K/yr
Own observability and monitoring , ensuring robust alerting, metrics, and tracing across ... remote
At Grafana Labs, we build observability tools that help users understand, respond to, and improve ... Engineers are empowered to make decisions, move quickly, and validate ideas early - while being ...
At Grafana Labs, we build observability tools that help users understand, respond to, and improve ... Engineers are empowered to make decisions, move quickly, and validate ideas early - while being ...
Senior Software Engineer, Site Reliability
OR ยท On-site +1
$57 - $75.75/hr
... observability platform * Experience leveraging LLM/GenAI to improve SRE efficiency and processes Position Location - This role is available in the following locations: Remote, San Mateo, Columbus ...
Senior Software Engineer, Site Reliability
OR ยท On-site +1
$57 - $75.75/hr
... observability platform * Experience leveraging LLM/GenAI to improve SRE efficiency and processes Position Location - This role is available in the following locations: Remote, San Mateo, Columbus ...
Solutions Engineer - US Remote
OR ยท Remote
$130K - $160K/yr
At Embrace, we're the only user-focused observability solution built on OpenTelemetry, providing ... Communicate directly with customers, product and engineering teams, and work together towards a ...
Solutions Engineer - US Remote
OR ยท Remote
$130K - $160K/yr
At Embrace, we're the only user-focused observability solution built on OpenTelemetry, providing ... Communicate directly with customers, product and engineering teams, and work together towards a ...
Senior DevOps Engineer, Ephemeral Platform
OR ยท On-site +1
$129K - $166K/yr
Enhance observability, incident response, and operational practices to improve platform health and ... Remote Time zone requirements The team operates on the East/West coast time zones. Travel ...
Senior DevOps Engineer, Ephemeral Platform
OR ยท On-site +1
$129K - $166K/yr
Enhance observability, incident response, and operational practices to improve platform health and ... Remote Time zone requirements The team operates on the East/West coast time zones. Travel ...
(Sr/Staff) Platform Engineer
OR ยท On-site +1
US - Remote Level : Senior Individual Contributor Team : Engineering The Opportunity Terzo ... Comfort with observability systems including monitoring, alerting, logging, tracing * Strong SRE ...
(Sr/Staff) Platform Engineer
OR ยท On-site +1
US - Remote Level : Senior Individual Contributor Team : Engineering The Opportunity Terzo ... Comfort with observability systems including monitoring, alerting, logging, tracing * Strong SRE ...
Staff AI Engineer | US | Remote
OR ยท Remote
This is a remote opportunity and we are looking for candidates from the U.S. The Opportunity ... Implement observability and feedback loops including logging, performance metrics, prompt iteration ...
Staff AI Engineer | US | Remote
OR ยท Remote
This is a remote opportunity and we are looking for candidates from the U.S. The Opportunity ... Implement observability and feedback loops including logging, performance metrics, prompt iteration ...
Senior Platform Engineer
OR ยท Remote
$165K - $210K/yr
Improve production observability: work with product teams to ensure they are equipped to ... We are remote-first with a dedicated NYC office and reimbursement options for co-working spaces.
Senior Platform Engineer
OR ยท Remote
$165K - $210K/yr
Improve production observability: work with product teams to ensure they are equipped to ... We are remote-first with a dedicated NYC office and reimbursement options for co-working spaces.
Senior AI Engineer | US | Remote
OR ยท Remote
$55.25 - $71.25/hr
This is a remote opportunity and we are looking for candidates from the U.S. The Opportunity ... Implement observability and feedback loops including logging, performance metrics, prompt iteration ...
Senior AI Engineer | US | Remote
OR ยท Remote
$55.25 - $71.25/hr
This is a remote opportunity and we are looking for candidates from the U.S. The Opportunity ... Implement observability and feedback loops including logging, performance metrics, prompt iteration ...
Staff Software Engineer, Applied AI
OR ยท Remote
$200K - $260K/yr
Improve production observability: accuracy signals and feedback loops. * Contribute to architecture ... We are remote-first with a dedicated NYC office and reimbursement options for co-working spaces.
Staff Software Engineer, Applied AI
OR ยท Remote
$200K - $260K/yr
Improve production observability: accuracy signals and feedback loops. * Contribute to architecture ... We are remote-first with a dedicated NYC office and reimbursement options for co-working spaces.
Work closely with PMs and with App Observability, Asserts, Drilldown, and Grafana Assistant teams ... Raise the engineering bar through code review, design feedback, pairing on hard problems, and ...
Work closely with PMs and with App Observability, Asserts, Drilldown, and Grafana Assistant teams ... Raise the engineering bar through code review, design feedback, pairing on hard problems, and ...
Remote Observability Engineer information
What are the typical collaboration patterns for a Remote Observability Engineer working with distributed teams?
What are the key skills and qualifications needed to thrive as a Remote Observability Engineer, and why are they important?
What is the difference between Remote Observability Engineer vs Site Reliability Engineer?
| Aspect | Remote Observability Engineer | Site Reliability Engineer |
|---|---|---|
| Credentials | Knowledge of monitoring tools, scripting, cloud platforms | Same as Observability Engineer, plus SRE certifications often preferred |
| Work Environment | Focus on monitoring, logging, and tracing systems remotely | Broader scope including system reliability, incident response, and automation |
| Industry Usage | Primarily in tech, SaaS, cloud services | Widely in tech, finance, and large-scale online services |
The Remote Observability Engineer specializes in monitoring and analyzing system performance remotely, focusing on tools like logs and metrics. In contrast, the Site Reliability Engineer has a broader role, ensuring overall system reliability, automation, and incident management. While both roles require similar technical skills, SREs often have additional responsibilities related to system resilience and scalability.
What is a Remote Observability Engineer?
Job description
ClickHouse is looking for an experienced engineer to join our Observability organization. We are hiring across two closely aligned teams: Observability Platform and Internal Observability.
Together, these teams build and operate the systems behind ClickHouse's internal observability and ClickStack Cloud. Our platforms process trillions of events per day, sustaining throughput in the hundreds of millions of events per second. The experience we gain operating observability at ClickHouse scale feeds directly back into the platform and product we provide to customers.
The Observability Platform team builds shared systems for telemetry ingestion, durable buffering, processing, storage, autoscaling, and service provisioning. The Internal Observability team operates ClickHouse's company-wide observability platform and works with engineering teams to improve reliability, debugging, and operational efficiency.
This is a software engineering role at the intersection of distributed systems, cloud infrastructure, and production operations. You will build systems, operate what you build, respond to incidents, and turn recurring operational problems into durable software and automation.
Your background may be in backend engineering, infrastructure, SRE, or systems engineering. What matters most is your ability to solve ambiguous production problems, make sound engineering tradeoffs, and take ownership from design through operation.
What you'll do- Design, build, and operate distributed systems that ingest, process, and store telemetry at very high scale.
- Own the reliability, performance, capacity, and cost-efficiency of telemetry pipelines and storage systems.
- Participate in the on-call rotation, help resolve production incidents, and drive root-cause fixes through to completion.
- Build software and automation that eliminate repetitive operational work and make the platform easier to operate.
- Identify architectural bottlenecks and help shape the roadmap for the next stage of scale.
- Work closely with product, infrastructure, and service teams across ClickHouse.
- Contribute to architecture and design reviews and help raise engineering quality across the team.
- You take ownership and proactively improve the systems around you.
- You are comfortable debugging unfamiliar distributed systems in production.
- You make pragmatic tradeoffs between reliability, performance, delivery speed, and cost.
- You communicate clearly in a remote, async-friendly environment.
- You prefer incremental delivery: ship a focused solution, validate it in production, and improve it based on evidence.
- You address the underlying cause of operational problems rather than repeatedly treating the symptoms.
- 5+ years of experience building and operating production systems at scale.
- Strong proficiency in Go.
- Experience building and operating services on Kubernetes.
- Experience with infrastructure-as-code and GitOps tooling such as Terraform, Helm, and Argo CD.
- Production experience with at least one major cloud provider: AWS, GCP, or Azure.
- Hands-on experience with telemetry systems such as OpenTelemetry, Prometheus, Grafana, or comparable technologies.
- Experience with ClickHouse.
- Experience with high-throughput ingestion, streaming, queueing, or storage systems.
- Experience building multi-tenant cloud services.
- Experience optimizing infrastructure for both performance and cost.
- Experience with TypeScript.
#LI-Remote
About ClickHouse
Sourced by ZipRecruiter
Industry
Software development
Company size
51 - 200 Employees
Headquarters location
San Francisco, CA, US
Year founded
2016