Skip to content

How it works

Connect once, then let Ketl read every alert

Point Ketl at the telemetry you already collect. The moment an alert fires, it reads across every signal, names the root cause with the evidence behind it, and proposes the exact fix or rollback.

01Connect

Three steps, then it runs itself

Most teams connect their first source and run a live diagnosis in under fifteen minutes. After that, Ketl fires automatically on every alert.

01

Connect your telemetry

Point Ketl at the places your signal already lives: logs, traces, recent deploys, and your past incidents. No new agent to run, no data to move. Read-only by default.

ketl connect datadog pagerduty github sentry
02

Ketl reads and reasons at each alert

The moment an alert fires, Ketl pulls the relevant window across every signal and reasons about cause the way a senior on-call would, in seconds rather than an hour.

alert → reading 2,140 log lines · 3 deploys · 11 traces
03

You get root cause plus a proposed fix

A ranked root cause, each with the exact lines Ketl reasoned from, and one concrete next step: a fix or a rollback, with the command and the risk. You decide, then act.

cause: N+1 from v2.48.0 · fix: rollback → v2.47.3
02What you connect

The tools your signal already lives in

No new agent to run, no data to move. Ketl reads the sources you already run, read-only by default.

PagerDuty

Alerts

Trigger a diagnosis the moment a page fires.

Datadog

Logs & metrics

Read the log and metric window around an alert.

Grafana

Logs & dashboards

Pull Loki logs and dashboard context.

Sentry

Errors

Attach the exact exception and stack.

GitHub

Deploys

Correlate the alert with recent releases and PRs.

Slack

Delivery

Post the ranked cause and fix into the incident channel.

All connections are read-only by default. Ketl never writes to your systems. See the integration docs for setup guides and required permission scopes.

03What you get back

Root cause, evidence, and a concrete next step

Every diagnosis Ketl returns has the same three parts. Together they give you enough to act in minutes, not an hour of blind digging.

Ranked root cause

A prioritized list of likely causes, ordered by how strongly the telemetry supports each one. Each carries an honest confidence rather than a false claim of certainty.

Cited evidence

Every cause cites the exact log lines, deploy record, or past incident it reasoned from, quoted verbatim. You confirm the read or overrule it in seconds.

Proposed fix or rollback

One concrete next step for the leading cause: the exact command, the blast radius, and how to verify the fix worked. Ketl proposes; you decide; you act.

Run a diagnosis on your next alert

Start free and run a real diagnosis before you connect anything. Paste an alert and its logs, and Ketl returns a ranked root cause with evidence and a proposed fix.