All templates

Automated Imply Ingestion Lag Alerts

Every few minutes, WebRun checks your Imply cluster's ingestion lag and query latency against the limits you set, and the moment either crosses the line it pings the on-call engineer on WhatsApp with the datasource and the numbers, then books a short incident review on Google Calendar for the team.

Runs on WebRun · Strict Lockdown policy
Every few minutes, around the clock WebRunorchestrates each step
1 Imply check ingestion lag and query latency
2 WhatsApp ping the on-call engineer
3 Google Calendar book an incident review
In short

How do I get alerted when Imply ingestion lag or query latency breaches a limit?

WebRun checks your Imply cluster's ingestion lag and query latency every few minutes against the limits you set. The moment a datasource breaches either one, it pings the on-call engineer on WhatsApp with the exact figures, then books a short incident review on Google Calendar, so a lagging pipeline gets caught before queries turn stale.

  • Ingestion lag gets caught within minutes instead of when a dashboard goes stale
  • On-call engineers get the exact breached metric, not just a generic alert
  • Every breach comes with a booked review so it gets a proper follow-up

Built for Data platform engineers · Site reliability teams · Analytics engineers · On-call engineers

Step by step

What does WebRun do on every run?

The exact actions WebRun takes, in order - in plain language, so you can adjust anything.

  1. WebRun signs in and gets to work

    Opens imply.io in a real browser with your saved login - no setup, no API keys.

  2. 1
    Imply - check ingestion lag and query latency
    imply.io
    WebRun in Imply: check ingestion lag and query latency
    WebRun opens Imply to check ingestion lag and query latency.
    • Open the Imply console and read ingestion lag for each datasource
    • Read query latency across recent queries
    • Compare both against the limits you set

    Done when Every datasource has been checked against its lag and latency limits.

  3. 2
    WhatsApp - ping the on-call engineer
    whatsapp.com
    WebRun in WhatsApp: ping the on-call engineer
    WebRun opens WhatsApp to ping the on-call engineer.
    • Send a message to the on-call engineer naming the datasource and the breached metric
    • Include the current lag or latency figure against the limit
    • Send only while the breach is still active

    Done when The on-call engineer has been pinged for this breach.

  4. 3
    Google Calendar - book an incident review
    calendar.google.com
    WebRun in Google Calendar: book an incident review
    WebRun opens Google Calendar to book an incident review.
    • Create a short incident review slot for the next available hour
    • Title it with the datasource name
    • Leave it for the on-call engineer to accept or move

    Done when A review slot exists on the calendar for the breach.

Run settings

How is each run configured?

Starting pageWhere Chrome opens at the start of each run
imply.io
ScheduleRuns automatically on this cadence
Every few minutes, around the clock
DeliveryHow each run's result reaches you
Ingestion lag alert · WhatsApp
OutputWhat each run produces - A WhatsApp ping naming the datasource and breached metric, plus a booked incident review.
Text
Setup & safety

Secure by default

Connect once, stays signed in

WebRun signs in once and keeps each session in a persistent environment, so every run picks up right where it left off.

Your credentials stay in your own private environment - WebRun never stores your passwords.
Strict Lockdown

Every action is checked against this policy before it runs.

Domains ALLOWLIST
Typed input ALLOW
Shell command BLOCK
File uploads BLOCK
Runs in a contained environment More on policies
Good to know

Questions, answered

What counts as a lag or latency breach?

Whatever limits you set for ingestion lag and query latency per datasource. Nothing is flagged until a reading actually crosses your line.

Does WebRun restart nodes or change cluster settings?

No. It only reads metrics from the Imply console and sends alerts. Cluster configuration and scaling stay in your engineers' hands.

Will we get pinged again if the lag keeps climbing?

It pings once per breach, then again only if the metric recovers and breaches a second time, so WhatsApp does not fill up with repeats of the same open issue.

Put this on autopilot.

Turn it on in minutes - or have our team set it up for you.