Monitor the numbers that run your business.
Not just CPU and uptime.
Track business-specific and technical metrics with your own thresholds: queue depth, throughput, conversion, failed orders. Define incident triggers around your SLOs and catch async slowdowns before users feel them.
Your stack, your rules.
Stop monitoring only what your infrastructure exposes. Push any number (from any source) and define exactly when it means something is wrong.
Your metrics, your thresholds
Define exactly what healthy looks like for your business. Push any numeric value (throughput, queue depth, conversion rate, error counts) and set the precise threshold that triggers an incident.
- Alert when conversion, queue depth or throughput shifts
- Column-value / count / threshold criteria
- Units and human-readable labels
SLO-aligned incidents
Tie every alert to an objective. Sustained-breach windows let you ignore transient spikes and fire only when a threshold stays violated for the window you care about.
- Define triggers around your reliability targets
- Sustained-breach windows to avoid flapping
- Route to the owning team
SLO Breach Window
Catch async slowdowns
Silent pipeline failures are the hardest to detect. Jobs back up, queues stop draining, and users see nothing, until it is too late. Track the numbers that reveal what is happening behind your HTTP layer.
- Detect when queues stop draining
- Catch silent pipeline backups
- Correlate with incidents and AI root cause
Alert on what actually matters to your business.
Define your first custom metric in under a minute.
Create Your First Custom Metric