AI Outage Analysis

AI Outage Analysis turns monitoring data into incident reports. When a service goes down or begins experiencing problems, Layeredy AI analyzes signals across your whole stack to help you determine what happened and how to resolve it.

The analysis combines HTTP(s), ping results, server metrics such as CPU, RAM, disk usage, and system load, and application logs including errors, exceptions, timeouts, and events. These collected around the incident start help Layeredy AI to identify failures and why they happened.

For example, a HTTP failure tells you that a service is unavailable, but when combined with the knowledge of a sudden increase in memory usage and application logs showing an out-of-memory error, Layeredy AI can identify memory exhaustion as the likely cause.

image