Check That AI SRE Is Enabled
AI SRE has no self-service switch. Rootly checks it on the team you have selected, so AI SRE can be on for one team in your workspace and off for another. The Opt in to Rootly AI capabilities setting under AI SRE → Atlas → Global doesn’t turn AI SRE on or off. To check, open AI SRE → Atlas. The AI SRE card shows Enabled or Off. When AI SRE is enabled, alerts and incidents also show an AI SRE tab. Scheduled maintenance incidents never get the tab. If your sidebar has no AI SRE item, open AI & Agents → AI SRE instead: either AI SRE is off for the selected team, or the consolidated navigation isn’t enabled for your team yet. When AI SRE is off, AI & Agents → AI SRE shows a Your AI SRE Teammate page with a Join waitlist button. Joining records your interest but doesn’t turn anything on. Contact your Rootly account team to enable AI SRE.Capabilities Rolled Out Separately
Some capabilities have their own enablement and may not be on for every team. Your Rootly account team turns them on:- Instructions, the team-wide guidance page
- Knowledge graph and Memory
- Generate with AI, which drafts Instructions for you
- Rerun investigation on a finished investigation
- The Investigations page, which lists AI SRE investigations for the selected team
- Posting investigation results to Slack
- Automatic investigation of new alerts and incidents without a rule
- Automatic activation of Memory notes
- Private agents
Where Settings Live
AI SRE investigates through Atlas, Rootly’s AI layer. Investigations can use AI connectors, the Knowledge graph and Memory where enabled. Private Agent is an Early Preview for approved customers; those teams can also use it during investigations. Each part has its own card on the Atlas page./account/private-connect/agents to the Rootly app URL, when Private Agent and AI SRE are enabled. Connecting AI connectors and managing investigation rules are open to Incident Response Owners and Admins and to On-Call Admins. Editing Instructions takes one of those roles plus an Incident Response seat. Starting an investigation needs write access to the alert or incident, and anyone who can read it can view the result. The permissions matrix has the full detail.
Run Your First Investigation
Work through these steps once, in order. Calibrating on an alert whose cause you already know tells you whether AI SRE reaches your team’s answer and which sources it used to get there.Connect the Sources That Hold Your Evidence
- A code or deploy source, such as GitHub. It answers what changed before the alert fired. Without one, AI SRE doesn’t raise change-related causes and reports that it couldn’t examine them.
- Your main observability provider, such as Datadog, Grafana, New Relic, Honeycomb, Sentry or Dash0. It supplies the metrics, logs, traces and monitors behind the alert.
Write Your Instructions
Create an Investigation Rule in Manual
- Add conditions that describe the alert class. All conditions must match.
- Check the Matches · 90d count beside each condition and the All conditions total. Hover a count and select View alerts to see which alerts match.
- Optionally add rule Instructions for this alert class.
- Leave Run mode on Manual, the default, and save.
Start an Investigation on a Known Alert
Read the Result Critically
- The outcome. Root cause identified comes with a confidence tier, such as High confidence. Ask whether the evidence justifies it. Inconclusive — needs human means no root cause was confirmed from the evidence AI SRE could reach. Contributing factor identified and Could not investigate are covered in Reading the Result.
- The decisive evidence. Open the Decisive items under Evidence with View and check that each source says what the report claims.
- What was ruled out. Lines marked ruled out in the Investigation path show which explanations AI SRE rejected.
- Where the evidence ran out. Checks marked inconclusive or no evidence, and the next step on a Blocked outcome, can point to a source to connect or re-authorize. A Suspected area names the leading explanation AI SRE couldn’t confirm.
Promote the Rule to Auto-run
Widen Coverage Deliberately
With one rule working, expand along whichever gap your results point to:Troubleshooting
The AI SRE card shows Off, or the page shows Your AI SRE Teammate
The AI SRE card shows Off, or the page shows Your AI SRE Teammate
The AI SRE tab appears on alerts, but you can't open AI & Agents
The AI SRE tab appears on alerts, but you can't open AI & Agents
An incident has no AI SRE tab
An incident has no AI SRE tab
The tab says you don't have permission to start an investigation
The tab says you don't have permission to start an investigation
The Instructions, Knowledge graph or Memory card is missing
The Instructions, Knowledge graph or Memory card is missing
Alerts are investigated automatically although the rule is Manual
Alerts are investigated automatically although the rule is Manual
Frequently Asked Questions
Do I need an investigation rule to use AI SRE?
Do I need an investigation rule to use AI SRE?
Is it safe to calibrate on a resolved alert?
Is it safe to calibrate on a resolved alert?
How many AI connectors do I need before results are useful?
How many AI connectors do I need before results are useful?
What if my observability data has no built-in connector?
What if my observability data has no built-in connector?
How do I re-test the calibration alert after connecting more sources?
How do I re-test the calibration alert after connecting more sources?
Why can't I find Atlas in the product?
Why can't I find Atlas in the product?