> For the complete documentation index, see [llms.txt](https://developer.harness.io/llms.txt). Markdown versions of documentation pages are available by appending `.md` to page URLs; this page is available as [Markdown](https://developer.harness.io/ai-sre/ai-sre-for-incident-responders/onboarding-guide-users.md).

# AI SRE Onboarding Guide for Incident Responders

This guide walks you through the essentials of using Harness AI SRE as a responder or engineer.

You will learn how to navigate the dashboard, respond to incidents, collaborate with your team, and use runbooks and AI-powered tools to resolve issues faster.

Your administrator has already configured the integrations and incident types. This guide focuses on what you need to know to be effective as an incident responder from day one.

### Prerequisites <a href="#prerequisites" id="prerequisites"></a>

Before getting started, confirm the following with your administrator:

| Item                             | Details                                                                                         |
| -------------------------------- | ----------------------------------------------------------------------------------------------- |
| Harness account access           | You have been added to your organization's Harness account with appropriate permissions         |
| Collaboration tools connected    | The Harness AI SRE bot is installed in your team's Slack workspace or Google Chat               |
| Monitoring tools configured      | Your organization's monitoring tools (Datadog, New Relic, Grafana, etc.) are already integrated |
| On-call schedule (if applicable) | You have been added to your team's on-call rotation in PagerDuty, OpsGenie, or a similar tool   |

{% hint style="info" %}
**NEED ADMIN SETUP FIRST?**

If your organization has not configured AI SRE yet, share the [AI SRE Onboarding Guide for Administrators](/ai-sre/ai-sre-for-administrators/get-started/overview.md) with your platform team to get started.
{% endhint %}

***

### 1. Explore the AI SRE dashboard <a href="#id-1-explore-the-ai-sre-dashboard" id="id-1-explore-the-ai-sre-dashboard"></a>

{% tabs %}
{% tab title="Step by Step" %}
The AI SRE dashboard is your central hub during on-call shifts and day-to-day operations.

1. Log in to your **Harness account**.
2. Navigate to **AI SRE** from the left navigation panel, then click **Overview**.

   <figure><img src="/files/0AJPeboZUnZASwYcXCzE" alt="Harness left navigation with AI SRE highlighted"><figcaption></figcaption></figure>
3. The dashboard opens, shown below.

   ![AI SRE dashboard overview](/files/Y8W4Nb5j12MDfRKC3R5o)

   On the dashboard, review the following:

   * **Active Incidents:** Any ongoing incidents that need attention.
   * **Recent Alerts:** The latest alerts from your monitoring tools.
   * **Metrics and Trends:** Key reliability metrics like MTTR and incident volume.
4. Use the filters at the top to narrow by **incident type**, **severity**, **status**, or **assigned team**.

{% hint style="info" %}
**QUICK ORIENTATION**

Bookmark the AI SRE dashboard for quick access during on-call shifts. The active incidents panel updates in real time.
{% endhint %}
{% endtab %}

{% tab title="Interactive Guide" %}
{% embed url="<https://app.tango.us/app/embed/c55a8b8f-bce1-487c-b8ea-5d178a844682?skipCover=true&defaultListView=false&skipBranding=false&makeViewOnly=false&hideAuthorAndDetails=true>" %}
Explore the Harness AI SRE Dashboard
{% endembed %}

Get familiar with the dashboard layout, active incidents, alerts, and key metrics at a glance.
{% endtab %}
{% endtabs %}

Explore the dashboard further with these resources:

* Go to [Understanding Incident Types](/ai-sre/ai-sre-for-administrators/set-up-incident-management/incidents.md) to understand how incident types map to severity levels and responder teams.
* Go to [Integration Overview](/ai-test-automation/integrations/integrations.md) to review which monitoring tools are connected to your environment.

***

### 2. Respond to an incident <a href="#id-2-respond-to-an-incident" id="id-2-respond-to-an-incident"></a>

{% tabs %}
{% tab title="Step by Step" %}
When an incident is created, automatically from a monitoring alert or manually by a teammate, here is how to respond.

1. You will receive a notification via [**Harness On-Call**](/ai-sre/ai-sre-for-incident-responders/handle-on-call.md), **Slack**, **Google Chat**, or your on-call tool.

   ![Slack incident notification](/files/9duoKq27h6vBeWgF7MMU)
2. Click the notification link to open the **incident detail page** in Harness.

   ![Incident detail page](/files/H7OYCfYLplWKk7wLewWT)
3. Review the incident summary:
   * **Severity** and **incident type:** Understand the scope and priority.
   * **Timeline:** The sequence of alerts and events that triggered the incident.
   * **Related alerts:** Correlated monitoring data and affected services.
4. If you have been paged about the incident, **acknowledge** the incident to let your team know you are on it.
5. Update the **status** as you work through it: **Investigating**, **Fixing**, **Monitoring**, **Closed**.

   ![Incident status dropdown](/files/fFmFu4RisTlhNx72ZyqZ)
6. Use the **incident channel** in Slack or Google Chat to collaborate with other responders in real time.
7. Add **notes and updates** to the incident timeline to keep a clear record of actions taken.

{% hint style="info" %}
**SLACK COMMANDS**

You can manage incidents without leaving Slack. Use `/harness` slash commands to acknowledge, update status, add notes, and more.
{% endhint %}
{% endtab %}

{% tab title="Interactive Guide" %}
{% embed url="<https://app.tango.us/app/embed/50543ebc-97c8-4b92-86c2-bc19cd4fc230?skipCover=true&defaultListView=false&skipBranding=false&makeViewOnly=false&hideAuthorAndDetails=true>" %}
Respond to an incident in Harness AI SRE
{% endembed %}

Learn how to acknowledge, triage, and begin working on an incident when you are paged or alerted.
{% endtab %}
{% endtabs %}

Extend your incident response skills with these resources:

* Go to [Slack Commands Reference](/ai-sre/ai-sre-for-incident-responders/slack-commands.md) to manage incidents directly from Slack without switching to the UI.
* Go to [AI Scribe Agent](/ai-sre/ai-sre-for-incident-responders/use-ai-agents/ai-agent.md) to understand how the Scribe captures your incident activity automatically.

***

### 3. Create an incident manually <a href="#id-3-create-an-incident-manually" id="id-3-create-an-incident-manually"></a>

{% tabs %}
{% tab title="Step by Step" %}
Not every incident starts from an automated alert. If you notice a problem, customer reports, degraded performance, or a teammate flagging something, you can create an incident manually.

1. Navigate to **Incidents** from the left panel.

   ![Incident list view](/files/zAG2l3DdRwx6OE88974W)
2. Click **New Incident**, or select an incident type from the **New Incident** dropdown.
3. The **Create a New Incident** form appears.

   <div data-with-frame="true"><figure><img src="/files/lhWPS6oSIhVlvkQMj13b" alt="Harness left navigation with AI SRE highlighted"><figcaption></figcaption></figure></div>
4. Fill in the incident details:
   * **Title:** A clear, concise summary (for example, "Elevated error rates on checkout API").
   * **Severity:** Choose the appropriate level based on impact.
   * **Description:** What you are observing, when it started, and any initial hypotheses.
   * Any additional **required fields** specific to your incident type.
5. Click **Save**.

An incident channel is created in your communication tool and relevant team members are notified.

{% hint style="info" %}
**FROM SLACK**

You can also create incidents directly from Slack using the `/harness new` command. This is useful during on-call when you want to stay in your communication tool.
{% endhint %}
{% endtab %}

{% tab title="Interactive Guide" %}
{% embed url="<https://app.tango.us/app/embed/f14f004b-3405-4384-baae-48a035a8eb12?skipCover=true&defaultListView=false&skipBranding=false&makeViewOnly=false&hideAuthorAndDetails=true>" %}
Create a new incident in Harness AI SRE
{% endembed %}

Sometimes you will spot an issue before automated monitoring catches it. Learn how to declare an incident manually.
{% endtab %}
{% endtabs %}

Create incidents more efficiently with these resources:

* Go to [Slack Commands Reference](/ai-sre/ai-sre-for-incident-responders/slack-commands.md) to use `/harness new` and other commands to create and manage incidents from Slack.
* Go to [Understanding Incident Types](/ai-sre/ai-sre-for-administrators/set-up-incident-management/incidents.md) to understand what incident types are available and how they affect notifications and runbooks.

***

### 4. Use runbooks during an incident <a href="#id-4-use-runbooks-during-an-incident" id="id-4-use-runbooks-during-an-incident"></a>

{% tabs %}
{% tab title="Step by Step" %}
Runbooks are predefined playbooks that guide you through incident response.

Some run automatically when certain conditions are met; others can be triggered manually.

1. Navigate to **Incidents** from the left panel.

   ![Incident list view](/files/zAG2l3DdRwx6OE88974W)
2. Click the **Incident ID** of the relevant incident to open the **Details** tab for an active incident.

   ![Runbooks tab on incident detail page](/files/KdBwUi4mbbC5WFlPaOfR)
3. Click the **Runbooks** tab.

   ![Runbooks tab on incident detail page](/files/zSCxonPib7IU6q7r4VOU)
4. Review any runbooks that have been **auto-attached** based on the incident type.
5. To manually attach a runbook, click **Add Runbook**, search for the one you need, and confirm.
6. Work through the runbook step by step:

   * **Automated steps:** Run and report results without any action from you.
   * **Manual steps:** Show instructions for you to follow. Mark each one complete as you go.

   Runbook execution is logged in the incident timeline.

{% hint style="info" %}
**WHEN TO USE RUNBOOKS**

If you are unsure which runbook applies, check the incident type. Your administrator has likely associated recommended runbooks with each type. You can also browse all available runbooks under **Runbooks** in the left navigation.
{% endhint %}
{% endtab %}

{% tab title="Interactive Guide" %}
{% embed url="<https://app.tango.us/app/embed/48a2f0ca-d07f-4395-aa7b-9b5c2c7b9018?skipCover=true&defaultListView=false&skipBranding=false&makeViewOnly=false&hideAuthorAndDetails=true>" %}
Use runbooks during an incident
{% endembed %}

Runbooks guide you through predefined response steps and can automate common actions during an incident.
{% endtab %}
{% endtabs %}

Get more out of runbooks with these resources:

* Go to [Browsing Runbooks](/ai-sre/ai-sre-for-administrators/set-up-runbook-management/create-runbook.md) to explore the runbook library and see what playbooks are available to you.
* Go to [Understanding Incident Types](/ai-sre/ai-sre-for-administrators/set-up-incident-management/incidents.md) to review which runbooks are associated with each incident type.

***

### 5. Use the AI Scribe Agent <a href="#id-5-use-the-ai-scribe-agent" id="id-5-use-the-ai-scribe-agent"></a>

The AI Scribe Agent works alongside you during incidents to reduce manual overhead.

* **Automatic summaries:** The Scribe monitors your incident channel and picks out key decisions, actions, and findings as they happen.
* **Timeline generation:** It builds a structured timeline from channel activity, status changes, and runbook execution.
* **Post-incident reports:** After resolution, the Scribe drafts a post-incident report from the timeline and channel discussions, giving you a head start on the retrospective.

To access Scribe outputs, open the **Details** page and look for the AI-generated **Incident Summary**.

![AI Summary section on incident detail page](/files/eKSITBpa7kYDNHuBSQo2)

Also, the **Timeline** tab shows updates generated by the Scribe.

![AI Summary section on incident detail page](/files/JVGOVrycbpzsG4lQA9u5)

Dive deeper into the AI agents with these resources:

* Go to [AI Scribe Agent](/ai-sre/ai-sre-for-incident-responders/use-ai-agents/ai-agent.md) to read the full documentation on how the Scribe works and how to get the most out of it.
* Go to [RCA Change Agent](/ai-sre/ai-sre-for-incident-responders/use-ai-agents/rca-change-agent.md) to understand how AI-powered root cause analysis works alongside the Scribe during an incident.

***

### Next steps <a href="#next-steps-ai-sre-user-next-steps" id="next-steps-ai-sre-user-next-steps"></a>

Continue with these resources to deepen your AI SRE knowledge:

* Go to [Slack Commands Reference](/ai-sre/ai-sre-for-incident-responders/slack-commands.md) to use the full set of slash commands for managing incidents from Slack.
* Go to [Understanding Incident Types](/ai-sre/ai-sre-for-administrators/set-up-incident-management/incidents.md) to understand how incident types map to severity levels, responder teams, and escalation paths.
* Go to [Browsing Runbooks](/ai-sre/ai-sre-for-administrators/set-up-runbook-management/create-runbook.md) to explore the automated playbooks available to you.
* Go to [Integration Overview](/ai-test-automation/integrations/integrations.md) to see which monitoring, communication, and ITSM tools are connected to your environment.
* Go to [AI Scribe Agent](/ai-sre/ai-sre-for-incident-responders/use-ai-agents/ai-agent.md) for deeper documentation on AI-powered incident documentation and insights.
