---
sourceDocument: Australia IT Operations Management
sourceDocumentLink: https://servicenow-prod.fluidtopics.net/r/de-DE/it-operations-management

 Release :

    - australia

ft:locale :

    - de-DE

ft:publication_title :

    - Australia IT Operations Management

ft:clusterId :

    - itom

bundleId :

    - itom

workflow :

    - Technology


---

# Exploring Service Reliability Management

# Exploring Service Reliability Management {#ariaid-title1}

* Freigeben Version: Australia
* 
* Aktualisiert 12. März 2026
* 
* ![](https://www.servicenow.com/docs/portal-asset/ico-clock) 3 Minuten Lesedauer

Service Reliability Management (SRM) provides a self-serve, guided experience for teams to manage service health. The experience is built using the Service Operations Workspace application and combines ITOM and ITSM
capabilities into a single workflow.

## SRM overview {#exploring-service-reliability-management__cf-exploring-parent-overview}

Optimize service health with site reliability engineering (SRE) practices. SRM is a single operations workspace that empowers teams to improve the reliability of digital services with SRE.

* Use on-call escalations to respond to issues in a timely manner.
* Reduce setup friction with guided self-service to onboard distributed teams with separated data, empowered access, and minimal governance from central IT.
{#exploring-service-reliability-management__ul_ahx_ypj_4bc}

When SRM is installed, several plugins and applications are also activated. For more information, see [Plugins or applications installed with ITOM AIOps](https://servicenow-prod.fluidtopics.net/77F6fXXw64~htXxNArb_hw "Tables that list the plugins or applications that are installed with ITOM AIOps applications. When you update your application, any newly required application dependencies are installed.").

## SRM users {#exploring-service-reliability-management__cf-exploring-parent-users}

{#exploring-service-reliability-management__table_ayk_qqd_pwq__entry__3}

| Users | Description | Contains Roles |
|-|-|-|
| admin | A ServiceNow administrator is responsible for the administration, development, operation, education, and maintenance of the ServiceNow platform. Responsible for installation and can perform Service Operations Workspace Admin Center configuration of SRM. ServiceNow administrators manage, configure, and maintain the ServiceNow platform. In SRM, they can access and work in the Service Operations Workspace Admin Center. Only administrators can do the following: * Install SRM. * Add and manage SRM administrators. * Create and manage integration users. {#exploring-service-reliability-management__ul_iws_f3s_1hc} | All |
| SRM administrator \[srm_admin\] Hinweis: This role differs from the ServiceNow admin role. | SRM administrators can manage account settings, configurations, and users. Administrators can perform the following actions: * Access, create, edit, or delete all SRM configurations. * Add or manage integrations. * Create integrations with Application Performance Monitoring (APM) tools. * Set up and maintain reliability metrics. * Set up and maintain error budget policies. {#exploring-service-reliability-management__ul_byk_qqd_pwb} | * Manager * Responder {#exploring-service-reliability-management__ul_cyk_qqd_pwb} |
| SRM manager \[srm_manager\] | Managers oversee a team of SREs. Managers assign SREs to the team on-call schedule, monitor their performance, and create procedures to handle incidents and develop solutions. Managers promote resilience across all the systems and the DevOps workflows. Managers can perform the following actions within the context of their teams: * Define and set up teams, on-call schedules, and services. * Add and delete users such as responders and managers for the teams they're a part of. * Add or manage integrations. * Create Integrations with Application Performance Monitoring (APM) tools. * Set up and maintain reliability metrics. * Set up and maintain error budget policies. {#exploring-service-reliability-management__ul_dyk_qqd_pwb} | Responder |
| SRM responder \[srm_responder\] | A Service Reliability Engineer (SRE) that uses SRM to perform everyday tasks. Responders are the individuals who are on call and diagnose and remediate incidents. Responders can only access configurations that they're a part of. They can only access the alerts or incidents for which they have permissions. SREs can perform the following actions, within the context of their teams: * Set up services, teams, and integrations. * Confirm their on-call schedules. * Manage incident and alert records. * Update teams that they've created. * Add other responders. * Create integrations with Application Performance Monitoring (APM) tools. * Set up and maintain reliability metrics. * Set up and maintain error budget actions. {#exploring-service-reliability-management__ul_eyk_qqd_pwb} | Inherits 17 roles including the following: * cmdb_read * sn_sow.sow_user * sn_sow_srm.srm_responder * workspace_user * slo_operator {#exploring-service-reliability-management__ul_fyk_qqd_pwb} |
[Tabelle : 1. Users]

{#exploring-service-reliability-management__table_ayk_qqd_pwq}

For more information, see [SRM roles and responsibilities](https://servicenow-prod.fluidtopics.net/Kc0aTrAdFKZ~j8Hri26bSg "Roles grant users access to different parts of the SRM console. Roles determine the actions that users can or can't perform in Service Reliability Management.").

## SRM workflow {#exploring-service-reliability-management__cf-exploring-parent-workflow}

1. Product teams in IT or Lines of Business continuously deliver new service instances and technology management services technical and application services. Example: New customer billing portal.
2. Along with SLO Management, teams can register services and define service level objectives (SLOs), helping them reach business outcomes. Example: 95% monthly availability for the billing portal.
3. Monitoring integrations are set up by the teams to collect the real-time health of these services. Example: Cloud Observability.
4. Monitoring creates service level indicators (SLIs) impacting alerts when services are underperforming. Automation groups and enriches. Example: Billing portal latency is exceeding 7 s.
5. When the alerts indicate an outage or customer-impacting degradation, incidents are created and on-call notifications notify appropriate team resources. Example: A Billing SRE team is notified via phone of a latency issue on the billing portal.
6. After teams collaboratively diagnose and remediate incidents, they identify action items for improving the system's resilience. Example: The Billing team decides to add additional web server capacity.
7. Management continually reviews SLO performance, helps to prevent changes when the error budget is exhausted, and prioritizes improvement initiatives for underperforming services.
{#exploring-service-reliability-management__cf-exploring-parent-workflow-ol}

## SRM benefits {#exploring-service-reliability-management__cf-exploring-parent-benefits}

{#exploring-service-reliability-management__table_mlp_szc_4bc__entry__3}

| Benefit | Feature | Users |
|-|-|-|
| Team-based experience | [Working with SRM teams](https://servicenow-prod.fluidtopics.net/n3KW39YiZ4HCMwWIe_GLTA "Manage schedules and define escalation policies for your team. That way, your team sees who is on call and accountable and can have the confidence that critical alerts or incidents are acknowledged in a timely manner.") | SRM administrators, managers, and responders |
| Service registration | [Working with SRM services](https://servicenow-prod.fluidtopics.net/BdAtdC1wVQ6dOeiIPojPTA "A service represents a functional outcome like networking, payments, or HR services, that is owned by a team. To deliver that outcome, a service can contain one or more technical components like a user authentication service, or a piece of shared infrastructure like a database.") | SRM administrators, managers, and responders |
| Prebuilt integrations | [Working with integrations in SRM](https://servicenow-prod.fluidtopics.net/j_~m3EloCA7~IYzZQNbZyw "Connect your services to monitoring tools using the Integrations Launchpad integrations. Integrations send information to Service Reliability Management (SRM), helping you track alerts, manage incidents, and maintain service health.") | SRM administrators, managers, and responders |
| Measure service health | [Working with reliability metrics](https://servicenow-prod.fluidtopics.net/sixwTtdyj9W5yPw_gnhWRA "Learn about the reliability metrics and features that can help you track service health, respond to issues, and support business goals.") | SRM administrators, managers, and responders |
| On-call coverage | [Create an SRM on-call schedule](https://servicenow-prod.fluidtopics.net/tR3bn~_6TSbpgBN~tJrlvg "Set up an on-call schedule to make sure that someone is available to respond to incidents and critical alerts.") | SRM administrators, managers, and responders |
| Remediate high severity alerts and incidents | [Working with SRM reliability tasks](https://servicenow-prod.fluidtopics.net/O4kIzlXiHNLUD2ksEBmiwg "Alerts, incidents, and change requests are reliability tasks. From creation to resolution, SRM helps you manage your alerts throughout the response life cycle.") | SRM administrators, managers, and responders |
[ ]

{#exploring-service-reliability-management__table_mlp_szc_4bc}

## What to explore next {#exploring-service-reliability-management__cf-exploring-parent-links}

To learn more about configuring and using SRM, see:

* [Configuring Service Reliability Management](https://servicenow-prod.fluidtopics.net/nJX1UAWH7f~sv8GNh7966g "Configuring Service Reliability Management (SRM) involves installing it from the ServiceNow Store or Service Operations Workspace Admin Center. You can also assign administrators and import services and teams.")
* [Using Service Reliability Management](https://servicenow-prod.fluidtopics.net/RaQZpSjMugdBRAzs7Jk3UA "Service Reliability Management (SRM) enables you to register services, monitor service health, respond to service degradations with on-call shifts and escalation policies and triggers, and onboard distributed teams with minimal governance from central IT.")
* [Service Reliability Management reference](https://servicenow-prod.fluidtopics.net/Z8B0mMTxIY1XVVjy4m~_Tg "Reference topics provide additional information about administering Service Reliability Management.")
{#exploring-service-reliability-management__ul_nlp_szc_4bc}
* **[Get started with Service Reliability Management](https://servicenow-prod.fluidtopics.net/R2YmQTEl71O5Je8xCtDIIw)**   
  Service Reliability Management (SRM) accelerates your path to viewing service health in the context of service level objectives and incident resolution. Helps IT Operations and DevOps teams deliver on the promise of agility, performance, and uptime.
* **[SRM alert workspace](https://servicenow-prod.fluidtopics.net/7qc6ONvHZIBnHy5HahCE9A)**   
  The alert workspace contains various areas containing alert details and possible actions.
* **[SRM incidents](https://servicenow-prod.fluidtopics.net/2jHa5WfqwnHHOWkLmEk5Ag)**   
  Track and collaborate on incidents in the Incidents tab, helping you and your teams resolve issues efficiently.

