AI-Powered Service Operations Part 2 - collaborate, deflect and resolve
hi this is evans nicholson from technical marketing at servicenow welcome to part two of the mini-series ai-powered service operations starring servicenow itsm and itom at servicenow we understand the amount of data coming in and the complexity of today's hybrid cloud environments we know many times lack of data and information is not the issue ai powered service operations enables the move away from an incident creation culture to more of an alert based actionable mindset this helps you find problems and solve them in a more predictive fashion versus being stuck in problem change and incident management the key is being proactive versus reactive and having the right supporting system to be able to achieve that in part one we saw how an issue with a company's mission critical internal ordering service was causing a flood of new tickets in our service engineer's queue we then showed how in parallel the operations engineer was able to quickly determine probable root cause of the outage and trigger a service degradation so let's pick the demo up from there nora griffith is an employee who uses the order status application called porsche online in her day-to-day work she's having issues completing her work and so she's going to check in and try to get some help remember even though she's using this from the servicenow web interface she could also use microsoft teams skype facebook workplace or other collaboration platforms let's just type in the summary of the issue and thanks to the natural language understanding model virtual agent can pick out keywords and phrases and suggest the right workflow in this case nora's team has configured a set of additional questions to help troubleshoot the different areas related to the ordering system now nora has confirmed that there is an issue with the order status process which is used in the porsche online service she doesn't need to open up a new incident because she can already see that there's a ticket open this is what we mean by deflection alerting employees that there's an ongoing issue help stop the inflow of new tickets allowing service operations teams to concentrate on troubleshooting and fixing the issue instead of managing more and more tickets now let's continue looking at things as roberto hopper a member of nora's it operations team from operator workspace we can see the order status service is in a critical state as we saw in the previous video related alerts have been grouped together by itom event management let's just go back into the group alert so we can continue troubleshooting as we look for a way to resolve the issue with the order status service the first alert that came in was from health log analytics this hla alert is letting us know there are some anomalies with the logs basically unusual patterns or something that happened recently that hasn't happened before if i open up this option we can do a deep dive into the logs on the first line we can see details on a suspect log associated with the windows host x42 opening up log viewer allows us to really dive deep into these logs and understand what we're seeing based on what was triggered by hla we can see more about pattern recognition both out of the box and patterns that are learned over time thanks to the machine learning of health log analytics now let's go back to the group alert speaking of machines recognizing patterns and helping us troubleshoot let's click on probable root causes and we can see another example of this two recent changes are flagged as possible causes of the issue we're seeing with the order status service the bottom one appears to be a scheduled change by roberto's devops team but the top one is unauthorized so let's take a look at that scrolling down we can see that in this emergency change someone has changed the sql trace flag to true from false this is very helpful when troubleshooting sql but sometimes this can cause disk space issues if you forget to turn it off because logs are much larger when this is enabled there's one additional step we could perform to take a look at the logs but in the essence of time i'm going to skip to roberto's conclusion that disk space is filling on the sql server and that's what's causing a cascading effect to the order status service roberto is using all the information within these pages within the servicenow platform to reach this conclusion let's go back to the group alert and i'll show you how we can trigger the fix through flow designer directly from the servicenow platform first roberto decides to bring up the agent assist function to look for helpful articles related to disk space here's one related to clearing logs on a windows server seems appropriate right so let's attach that to the ticket roberto will put in a note about his findings giving some insight to the team into what he's thinking now let's open up actions so we can actually fix the issue that's the whole point right we're trying to fix the problem now this is one of my favorite things about the servicenow platform in this scenario in this case with flow designer we actually have the tools to reach out and fix issues on affected ci's which is great without such functionality we'd still have to manually remove logs expand disk space things like that now clicking on the run remediation button the platform actually has a pre-built workflow to handle this situation as i scroll down in the main group alert you can see all of the associated alerts have been closed because the issue has been addressed you can see we cleared out about two and a half gigs of space by running the remediation and here's roberto's note with the kb article attachment now let's go back to the operator workspace and check the status of the service order status is no longer in a critical state we've just seen how servicenow can help it teams solve complex issues we can help identify if all events coming in are tied to a single problem with event correlation we can find the root cause of the issue faster we can notify employees of ongoing issues and of course we can provide suggestions but most importantly remediation actions to fix problems we've also seen how operations and service teams can work on the same data in the same system instead of working in silos this moves organizations towards a more modern and cooperative approach an approach we call service operations for more information check out the main itom product page thanks a lot and i hope to talk to you soon you
https://www.youtube.com/watch?v=WyQTP0AA1VU