The New Normal for IT Ops Deepens Need for AI - Part 1
May 05, 2020

Will Cappelli
Moogsoft

Share this

The global pandemic has radically changed how enterprise IT services are consumed, both in the short and long term. Here's how AIOps can help IT Ops teams.

The current crisis has upended all aspects of our personal and work lives, and IT Ops pros aren't the exception. The abrupt shift to remote work has created unprecedented challenges for IT Ops teams, while increasing pressure on them to prevent outages and provide service assurance.

Specifically, new consumption patterns of enterprise IT services have put stress on systems, architectures and topologies at all stack layers. In response, IT Ops teams must rapidly implement structural and management changes to address both temporary and permanent shifts.

In this turmoil, AIOps has emerged as a lifeline. By streamlining and automating IT operations, AIOps helps IT leaders collaborate remotely and act quickly and precisely to maintain business-critical digital services — during the pandemic and beyond.

Let's look in more detail at these challenges and at how AIOps can help IT Ops teams cope and succeed.

AIOps: A Definition

An AIOps solution must have these five types of algorithms that fully automate and streamline five key dimensions of IT operations monitoring:

■ Data selection: Identifying and surfacing the most relevant information.

■ Pattern discovery: Correlating and finding relationships between events across your tool stack.

■ Inference: Identifying root causes and recurring issues.

■ Collaboration: Notifying appropriate operators, and facilitating collaboration.

■ Automation: Automating remediation

In a real world setting, an AIOps solution ingests heterogeneous data from many different sources. Using entropy algorithms, it removes noise and duplication, and selects only the truly relevant data. It then groups and correlates this relevant information using various criteria, like text, time and topology.

Next, it discovers patterns in the data, and infers which data items signify causes, and which signify events. It then communicates the result of that analysis to a collaborative environment, which will support automated responses to what has been discovered.

As such, an AIOps solution plays the role of organizing and integrating what an organization's domain-specific IT monitoring and management tools do, intelligently integrating the stack's functionalities. AIOps should act as the brain that brings together these tools, and becomes a coordinating, central layer.

Transitioning to the New Normal

As the workforce shifts to remote work, user behaviors will change and different elements of the IT infrastructure, both in-house and publicly sourced, will be stressed. This will result in new, quickly-evolving types of incidents and outages. With AIOps, IT Ops teams can detect and analyze genuinely novel anomalies which can cause incidents and outages rapidly and stealthily.

Cross-regional and intra-regional team collaboration among IT operations and NOC organizations will need to be reinforced virtually as the implicit supports derived from physical co-presence are removed. AIOps can enable and guide virtual collaborative observation, analysis and response efforts, helping IT Ops teams collaborate and communicate despite being physically dispersed.

Sharp and unpredictable levels of staff reduction due to illness and self-isolation will force IT operations and NOC organizations to "do more with less" on both the side of signal observation and the side of signal response. Here again AIOps can help IT Ops teams to respond by both dynamically filtering noisy alert streams, and integrating and automating platforms that support various aspects of incident and problem management.

Go to The New Normal for IT Ops Deepens Need for AI - Part 2

Will Cappelli is Field CTO at Moogsoft
Share this

The Latest

November 21, 2024

Broad proliferation of cloud infrastructure combined with continued support for remote workers is driving increased complexity and visibility challenges for network operations teams, according to new research conducted by Dimensional Research and sponsored by Broadcom ...

November 20, 2024

New research from ServiceNow and ThoughtLab reveals that less than 30% of banks feel their transformation efforts are meeting evolving customer digital needs. Additionally, 52% say they must revamp their strategy to counter competition from outside the sector. Adapting to these challenges isn't just about staying competitive — it's about staying in business ...

November 19, 2024

Leaders in the financial services sector are bullish on AI, with 95% of business and IT decision makers saying that AI is a top C-Suite priority, and 96% of respondents believing it provides their business a competitive advantage, according to Riverbed's Global AI and Digital Experience Survey ...

November 18, 2024

SLOs have long been a staple for DevOps teams to monitor the health of their applications and infrastructure ... Now, as digital trends have shifted, more and more teams are looking to adapt this model for the mobile environment. This, however, is not without its challenges ...

November 14, 2024

Modernizing IT infrastructure has become essential for organizations striving to remain competitive. This modernization extends beyond merely upgrading hardware or software; it involves strategically leveraging new technologies like AI and cloud computing to enhance operational efficiency, increase data accessibility, and improve the end-user experience ...

November 13, 2024

AI sure grew fast in popularity, but are AI apps any good? ... If companies are going to keep integrating AI applications into their tech stack at the rate they are, then they need to be aware of AI's limitations. More importantly, they need to evolve their testing regiment ...

November 12, 2024

If you were lucky, you found out about the massive CrowdStrike/Microsoft outage last July by reading about it over coffee. Those less fortunate were awoken hours earlier by frantic calls from work ... Whether you were directly affected or not, there's an important lesson: all organizations should be conducting in-depth reviews of testing and change management ...

November 08, 2024

In MEAN TIME TO INSIGHT Episode 11, Shamus McGillicuddy, VP of Research, Network Infrastructure and Operations, at EMA discusses Secure Access Service Edge (SASE) ...

November 07, 2024

On average, only 48% of digital initiatives enterprise-wide meet or exceed their business outcome targets according to Gartner's annual global survey of CIOs and technology executives ...

November 06, 2024

Artificial intelligence (AI) is rapidly reshaping industries around the world. From optimizing business processes to unlocking new levels of innovation, AI is a critical driver of success for modern enterprises. As a result, business leaders — from DevOps engineers to CTOs — are under pressure to incorporate AI into their workflows to stay competitive. But the question isn't whether AI should be adopted — it's how ...