ip-label is now officially part of ITRS. Read the press release.

ip-label is now officially part of ITRS. Read the press release.

Digital Monitoring: Limitations and Best Practices for Effective Monitoring

01
Monitoring
& Digital Experience
Digital Monitoring

Monitoring more does not necessarily mean monitoring better.

Business applications, websites, SaaS platforms, mobile services, APIs, cloud infrastructures… Information systems are now made up of a growing number of interdependent components.

An increasingly complex digital ecosystem
Business applications
Websites
SaaS platforms
Mobile services
API
Cloud infrastructures
01 / Monitoring

Digital monitoring has become essential

In this context, digital monitoring has become essential for monitoring the availability and performance of digital services, detecting anomalies and enabling IT teams to respond quickly.

But monitoring more does not necessarily mean monitoring better.

Organizations now have access to a considerable volume of metrics, logs, alerts and dashboards. Yet they may still discover certain issues only after their users do.

01
Metrics
Continuously measure system health and performance.
02
Logs
Understand what is happening within applications and infrastructure.
03
Alerts
Quickly identify anomalies and incidents.
04
Dashboards
Centralize the indicators needed for effective monitoring.
The paradox of modern monitoring

We have never had so much data about our systems, yet data alone does not guarantee a clear understanding of the experience actually delivered to users.

The question is changing
Yesterday
“Is my system working?”
Today
“Is my digital service actually working as intended for the user?”
02 / Experience

When technical monitoring is no longer enough

For a long time, monitoring focused primarily on the health of infrastructure and technical components: CPU, memory, servers, network availability, response times and error rates.

These indicators remain essential, but they tell only part of the story.

CPU
Memory
Servers
Network
Response time
Error rate
Technical view

Everything appears operational.

A server may be available. An API may respond. An application may show an operational status.

User view

Yet the service may still not work.

A user may be unable to complete an order, sign in or access an essential feature.

Consider an e-commerce journey:

Example user journey
01
Login
02
Product search
03
Add to cart
04
Payment
05
Confirmation

Each component may appear to work when monitored individually. But if a dependency slows down, a button stops responding or a payment step fails, the user journey is compromised.

The key point

Component availability does not guarantee that the service is truly available to the user.

An end-to-end view

An effective monitoring strategy must go beyond basic technical monitoring.

It must consider the digital service as a whole and the experience actually delivered to the user.

03 / Limitations

The main limitations of traditional monitoring

01
Fragmentation

A fragmented view of the information system

Infrastructure, network, applications, cloud, APIs, third-party services, mobile… each environment may have its own monitoring tool.

The problem is no longer necessarily a lack of data, but their fragmentation.

Infrastructure
Network
Applications
Cloud
API
Third-party services
Mobile

When an incident occurs, teams may have to reconstruct the chain of events across several tools to understand what actually happened.

This fragmentation makes it more difficult to identify the root cause and can increase resolution time.

→
The more fragmented the data, the more complex it can become to reconstruct an incident from end to end.
02
Alerts

Too many alerts, not enough context

Monitoring relies heavily on alerts. But when thresholds are poorly defined or monitoring tools multiply, teams can quickly face alert overload.

Yet not all alerts have the same level of importance.

Not all alerts have the same impact
Minor degradation on a secondary page
Temporary increase in response time
Users unable to log in
Payment journey failure

A degradation on a secondary page does not have the same impact as being unable to log in or complete a payment.

The goal should therefore not be to generate more and more alerts, but to produce relevant, contextualized and actionable alerts.
03
Real experience

Technical metrics disconnected from the real user experience

A satisfactory average response time can hide very different situations.

Most users may enjoy a smooth experience while a specific region, browser, device or user segment experiences significant slowdowns.

A global average can hide very different experiences
Global view
Average indicators may suggest that the service is working properly.
Individual experiences
Yet a specific region, browser, device or user segment may still experience slowdowns.

To understand the true quality of a service, it is necessary to observe what happens under real-world usage conditions.

This is where Real User Monitoring (RUM) and Synthetic Monitoring become complementary.

Two complementary approaches
01

Real User Monitoring

Observe the real user experience

RUM makes it possible to analyze the experience actually lived by users under real-world usage conditions.

02

Synthetic Monitoring

Test proactively

Synthetic Monitoring automatically reproduces critical user journeys to identify issues, sometimes even before users encounter them.

+

RUM shows what users are actually experiencing. Synthetic Monitoring proactively verifies that critical user journeys continue to work.

Towards a more complete view

Truly understanding the performance of a digital service requires connecting technical data with context and the experience actually lived by users.

04 / Experience

From technical monitoring to experience-driven monitoring

This evolution requires a shift in perspective.

Historically, monitoring was primarily organized around the health of systems and technical components. Today, a more mature approach puts the user and their ability to complete an action at the heart of the monitoring strategy.

A shift in perspective
Traditional approach

“Is the server working?”

Experience-driven approach

“Can the user accomplish what they came to do?”

01
E-commerce
Verify that a user can search for a product, add it to their cart and complete the payment.
02
Banking
Ensure that a customer can successfully authenticate, view their accounts or complete a transaction.
03
SaaS platform
Verify that critical features remain available and perform well for users.
A new unit of measurement

The user journey thus becomes an essential measure of digital performance.

05 / Best practices

5 best practices for truly effective monitoring

01
Business priorities

Start with critical user journeys

A monitoring strategy should not start with: “What data can we collect?”

Instead, it should start with a question much more directly related to the organization’s business.

“What data can we collect?”
→ “Which services and user journeys cannot afford to fail?”
Login
Payment
Search
Subscription
Appointment booking
Customer portal

Every organization has user journeys whose availability and performance have a direct impact on its business. These journeys should guide the monitoring strategy and the definition of priority indicators.

02
Holistic view

Combine multiple levels of observation

No monitoring technology can provide a complete view of the system on its own.

01
Infrastructure
Monitor technical resources and components.
02
APM
Analyze application behavior.
03
RUM
Observe the experience of real users.
04
Synthetic Monitoring
Automatically test predefined user journeys.
Infrastructure + APM + RUM + Synthetic Monitoring = une vision plus complète du service numérique

These approaches are not competing: they are complementary.

Combining them makes it possible to move from an accumulation of metrics to a holistic understanding of digital service quality.

03
User perspective

Observe the service from the outside

There is a fundamental difference between knowing that a system is working internally and verifying that it actually works from the outside.

From the inside

The system appears to be working.

Servers, applications and technical components may all show normal indicators.

From the outside

Verify real-world access conditions.

The service is tested from a perspective closer to that of the user.

User
→
Internet
→
Browser / Mobile
→
CDN
→
API
→
Third-party services

Users rely on the Internet, browsers, mobile applications, CDNs, APIs and many third-party services before accessing a service.

Regularly testing applications from different monitoring locations makes it possible to better reflect real-world access conditions and detect issues that would remain invisible with internal monitoring alone.

04
Prioritization

Prioritize alerts based on their impact

Not all anomalies require the same level of urgency.

Alerts should be prioritized across several dimensions to reflect their actual impact on the service.

Prioritization criteria
01
Affected user journey
02
Users impacted
03
Incident duration
04
Business criticality
Monitoring no longer only tells teams what is wrong : il aide les équipes à comprendre what should be addressed first.
05
Automation & AI

Use automation and AI to improve understanding

Artificial intelligence opens up new possibilities for monitoring: anomaly detection, event correlation, trend analysis and diagnostic assistance.

But its purpose should not simply be to generate more data or alerts.

What AI can bring to monitoring
Anomaly detection
Event correlation
Trend analysis
Diagnostic assistance
Its value lies primarily in its ability to turn a large volume of signals into actionable information more quickly.
Artificial intelligence
Reduce noise, analyze signals, identify anomalies and accelerate analysis.
Human expertise
Interpret context, prioritize issues and make appropriate decisions.

AI can help reduce noise and accelerate analysis, while teams retain an essential role in interpreting context, prioritizing issues and making decisions.

Truly effective monitoring

Move from monitoring components to understanding the digital service as a whole.

Critical user journeys, infrastructure, applications, real user experience, synthetic testing, business context and artificial intelligence: combining these different levels of observation is what makes it possible to build truly experience-driven monitoring.

06 / Understanding

From visibility to understanding

The next challenge in digital monitoring is probably not to collect ever more data.

It is about making better use of it.

Logs, metrics, traces, user experience data and synthetic scenarios all provide valuable signals. Taken separately, they offer only a partial view.

The different monitoring signals
01
Logs
02
Metrics
03
Traces
04
User experience
05
Synthetic scenarios
From isolated signals to a holistic understanding
01
Collect
Observe the different signals generated by systems and users.
02
Correlate & contextualize
Connect data points and place them in the context of the user journey.
03
Understand
Transform technical signals into a clearer view of the digital service.

When correlated and placed in the context of the user journey, these signals become much more powerful.

This evolution gradually marks the shift from component monitoring to understanding the digital service.

Monitoring is evolving
Monitoring

Monitor components

Individually check the health of servers, applications, networks and infrastructure.

Understanding

Understand the digital service

Connect technical signals with the experience and journey actually delivered to the user.

→
Monitoring & observability
This shift toward greater context, correlation and understanding is also what is gradually bringing monitoring and observability closer together.
07 / Tomorrow

Tomorrow’s monitoring will be user-centric

As digital architectures become more complex, the very definition of availability must evolve.

A service is not truly available simply because its servers are responding.

Rethinking availability
Technical availability

The servers are responding.

Technical components appear to be available and indicators remain operational.

Real availability

The user can accomplish their goal.

The user journey works end to end and under good usage conditions.

A service is truly available when the user can accomplish their goal under good conditions.

The Ekara approach

Put user experience at the heart of digital monitoring.

At Ekara, cette approche consiste notamment à surveiller les services et les parcours critiques dans des conditions représentatives de leur utilisation réelle.

Beyond technology, this evolution is above all methodological.

Modern monitoring should no longer simply answer a technical question. It should make it possible to understand service quality as it is actually experienced.

The key question is evolving
Traditional monitoring

“Is everything working?”

User-centric monitoring

“Is everything working properly for our users?”

From monitoring to performance

Understanding what users actually experience changes the value of monitoring.

This difference transforms monitoring from a simple monitoring tool into a true driver of digital performance.

Previous Post

Leave a Reply

Discover more from Ekara by ip-label

Subscribe now to keep reading and get access to the full archive.

Continue reading