Skip to main content

CUSTOMER STORY

Connect procurement systems
Investigate faults faster

ABOUT STAILBUR

About Stailbur

Stailbur provides office supplies, MRO industrial products, corporate gifts and enterprise services, meeting a broad range of business procurement needs. Its production environment spans procurement, supplier and product master-data systems, supported by H5 frontends, business applications, databases, caches and message queues.

Enterprise procurement Industry
60+ Linux hosts in the initial rollout
Procurement & supply chain Core business context

Guance helped us bring infrastructure, application traces and frontend experience into one view for the first time. Problems that once took hours of cross-team investigation can now be traced to their root cause in around fifteen minutes.

Stailbur · Operations team lead

THE CHALLENGE

Complex dependencies demand complete investigation context

As the business grows, a single request can cross multiple applications, databases, caches and message queues. Stailbur needs to determine impact quickly and turn fragmented telemetry into a clear investigation path.

Complex dependencies make fault impact difficult to scope. Teams have to work through several systems, spending about two hours on average just determining the affected area.

Infrastructure, traces, frontend experience, errors and events are spread across tools and teams. Switching systems and manually joining information slows investigation and risks missing key evidence.

Resource anomalies and slow calls require investigation down to individual services, dependencies and spans. Without standard procedures, resolution depends heavily on individual experience.

KEY RESULTS

85%+

Less time to scope cross-system faults

~45%

Reduction in mean time to repair (MTTR)

~60%

Fewer non-actionable alerts

SOLUTION

Stailbur × Guance — Solution

01

Bring telemetry into a shared investigation context

Using DataKit, APM and RUM, Stailbur brings infrastructure, application traces, frontend experience and events into one Guance workspace, connecting services with upstream and downstream dependencies. The initial rollout covered more than 60 Linux hosts and core applications. Shared context supports investigation and recovery, with MTTR falling by approximately 45%.

~45%

Reduction in mean time to repair (MTTR)

02

Follow service dependencies from anomaly to root cause

Service topology, dependency relationships, trace waterfalls, error tracking and performance analysis let the team investigate individual services and spans. In one database incident, a honeycomb view revealed memory usage near 90% and elevated load. Following the topology into related traces took the team from discovery to root cause in about 15 minutes. Average fault scoping fell from about two hours to around 15 minutes.

85%+

Less time to scope cross-system faults

03

Make alert handling and investigation repeatable

Tiered alerting and alert consolidation reduced non-actionable alerts by approximately 60%, helping the team focus on incidents that require intervention. Building on validated POC capabilities, Stailbur will further develop critical-path dashboards, investigation SOPs, and log, network and Kubernetes observability to establish repeatable operations practices.

~60%

Fewer non-actionable alerts

BUSINESS IMPACT

Restore services faster and make room for improvement

01

Faster core procurement requests

P95 response time for core procurement APIs improved by approximately 30%. Slow-call remediation became more targeted, with fewer user-experience complaints.

02

Nearly half as many cross-team handoffs

Operations and development share the same views of services, dependencies and traces, reducing the communication needed to investigate problems.

03

More attention for actionable incidents

Alert noise reduction and faster recovery shorten disruption to critical business paths, replacing repeated information gathering with evidence-based decisions.

Looking ahead

Stailbur and Guance will extend log, network and Kubernetes coverage for full-stack visibility into containerized environments. Standard dashboards, tiered alerts and refined investigation SOPs will support critical business paths as the business continues to grow.

Bring clarity to complex procurement systems

Connect frontend experience, application services and infrastructure in a shared observability and troubleshooting workflow.