NetRadar: Monitoring the Datacenter Network

Wednesday, June 12, 2019 - 5:00 pm5:30 pm

Yun Chen, Baidu

Abstract: 

The quality of a datacenter network directly affects the stability and performance of production systems. A network outage can happen on various devices and influence different scope. The SRE needs to quickly identify the scope of the outage to determine the remediation actions.

At Baidu, we built NetRadar, a datacenter monitoring system, for this purpose. NetRadar applies multi-dimensional analysis algorithm on various kinds of network quality data to identify the scope of the network outages. In this talk, we will introduce our consideration in designing the monitoring system, as well as the analysis algorithm.

Yun Chen, Baidu

Yun Chen is a Senior Software Engineer at Baidu. Yun's work focuses on datacenter network monitoring and operation data analysis, including time-series anomaly detection and service diagnosis.

Open Access Media

USENIX is committed to Open Access to the research presented at our events. Papers and proceedings are freely available to everyone once the event begins. Any video, audio, and/or slides that are posted after the event are also free and open to everyone. Support USENIX and our commitment to Open Access.

BibTeX
@conference {233219,
author = {Yun Chen},
title = {{NetRadar}: Monitoring the Datacenter Network},
year = {2019},
address = {Singapore},
publisher = {USENIX Association},
month = jun
}

Presentation Video