欢迎访问中国科学院大学学报,今天是
简报

一种监控网络程序的信息收集方法

  • 刘伟 ,
  • 罗铁坚 ,
  • 陈肃
展开
  • 中国科学院研究生院,北京 100049

收稿日期: 2009-01-04

  修回日期: 2009-05-08

  网络出版日期: 2009-11-15

Brief Repoot An information collection approach for network monitoring

  • LIU Wei ,
  • LUO Tie-Jian ,
  • CHEN Su
Expand
  • Graduate University of the Chinese Academy of Sciences, Beijing 100049, China

Received date: 2009-01-04

  Revised date: 2009-05-08

  Online published: 2009-11-15

摘要

虽然多数的现有监控工具可以提供关于系统整体性能的宏观信息,但是这样的数据并不能帮助用户理解系统内部的运行状况.为了揭示内部行为,提出了一种基于监控元数据的方法来收集系统执行过程中的内部组件状态和通信.新方法能应用于许多网络通讯协议和应用程序类型,通过信息跟踪机制记录系统内部发生的关键性能事件,并捕获它们之间的时序和因果关系,为系统诊断提供支持.

关键词: 监控; 跟踪; 元数据; 网络

本文引用格式

刘伟 , 罗铁坚 , 陈肃 . 一种监控网络程序的信息收集方法[J]. 中国科学院大学学报, 2009 , 26(6) : 850 -854 . DOI: 10.7523/j.issn.2095-6134.2009.6.018

Abstract

Large network systems need monitoring to detect flaws and bottlenecks. Many existing monitoring tools only provide overall statistics about global performance metrics. In order to expose the systems internal behavior, we propose a metadata based approach to record the actual execution paths. We have designed proper data structure and information propagation mechanisms to trace the applications internal critical performance-related events. This approach can apply to many network protocols and applications,and it helps capture the sequential and causal relations across domains. Thus it opens a way to diagnose large and complex network systems.

Key words: monitor; tracing; metadata; network

参考文献


[1] Massie M L, Chun B N, Culler D E. The ganglia distributed monitoring system: design, implementation, and experience
[M]. Parallel Computing, 2004.

[2] Gerndt M, Wismueller R, Balaton Z, et al. Performance tools for the grid: State of the art and future. APART-2 White Paper on Grid performance analysis
[M]. APART WP3, 2004.

[3] Hussain A, Bartlett G, Pryadkin Y, et al. Experiences with a continuous network tracing infrastructure //The 2005 ACM SIGCOMM workshop on Mining network data. 2005.

[4] Chen M, Kiciman E, Fratkin E, et al. Pinpoint: Problem determination in large, dynamic internet services //Proceedings of the International Conference on Dependable Systems and Networks (DSN02). 2002.

[5] Aguilera M K, Mogul J C, Wiener J L, et al. Performance debugging for distributed systems of black boxes //Proceedings of the Nineteenth ACM Symposium on Operating Systems Principles. 2003.

文章导航

/