from    
to    
search  

 


天文系 Colloquium: Probing Planet Formation with the Most Extreme Cases
清华大学材料科学与工程研究院《材料科学论坛》:Magnetocaloric effect: from the...
Data-driven Discovery of Optimization-based Decision-making Models
Toward the Efficient Operation of an Electrified Chemical Industry
报告题目:
TOP-K Phi Correlation Computation
 报告人:
熊辉(Hui Xiong)
Management Science and Information Systems Department
Rutgers - the State University of New Jersey, USA. 
(More on: http://cimic.rutgers.edu/~hui/ )
报告时间:
2007-05-29 10:30
报告地点:
清华大学经济管理学院舜德楼302室
主办单位:
清华大学经济管理学院现代管理研究中心
  简介:
Abstract:
 
  The problem of association pattern mining is to develop techniques for finding groups of highly-correlated objects from massive data. This problem is important for various application domains, such as homeland security, market basket study, and biomedical data analysis. A large body of association mining work was motivated by the difficulty of efficiently identifying highly correlated objects using traditional statistical correlation measures. This has led to the use of alternative interest measures, such as support and confidence, despite the lack of a precise relationship between these new interest measures and statistical correlation measures. However, this approach tends to generate too many spurious patterns involving objects which are poorly correlated. In this talk, we provide a precise relationship between Phi correlation coefficient and the support measure. We also identify a 2-D monotone property of an upper bound of Phi correlation coefficient and develop an efficient algorithm, called TOP-COP to exploit this property to effectively prune many pairs even without computing their correlation coefficients. Our experimental results show that TOP-COP can be an order of magnitude faster than alternative approaches for mining the top-k strongly correlated pairs. Finally, we show that the performance of the TOP-COP algorithm is tightly related to the degree of data dispersion. Indeed, the higher the degree of data dispersion, the larger the computational savings achieved by the TOP-COP algorithm.
 
Brief Biography:
 
  Hui Xiong is currently an Assistant Professor in the Management Science and Information Systems Department at Rutgers - the State University of New Jersey, USA. He received the Ph.D. degree in Computer Science from the University of Minnesota, USA, in 2005, the B.E. degree in Automation from the University of Science and Technology of China, and the M.S. degree in Computer Science from the National University of Singapore. His research interests include data mining, spatial databases, statistical computing, and Geographic Information Systems (GIS) with applications in business, database security, self-managing systems, and bio-medical informatics. He has published over 30 papers in the refereed journals and conference proceedings, such as IEEE Transactions on Knowledge and Data Engineering, VLDB Journal, Data Mining and Knowledge Discovery Journal, ACM SIGKDD, SIAM SDM, IEEE ICDM, ACM CIKM, ACM GIS, and PSB. He is the co-editor of the book entitled "Clustering and Information Retrieval", the author of a monograph entitled "Hyperclique pattern discovery: Algorithms and applications", and the co-Editor-in-Chief of Encyclopedia of Geographical Information Science. He has also served on the organization committees and the program committees of a number of conferences, such as ACM SIGKDD, SIAM SDM, IEEE ICDM, IEEE ICTAI, ACM CIKM, and IEEE ICDE. Dr. Xiong is a member of the IEEE Computer Society, the ACM, and the Sigma Xi.

 

 

 


今日相关信息
Why Neural Nets are Still Our Best Ho...
Explorations in Organic Synthesis: Fr...
新经典:西方现代艺术的魅力—剑桥大学访问...
 
同类别相关信息
技术的本质及其进化机制
互联网是中国经济转型升级的历史性机遇(...
清华五道口全球金融论坛—改革:发展新征...
自西方而返:公共关系研究的挑战与前瞻
探讨危机传播的理论取径
学术活动