from    
to    
search  

 


全球变化科学紫荆论坛第440期:Tropical forest growth and mortality under globa...
清华大学材料科学与工程研究院《材料科学论坛》:弛豫铁电相和普通铁电相之间的“模糊...
Pair density wave state and higher charge superconductivity in lowdimensional...
“清芬”学术论坛-基于CRISPR生物探针的基因突变成像分析
报告题目:
TOP-K Phi Correlation Computation
 报告人:
熊辉(Hui Xiong)
Management Science and Information Systems Department
Rutgers - the State University of New Jersey, USA. 
(More on: http://cimic.rutgers.edu/~hui/ )
报告时间:
2007-05-29 10:30
报告地点:
清华大学经济管理学院舜德楼302室
主办单位:
清华大学经济管理学院现代管理研究中心
  简介:
Abstract:
 
  The problem of association pattern mining is to develop techniques for finding groups of highly-correlated objects from massive data. This problem is important for various application domains, such as homeland security, market basket study, and biomedical data analysis. A large body of association mining work was motivated by the difficulty of efficiently identifying highly correlated objects using traditional statistical correlation measures. This has led to the use of alternative interest measures, such as support and confidence, despite the lack of a precise relationship between these new interest measures and statistical correlation measures. However, this approach tends to generate too many spurious patterns involving objects which are poorly correlated. In this talk, we provide a precise relationship between Phi correlation coefficient and the support measure. We also identify a 2-D monotone property of an upper bound of Phi correlation coefficient and develop an efficient algorithm, called TOP-COP to exploit this property to effectively prune many pairs even without computing their correlation coefficients. Our experimental results show that TOP-COP can be an order of magnitude faster than alternative approaches for mining the top-k strongly correlated pairs. Finally, we show that the performance of the TOP-COP algorithm is tightly related to the degree of data dispersion. Indeed, the higher the degree of data dispersion, the larger the computational savings achieved by the TOP-COP algorithm.
 
Brief Biography:
 
  Hui Xiong is currently an Assistant Professor in the Management Science and Information Systems Department at Rutgers - the State University of New Jersey, USA. He received the Ph.D. degree in Computer Science from the University of Minnesota, USA, in 2005, the B.E. degree in Automation from the University of Science and Technology of China, and the M.S. degree in Computer Science from the National University of Singapore. His research interests include data mining, spatial databases, statistical computing, and Geographic Information Systems (GIS) with applications in business, database security, self-managing systems, and bio-medical informatics. He has published over 30 papers in the refereed journals and conference proceedings, such as IEEE Transactions on Knowledge and Data Engineering, VLDB Journal, Data Mining and Knowledge Discovery Journal, ACM SIGKDD, SIAM SDM, IEEE ICDM, ACM CIKM, ACM GIS, and PSB. He is the co-editor of the book entitled "Clustering and Information Retrieval", the author of a monograph entitled "Hyperclique pattern discovery: Algorithms and applications", and the co-Editor-in-Chief of Encyclopedia of Geographical Information Science. He has also served on the organization committees and the program committees of a number of conferences, such as ACM SIGKDD, SIAM SDM, IEEE ICDM, IEEE ICTAI, ACM CIKM, and IEEE ICDE. Dr. Xiong is a member of the IEEE Computer Society, the ACM, and the Sigma Xi.

 

 

 


今日相关信息
Why Neural Nets are Still Our Best Ho...
Explorations in Organic Synthesis: Fr...
新经典:西方现代艺术的魅力—剑桥大学访问...
 
同类别相关信息
如何克服中国公共外交悖论?
用照片为历史存档——一名红色新闻兵的摄...
TNC经济社会讲座:企业发展目标与塑造...
长期滞胀问题——一位日本经济学家的观点
日本的存款保险制度和金融机构破产处置
学术活动