from    
to    
search  

 


第478期“工物学术论坛”:X射线探测器领域的行业发展情况和机遇
天文系 Colloquium: Interstellar X-ray Dust Scattering: Current Research andFu...
【图书馆系列讲座】开题与立项前的文献调研概述(理工类)
【图书馆系列讲座】开题与立项前的文献调研概述(社科类)
报告题目:
Document Image Classification and Retrieval
 报告人:
Dr. David Doermann
University of Maryland College Park
报告时间:
2013-09-23 09:30
报告地点:
罗姆楼(ROHM)8-208
主办单位:
电子工程系
  简介:
Abstract
Traditional approaches to document retrieval focus on conversion to electronic text followed by indexing of the text content.  Recently some work in the community has focused on indexing document image content directly.  In this talk, we will overview work at Maryland on Classification and Indexing that scales to millions of documents.  First we present a learning based approach for computing structural similarities among document images for unsupervised exploration in large document collections. The approach is based on multiple levels of content and structure. At a local level, a bag-of-visual words based on SURF features provides an effective way of computing content similarity. The document is then recursively partitioned and a histogram of codewords is computed for each partition. Structural similarity is computed using a random forest classifier trained with these histogram features. We experiment with three diverse datasets of document images varying in size, degree of structural similarity, and types of document images. Second, we present a scalable algorithm for segmentation free content retrieval in document images. The contributions of this paper include the use of the SURF feature for image passage retrieval, a novel indexing algorithm for efficient retrieval of SURF features and a method to filter results using the orientation of local features and geometric constraints. Results demonstrate that logo, signature block and stamp retrieval can be performed with high accurately and efficiently scaled to a large datasets.
 
Dr. Doermann will be available to meet with students.  He will also highlight the University of Maryland graduate program as part of his talk, so students considering graduate school in the US are encouraged to attend.
 
 
Biography:
Dr. David Doermann is a senior research scientist in UMIACS.  He received a B.Sc. degree in Computer Science and Mathematics from Bloomsburg University in 1987, and a M.Sc. degree in 1989 in the Department of Computer Science at the University of Maryland, College Park. He continued his studies in the Computer Vision Laboratory, where he earned a Ph.D. 1993. Since 1993, he has served as co-director of the Laboratory for Language and Media Processing in the University of Maryland's Institute for Advanced Computer Studies and as an adjunct member of the graduate faculty.
His team of researchers focuses on topics related to document image analysis and multimedia information processing. Recent intelligent document image analysis projects include page decomposition, structural analysis and classification, page segmentation, logo recognition, document image compression, duplicate document image detection, image based retrieval, character recognition, generation of synthetic OCR data, and signature verification. In video processing, projects have centered on the segmentation of compressed domain video sequences, structural representation and classification of video, detection of reformatted video sequences and the performance evaluation of automated video analysis algorithms.
In 2002 he received an Honorary Doctorate of Technology Sciences from the University of Oulu for his contributions to digital media processing and document analysis research. He is a founding co-editor of the International Journal on Document Analysis and Recognition, has the General Chair or Co-Chair of over a half dozen international conferences and workshops and was the General Chair of the International Conference on Document Analysis and Recognition (ICDAR)  held in Washington DC in 2013.  He has over 30 journal publications and over 160 refereed conference papers.
今日相关信息
Dynamics and conductivity near quantu...
Spintronics based Computing
清华大学新人文讲座-剑桥大学系列专场: ...
Localization and topology protected q...
 
同类别相关信息
人工智能在各行业应用的思考
博士创“芯”说第五期:面向自动驾驶的高...
学堂班系列讲座:2D materials for ne...
第003期“水木谈芯”集成电路技术与产...
北京信息科学与技术国家研究中心系列交叉...
学术活动