Improving Access to Digital Video Archives through Informedia Technology
Journal Article, Journal of the Audio Engineering Society, Vol. 49, No. 7, pp. 595 - 600, July, 2001
Abstract
Informedia research at Carnegie Mellon University combines speech recognition, image processing, and natural language processing to automatically index a digital video library. This engineering report focuses on the contribution of speech analysis for transcript generation and alignment, and the use of these features in library interface development. By deepening the automated analysis, such as using named entity extraction to identify people and place names in the audio transcript, better summaries and visualizations can be produced to navigate through video libraries holding thousands of hours of material.
BibTeX
@article{Christel-2001-16827,author = {Michael Christel and Alex Hauptmann and Howard Wactlar},
title = {Improving Access to Digital Video Archives through Informedia Technology},
journal = {Journal of the Audio Engineering Society},
year = {2001},
month = {July},
volume = {49},
number = {7},
pages = {595 - 600},
}
Copyright notice: This material is presented to ensure timely dissemination of scholarly and technical work. Copyright and all rights therein are retained by authors or by other copyright holders. All persons copying this information are expected to adhere to the terms and constraints invoked by each author's copyright. These works may not be reposted without the explicit permission of the copyright holder.