The definitive guide to the state of the art of multimedia information extraction
Government analysts, think tank researchers, managers at top websites-basically everyone-is searching for the best ways to access and exploit the vast amounts of multimedia data made available over large networks every day. Written by an international team of experts, Multimedia Information Extraction provides a detailed road map to how that's done.
The first book to address not only multimedia retrieval but also information extraction from and across media, it offers diverse perspectives on how this emerging technology can help meet the growing demand in industry and government for stock media access, media preservation, broadcast news retrieval, identity management, video surveillance, and more.
Including a Foreword by Professor Alan Smeaton, founding coordinator of the international TRECVid, Multimedia Information Extraction covers:
* The fundamental issues in processing and multimedia source extraction
* The history and state of the art of multimedia information extraction
* Image and video extraction, with tools ranging from visual feature localization to social redundancy
* Affect extraction in audio and imagery, from paralinguistic information retrieval to affect-based indexing
* Multimedia annotation and authoring
An inspiring, much-needed resource for researchers and developers in government, industry, and academia, this book also offers guidance on using the material in the core curriculum of ACM SIGCHI, ACM/IEEE Computer Science, and ACM/IEEE Information Technology.
The advent of increasingly large consumer collections of audio (e.g., iTunes), imagery (e.g., Flickr), and video (e.g., YouTube) is driving a need not only for multimedia retrieval but also information extraction from and across media. Furthermore, industrial and government collections fuel requirements for stock media access, media preservation, broadcast news retrieval, identity management, and video surveillance. While significant advances have been made in language processing for information extraction from unstructured multilingual text and extraction of objects from imagery and video, these advances have been explored in largely independent research communities who have addressed extracting information from single media (e.g., text, imagery, audio). And yet users need to search for concepts across individual media, author multimedia artifacts, and perform multimedia analysis in many domains.
This collection is intended to serve several purposes, including reporting the current state of the art, stimulating novel research, and encouraging cross-fertilization of distinct research disciplines. The collection and integration of a common base of intellectual material will provide an invaluable service from which to teach a future generation of cross disciplinary media scientists and engineers.
Mark T. Maybury
Electrical & Electronics Engineering Elektrotechnik u. Elektronik Intelligent Systems & Agents Intelligente Systeme u. Agenten Multimedia Mustererkennung Signal Processing Signalverarbeitung