Experimental Results

Zhang, Tong; Kuo, C.-C. Jay

doi:10.1007/978-1-4757-3339-6_7

Tong Zhang³ &
C.-C. Jay Kuo⁴

Part of the book series: The Springer International Series in Engineering and Computer Science ((SECS,volume 606))

190 Accesses

Abstract

We have built a generic audio database to be used as the testbed of the proposed algorithms, which consists of the following contents: 1000 clips of environmental audio including the sounds of applause, animal, footstep, raining, explosion, knocking, vehicles and so on; 100 pieces of classical music played with 10 kinds of instruments, 100 other music pieces of different styles (classic, jazz, blues, light music, Chinese and Indian folk music, etc.); 50 clips of songs sung by male, female, or children, with or without musical instrument accompaniment; 200 speech pieces in different languages (English, German, French, Spanish, Japanese, Chinese, etc.) and with different levels of noise; 50 clips of speech with the music background; 40 clips of environmental sound with the music background; and 20 samples of silence segment with different types of low-volume noise (clicks, brown noise, pink noise and white noise). These short pieces of sound clips (with duration from several seconds to more than one minute) are used to test the audio classification performances. We also collected dozens of longer audio clips recorded from movies or video programs. These pieces last from several minutes to half an hour, and contain various types of audio. They are used to test the performances for audiovisual data segmentation and indexing.

This is a preview of subscription content, log in via an institution to check access.

Access this chapter

Log in via an institution

Chapter: USD 29.95; Price excludes VAT (USA)

eBook: USD 84.99; Price excludes VAT (USA)

Softcover Book: USD 109.99; Price excludes VAT (USA)

Hardcover Book: USD 109.99; Price excludes VAT (USA)

Tax calculation will be finalised at checkout

Purchases are for personal use only

Institutional subscriptions

Preview

Unable to display preview. Download preview PDF.

Author information

Authors and Affiliations

Integrated Media Systems Center, University of Southern California, 90089-2564, Los Angeles, CA, USA
Tong Zhang
Department of Electrical Engineering — Systems, University of Southern California, 90089-2564, Los Angeles, CA, USA
C.-C. Jay Kuo

Authors

Tong Zhang
View author publications
You can also search for this author in PubMed Google Scholar
C.-C. Jay Kuo
View author publications
You can also search for this author in PubMed Google Scholar

Rights and permissions

Reprints and permissions

Copyright information

About this chapter

Cite this chapter

Zhang, T., Kuo, CC.J. (2001). Experimental Results. In: Content-Based Audio Classification and Retrieval for Audiovisual Data Parsing. The Springer International Series in Engineering and Computer Science, vol 606. Springer, Boston, MA. https://doi.org/10.1007/978-1-4757-3339-6_7

Download citation

DOI: https://doi.org/10.1007/978-1-4757-3339-6_7
Publisher Name: Springer, Boston, MA
Print ISBN: 978-1-4419-4878-6
Online ISBN: 978-1-4757-3339-6
eBook Packages: Springer Book Archive

Publish with us

Policies and ethics