The NIGENS General Sound Events Database
This provides a resource for researchers in computational auditory scene analysis, though it is incremental as it builds on existing database efforts.
The authors tackled the lack of suitable data for general sound event detection by releasing the NIGENS database, which includes 714 wav files with isolated sound events of 14 types and 303 general files, all strongly labeled with perceptual on- and offset times.
Computational auditory scene analysis is gaining interest in the last years. Trailing behind the more mature field of speech recognition, it is particularly general sound event detection that is attracting increasing attention. Crucial for training and testing reasonable models is having available enough suitable data -- until recently, general sound event databases were hardly found. We release and present a database with 714 wav files containing isolated high quality sound events of 14 different types, plus 303 `general' wav files of anything else but these 14 types. All sound events are strongly labeled with perceptual on- and offset times, paying attention to omitting in-between silences. The amount of isolated sound events, the quality of annotations, and the particular general sound class distinguish NIGENS from other databases.