Published record

videssonus/mydataset

A working corpus of urban soundscape excerpts collected for surrogate background sound generation. Each clip carries a location context, a coarse source label, and perceptual ratings used to condition text-to-audio retrieval and mixing experiments.

Open on Hugging Face

Au-Yeung, H. H. (2026). videssonus/mydataset [Data set]. Hugging Face.

Host
HuggingFace
Version
v0.1
Records
1,240 clips
Size
3.8 GB
Format
WAV 48 kHz / Parquet
Modality
Audio + tabular annotations
Licence
CC BY 4.0
Updated
Aug 2026
datasetslibrosapandasChromaDB
10 sample rows
SC-0011CourtyardBirds1048.24.3
SC-0042Street canyonTraffic1068.71.9
SC-0067SquareVoices1061.43.1
SC-0104CourtyardVoices10553.6
SC-0138ParkBirds1044.14.6
SC-0159Street canyonConstruction1074.31.4
SC-0201SquareTraffic1066.22.2
SC-0246ParkWater1052.84.4
SC-0288CourtyardTraffic1059.92.6
SC-0311SquareMusic1063.53.9

Click a column header to sort the sample

Sample distribution

Clips in sample · computed from the bundled sample

(01) Collection

Binaural and ambisonic recordings captured across Stuttgart courtyards, squares and street canyons at varied times of day.

(02) Annotation

Each clip is tagged with a dominant source class, a context string, and listener ratings for pleasantness and eventfulness.

(03) Intended use

Conditioning and evaluation of surrogate background sound generators; not a calibrated noise-level reference.

Adjacent records