Hi,
I cloned the repisotory and performed all the preparation steps using a huggingface hosted Audioset (~2m Samples):
https://huggingface.co/datasets/confit/audioset-full
When performing a training I am only getting ESC-50 results of roughly 60%! Why is that and has anyone else managed to replicate the performance from the paper?
Are the default pipeline settings the same as the paper?
Thanks in advance :)
Hi,
I cloned the repisotory and performed all the preparation steps using a huggingface hosted Audioset (~2m Samples):
https://huggingface.co/datasets/confit/audioset-full
When performing a training I am only getting ESC-50 results of roughly 60%! Why is that and has anyone else managed to replicate the performance from the paper?
Are the default pipeline settings the same as the paper?
Thanks in advance :)