If I had to do it synthetically, take single subjects with a single sound and combine them together. Then train a model to separate them again.
If I had to do it synthetically, take single subjects with a single sound and combine them together. Then train a model to separate them again.