arXiv AI By Simon Dahl Jepsen, Mads Gr{\ae}sb{\o}ll Christensen, Jesper Rindom Jensen

A Study of the Scale Invariant Signal to Distortion Ratio in Speech Separation with Noisy References

Read the original on arXiv AI →

arXiv:2508. 14623v2 Announce Type: replace-cross Abstract: This paper examines the implications of using the Scale-Invariant Signal-to-Distortion Ratio (SI-SDR) as both evaluation and training objective in supervised speech separation, when the training references contain noise, as is the case with the de facto benchmark WSJ0-2Mix.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.