← Glossary

DNSMOS (Deep Noise Suppression Mean Opinion Score)

A model that predicts how human listeners would rate the quality of a speech recording, on a scale of 1 to 5.

DNSMOS is a neural network trained on human listening-test data to predict mean opinion scores for speech recordings. It was developed for the Deep Noise Suppression Challenge and reports an overall score alongside sub-scores for speech quality and for background noise. A personalized variant judges quality relative to a target speaker.

Because it needs no clean reference signal, DNSMOS can be run on any recording, including material for which no ground truth exists. That same property is its main limitation: it rates what it hears, so a track can score well while containing the wrong speaker's words. It is normally reported alongside reference-based measures rather than on its own.

Where this comes up at AudioShake
No items found.