SDAILGJun 17

PrefSQA: Pairwise Preference Prediction for Speech Quality Assessment and the Critical Role of High Quality Datasets

arXiv:2606.195976.4
Predicted impact top 68% in SD · last 90 daysOriginality Incremental advance
AI Analysis

For speech quality assessment researchers, this work addresses label noise in MOS by proposing a preference-based method, though improvements are incremental on standard MOS data.

PrefSQA improves speech quality assessment by predicting pairwise preferences instead of MOS, reducing labeling noise. It achieves clear improvements on high-quality preference datasets but only small gains on MOS-derived data.

Mean opinion scores (MOS) are widely used for speech quality assessment, yet scalar labels are sensitive to rater variability and listening test differences. This introduces labeling noise, which limits the reliability of MOS prediction. Preference prediction reduces this variability as listeners compare signals directly, producing cleaner labels. We study MOS-free preference prediction and propose PrefSQA, which incorporates uncertainty-aware logits, an impairment attention head, and a module based on non-matching-reference comparisons. We use and refine five datasets, including MOS-derived and low-noise simulated sets with matching and non-matching content, experiment with human preference sets, and test on unseen data. Experiments show small improvements on MOS-derived data, while other sets reveal clear improvement over the baselines, highlighting the value of high-quality preference data and demonstrating the effectiveness of the proposed method.

Foundations

The foundational work for this paper's niche, ranked by how specifically the neighbourhood builds on it — not by global fame.

Your Notes