Found an issue? Give us feedback

https://doi.org/10.5...arrow_drop_down

https://doi.org/10.5281/zenodo...

Other ORP type . 2021

License: CC BY

Data sources: Sygma

Select content type to embed

All Research products

arrow_drop_down

<script type="text/javascript">
<!--
document.write('<div id="oa_widget"></div>');
document.write('<script type="text/javascript" src="https://www.openaire.eu/index.php?option=com_openaire&view=widget&format=raw&projectId=undefined&type=result"></script>');
-->
</script>

COPY SCRIPT

For further information contact us at helpdesk@openaire.eu

Fairness and underspecification in acoustic scene classification: The case for disaggregated evaluations

Name: Fairness and underspecification in acoustic scene classification: The case for disaggregated evaluations
Keywords: transparency, evaluation, fairness, acoustic scene classification, ethics

appsOther research productkeyboard_double_arrow_right Other ORP type 06 Oct 2021 English Publisher:ZenodoFunded by:EC | EVERYSOUND, EC | MARVEL

Authors: Triantafyllopoulos, Andreas; Milling, Manuel; Drossos, Konstantinos; Schuller, Björn W.;

Fairness and underspecification in acoustic scene classification: The case for disaggregated evaluations

- Summary
- Subjects
- Related research
  (3)
- Metrics

Abstract

Underspecification and fairness in machine learning (ML) applications have recently become two prominent issues in the ML community. Acoustic scene classification (ASC) applications have so far remained unaffected by this discussion, but are now becoming increasingly used in real-world systems where fairness and reliability are critical aspects. In this work, we argue for the need of a more holistic evaluation process for ASC models through disaggregated evaluations. This entails taking into account performance differences across several factors, such as city, location, and recording device. Although these factors play a well-understood role in the performance of ASC models, most works report single evaluation metrics taking into account all different strata of a particular dataset. We argue that metrics computed on specific sub-populations of the underlying data contain valuable information about the expected real-world behaviour of proposed systems, and their reporting could improve the transparency and trustability of such systems. We demonstrate the effectiveness of the proposed evaluation process in uncovering underspecification and fairness problems exhibited by several standard ML architectures when trained on two widely-used ASC datasets. Our evaluation shows that all examined architectures exhibit large biases across all factors taken into consideration, and in particular with respect to the recording location. Additionally, different architectures exhibit different biases even though they are trained with the same experimental configurations.

Related Organizations

Keywords

transparency, evaluation, fairness, acoustic scene classification, ethics

Filter by relation

All relations

arrow_drop_down

3 Research products, page 1 of 1

Fairness and underspecification in acoustic scene classification: The case for disaggregated evaluations
2022IsVersionOf
TUT Urban Acoustic Scenes 2018, Development dataset
2018IsSupplementedBy
TUT Urban Acoustic Scenes 2018 Mobile, Development dataset
2018IsSupplementedBy

Impact byBIP!

	citations This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	0
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Average
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Average
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Average

Found an issue? Give us feedback

Average

Funded by

EC| EVERYSOUND, EC| MARVEL

Related to Research communities

Energy Planning

Knowmad Institut

UArctic

Fairness and underspecification in acoustic scene classification: The case for disaggregated evaluations

Fairness and underspecification in acoustic scene classification: The case for disaggregated evaluations

3 Research products, page 1 of 1

Fairness and underspecification in acoustic scene classification: The case for disaggregated evaluations

TUT Urban Acoustic Scenes 2018, Development dataset

TUT Urban Acoustic Scenes 2018 Mobile, Development dataset