Explaining siamese Networks in few-shot learning for audio data

Fedele, A.; Guidotti, R.; Pedreschi, D.

doi:10.1007/978-3-031-18840-4_36

Machine learning models are not able to generalize correctly when queried on samples belonging to class distributions that were never seen during training. This is a critical issue, since real world applications might need to quickly adapt without the necessity of re-training. To overcome these limitations, few-shot learning frameworks have been proposed and their applicability has been studied widely for computer vision tasks. Siamese Networks learn pairs similarity in form of a metric that can be easily extended on new unseen classes. Unfortunately, the downside of such systems is the lack of explainability. We propose a method to explain the outcomes of Siamese Networks in the context of few-shot learning for audio data. This objective is pursued through a local perturbation-based approach that evaluates segments-weighted-average contributions to the final outcome considering the interplay between different areas of the audio spectrogram. Qualitative and quantitative results demonstrate that our method is able to show common intra-class characteristics and erroneous reliance on silent sections.

Explaining siamese Networks in few-shot learning for audio data

Fedele A.;Guidotti R.;Pedreschi D.

2022

Abstract

Machine learning models are not able to generalize correctly when queried on samples belonging to class distributions that were never seen during training. This is a critical issue, since real world applications might need to quickly adapt without the necessity of re-training. To overcome these limitations, few-shot learning frameworks have been proposed and their applicability has been studied widely for computer vision tasks. Siamese Networks learn pairs similarity in form of a metric that can be easily extended on new unseen classes. Unfortunately, the downside of such systems is the lack of explainability. We propose a method to explain the outcomes of Siamese Networks in the context of few-shot learning for audio data. This objective is pursued through a local perturbation-based approach that evaluates segments-weighted-average contributions to the final outcome considering the interplay between different areas of the audio spectrogram. Qualitative and quantitative results demonstrate that our method is able to show common intra-class characteristics and erroneous reliance on silent sections.

Scheda breve

Scheda completa

Scheda completa (DC)

	Anno
	
				2022
			
	Strutture organizzative
	
				Istituto di Scienza e Tecnologie dell'Informazione "Alessandro Faedo" - ISTI
			
	Lingua/e
	
				Inglese
			
	Titolo del Volume
	
				na
			
	Serie
	
				LECTURE NOTES IN ARTIFICIAL INTELLIGENCE
			
	Titolo del convegno
	
				DS 2022 - 25th International Conference on Discovery Science
			
	Volume
	
				13601
			
	Da pagina
	
				509
			
	A pagina
	
				524
			
	Numero di pagine
	
				16
			
	Codice ISBN
	
				9783031188398
			
	Codice DOI
	
				https://dx.doi.org/10.1007/978-3-031-18840-4_36
			
	URL
	
				https://dl.acm.org/doi/10.1007/978-3-031-18840-4_36
			
	Nome Editore
	
				ACM
			
	Nazione Editore
	
				STATI UNITI D'AMERICA
			
	Referee
	
				Sì, ma tipo non specificato
			
	Indicizzato
	
				Sì
			
	Periodo del Convegno
	
				10-12/10/2022
			
	Luogo del Convegno
	
				Montpellier, France
			
	Rilevanza
	
				Internazionale
			
	Parole chiave
	
				Audio Data
Explainable AI
Siamese Networks
			
	Codice Scopus
	
				2-s2.0-85142726229
			
	Codice Web of Science
	
				WOS:000897761100036
			
	Formato
	
				Elettronico
			
	Numero autori
	
				3
			
	Fulltext
	
				restricted
			
	Tutti gli autori
	
						Fedele, A.; Guidotti, R.; Pedreschi, D.
					
	Tipologia Login Miur
	
				273
			
	Tipologia
	
				info:eu-repo/semantics/conferenceObject
			
	Tipologia
	
				04 Contributo in convegno::04.01 Contributo in Atti di convegno
			
	Appare nelle tipologie:
	
				04.01 Contributo in Atti di convegno

File in questo prodotto:

File	Dimensione	Formato
Fedele-Guidotti-Pedreschi_Springer 2022.pdf solo utenti autorizzati Descrizione: Explaining Siamese Networks in Few-Shot Learning for Audio Data Tipologia: Versione Editoriale (PDF) Licenza: NON PUBBLICO - Accesso privato/ristretto Dimensione 2.26 MB Formato Adobe PDF Visualizza/Apri Richiedi una copia	2.26 MB	Adobe PDF	Visualizza/Apri Richiedi una copia

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/20.500.14243/457333

Citazioni

ND

10

8

CNR Institutional Research Information System

Explaining siamese Networks in few-shot learning for audio data

Fedele A.;Guidotti R.;Pedreschi D.

2022

Abstract

Scheda breve

Scheda completa

Scheda completa (DC)

Citazioni

social impact

CNR Institutional Research Information System

Explaining siamese Networks in few-shot learning for audio data

Fedele A.;Guidotti R.;Pedreschi D.

2022

Abstract

Scheda breve Scheda completa Scheda completa (DC)

Informazioni

Citazioni

social impact

Conferma cancellazione

Scheda breve

Scheda completa

Scheda completa (DC)