Repository logo

Infoscience

  • English
  • French
Log In
Logo EPFL, École polytechnique fédérale de Lausanne

Infoscience

  • English
  • French
Log In
  1. Home
  2. Academic and Research Output
  3. Conferences, Workshops, Symposiums, and Seminars
  4. SUPPRESSING NOISE DISPARITY IN TRAINING DATA FOR AUTOMATIC PATHOLOGICAL SPEECH DETECTION
 
conference paper

SUPPRESSING NOISE DISPARITY IN TRAINING DATA FOR AUTOMATIC PATHOLOGICAL SPEECH DETECTION

Amiri, Mahdi  
•
Kodrasi, Ina
January 1, 2024
2024 18Th International Workshop On Acoustic Signal Enhancement, Iwaenc 2024
18th International Workshop on Acoustic Signal Enhancement (IWAENC)

Although automatic pathological speech detection approaches show promising results when clean recordings are available, they are vulnerable to additive noise. Recently it has been shown that databases commonly used to develop and evaluate such approaches are noisy, with the noise characteristics between healthy and pathological recordings being different. Consequently, automatic approaches trained on these databases often learn to discriminate noise rather than speech pathology. This paper introduces a method to mitigate this noise disparity in training data. Using noise estimates from recordings from one group of speakers to augment recordings from the other group, the noise characteristics become consistent across all recordings. Experimental results demonstrate the efficacy of this approach in mitigating noise disparity in training data, thereby enabling automatic pathological speech detection to focus on pathology-discriminant cues rather than noise-discriminant ones.

  • Details
  • Metrics
Type
conference paper
DOI
10.1109/IWAENC61483.2024.10694597
Web of Science ID

WOS:001337653100023

Author(s)
Amiri, Mahdi  

École Polytechnique Fédérale de Lausanne

Kodrasi, Ina

Idiap Res Inst

Date Issued

2024-01-01

Publisher

IEEE

Publisher place

New York

Published in
2024 18Th International Workshop On Acoustic Signal Enhancement, Iwaenc 2024
ISBN of the book

979-8-3503-6186-5

979-8-3503-6185-8

Series title/Series vol.

International Workshop on Acoustic Signal Enhancement

ISSN (of the series)

2639-4316

Start page

110

End page

114

Subjects

pathological speech detection

•

noise disparity

•

data augmentation

•

TORGO

•

UA-Speech

Editorial or Peer reviewed

REVIEWED

Written at

EPFL

EPFL units
LIDIAP  
Event nameEvent acronymEvent placeEvent date
18th International Workshop on Acoustic Signal Enhancement (IWAENC)

Aalborg, DENMARK

2024-09-09 - 2024-09-12

FunderFunding(s)Grant NumberGrant URL

Swiss National Science Foundation (SNSF)

CRSII5 202228

Available on Infoscience
January 31, 2025
Use this identifier to reference this record
https://infoscience.epfl.ch/handle/20.500.14299/246165
Logo EPFL, École polytechnique fédérale de Lausanne
  • Contact
  • infoscience@epfl.ch

  • Follow us on Facebook
  • Follow us on Instagram
  • Follow us on LinkedIn
  • Follow us on X
  • Follow us on Youtube
AccessibilityLegal noticePrivacy policyCookie settingsEnd User AgreementGet helpFeedback

Infoscience is a service managed and provided by the Library and IT Services of EPFL. © EPFL, tous droits réservés