Automatic recognition of environmental sound events using all-pole group delay features

    Tutkimustuotos: Conference contributionScientificvertaisarvioitu

    9 Sitaatiot (Scopus)

    Abstrakti

    A feature based on the group delay function from all-pole models (APGD) is proposed for environmental sound event recognition. The commonly used spectral features take into account merely the magnitude information, whereas the phase is overlooked due to the complications related to its interpretation. Additional information concealed in the phase is hypothesised to be beneficial for sound event recognition. The APGD is an approach to inferring phase information, which has shown applicability for analysis of speech and music signals and is now studied in environmental audio. The evaluation is performed within a multi-label deep neural network (DNN) framework on a diverse real-life dataset of environmental sounds. It shows performance improvement compared to the baseline log mel-band energy case. In combination with the magnitude-based features, APGD demonstrates further improvement.
    AlkuperäiskieliEnglanti
    OtsikkoProceedings of the 2015 European Signal Processing Conference (EUSIPCO)
    KustantajaIEEE
    Sivut734-738
    Sivumäärä5
    ISBN (painettu)978-0-9928626-4-0
    TilaJulkaistu - elok. 2015
    OKM-julkaisutyyppiA4 Artikkeli konferenssijulkaisussa
    TapahtumaEUROPEAN SIGNAL PROCESSING CONFERENCE -
    Kesto: 1 tammik. 1900 → …

    Conference

    ConferenceEUROPEAN SIGNAL PROCESSING CONFERENCE
    Ajanjakso1/01/00 → …

    Julkaisufoorumi-taso

    • Jufo-taso 1

    Sormenjälki

    Sukella tutkimusaiheisiin 'Automatic recognition of environmental sound events using all-pole group delay features'. Ne muodostavat yhdessä ainutlaatuisen sormenjäljen.

    Siteeraa tätä