Utilize este identificador para referenciar este registo: https://hdl.handle.net/1822/89936

Registo completo
Campo DCValorIdioma
dc.contributor.authorDurães, Dalilapor
dc.contributor.authorVeloso, Brunopor
dc.contributor.authorNovais, Paulopor
dc.date.accessioned2024-03-25T10:57:35Z-
dc.date.available2024-03-25T10:57:35Z-
dc.date.issued2023-
dc.identifier.issn1989-1660-
dc.identifier.urihttps://hdl.handle.net/1822/89936-
dc.description.abstractHuman nature is inherently intertwined with violence, impacting the lives of numerous individuals. Various forms of violence pervade our society, with physical violence being the most prevalent in our daily lives. The study of human actions has gained significant attention in recent years, with audio (captured by microphones) and video (captured by cameras) being the primary means to record instances of violence. While video requires substantial processing capacity and hardware-software performance, audio presents itself as a viable alternative, offering several advantages beyond these technical considerations. Therefore, it is crucial to represent audio data in a manner conducive to accurate classification. In the context of violence in a car, specific datasets dedicated to this domain are not readily available. As a result, we had to create a custom dataset tailored to this particular scenario. The purpose of curating this dataset was to assess whether it could enhance the detection of violence in car-related situations. Due to the imbalanced nature of the dataset, data augmentation techniques were implemented. Existing literature reveals that Deep Learning (DL) algorithms can effectively classify audio, with a commonly used approach involving the conversion of audio into a mel spectrogram image. Based on the results obtained for that dataset, the EfficientNetB1 neural network demonstrated the highest accuracy (95.06%) in detecting violence in audios, closely followed by EfficientNetB0 (94.19%). Conversely, MobileNetV2 proved to be less capable in classifying instances of violence.por
dc.description.sponsorshipFCT - Fundação para a Ciência e a Tecnologia(UIDB/00319/2020)por
dc.language.isoengpor
dc.publisherUniversidad Internacional de La Rioja (UNIR)por
dc.relationinfo:eu-repo/grantAgreement/FCT/6817 - DCRRNI ID/UIDB%2F00319%2F2020/PTpor
dc.rightsopenAccesspor
dc.subjectAudiopor
dc.subjectDeep learningpor
dc.subjectHuman action recognitionpor
dc.subjectMachine learningpor
dc.subjectTransfer learningpor
dc.subjectViolence detection in a carpor
dc.titleViolence detection in audio: evaluating the effectiveness of deep learning models and data augmentationpor
dc.typearticlepor
dc.peerreviewedyespor
oaire.citationStartPage72por
oaire.citationEndPage84por
oaire.citationIssue3por
oaire.citationVolume8por
dc.date.updated2024-03-15T12:01:22Z-
dc.identifier.doi10.9781/ijimai.2023.08.007por
sdum.export.identifier13512-
sdum.journalInternational Journal of Interactive Multimedia and Artificial Intelligencepor
Aparece nas coleções:CAlg - Artigos em revistas internacionais / Papers in international journals

Ficheiros deste registo:
Ficheiro Descrição TamanhoFormato 
IJIMAI.pdf2,09 MBAdobe PDFVer/Abrir

Partilhe no FacebookPartilhe no TwitterPartilhe no DeliciousPartilhe no LinkedInPartilhe no DiggAdicionar ao Google BookmarksPartilhe no MySpacePartilhe no Orkut
Exporte no formato BibTex mendeley Exporte no formato Endnote Adicione ao seu ORCID