Violence Detection in Audio: Evaluating the Effectiveness of Deep Learning Models and Data Augmentation
Autor:
Durães, Dalila
; Veloso, Bruno
; Novais, Paulo
Fecha:
09/2023Palabra clave:
Revista / editorial:
International Journal of Interactive Multimedia and Artificial IntelligenceTipo de Ítem:
articleDirección web:
https://www.ijimai.org/journal/bibcite/reference/3373Resumen:
Human nature is inherently intertwined with violence, impacting the lives of numerous individuals. Various forms of violence pervade our society, with physical violence being the most prevalent in our daily lives. The study of human actions has gained significant attention in recent years, with audio (captured by microphones) and video (captured by cameras) being the primary means to record instances of violence. While video requires substantial processing capacity and hardware-software performance, audio presents itself as a viable alternative, offering several advantages beyond these technical considerations. Therefore, it is crucial to represent audio data in a manner conducive to accurate classification. In the context of violence in a car, specific datasets dedicated to this domain are not readily available. As a result, we had to create a custom dataset tailored to this particular scenario. The purpose of curating this dataset was to assess whether it could enhance the detection of violence in car-related situations. Due to the imbalanced nature of the dataset, data augmentation techniques were implemented. Existing literature reveals that Deep Learning (DL) algorithms can effectively classify audio, with a commonly used approach involving the conversion of audio into a mel spectrogram image. Based on the results obtained for that dataset, the EfficientNetB1 neural network demonstrated the highest accuracy (95.06%) in detecting violence in audios, closely followed by EfficientNetB0 (94.19%). Conversely, MobileNetV2 proved to be less capable in classifying instances of violence.
Ficheros en el ítem
Este ítem aparece en la(s) siguiente(s) colección(es)
Estadísticas de uso
Año |
2012 |
2013 |
2014 |
2015 |
2016 |
2017 |
2018 |
2019 |
2020 |
2021 |
2022 |
2023 |
2024 |
Vistas |
0 |
0 |
0 |
0 |
0 |
0 |
0 |
0 |
0 |
0 |
0 |
1102 |
427 |
Descargas |
0 |
0 |
0 |
0 |
0 |
0 |
0 |
0 |
0 |
0 |
0 |
81 |
308 |
Ítems relacionados
Mostrando ítems relacionados por Título, autor o materia.
-
DEVELOP-FPS: a First Person Shooter Development Tool for Rule-based Scripts
Correia, Bruno; Urbano, Paulo; Moniz, Luís (International Journal of Interactive Multimedia and Artificial Intelligence (IJIMAI), 09/2012)We present DEVELOP-FPS, a software tool specially designed for the development of First Person Shooter (FPS) players controlled by Rule Based Scripts. DEVELOP-FPS may be used by FPS developers to create, debug, maintain ... -
Sphericall: A Human/Artificial Intelligence interaction experience
Gechter, Frack; Ronzani, Bruno; Rioli, Fabien (International Journal of Interactive Multimedia and Artificial Intelligence (IJIMAI), 12/2014)Multi-agent systems are now wide spread in scientific works and in industrial applications. Few applications deal with the Human/Multi-agent system interaction. Multi-agent systems are characterized by individual entities, ... -
The Mediator Role of Feelings of Guilt in the Process of Burnout and Psychosomatic Disorders: A Cross-Cultural Study
Figueiredo-Ferraz, Hugo; Gil-Monte, Pedro R.; Grau-Alberola, Ester ; Ribeiro do Couto, Bruno (Frontiers in psychology, 2021)Burnout was recently declared by WHO as an "occupational phenomenon" in the International Classification of Diseases 11th revision (ICD-11), recognizing burnout as a serious health issue. Earlier studies have shown that ...