article
This study investigates the effectiveness of Retrieval-based Voice Conversion (RVC) in detecting AI-generated Arabic speech across diverse linguistic contexts. The primary research questions address whether the RVC model can accurately differentiate between real and synthetic speech samples in various languages and its generalization capability across different linguistic contexts. Experimental evaluations show that the model achieves accuracies of 92% using Ensemble Learning (XGBoost and LightGBM), 93% with Meta Learning, and 90% with Neural Networks. Assessments encompass scenarios with real Arabic and English speech not included in training datasets, as well as different speakers across languages. The study contributes insights into the robustness and practical application of RVC for speech authentication.
This page summarises published work. The authoritative version sits with the publisher.
DOI: 10.1109/niles63360.2024.10753151
Is something wrong with this record? Report it or request removal.
Discussion
Have you built on this work, tried to replicate it, or seen it applied in practice? Share what you know. Verified researchers and MARATTO™ domain experts can open a discussion, and any member can reply. Contributions are reviewed before they appear.
No discussion yet. Open the first thread.
New to MARATTO™? Create a free account.