wav2vec vm finetune This model is a fine tuned version of facebook/wav2vec2 xls r 300m for voicemail detection . It is trained on a dataset of call recordings to distinguish between voicemail greetings and live human responses . Model description This model builds on wav2vec2 xls r 300m , a self supervised speech model trained on large scale multilingual data. We fine tuned it on the first two seconds of a call. Intended uses & limitations Automated voicemail detection in AI powered call assistants. Filtering voicemail responses in customer service and sales call automation. Only trianed on the English language. Assumes the voicemail track is isolated and contains no audio from the caller. Designed for the first two seconds of audio when calling a voicemail. Training and evaluation data The model was trained on a proprietary dataset of call recordings, labeled as: Live human responses Voicemail greetings The dataset includes diverse voicemail recordings across multiple types to improve generalization. Evaluation metrics The model achieved: 98% accuracy on voicemail detection. Training procedure Training hyperparameters The following hyperparameters were used during training: learni…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy