Lahja SA Huba V1 Lahja SA Huba V1 is a production ready Arabic Text to Speech (TTS) model developed by WasmAI and optimized for the Saudi dialect. Built on the VITS architecture, the model generates natural, expressive, and human like speech while preserving Saudi pronunciation, rhythm, and linguistic characteristics. Designed for both research and enterprise applications, Lahja SA Huba V1 enables developers to integrate realistic Arabic speech synthesis into conversational AI, voice assistants, accessibility technologies, educational platforms, robotics, customer service systems, and other intelligent applications. ✨ Features 🇸🇦 Natural Saudi Arabic speech synthesis 🎙️ Human like pronunciation and expressive intonation ⚡ Fast inference with low latency 🧠 End to end VITS architecture 🤖 Production ready deployment 🤗 Fully compatible with Hugging Face Transformers 🏗️ Architecture The model is based on Variational Inference with Adversarial Learning for End to End Text to Speech (VITS) . Its architecture combines: Transformer Text Encoder Variational Autoencoder (VAE) Flow based Prior Network Stochastic Duration Predictor HiFi GAN Neural Decoder Base model: wasmdashai/vits ar s…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy