NeuCodec ๐ง Click the image above to see NeuCodec in action on Youtube! Created by Neuphonic building faster, smaller, on device voice AI A lightweight neural codec that encodes audio at just 0.8 kbps perfect for researchers and builders who need something that just works for training high quality text to speech models. Key Features ๐ Low bit rate compression a speech codec that compresses and reconstructs audio with near inaudible reconstruction loss ๐ผ Upsamples from 16kHz โ 24kHz ๐ Ready for real world use train your own SpeechLMs without needing to build your own codec ๐ข Commercial use permitted use it in your own tools or products ๐ Released with large pre encoded datasets weโve compressed Emilia YODAS from 1.7TB to 41GB using NeuCodec, significantly reducing the compute requirements needed for training Model Details NeuCodec is a Finite Scalar Quantisation (FSQ) based 0.8kbps audio codec for speech tokenization. It takes advantage of the following features: FSQ quantisation resulting in a single codebook, making it ideal for downstream modeling with Speech Language Models. Trained with CC data such that there are no Non Commercial data restrictions. At 50 tokens/sec and 1โฆ
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy