🔥 MERaLiON 3 🔥 🚀 MERaLiON 3 10B 💻 Web Demo ⚙️ vLLM coming soon Introduction We are pleased to announce the release of our flagship speech text large language model, MERaLiON 3 10B . MERaLiON 3 10B demonstrates competitive performance across benchmark evaluations in Age Recognition, Gender Recognition, Spoken Question Answering (SQA), and Contextual Paralinguistic Question Answering (CPQA) in the Southeast Asian context as compared to the latest AudioLLMs, including Gemini 3 Flash and Qwen3 Omni Instruct. The benchmark contains speech and prompts in Malay, Indonesian, English, Chinese, Tamil, Thai and Vietnamese to better represent the Southeast Asian context. The following table presents task specific evaluation scores, assessed using the LLM as a Judge framework across multiple datasets. Higher scores indicate better performance. We will open source the benchmark separately as part of a paper. See the Evaluation section for detailed benchmarking. Benchmark MERaLiON 3 10B MERaLiON 2 10B Qwen3 Omni Gemini 3 Flash GPT 4o Audio : : : : : : : : : : : Age (commonvoice en, ta, th, vi, zh) 76.84 61.77 70.38 77.00 68.90 Gender (Multi dataset) 92.70 54.19 95.34 81.72 40.25 Spoken Q&A (S…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy