Team MCTE Resilient AI Submission Voxtral Mini 4B Realtime 2602 This is the WINNING Submission for the Audio to Text Model Compression Track of the Resilient AI Challenge, a first of its kind international open model compression challenge by the Govt of France & India, UNESCO & Coalition for Sustainable AI This open weights contribution has been done by the 🇮🇳 MILITARY COLLEGE OF TELECOMMUNICATION ENGINEERING (MCTE), INDIAN ARMY Category: audio to text Inference server: vLLM 0.19.1 or newer License: Apache 2.0, matching the original model Launch Prerequisites Linux x86 64 NVIDIA GPU with a compatible NVIDIA driver NVIDIA L4 with 16 GB VRAM for the target evaluation environment Conda Python 3.11 No custom vLLM source changes, plugins, CUDA extensions, Python entrypoints, or external serving wrappers are required. Environment Setup From a shell with Conda available, create a clean Python 3.11 environment: Install the serving dependencies from this repository: The validated requirements install vLLM 0.19.1 and include the audio and API dependencies required by Voxtral. pip check should report: Start the Server Run this command from the repository root: For an evaluator using a Huggi…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy