FLM Audio FLM Audio is a audio language subversion of RoboEgo/FLM Ego an omnimodal model with native full duplexity. It simultaneously listens, speaks, and composes internal monologue, delivering low‑latency, duplex conversational responses in both English and Chinese. FLM‑Audio is robust to noise and user interruptions, prioritizing responsiveness and naturalness. 📄 Model Card Language(s): Chinese; English; 📚 Technical Report Motivation & Survey: Toward Embodied AGI: A Review of Embodied AI and the Road Ahead FLM Audio Research Paper: FLM Audio: Natural Monologues Improves Native Full Duplex Chatbots via Dual Training Omnimodal System Card: RoboEgo System Card: An Omnimodal Model with Native Full Duplexity ⚠️ Bias, Risks, and Limitations Despite extensive data cleaning, FLM‑Audio may still produce undesired content (e.g., biased or offensive language). Users should not disseminate unsafe outputs. Project authors are not responsible for misuse or harmful consequences. 🚀 Quick Start Please refer to the repository of FLM Audio server to interact with FLM Audio via WebUI. ℹ️ Usage Notice This project is intended for research use only in compliance with applicable laws. For commerci…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy