litert community/gemma 4 E4B it litert lm Main Model Card: google/gemma 4 E4B it This model card provides the Gemma 4 E4B model in a way that is ready for deployment on Android, iOS, Desktop, IoT and Web. Gemma is a family of lightweight, state of the art open models from Google, built from the same research and technology used to create the Gemini models. This particular Gemma 4 model is small so it is ideal for on device use cases. By running this model on device, users can have private access to Generative AI technology without even requiring an internet connection. These models are provided in the .litertlm format for use with the LiteRT LM framework. LiteRT LM is a specialized orchestration layer built directly on top of LiteRT, Google’s high performance multi platform runtime trusted by millions of Android and edge developers. LiteRT provides the foundational hardware acceleration via XNNPack for CPU and ML Drift for GPU. LiteRT LM adds the specialized GenAI libraries and APIs, such as KV cache management, prompt templating, and function calling. This integrated stack is the same technology powering the Google AI Edge Gallery showcase app. The model file size is 3.66 GB, whic…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy