litert community/gemma 4 E2B it litert lm Main Model Card: google/gemma 4 E2B it This model card provides the Gemma 4 E2B model in a way that is ready for deployment on Android, iOS, Desktop, IoT and Web. Gemma is a family of lightweight, state of the art open models from Google, built from the same research and technology used to create the Gemini models. This particular Gemma 4 model is small so it is ideal for on device use cases. By running this model on device, users can have private access to Generative AI technology without even requiring an internet connection. These models are provided in the .litertlm format for use with the LiteRT LM framework. LiteRT LM is a specialized orchestration layer built directly on top of LiteRT, Google’s high performance multi platform runtime trusted by millions of Android and edge developers. LiteRT provides the foundational hardware acceleration via XNNPack for CPU and ML Drift for GPU. LiteRT LM adds the specialized GenAI libraries and APIs, such as KV cache management, prompt templating, and function calling. This integrated stack is the same technology powering the Google AI Edge Gallery showcase app. LiteRT LM uses a state of the art Ge…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy