Table of Contents 0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation TL;DR Model Details Model Description Developed by: https://www.tii.ae Model type: Causal decoder only Architecture: Hybrid Transformers + Mamba architecture Language(s) (NLP): English Number of Parameters: 90M License: Falcon LLM License Training details For more details about the training protocol of this model, please refer to the Falcon H1 Tiny technical blogpost. Usage Currently to use this model you can either rely on Hugging Face transformers , vLLM , sglang , llama.cpp , ollama or mlx library. Inference 🤗 transformers Refer to the snippet below to run H1 models using 🤗 transformers: or llama.cpp You can find all GGUF files compatible with llama.cpp under our official collection an example setup could be: ollama Apple mlx vLLM For vLLM, simply start a server by executing the command below: sglang Evaluation For detailed evaluation of Tiny H1 series, please refer to our technical blogpost Useful links View our release blogpost. Feel free to join our discord server if you have any questions or to interact with our researchers and developers. Citation If the Falcon H1 Tiny famil…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy