Table of Contents 0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation TL;DR Model Details Model Description Developed by: https://www.tii.ae Model type: Causal decoder only Architecture: Hybrid Transformers + Mamba architecture Language(s) (NLP): English License: Falcon LLM License Training details For more details about the training protocol of this model, please refer to the Falcon H1 technical blogpost and Technical Report. Usage Currently to use this model you can either rely on Hugging Face transformers , vLLM or llama.cpp library. Inference Make sure to install the latest version of transformers or vllm , eventually install these packages from source: For vLLM, make sure to install vllm =0.9.0 : 🤗 transformers Refer to the snippet below to run H1 models using 🤗 transformers: vLLM For vLLM, simply start a server by executing the command below: llama.cpp You can find all GGUF files compatible with llama.cpp under our official collection Evaluation Falcon H1 series perform very well on a variety of tasks, including reasoning tasks. Tasks Falcon H1 0.5B Qwen3 0.6B Qwen2.5 0.5B Gemma3 1B Llama3.2 1B Falcon3 1B General BBH 42.91 32.95 33.26 35.86 33.2…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy