Table of Contents 0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation TL;DR Model Details Model Description Developed by: https://www.tii.ae Model type: Causal decoder only Architecture: Hybrid Transformers + Mamba architecture Language(s) (NLP): English, Multilingual License: Falcon LLM License Training details For more details about the training protocol of this model, please refer to the Falcon H1 technical blogpost and Technical Report. Usage Currently to use this model you can either rely on Hugging Face transformers , vLLM or llama.cpp library. Inference Make sure to install the latest version of transformers or vllm , eventually install these packages from source: For vLLM, make sure to install vllm =0.9.0 : 🤗 transformers Refer to the snippet below to run H1 models using 🤗 transformers: vLLM For vLLM, simply start a server by executing the command below: llama.cpp You can find all GGUF files under our official collection Evaluation Falcon H1 series perform very well on a variety of tasks, including reasoning tasks. Tasks Falcon H1 3B Qwen3 4B Qwen2.5 3B Gemma3 4B Llama3.2 3B Falcon3 3B General BBH 53.69 51.07 46.55 50.01 41.47 45.02 ARC C 49.5…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy