This is Phi 4 mini instruct with our BUG FIXES. See our collection for versions of Phi 4 with our bug fixes including GGUF & 4 bit formats. Unsloth's Phi 4 Dynamic Quants is selectively quantized, greatly improving accuracy over standard 4 bit. Finetune your own Reasoning model like R1 with Unsloth! We have a free Google Colab notebook for turning Phi 4 into a reasoning model: https://colab.research.google.com/github/unslothai/notebooks/blob/main/nb/Phi 4 (14B) GRPO.ipynb Unsloth bug fixes: 1. Padding and EOS tokens are the same fixed this. 2. Chat template had extra EOS token removed this. Otherwise you will be during inference. 3. EOS token should be not . Otherwise it'll terminate at 4. Changed unk token to � from EOS. ✨ Finetune for Free All notebooks are beginner friendly ! Add your dataset, click "Run All", and you'll get a 2x faster finetuned model which can be exported to GGUF, vLLM or uploaded to Hugging Face. Unsloth supports Free Notebooks Performance Memory use GRPO with Phi 4 ▶️ Start on Colab GRPO.ipynb) 2x faster 80% less Llama 3.2 (3B) ▶️ Start on Colab Conversational.ipynb) 2.4x faster 58% less Llama 3.2 (11B vision) ▶️ Start on Colab Vision.ipynb) 2x faster 60% le…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy