YaRN 128K. 32K non extended GGUF here: link Finetune Llama 3.1, Gemma 2, Mistral 2 5x faster with 70% less memory via Unsloth! We have a Qwen 2.5 (all model sizes) free Google Colab Tesla T4 notebook. Also a Qwen 2.5 conversational style notebook. ✨ Finetune for Free All notebooks are beginner friendly ! Add your dataset, click "Run All", and you'll get a 2x faster finetuned model which can be exported to GGUF, vLLM or uploaded to Hugging Face. Unsloth supports Free Notebooks Performance Memory use Llama 3.1 8b ▶️ Start on Colab 2.4x faster 58% less Phi 3.5 (mini) ▶️ Start on Colab 2x faster 50% less Gemma 2 9b ▶️ Start on Colab 2.4x faster 58% less Mistral 7b ▶️ Start on Colab 2.2x faster 62% less TinyLlama ▶️ Start on Colab 3.9x faster 74% less DPO Zephyr ▶️ Start on Colab 1.9x faster 19% less This conversational notebook is useful for ShareGPT ChatML / Vicuna templates. This text completion notebook is for raw text. This DPO notebook replicates Zephyr. \ Kaggle has 2x T4s, but we use 1. Due to overhead, 1x T4 is 5x faster. unsloth/Qwen2.5 Coder 7B Instruct 128K GGUF Introduction Qwen2.5 Coder is the latest series of Code Specific Qwen large language models (formerly known as Cod…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy