Chat & support: my new Discord server Want to contribute? TheBloke's Patreon page WizardLM: An Instruction following LLM Using Evol Instruct These files are the result of merging the delta weights with the original Llama7B model. The code for merging is provided in the WizardLM official Github repo. The original WizardLM deltas are in float32, and this results in producing an HF repo that is also float32, and is much larger than a normal 7B Llama model. Therefore for this repo I converted the merged model to float16, to produce a standard size 7B model. This was achieved by running model = model.half() prior to saving. WizardLM 7B HF This repo contains the full unquantised model files in HF format for GPU inference and as a base for quantisation/conversion. Other repositories available 4bit GGML models for CPU inference 4bit GPTQ models for GPU inference Discord For further support, and discussions on these models and AI in general, join us at: TheBloke AI's Discord server Thanks, and how to contribute. Thanks to the chirper.ai team! I've had a lot of people ask if they can contribute. I enjoy providing models and helping people, and would love to be able to spend even more time do…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy