RWKV7 G1 "GooseOne" pure RNN reasoning model These are BASE models (pretrained with web/code/synthetic + instruction/chat/reasoning data), suitable for post training and fine tuning (check https://huggingface.co/spaces/Jellyfish042/UncheatableEval to see their performance at language modeling). More info & Gradio demo: https://rwkv.com/ Search "RWKV Chat" in play store / app store for our local inference app RWKV Chat: https://rwkv.halowang.cloud/ (local inference for mobile/desktop) and https://github.com/RWKV APP/RWKV APP GGUF: https://huggingface.co/collections/shoumenchougou/rwkv7 gxx gguf More GGUF & mobile models: https://huggingface.co/mollysama/rwkv mobile models Ollama GGUF: https://ollama.com/mollysama RWKV 7 pth = GGUF script: https://github.com/MollySophia/rwkv mobile/blob/master/converter/convert rwkv pth to gguf.py Training: https://github.com/BlinkDL/RWKV LM and https://github.com/Joluck/RWKV PEFT Note: rwkv7a has DeepEmbed Efficient inference: https://github.com/BlinkDL/Albatross 145+ token/s RWKV 7 7.2B fp16 bsz1 decoding @ RTX5090 (always const speed & vram) 10250+ token/s RWKV 7 7.2B fp16 bsz960 decoding @ RTX5090 (always const speed & vram) 9650+ token/s RWKV 7…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy