HRM Text 1B A 1 B parameter language model checkpoint built on the Hierarchical Reasoning Model (HRM) architecture, trained by Sapient Intelligence from scratch on structured public datasets. HRM is a dual timescale recurrent architecture: two Transformer modules (H = high level / slow, L = low level / fast) iterate over the same input embeddings for H cycles × (L cycles + 1) steps, with additive state injection ( z L + z H ). This gives effectively unbounded compute depth at bounded parameter count. Disclaimer This is a pre alignment model checkpoint, not a chat or instruction following assistant. It is pre trained on a PrefixLM objective with condition prefix tokens and has not been multi turn dialogue tuned, long context adapted, instruction tuned, RLHF trained, or otherwise aligned for assistant style use. If you want to use HRM Text like a chat model, you would need to perform further alignment, such as SFT and/or RL, on task specific data. This checkpoint is meant to serve as a starting point, not a finished assistant. Practical guidance for prompting the raw checkpoint: NLP tasks (classification, extraction, structured output, short form QA) : use the direct condition with 2…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy