FlexOlmo is a new kind of LM that unlocks a new paradigm of data collaboration. With FlexOlmo, data owners can contribute to the development of open language models without giving up control of their data. There is no need to share raw data directly, and data contributors can decide when their data is active in the model, deactivate it at any time, and receive attributions whenever it's used for inference. Model Summary FlexOlmo 7x7B 1T (without router training) is a Mixture of Experts with 33B total parameters, combining independently trained experts on public mix, news, math, code, academic texts, creative writing, and Reddit data. The public mix expert is trained on 1T tokens of public data while the other experts are branched from the public mix expert and trained on 50B tokens of their respective data. This information and more can also be found: Paper : https://allenai.org/papers/flexolmo Code : https://github.com/allenai/FlexOlmo Blog : https://allenai.org/blog/flexolmo Data and corresponding models : Corpus Public Math News Academic Code Creative Writing Reddit Model Flex public 7B 1T Flex math 2x7B 1T Flex news 2x7B 1T Flex pes2o 2x7B 1T Flex code 2x7B 1T Flex creative 2x7…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy