propella 1 propel your data curation to the next level. propella 1 is a family of small multilingual LLMs that annotate text documents across six categories: core content, classification, quality & value, audience & purpose, safety & compliance, and geographic relevance. The annotations can be used to filter, select, and curate LLM training data at scale. Disclaimer: This is a research project, not an official ellamind product. For production ready evaluation solutions, check out elluminate. Highlights Annotate 18 properties : Covers well established dimensions like content quality and educational value, plus underexplored ones like reasoning indicators and time sensitivity. Fast & accurate : Small models (0.6B, 1.7B, 4B) that punch above their weight. Trained in fp8, ready for high throughput inference. Any text, any format : Handles web pages, PDFs, code, math, post training data and more. Highly multilingual : Supports 57 languages. The propella 1 family of models Model Parameters Performance Docs/s (A100/H100) : : : : : : propella 1 4b 4B 0.779 10.3 / 27.0 propella 1 1.7b 1.7B 0.737 17.8 / 39.1 propella 1 0.6b 0.6B 0.729 21.5 / 39.9 Properties propella 1 models evaluate documen…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy