OpenMed PII Spanish BiomedELECTRA Base 110M v1 Spanish PII Detection Model 110M Parameters Open Source Model Description OpenMed PII Spanish BiomedELECTRA Base 110M v1 is a transformer based token classification model fine tuned for Personally Identifiable Information (PII) detection in Spanish text . This model identifies and classifies 54 types of sensitive information including names, addresses, social security numbers, medical record numbers, and more. Key Features Spanish Optimized : Specifically trained on Spanish text for optimal performance High Accuracy : Achieves strong F1 scores across diverse PII categories Comprehensive Coverage : Detects 55+ entity types spanning personal, financial, medical, and contact information Privacy Focused : Designed for de identification and compliance with GDPR and other privacy regulations Production Ready : Optimized for real world text processing pipelines Performance Evaluated on the Spanish subset of AI4Privacy dataset: Metric Score : : : Micro F1 0.8375 Precision 0.8250 Recall 0.8503 Macro F1 0.8617 Weighted F1 0.8367 Accuracy 0.9859 Top 10 Spanish PII Models Rank Model F1 Precision Recall : : : : : : : : : 1 OpenMed PII Spanish Snowf…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy