Token Classification
Transformers
ONNX
Safetensors
English
Irish
distilbert
pii
de-identification
ireland
ppsn
eircode
finance
passport
phone-number
multilingual
Eval Results (legacy)
Instructions to use temsa/OpenMed-mLiteClinical-IrishCorePII-135M-v1 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use temsa/OpenMed-mLiteClinical-IrishCorePII-135M-v1 with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("token-classification", model="temsa/OpenMed-mLiteClinical-IrishCorePII-135M-v1")# Load model directly from transformers import AutoTokenizer, AutoModelForTokenClassification tokenizer = AutoTokenizer.from_pretrained("temsa/OpenMed-mLiteClinical-IrishCorePII-135M-v1") model = AutoModelForTokenClassification.from_pretrained("temsa/OpenMed-mLiteClinical-IrishCorePII-135M-v1", device_map="auto") - Notebooks
- Google Colab
- Kaggle
| OpenMed-mLiteClinical-IrishCorePII-135M-v1 | |
| This release is derived from: | |
| - OpenMed/OpenMed-PII-mLiteClinical-Base-135M-v1 (Apache-2.0) | |
| Training/evaluation data used for this derivative included: | |
| - temsa/OpenMed-Irish-PPSN-Eircode-Spec-v1 (Apache-2.0 synthetic dataset) | |
| - temsa/OpenMed-Irish-CorePII-TrainMix-v1 (composite training mix) | |
| - joelniklaus/mapa (CC-BY-4.0) | |
| - gretelai/synthetic_pii_finance_multilingual (Apache-2.0) | |
| Please review the dataset cards and upstream licenses before redistributing derivative datasets. | |