What problem does it solve?
This Skill removes personally identifiable information from non-English clinical text while keeping sensitive data on the device and preserving language-appropriate formatting.
Core Features & Use Cases
- Multilingual PII detection: Identify sensitive information in Spanish, German, French, Italian, Portuguese, Dutch, Hindi, Telugu, Arabic, Japanese, Turkish, and other supported languages.
- Locale-aware replacement: Generate realistic surrogate names, dates, phone numbers, and national identifiers for the source language and regional locale.
- Language-specific validation: Detect structured identifiers such as DNI, NIR, Steuer-ID, codice fiscale, BSN, CPF, TCKN, and Aadhaar using language-aware patterns and validators.
- Use Case: De-identify a German medical note with a German PII model, validate and replace a Steuer-ID, and generate consistent German-looking surrogates without sending the note to a cloud service.
Quick Start
Use the OpenMed de-identification skill with the appropriate language code and locale to replace personally identifiable information in the provided multilingual clinical note with consistent, native-looking surrogates.