Tag: Text Anonymization

SIESTA Concepts #5: Named Entity Recognition
Sharing cyber incident reports is vital for threat detection, but they’re full of personal data. Anonymisation pipelines rely on Named Entity Recognition models to catch sensitive mentions first — and these are usually trained in English. Researchers within EOSC SIESTA tested whether that holds up in Spanish, and found that multilingual models outperform English cybersecurity…

SIESTA Tool: Text Anonymization on Sensitive Data
EOSC-SIESTA has developed a text anonymization tool prototype to securely share sensitive cyber incident reports. Created by Universidad de León, the tool uses AI and four anonymization techniques to mask confidential data, enabling safe data reuse for research, machine learning, and collaborative cyber defence.





