conference-paper
Open access
SpeeDF - A Speech De-Identification Framework
Research footprint
At a glance
- Citations
- 2
- References
- 17
- Comments
- 0
Paper overview
Abstract
This paper proposes SpeeDF, a novel three-step framework for anonymizing speech data, particularly focusing on Singaporean English (Singlish). SpeeDF tackles the challenge of protecting less-studied Personally Identifiable Information (PII) like NRIC and passport numbers, which often go overlooked by traditional de-identification methods. Unlike approaches focused solely on entity extraction, SpeeDF leverages a combination of automatic speech recognition (ASR), named entity recognition (NER), and information anonymization. This comprehensive approach ensures thorough PII redaction while preserving the naturalness and usability of the anonymized speech data for research and various downstream applications.
Record transparency
Publication details
- DOI
- 10.1109/tencon61640.2024.10902957
- OpenAlex
- W4408258559
- Document type
- conference-paper
- Language
- EN
- Last metadata update
Comments
Log in to join the discussion.