**arXiv ID:** 2504.14963 **Authors:** Rui Ribeiro, Luísa Coheur, Joao P. Carvalho **Published:** 2025-04-21T08:44:33Z **Abstract:** Speaker identification using voice recordings leverages unique acoustic features, but this approach fails when only textual data is available. Few approaches have attempted to tackle the problem of identifying speakers solely from text, and the existing ones have primarily relied on traditional methods. In this work, we explore the use of fuzzy fingerprints from ...
Scanned 9/11/2026
Install to Claude Code
npx -y skills add hiyenwong/ai_collection --skill speaker-fuzzy-fingerprints-benchmarking-textbased-identification-in-multiparty-dialogues --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Speaker Fuzzy Fingerprints Benchmarking Textbased Identification In Multiparty Dialogues?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/hiyenwong-speaker-fuzzy-fingerprints-benchmarking-textbased)More formats (shields.io, HTML) on the badges page.
# Speaker Fuzzy Fingerprints: Benchmarking Text-Based Identification in Multiparty Dialogues
**arXiv ID:** 2504.14963
**Authors:** Rui Ribeiro, Luísa Coheur, Joao P. Carvalho
**Published:** 2025-04-21T08:44:33Z
**Abstract:**
Speaker identification using voice recordings leverages unique acoustic features, but this approach fails when only textual data is available. Few approaches have attempted to tackle the problem of identifying speakers solely from text, and the existing ones have primarily relied on traditional methods. In this work, we explore the use of fuzzy fingerprints from large pre-trained models to improve text-based speaker identification. We integrate speaker-specific tokens and context-aware modeling, demonstrating that conversational context significantly boosts accuracy, reaching 70.6% on the Friends dataset and 67.7% on the Big Bang Theory dataset. Additionally, we show that fuzzy fingerprints can approximate full fine-tuning performance with fewer hidden units, offering improved interpretability. Finally, we analyze ambiguous utterances and propose a mechanism to detect speaker-agnostic lines. Our findings highlight key challenges and provide insights for future improvements in text-based speaker identification.
## Skill Description
This skill is generated from the arXiv paper: Speaker Fuzzy Fingerprints: Benchmarking Text-Based Identification in Multiparty Dialogues (2504.14963).
## How to Use
[To be filled in by the user or by future automation]
## References
- [arXiv:2504.14963](http://arxiv.org/abs/2504.14963v1)
Is this your skill, or is something wrong with this listing? . Author removals are honored within 72 hours.
No comments yet. Be the first to comment!