speaker-diarization-3.1
Hugging Face model with 3325 likes. Tags: pyannote-audio, pyannote, pyannote-audio-pipeline, audio, voice, speech, speaker, speaker-diarization, speaker-change-detection, voice-activity-detection
Papers6
Empathetic social robots should respond not only to what users say, but also to how their emotions dynamically evolve du
paperAMR: Adaptive Modality Routing for Multimodal Polyglot Speaker IdentificationMultimodal speaker identification systems face two key challenges in real-world deployment: missing modalities and langu
paperDisentangling Speaker and Language Effects in Cross-Lingual Speaker Verification for Iberian LanguagesCross-lingual speaker verification (SV) systems typically exhibit performance degradation when enrollment and test utter
paperDG^VoiC: Speaker Clustering for Fraud Investigation under Real Call-Centre ConditionsInsurance fraud remains costly and operationally difficult, particularly in call-centre workflows where many customer in
paperChoosing a PEFT Variant for Per-Patient Dysarthric ASR: A Single-Speaker Case Study on Two ASR BasesPer-patient adapters are the preferred production architecture for dysarthric automatic speech recognition (ASR), yet pa
paperSpEmoC: A Balanced Speaker-Segment Multimodal Emotion BenchmarkUnderstanding human emotions in spoken conversations is a key challenge in affective computing, with applications in emp
