Lukáš Burget
5 papers in the PaperMetrix corpus
Papers by this author
-
Analysis of the but Diarization System for Voxconverse Challenge
2021
This paper describes the system developed by the BUT team for the fourth track of the VoxCeleb Speaker Recognition Challenge, focusing on diarization on the VoxConverse dataset. The system consists of signal pre-processing, voice activity …
-
Eat: Enhanced ASR-TTS for Self-Supervised Speech Recognition
2021
Self-supervised ASR-TTS models suffer in out-of-domain data conditions. Here we propose an enhanced ASR-TTS (EAT) model that incorporates two main features: 1) The ASR→TTS direction is equipped with a language model reward to penalize the …
-
From Simulated Mixtures to Simulated Conversations as Training Data for End-to-End Neural Diarization
2022 · arXiv (Cornell University)
End-to-end neural diarization (EEND) is nowadays one of the most prominent research topics in speaker diarization. EEND presents an attractive alternative to standard cascaded diarization systems since a single system is trained at once to …
-
Bayesian joint-sequence models for grapheme-to-phoneme conversion
2017 · HAL (Le Centre pour la Communication Scientifique Directe)
International audience
-
Adapting Diarization-Conditioned Whisper for End-to-End Multi-Talker Speech Recognition
2026
We propose a speaker-attributed (SA) Whisper-based model for multi-talker speech recognition that combines target-speaker modeling with serialized output training (SOT). Our approach leverages a Diarization-Conditioned Whisper (DiCoW) encoder to extract target-speaker embeddings, which are concatenated …