Yayın: Speaker Diarization using Embedding Vectors
| dc.contributor.author | Toruk, Mesut | |
| dc.contributor.author | Bilgin, Gokhan | |
| dc.contributor.author | Serbes, Ahmet | |
| dc.date.accessioned | 2026-06-27T14:28:27Z | |
| dc.date.issued | 2020 | |
| dc.description.abstract | In recent years, with the rapid increase of voice data, solutions are being sought extensively for examination and indexing in the field of speech processing. One of these solutions is the speaker diarization, which is used to examine speech records that include multi-speaker. The speaker diarization system basically splits the speech file into segments using the speech file's silence fields and examines the similarity between the segments. In this study, deep learning based embedding vectors are used for speaker representation. The vector embedding proposed for speaker representation is performed with x-vectors extracted using time-delayed deep neural network and d-vectors extracted using LSTM. Then the system is tested with xd-vectors consisting of the combination of these two vectors. As a result, the effect of representative embedding vectors on the performance of the diarization system is examined in the scope of this paper. | en |
| dc.description.uri | https://doi.org/10.1109/siu49456.2020.9302162 | |
| dc.identifier.doi | 10.1109/siu49456.2020.9302162 | |
| dc.identifier.isbn | 978-1-7281-7206-4 | |
| dc.identifier.issn | 2165-0608 | |
| dc.identifier.uri | https://hdl.handle.net/20.500.14981/61013 | |
| dc.identifier.wos | 000653136100136 | |
| dc.language.iso | tur | |
| dc.publisher | IEEE | |
| dc.relation.conference | 28th Signal Processing and Communications Applications Conference (SIU) | |
| dc.relation.ispartof | 2020 28TH SIGNAL PROCESSING AND COMMUNICATIONS APPLICATIONS CONFERENCE (SIU) | |
| dc.subject | Speaker diarization | |
| dc.subject | x-vector | |
| dc.subject | time-delay neural network | |
| dc.subject | d-vector | |
| dc.subject | LSTM | |
| dc.subject | Engineering | |
| dc.subject | Telecommunications | |
| dc.title | Speaker Diarization using Embedding Vectors | |
| dc.type | Proceedings Paper | |
| dspace.entity.type | Publication | |
| local.import.source | WOS |