An Iterative Framework for Self-Supervised Deep Speaker Representation Learning

机译：自我监督深层扬声器代表学习的迭代框架

获取原文

页面导航

摘要
著录项
相似文献
相关主题

摘要

In this paper, we propose an iterative framework for self-supervised speaker representation learning based on a deep neural network (DNN). The framework starts with training a self-supervision speaker embedding network by maximizing agreement between different segments within an utterance via a contrastive loss. Taking advantage of DNN’s ability to learn from data with label noise, we propose to cluster the speaker embedding obtained from the previous speaker network and use the subsequent class assignments as pseudo labels to train a new DNN. Moreover, we iteratively train the speaker network with pseudo labels generated from the previous step to bootstrap the discriminative power of a DNN. Speaker verification experiments are conducted on the VoxCeleb dataset. The results show that our proposed iterative self-supervised learning framework outperformed previous works using self-supervision. The speaker network after 5 iterations obtains a 61% performance gain over the speaker embedding model trained with contrastive loss.

机译：在本文中，我们提出了一种基于深神经网络（DNN）的自我监督扬声器表示学习的迭代框架。该框架首先通过通过对比损失最大化不同段之间的不同段之间的协议来培训自我监督扬声器嵌入网络。利用DNN从带有Label噪声的数据学习的能力，我们建议培养从先前扬声器网络获得的扬声器嵌入，并使用后续类分配作为伪标签来培训新的DNN。此外，我们迭代地将扬声器网络与从上一步产生的伪标签训练，以引导DNN的识别力。扬声器验证实验是在VoxceleB数据集上进行的。结果表明，我们建议的迭代自我监督的学习框架优于以前的自我监督工作。 5次迭代后的扬声器网络在培训具有对比损失的扬声器嵌入模型上获得61％的性能增益。

著录项

来源
《IEEE International Conference on Acoustics, Speech and Signal Processing》|2021年|6728-6732|共5页
会议地点
作者
Danwei Cai; Weiqing Wang; Ming Li;
展开▼
作者单位

展开▼
会议组织
原文格式 PDF
正文语种
中图分类
关键词
Training; Neural networks; Signal processing algorithms; Speech recognition; Signal processing; Performance gain; Data models;

机译：培训;神经网络;信号处理算法;语音识别;信号处理;性能增益;数据模型;

相似文献

外文文献
中文文献
专利

1. A Self-Supervised Deep Learning Framework for Unsupervised Few-Shot Learning and Clustering [J] . Zhang Hongjing, Zhan Tianyang, Davidson Ian Pattern recognition letters . 2021,第Auga期

机译：无监督的少量学习和聚类自我监督的深度学习框架
2. Deep Self-Supervised Representation Learning for Free-Hand Sketch [J] . Xu Peng, Song Zeyu, Yin Qiyue, IEEE Transactions on Circuits and Systems for Video Technology . 2021,第4期

机译：自由素描的深度自我监督的代表学习
3. An Investigation of Deep-Learning Frameworks for Speaker Verification Antispoofing [J] . Chunlei Zhang, Chengzhu Yu, John H. L. Hansen Selected Topics in Signal Processing, IEEE Journal of . 2017,第4期

机译：说话人验证反欺骗的深度学习框架研究
4. Self-supervised learning framework for speaker localisation with a humanoid robot [C] . Jonas Gonzalez-Billandon, Giulia Belgiovine, Matthew Tata, International Conference on Development and Learning . 2021

机译：具有人形机器人的扬声器本地化自我监督的学习框架
5. Self-Supervised Representation Learning Via Image Out-Painting for Medical Image Analysis [D] . ?Sodha, Vatsal Arvindkumar 2020

机译：通过图像外绘画进行自我监督的代表学习医学图像分析
6. Representation Learning: A Unified Deep Learning Framework for Automatic Prostate MR Segmentation [O] . Shu Liao, Yaozong Gao, Aytekin Oto, -1

机译：代表学习：一个统一的深学习框架自动前列腺mR分割
7. Endmember-Guided Unmixing Network (EGU-Net): A General Deep Learning Framework for Self-Supervised Hyperspectral Unmixing [O] . Danfeng Hong, Lianru Gao, Jing Yao, 2021

机译：EndMember-Buided Unmixing Network（egu-net）：自我监督高光谱解密的一般深入学习框架

An Iterative Framework for Self-Supervised Deep Speaker Representation Learning

摘要

著录项

相似文献

相关主题

期刊订阅