首页> 外文会议>Workshop on Automatic Speech Recognition and Understanding >AN SVD-BASED SCHEME FOR MFCC COMPRESSION IN DISTRIBUTED SPEECH RECOGNITION SYSTEM
【24h】

AN SVD-BASED SCHEME FOR MFCC COMPRESSION IN DISTRIBUTED SPEECH RECOGNITION SYSTEM

机译:分布式语音识别系统中的MFCC压缩基于SVD的方案

获取原文

摘要

This paper proposes a new scheme for low bit-rate source coding of Mel Frequency Cepstral Coefficients (MFCCs) in Distributed Speech Recognition (DSR) system. The method uses the compressed ETSI Advanced Front-End (ETSI-AFE) features factorized into SVD components. By investigating the correlation property between successive MFCC frames, the odd ones are encoded using ETSI-AFE, while only the singular values and the nearest left singular vectors index are encoded and transmitted for the even frames. At the server side, the non-transmitted MFCCs are evaluated through their quantized singular values and the nearest left singular vectors. The system provides a compression bit-rate of 2.7 kbps. The recognition experiments were carried out on the Aurora-2 database for clean and multi-condition training modes. The simulation results show good recognition performance without significant degradation, with respect to the ETSI-AFE encoder.
机译:本文提出了分布式语音识别(DSR)系统中MEL频率谱系数(MFCC)的低比特率源编码的新方案。该方法使用压缩的ETSI高级前端(ETSI-AFE)分为SVD组件。通过研究连续MFCC帧之间的相关性,使用ETSI-AFE对奇数进行编码,而仅对偶数帧进行编码和传输奇异值和最近的左字奇异矢量索引。在服务器端,通过它们量化的奇异值和最近的左奇异向量来评估未发送的MFCC。该系统提供了2.7 kbps的压缩比特率。识别实验是在Aurora-2数据库上进行的,以进行清洁和多条件培训模式。仿真结果表明,对于ETSI-AFE编码器,良好的识别性能而无需显着降级。

著录项

相似文献

  • 外文文献
  • 中文文献
  • 专利
获取原文

客服邮箱:kefu@zhangqiaokeyan.com

京公网安备:11010802029741号 ICP备案号:京ICP备15016152号-6 六维联合信息科技 (北京) 有限公司©版权所有
  • 客服微信

  • 服务号