当前位置:网站首页>声纹技术(六):声纹技术的其他应用
声纹技术(六):声纹技术的其他应用
2022-06-25 07:36:00 【u013250861】
6.1 声纹的力量
前面几章介绍的声纹识别与声纹分割聚类都属于声纹技术在音频信号处理中的最直接的应用。而除了这些直接应用,由于声纹本身包含着与说话人身份相关的信息,其在其他领域也能发挥出重要作用。
声纹信息在其他领域中发挥作用有很多种方式,其中一种比较经典的架构便是通过声纹嵌入码,将特定说话人的身份信息,作为该领域传统模型的辅助输入,融合到模型的训练过程中,如图6.1 所示。该架构中的辅助音频,来自该任务所对应的具体说话人。而基于从该辅助音频中提取的声纹嵌入码,能够让传统模型更精准地针对该说话人完成相应的任务。这里的声纹编码器可以采用第3 章介绍过的各种模型,不过现在一般都采用基于神经网络的声纹编码器。而架构中的输入与输出可以有很多种形式,既可以是音频,也可以是时频谱、文字、类别或其他信息,具体依应用而异。

6.2 用于语音识别
6.2.1 语音识别技术概述
5.5.7 节介绍声纹分割聚类与语音识别的联合训练时,简单介绍了一些关于语音识别的概念。语音识别本身可以算是音频信号处理领域下最庞大、最重要的一门学科。由于本书主要以介绍声纹技术及相关应用为重点,不可能单独对语音识别技术进行详尽的介绍。为了更好地描述将声纹信息应用于语音识别领域的方法,我们还是简略介绍一下语音识别中的一些常用架构。对语
边栏推荐
- Bluecmsv1.6-代码审计
- Easyplayer streaming media player plays HLS video. Technical optimization of slow starting speed
- 【操作教程】TSINGSEE青犀视频平台如何将旧数据库导入到新数据库?
- Find out the possible memory leaks caused by the handler and the solutions
- Sharepoint:sharepoint server 2013 and adrms Integration Guide
- openid是什么意思?token是什么意思?
- Unity addressable batch management
- Summary of NLP data enhancement methods
- [summary] 1361- package JSON and package lock JSON relationship
- 城链科技平台,正在实现真正意义上的价值互联网重构!
猜你喜欢

Exchange: manage calendar permissions

Prepare these before the interview. The offer is soft. The general will not fight unprepared battles

Check whether the point is within the polygon

How to calculate the distance between texts: WMD

Index analysis of DEMATEL model

如何设计测试用例

Various synchronous learning notes

Similarity calculation method
How to calculate the characteristic vector, weight value, CI value and other indicators in AHP?

What are the indicators of entropy weight TOPSIS method?
随机推荐
VOCALOID notes
Swiperefreshlayout+recyclerview failed to pull down troubleshooting
Various synchronous learning notes
某视频网站m3u8非感知加密分析
OpenFOAM:底层
Is the securities account given by Qiantang education business school safe? Can I open an account?
Want to open an account, is it safe to open an online stock account?
Easyplayer streaming media player plays HLS video. Technical optimization of slow starting speed
开户券商怎么选择?在线开户是安全么?
Nodehandle common member functions
Prepare these before the interview. The offer is soft. The general will not fight unprepared battles
How to calculate the D value and W value of statistics in normality test?
C language "recursive series": recursive implementation of 1+2+3++ n
RTOS 多线程下hardfault问题总结
Almost taken away by this wave of handler interview cannons~
Trendmicro:apex one server tools folder
紧急行政中止令下达 Juul暂时可以继续在美国销售电子烟产品
Scanpy (VII) spatial data analysis based on scanorama integrated scrna seq
Bluecmsv1.6- code audit
Bluecmsv1.6-代码审计