| International Journal of Advanced Robotic Systems | |
| Spoken Document Retrieval Based on Confusion Network with Syllable Fragments | |
| 关键词: spoken document retrieval; lattice; confusion network; syllable fragment; | |
| DOI : 10.5772/52454 | |
| 学科分类:自动化工程 | |
| 来源: InTech | |
PDF
|
|
【 摘 要 】
This paper addresses the problem of spoken document retrieval under noisy conditions by incorporating sound selection of a basic unit and an output form of a speech recognition system. Syllable fragment is combined with a confusion network in a spoken document retrieval task. After selecting an appropriate syllable fragment, a lattice is converted into a confusion network that is able to minimize the word error rate instead of maximizing the whole sentence recognition rate. A vector space model is adopted in the retrieval task where tf-idf weights are derived from the posterior probability. The confusion network with syllable fragments is able to improve the mean of average precision (MAP) score by 0.342 and 0.066 over one-best scheme and the lattice.
【 授权许可】
CC BY
【 预 览 】
| Files | Size | Format | View |
|---|---|---|---|
| RO201902188525543ZK.pdf | 831KB |
PDF