ImpactU - Detalle del Producto

A hybrid language model based on a combination of N -grams and stochastic context-free grammars

Acceso Cerrado

Idioma: Inglés

Publicado: 01/06/2004

APC (est): No disponible

JSON

HTML

BibTeX

Abstract:

In this paper, a hybrid language model is defined as a combination of a word-based <i>n</i>-gram, which is used to capture the local relations between words, and a category-based stochastic context-free grammar (SCFG) with a word distribution into categories, which is defined to represent the long-term relations between these categories. The problem of unsupervised learning of a SCFG in General Format and in Chomsky Normal Form by means of estimation algorithms is studied. Moreover, a bracketed version of the classical estimation algorithm based on the Earley algorithm is proposed. This paper also explores the use of SCFGs obtained from a treebank corpus as initial models for the estimation algorithms. Experiments on the UPenn Treebank corpus are reported. These experiments have been carried out in terms of the test set perplexity and the word error rate in a speech recognition experiment.

Tópico:

Speech Recognition and Synthesis

Citaciones:

Citaciones por año:

Altmétricas:

Información de la Fuente:

FuenteACM Transactions on Asian Language Information Processing	Cuartil año de publicaciónNo disponible	Volumen3
Issue2	Páginas113 - 127	pISSNNo disponible
ISSN1558-3430	Perfil OpenAlexhttps://openalex.org/S56575750

Enlaces e Identificadores:

Scholar citations URL	https://scholar.google.com/scholar?cites=17268345163696971748&as_sdt=2005&sciodt=0,5&hl=en	Doi URL	https://doi.org/10.1145/1034780.1034783	Pdf URL	https://dl.acm.org/doi/pdf/10.1145/1034780.1034783
Openalex URL	https://openalex.org/W1974043298	Scholar URL	https://scholar.google.com/scholar?hl=en&as_sdt=0%2C5&q=info%3A5B_-TG98pe8J%3Ascholar.google.com&btnG=

Artículo de revista