paper-with-me

홈 › Papers

Comparative Analysis of Viterbi Training and Maximum Likelihood Estimation for HMMs

2013-12-16 · NeurIPS 2011 12 · Armen E. Allahverdyan, Aram Galstyan

We present an asymptotic analysis of Viterbi Training (VT) and contrast it with a more conventional Maximum Likelihood (ML) approach to parameter estimation in Hidden Markov Models. While ML estimator works by (locally) maximizing the likelihood of the observed data, VT seeks to maximize the probability of the most likely hidden state sequence. We develop an analytical framework based on a generating function formalism and illustrate it on an exactly solvable model of HMM with one unambiguous symbol. For this particular model the ML objective function is continuously degenerate. VT objective, in contrast, is shown to have only finite degeneracy. Furthermore, VT converges faster and results in sparser (simpler) models, thus realizing an automatic Occam's razor for HMM learning. For more general scenario VT can be worse compared to ML but still capable of correctly recovering most of the parameters.

📄 PDF Abstract BibTeX arXiv:1312.4551

Code (0)

등록된 구현이 없습니다.

Tasks

parameter estimation

Similar Papers 제목 키워드 기반

On the accuracy of the Viterbi alignment

2013-07-30 · Kristi Kuljus, Jüri Lember

In a hidden Markov model, the underlying Markov chain is usually hidden. Often, the maximum likelihood alignment (Viterbi alignment) is used as its estimate. Although having the biggest likelihood, the Viterbi alignment …

Optimal word order for non-causal text generation with Large Language Models: the Spanish case

2025-02-20 · Andrea Busto-Castiñeira, Silvia García-Méndez, Francisco de Arriba-Pérez, Francisco J. González-Castaño

Natural Language Generation (NLG) popularity has increased owing to the progress in Large Language Models (LLMs), with zero-shot inference capabilities. However, most neural systems utilize decoder-only causal (unidirect…

DecoderSentenceText Generation

Online Learning of Trellis Diagram Using Neural Network for Robust Detection and Decoding

2022-02-22 · Jie Yang, Qinghe Du, Yi Jiang

This paper studies machine learning-assisted maximum likelihood (ML) and maximum a posteriori (MAP) receivers for a communication system with memory, which can be modelled by a trellis diagram. The prerequisite of the ML…

PR-NN: RNN-based Detection for Coded Partial-Response Channels

2020-07-30 · Simeng Zheng, Yi Liu, Paul H. Siegel

In this paper, we investigate the use of recurrent neural network (RNN)-based detection of magnetic recording channels with inter-symbol interference (ISI). We refer to the proposed detection method, which is intended fo…

The infinite Viterbi alignment and decay-convexity

2018-10-08 · Nick Whiteley, Matt W. Jones, Aleks P. F. Domanski

The infinite Viterbi alignment is the limiting maximum a-posteriori estimate of the unobserved path in a hidden Markov model as the length of the time horizon grows. For models on state-space $\mathbb{R}^{d}$ satisfying …