A Convergence Analysis of Log-Linear Training

2011-12-01NeurIPS 2011Unverified0· sign in to hype

Simon Wiesler, Hermann Ney

Unverified — Be the first to reproduce this paper.

Abstract

Log-linear models are widely used probability models for statistical pattern recognition. Typically, log-linear models are trained according to a convex criterion. In recent years, the interest in log-linear models has greatly increased. The optimization of log-linear model parameters is costly and therefore an important topic, in particular for large-scale applications. Different optimization algorithms have been evaluated empirically in many papers. In this work, we analyze the optimization problem analytically and show that the training of log-linear models can be highly ill-conditioned. We verify our findings on two handwriting tasks. By making use of our convergence analysis, we obtain good results on a large-scale continuous handwriting recognition task with a simple and generic approach.

Tasks

Handwriting Recognition

A Convergence Analysis of Log-Linear Training

Abstract

Tasks

Reproductions