PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
April 27, 2026International Journal on Document Analysis and Recognition (IJDAR)0 citationsOpen Access

Predicting text recognition word error rate of image documents without ground truth transcripts

View Full Paper
EVEnrique VidalATAlejandro H. Toselli

Key Points

  • This research aims to develop methods for predicting the word error rate (WER) of text image recognizers without needing ground truth transcripts.
  • Examined unsupervised error estimation using statistical decision theory.
  • Developed naive estimates of classifier error and calibrated them with class-labeled data.
  • Conducted experiments on three large handwritten text datasets.
  • Predicted absolute deviations from real WER were under 1.7% across all datasets.
  • Relative deviations from the WER were recorded at 5.6%, 4.2%, and 6.1% for the respective datasets.

Abstract

Abstract Unsupervised error estimation is examined under the framework of the statistical decision theory. The error probability of a classifier trained with a training set of a finite, fixed size is considered, along with an unsupervised naive estimate of this error. In general, the naive estimates tend to be significantly smaller than the empirical errors measured using ground truth class labels, but the estimated values can be easily calibrated with a function whose parameters can be trained using moderate amounts of class-labelled data. This way, true error rates can be accurately predicted for new unlabelled data. These ideas are applied to the essential classification problem which underlies the automatic transcription of text images. As a result, various methods are developed to predict the Word Error Rate (WER) of a text image recognizer, on unseen sets of images for which no ground truth transcripts are available. Experiments on three large handwritten text datasets show that the error rates predicted by some of these methods are sufficiently accurate for practical applications. More specifically, absolute deviations of the predicted word error percentages from the corresponding real WER are lower than 1.7% for the three datasets, and the relative values of these deviations with respect to the WER of each dataset, are 5.6%, 4.2% and 6.1%.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Vidal et al. (2026) studied this question.

synapsesocial.com/papers/69eefd82fede9185760d42f4https://doi.org/10.1007/s10032-026-00578-6
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1An Algorithm for Least-Squares Estimation of Nonlinear Parameters1963 · 30,629 citations
  2. 2Confidence measures for large vocabulary continuous speech recognition2001 · 432 citations
  3. 3Computation of normalized edit distance and applications1993 · 339 citations
  4. 4Obtaining Well Calibrated Probabilities Using Bayesian Binning2015 · 979 citations
  5. 5Note on the R2 measure of goodness of fit for nonlinear models1983 · 75 citations