Bayesian OCR methods can use document context like language to improve character recognition accuracy, but this approach has limitations. When applied at scale to large databases like Google's Ngram, contextual priors can introduce systematic errors—for example, incorrectly recognizing ambiguous characters as English words in English texts, as happened with the word 'grok' appearing before its 1961 coinage.