EP0352011B1

Method for establishing pixel colour probabilities for use in OCR logic

Abstract

This record has no abstract on file.

EP0352011B1, drawing sheet 1
Sheet 1 of 5

Term

Term ended

Expired 12 July 2009, 17.2 years ago.

  1. Priority
  2. Filed
  3. Granted
  4. Expired
  5. Today

14 claims: 4 independent, 10 dependent

  1. 1
    A method of operating an optical character recognition device which uses a decision tree for the recognition of characters in a first font, the device having a memory storing a probability table providing the probability, in a second font different from said first font, that a pixel is a first distinguishable colour as a function of its neighbourhood state, the method comprising the steps of:(i) scanning a document having characters in said first font to generate an array of pixels, each having a neighbourhood state;(ii) identifying in said array a cluster of pixels representing a character;(iii) selecting a pixel in said cluster;(iv) determining the neighbourhood state of said pixel in the exact same way as the neighbourhood state was determined for each pixel in the second font;(v) addressing said memory with the neighbourhood state of said selected pixel;(vi) reading from said memory the probability of said selected pixel being said distinguishable colour, associated with the neighbourhood state of said selected pixel, and assigning said probability to said selected pixel;(vii) repeating steps (iii) to (vi) for further selected pixels in said cluster;(viii)generating a new decision tree for said first font using said assigned probabilities of said distinguishable colour without reliance on pre-existing decision trees, and ;(ix) using the new decision tree for interpreting pixels to recognise characters.
  2. 8
    An optical character recognition device having means for generating a decision tree for interpreting pixels to recognise characters in a first font, the device having a memory storing a probability table providing the probability, in a second font different from said first font, that a pixel is a first distinguishable colour as a function of its neighbourhood state, said decision tree generation means functioning according to the following steps:(i) scanning a document having characters in said first font to generate an array of pixels, each having a neighbourhood state;(ii) identifying in said array a cluster of pixels representing a character;(iii) selecting a pixel in said cluster;(iv) determining the neighbourhood state of said pixel in the exact same way as the neighbourhood state was determined for each pixel in the second font;(v) addressing said memory with the neighbourhood state of said selected pixel;(vi) reading from said memory the probability of said selected pixel being said distinguishable colour, associated with the neighbourhood state of said selected pixel, and assigning said probability to said selected pixel;(vii) repeating steps (iii) to (vi) for further selected pixels in said cluster;and (viii) generating a new decision tree for said first font using said assigned probabilities of said distinguishable colour, without reliance on pre-existing decision trees.
  3. 9
    A method for the recognition of characters in an unknown font by means of an optical character recognition device which uses a decision tree to interpret pixels, the decision tree being created according to the following method steps:scanning a document in a first font to generated an array of pixels;for each pixel in said array, identifying a preselected matrix of neighbour pixels as the pixel's neighbourhood;determining from said pixel array the probability that a selected pixel is a first distinguishable colour, given the neighbourhood state of the pixel;storing a probability table providing, for each pixel neighbourhood state, the probability the selected pixel is said distinguishable colour;scanning a document in an unknown font to generate a second array of pixels;identifying in said array a plurality of clusters of pixels, each cluster being representative of a character;for each of a plurality of pixels in each of said clusters, identifying said preselected matrix of neighbour pixels as the pixel's neighbourhood, establishing the neighbourhood state of that pixel based on said pixel neighbourhood, addressing said table by the neighbourhood state of that pixel, and reading from said table the probability said pixel is said distinguishable colour;and generating a new decision tree for said unknown font using the probabilities from said table for selected pixels, without reliance on pre-existing decision trees.
  4. 12
    A method of controlling an OCR device to recognise characters in an unknown font using a decision tree to interpret pixels for said recognition, the decision tree being generated using probabilities that pixels have a first distinguishable colour, said probabilities being assigned to pixels by the following sequence of steps:(a) scanning a reference document in a first font and converting said scanned reference document into an array of pixels, wherein: (i) each converted pixel has either a first colour or a second colour: and (ii) each converted pixel has a neighbourhood comprising a predetermined number of neighbour pixels of said converted pixel;and (iii) said neighbourhood has a neighbourhood state determined by the colours of the neighbour pixels;(b) determining, for each pixel of said scanned reference document, the colour of that pixel;(c) determining, for each pixel of said scanned reference document, its neighbourhood state;(d) dividing the number of times a pixel in a neighbourhood having a first neighbourhood state has said first colour by the number of times in said scanned document said first neighbourhood state appears in said reference document to obtain a probability that a pixel has said first colour given that its neighbourhood has said first neighbourhood state;(e) repeating step (d) for each possible neighbourhood state to generate a probability table comprising a list of all the possible neighbourhood states and the associated probability of each selected pixel having said first colour;(f) storing said probability table in said OCR device;and (g) scanning a second document in an unknown second font and converting said second document into an array of pixels, wherein: (i) each converted pixel has either a first colour or a second colour;(ii) each converted pixel has a neighbourhood comprising a predetermined number of neighbour pixels of said converted pixel, wherein said neighbourhood comprises neighbour pixels corresponding to the neighbourhood of the pixels of said reference document;and (iii) said neighbourhood has a neighbourhood state determined by the colours of said neighbourhood pixels;(h) determining, for each of a selected plurality of converted pixels in said cluster, the state of its neighbourhood;and (i) for each selected converted pixel, looking up the corresponding neighbourhood state in said stored probability table and reading out from said probability table the probability said pixel has said first colour.