摘要 |
A method of standardizing address data in a database using a word dictionary and a pattern dictionary. The method includes the steps of a) breaking up a set of address data into lines; b) breaking up each line into words; c) looking up each word in the word dictionary for identifying the field type of the word; d) forming a line pattern from the field type of the words in the line; e) looking up the line pattern in the pattern dictionary; and f) returning a line pattern to each of the lines in the address data. With each word in each line in a set of address data having a field type assigned thereto, the address components can be easily identified by a machine reading the address data. With the standardized address data, for example, a machine can identify which component of an address is the street name, and which component of a name line is the title of the addressee.
|