摘要 |
A system for indexing unstructured or semi-structured data is disclosed. The system may identify regions within the data, such as "Abstract" or "References". The system may identify linguistic units such as sentences, noun groups, verb groups. The system may also identify concepts such as companies, people, diseases, amounts, and so forth. The query results may be formatted so that similar results from different documents, or from the same document, are clustered together.
|