2002
Conference article  Restricted

A Metric Index for Approximate Text Management

Dohnal V., Gennaro C., Zezula P.

Text management  Information retrieval  Metric space 

Text collections of data need not only search support for identical objects, but the approximate matching is even more important. A suitable metric to such a task is the edit distance measure. However, the quadratic computational complexity of edit distance prevents from applying naive storage organizations, such as the sequential search, and more sophisticated search structures must be applied. We have investigated the properties of the D-index to approximate searching and matching in text databases. The experiments confirm a very good performance for retrieving close objects and sub-linear scalability to process large files. Even the similarity joins can be performed efficiently.

Source: International Conference Information Systems and Databases (ISDB 2002), pp. 37–42, Tokyo, Japan, 5-27 September 2002



Back to previous page
BibTeX entry
@inproceedings{oai:it.cnr:prodotti:91518,
	title = {A Metric Index for Approximate Text Management},
	author = {Dohnal V. and Gennaro C. and Zezula P.},
	booktitle = {International Conference Information Systems and Databases (ISDB 2002), pp. 37–42, Tokyo, Japan, 5-27 September 2002},
	year = {2002}
}
CNR ExploRA

Bibliographic record

Also available from

dblp.uni-trier.deRestricted