arXiv · 1001.0833
Random Indexing K-tree
Abstract
Random Indexing (RI) K-tree is the combination of two algorithms for clustering. Many large scale problems exist in document clustering. RI K-tree scales well with large inputs due to its low complexity. It also exhibits features that are useful for managing a changing collection. Furthermore, it solves previous issues with sparse document vectors when using K-tree. The algorithms and data structures are defined, explained and motivated. Specific modifications to K-tree are made for use with RI. Experiments have been executed to measure quality. The results indicate that RI K-tree improves document cluster quality over the original K-tree algorithm.
Explore related subjects
Keep this discovery
Christopher M. De Vries, Lance De Vine, Shlomo Geva. 2010-02-02. Random Indexing K-tree. https://arxiv.org/abs/1001.0833
Cite the original work for its findings. Save a collection to share your selection of sources.