arXiv
Open Access
2016
Frequent-Itemset Mining using Locality-Sensitive Hashing
Debajyoti Bera
Rameshwar Pratap
Abstrak
The Apriori algorithm is a classical algorithm for the frequent itemset mining problem. A significant bottleneck in Apriori is the number of I/O operation involved, and the number of candidates it generates. We investigate the role of LSH techniques to overcome these problems, without adding much computational overhead. We propose randomized variations of Apriori that are based on asymmetric LSH defined over Hamming distance and Jaccard similarity.
Topik & Kata Kunci
Penulis (2)
D
Debajyoti Bera
R
Rameshwar Pratap
Akses Cepat
Informasi Jurnal
- Tahun Terbit
- 2016
- Bahasa
- en
- Sumber Database
- arXiv
- Akses
- Open Access ✓