Linear and Sublinear Time Algorithms for Mining Frequent Traversal Path Patterns from Very Large Web Logs

Zhixiang Chen; Richard H. Fowler; Ada Wai-Chee Fu; Chunyue Wang

doi:10.1109/IDEAS.2003.1214918

Database Engineering and Applications Symposium, International

Linear and Sublinear Time Algorithms for Mining Frequent Traversal Path Patterns from Very Large Web Logs

Year: 2003, Pages: 117

DOI Bookmark: 10.1109/IDEAS.2003.1214918

Authors

Zhixiang Chen, University of Texas-Pan American
Richard H. Fowler, University of Texas-Pan American
Ada Wai-Chee Fu, Chinese University of Hong Kong
Chunyue Wang, University of Texas-Pan American

Abstract

This paper aims for designing algorithms for the problem of mining frequent traversal path patterns from very large Web logs with best possible efficiency. We devise two algorithms for this problem with the help of fast construction of "shallow" generalized suffix trees over a very large alphabet. These two algorithms have respectively provable linear time and sublinear complexity, and their performance is analyzed in comparison with the two apriori-like algorithms in [4] and the well-known Ukkonen algorithm for on-line suffix tree construction [13]. It is shown that these two algorithms are substantially efficient than the two apriori-like algorithms and the Ukkonen algorithm. The linear time algorithm has optimal performance in theory, while the sublinear time algorithm has better empirical performance.

Like what you’re reading?

Already a member?

Get this article FREE with a new membership!

Sublinear Recovery of Sparse Wavelet Signals
2008 Data Compression Conference
Visual Interface for Online Watching of Frequent Itemset Generation in Apriori and Eclat
Fourth International Conference on Machine Learning and Applications
Linear Algorithms in Sublinear Timeâ€”a Tutorial on Statistical Estimation
IEEE Computer Graphics and Applications
Metric Sublinear Algorithms via Linear Sampling
2018 IEEE 59th Annual Symposium on Foundations of Computer Science (FOCS)
Efficient Discovery of Weighted Frequent Itemsets in Very Large Transactional Databases: A Re-visit
2018 IEEE International Conference on Big Data (Big Data)
SubLinearForce: Fully Sublinear-Time Force Computation for Large Complex Graph Drawing
IEEE Transactions on Visualization & Computer Graphics
A Novel Method to Generate Frequent Itemsets in Distributed Environment
2018 IEEE 37th International Performance Computing and Communications Conference (IPCCC)
Sublinear Algorithms for Gap Edit Distance
2019 IEEE 60th Annual Symposium on Foundations of Computer Science (FOCS)
Sublinear-time Algorithms for Stress Minimization in Graph Drawing
2021 IEEE 14th Pacific Visualization Symposium (PacificVis)
Sublinear-Time Attraction Force Computation for Large Complex Graph Drawing
2021 IEEE 14th Pacific Visualization Symposium (PacificVis)

Linear and Sublinear Time Algorithms for Mining Frequent Traversal Path Patterns from Very Large Web Logs

Authors

Abstract

Related Articles