Efficient data mining for path traversal patterns

Ming-Syan Chen Ming-Syan Chen,P.S Yu,Jong Soo Park Jong Soo Park

doi:10.1109/69.683753

Abstract

The authors explore a new data mining capability that involves mining path traversal patterns in a distributed information-providing environment where documents or objects are linked together to facilitate interactive access. The solution procedure consists of two steps. First, they derive an algorithm to convert the original sequence of log data into a set of maximal forward references. By doing so, one can filter out the effect of some backward references, which are mainly made for ease of traveling and concentrate on mining meaningful user access sequences. Second, they derive algorithms to determine the frequent traversal patterns-i.e., large reference sequences-from the maximal forward references obtained. Two algorithms are devised for determining large reference sequences; one is based on some hashing and pruning techniques, and the other is further improved with the option of determining large reference sequences in batch so as to reduce the number of database scans required. Performance of these two methods is comparatively analyzed. It is shown that the option of selective scan is very advantageous and can lead to prominent performance improvement. Sensitivity analysis on various parameters is conducted.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Efficient data mining for path traversal patterns

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Knowledge and Data Engineering

Lead the way for us

Journal: IEEE Transactions on Knowledge and Data Engineering	Publication Date: Jan 1, 1998
Citations: 526

Similar Papers

Data mining for path traversal patterns in a web environment
Ming-Syan Chen ... P.S Yu
-
Ming-Syan Chen, et. al. Ming-Syan Chen ... P.S Yu
27 May 1996
27 May 1996

DSM-PLW: Single-pass mining of path traversal patterns over streaming Web click-sequences
Hua-Fu Li ... Man-Kwan Shan
Computer Networks | VOL. 50
Hua-Fu Li, et. al.Hua-Fu Li ... Man-Kwan Shan
20 Dec 2005
Computer Networks | VOL. 50

A sliding window method for finding top- k path traversal patterns over streaming Web click-sequences
Hua-Fu Li
Expert Systems With Applications | VOL. 36
Hua-Fu LiHua-Fu Li
13 May 2008
Expert Systems With Applications | VOL. 36

Efficient Web Mining for Traversal Path Patterns
Zhixiang Chen ... Chunyue Wang
-
Zhixiang Chen, et. al.Zhixiang Chen ... Chunyue Wang
01 Jan 2004
01 Jan 2004

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Efficient data mining for path traversal patterns

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Knowledge and Data Engineering