Date of Award

5-2018

Degree Type

Dissertation

Degree Name

Doctor of Philosophy (PhD)

Department

Computer Science

Committee Chair

David F. Gleich

Committee Member 1

Ananth Grama

Committee Member 2

Bruno Ribeiro

Committee Member 3

Hemanta K. Maji

Abstract

Markov random walk models are powerful analytical tools for multiple areas in machine learning, numerical optimizations and data mining tasks. The key assumption of a frst-order Markov chain is memorylessness, which restricts the dependence of the transition distribution to the current state only. However in many applications, this assumption is not appropriate. We propose a set of higher-order random walk techniques and discuss their applications to tensor co-clustering, user trails modeling, and solving linear systems. First, we develop a new random walk model that we call the super-spacey random surfer, which simultaneously clusters the rows, columns, and slices of a nonnegative three-mode tensor. This algorithm generalizes to tensors with any number of modes. We partition the tensor by minimizing the exit probability between clusters when the super-spacey random walk is at stationary. The second application is user trails modeling, where user trails record sequences of activities when individuals interact with the Internet and the world. We propose the retrospective higher-order Markov process as a two-step process by frst choosing a state from the history and then transitioning as a frst-order chain conditional on that state. This way the total number of parameters is restricted and thus the model is protected from overftting. Lastly we propose to use a time-inhomogeneous Markov chain to approximate the solution of a linear system. Multiple simulations of the random walk are conducted to approximate the solution. By allowing the random walk to transition based on multiple matrices, we decrease the variance of the simulations, and thus increase the speed of the solver.

Recommended Citation

Wu, Tao, "Higher-order Random Walk Methods for Data Analysis" (2018). Open Access Dissertations. 1845.
https://docs.lib.purdue.edu/open_access_dissertations/1845

Download

COinS

Open Access Dissertations

Higher-order Random Walk Methods for Data Analysis

Date of Award

Degree Type

Degree Name

Department

Committee Chair

Committee Member 1

Committee Member 2

Committee Member 3

Abstract

Recommended Citation

Search

Links

Links for Authors

Browse

Open Access Dissertations

Higher-order Random Walk Methods for Data Analysis

Author

Date of Award

Degree Type

Degree Name

Department

Committee Chair

Committee Member 1

Committee Member 2

Committee Member 3

Abstract

Recommended Citation

Share

Search

Links

Links for Authors

Browse