>
Fa   |   Ar   |   En
   Analyzing state sequences with probabilistic suffix trees: the PST R package  
   
نویسنده gabadinho a. ,ritschard g.
منبع journal of statistical software - 2016 - دوره : 72 - شماره : 0
چکیده    This article presents the pst r package for categorical sequence analysis with probabilistic suffix trees (psts),i.e.,structures that store variable-length markov chains (vlmcs). vlmcs allow to model high-order dependencies in categorical sequences with parsimonious models based on simple estimation procedures. the package is specifically adapted to the field of social sciences,as it allows for vlmc models to be learned from sets of individual sequences possibly containing missing values; in addition,the package is extended to account for case weights. this article describes how a vlmc model is learned from one or more categorical sequences and stored in a pst. the pst can then be used for sequence prediction,i.e.,to assign a probability to whole observed or artificial sequences. this feature supports data mining applications such as the extraction of typical patterns and outliers. this article also introduces original visualization tools for both the model and the outcomes of sequence prediction. other features such as functions for pattern mining and artificial sequence generation are described as well. the pst package also allows for the computation of probabilistic divergence between two models and the fitting of segmented vlmcs,where sub-models fitted to distinct strata of the learning sample are stored in a single pst. © 2016,american statistical association. all rights reserved.
کلیدواژه Categorical sequences; Probabilistic suffix trees; R; Sequence data mining; Sequence visualization; State sequences; Variable-length Markov chains
آدرس nccr lives,institute for life course and demographic studies,university of geneva,geneva 4,ch-1211, Switzerland, nccr lives,institute for life course and demographic studies,university of geneva,geneva 4,ch-1211, Switzerland
 
     
   
Authors
  
 
 

Copyright 2023
Islamic World Science Citation Center
All Rights Reserved