The HPCLUS Procedure

SCORE Statement

  • SCORE <options>;

The SCORE statement causes the HPCLUS procedure to write the cluster membership information of each observation to the output data set. This information includes the variables that are specified in the ID statement and two new variables, _CLUSTER_ID_ and _DISTANCE_, which are the ID of the closest cluster and the distance between the observation and the centroid of that cluster, respectively. If you specify STANDARDIZE=RANGE or STANDARDIZE=STD in the PROC HPCLUS statement, then PROC HPCLUS adds another column called _STANDARDIZED_DISTANCE_ which contains the distance between the standardized values of the observation and the standardized values of cluster centroid.

OUT=<libref.>SAS-data-set

specifies the name of the output data set to contain the scored data.

Last updated: May 25, 2022