All,
I’m using Proc fastclus in SAS to perform a cluster analysis. I’m trying to figure out a way to determine the relative importance of variables within a cluster. So, what variables are the primary drivers within a cluster or variables have the most predictive power, so to speak. And rank order.
Here is an example .
Using the variables A, B, C, D, E, F I build a cluster model with 3 segments.
proc fastclus data= DATA_SET maxc=3 out=CUSTER_Results ;
var A B C D E F ;run;
I want to rank order the variables by relative importance within each cluster.
Cluster 1
Rank order of variable Importance:
- B – Primary Driver of segment (most predictive)
- D
- A
- C
- E
- F – Least predictive
Cluster 2
Rank order of variable Importance:
- A – Primary Driver of segment (most predictive
- C
- F
- E
- B
- D – Least predictive
Cluster 3
Rank order of variable Importance:
- D – Primary Driver of segment (most predictive
- A
- C
- B
- F
- E – Least predictive
Is there an option for proc fastclus which will do this automatically? In not, any recommendations on how to determine the predictive rank order?
Thanks.