BookmarkSubscribeRSS Feed
Ajay
Fluorite | Level 6

I'm using EM 7.1 and the variable clustering node gives an error if the dataset has over 100,000 observations.  I can sample the dataset, and then cluster my variables, but is there an easy way to pass along only the new clustered variables along with the full dataset (train, validate, test) to the models?

Thanks

1 REPLY 1
jlevine
Fluorite | Level 6

On the options for the node, look under "Stopping Criteria".  There is an option called "Suppress Sampling Warning".  Set it to "Yes" and your problem should be solved.

sas-innovate-2024.png

Don't miss out on SAS Innovate - Register now for the FREE Livestream!

Can't make it to Vegas? No problem! Watch our general sessions LIVE or on-demand starting April 17th. Hear from SAS execs, best-selling author Adam Grant, Hot Ones host Sean Evans, top tech journalist Kara Swisher, AI expert Cassie Kozyrkov, and the mind-blowing dance crew iLuminate! Plus, get access to over 20 breakout sessions.

 

Register now!

How to choose a machine learning algorithm

Use this tutorial as a handy guide to weigh the pros and cons of these commonly used machine learning algorithms.

Find more tutorials on the SAS Users YouTube channel.

Discussion stats
  • 1 reply
  • 1727 views
  • 0 likes
  • 2 in conversation