LIME Procedure
The PROC LIME statement invokes the procedure. Table 1 summarizes the options in this statement.
Table 1: PROC LIME Statement Options
|
Option | Description |
|---|
|
DATA= | Specifies the query input data table, which contains the query observation |
|
INCLUDEMISSING | Specifies that the missing values of nominal variables should be treated as a valid level |
|
LOGLEVEL= | Specifies a value to control the quantity and type of notes to print to the client log |
|
NTHREADS= | Specifies the number of threads to use in the computation |
|
REFERENCEDATA= | Specifies the reference data table |
|
SAMPLESIZE= | Specifies the number of observations to generate |
|
SEED= | Specifies the seed for generating the pseudorandom numbers |
You must specify the following options:
-
DATA=libref.data-table
-
names the input data table for PROC LIME to use. The default is the most recently created data table. libref.data-table is a two-level name, where
- libref
refers to a collection of information that is defined in the LIBNAME statement and includes the library, which includes a path to the data. For more information about libref, see the section Using SAS Viya Workbench.
- data-table
specifies the name of the input data table.
-
REFERENCEDATA=libref.data-table
specifies the reference data table. libref.data-table is a two-level name, where libref refers to the library, and data-table specifies the name of the input data table. For more information about this two-level name, see the DATA= option and the section Using SAS Viya Workbench.
You can also specify the following options:
-
INCLUDEMISSING
treats missing values of nominal variables as a valid level.
-
LOGLEVEL=0 | 1 | 2
-
controls the quantity and type of notes to print to the client log.
- 0
prints only warnings and errors.
- 1
prints some notes.
- 2
prints many notes.
By default, LOGLEVEL=1.
-
NTHREADS=number-of-threads
specifies the number of threads to use in the computation. The default value is the number of CPUs available on the machine.
-
SAMPLESIZE=number
specifies the number of observations to generate. The default is 500 times the number of variables that are listed in the INPUT statements or 1,000,000, whichever value is smaller.
-
SEED=number
specifies the seed value for pseudorandom number generation. If you do not specify a seed, or if you specify a value less than or equal to 0, the seed is generated by reading the time of day from the computer’s clock.
Last updated: September 23, 2026