LIME Procedure

PROC LIME Statement

  • PROC LIME DATA=libref.data-table REFERENCEDATA=libref.data-table <options>;

The PROC LIME statement invokes the procedure. Table 1 summarizes the options in this statement.

Table 1: PROC LIME Statement Options

Option Description
DATA= Specifies the query input data table, which contains the query observation
INCLUDEMISSING Specifies that the missing values of nominal variables should be treated as a valid level
LOGLEVEL= Specifies a value to control the quantity and type of notes to print to the client log
NTHREADS= Specifies the number of threads to use in the computation
REFERENCEDATA= Specifies the reference data table
SAMPLESIZE= Specifies the number of observations to generate
SEED= Specifies the seed for generating the pseudorandom numbers


You must specify the following options:

DATA=libref.data-table

names the input data table for PROC LIME to use. The default is the most recently created data table. libref.data-table is a two-level name, where

libref

refers to a collection of information that is defined in the LIBNAME statement and includes the library, which includes a path to the data. For more information about libref, see the section Using SAS Viya Workbench.

data-table

specifies the name of the input data table.

REFERENCEDATA=libref.data-table

specifies the reference data table. libref.data-table is a two-level name, where libref refers to the library, and data-table specifies the name of the input data table. For more information about this two-level name, see the DATA= option and the section Using SAS Viya Workbench.

You can also specify the following options:

INCLUDEMISSING

treats missing values of nominal variables as a valid level.

LOGLEVEL=0 | 1 | 2

controls the quantity and type of notes to print to the client log.

0

prints only warnings and errors.

1

prints some notes.

2

prints many notes.

By default, LOGLEVEL=1.

NTHREADS=number-of-threads

specifies the number of threads to use in the computation. The default value is the number of CPUs available on the machine.

SAMPLESIZE=number

specifies the number of observations to generate. The default is 500 times the number of variables that are listed in the INPUT statements or 1,000,000, whichever value is smaller.

SEED=number

specifies the seed value for pseudorandom number generation. If you do not specify a seed, or if you specify a value less than or equal to 0, the seed is generated by reading the time of day from the computer’s clock.

Last updated: September 23, 2026