Bayesian Network Classifier Action Set: Syntax
Provides actions for performing classification using Bayesian network models
bnet Action
Bayesian Network Classifier Action.
CASL Syntax
Parameter Descriptions
alpha={double-1 <, double-2, ...>}
specifies the significance level for independence tests by using chi-square or G-square statistics. If you want to choose the best model among several, you can specify up to five numbers, separated by spaces. If you specify multiple numbers but you do not specify the value True for the bestModel parameter, the action uses the first number and ignores the remaining numbers.
| Requirement | The specified values must be unique. |
attributes={{casinvardesc-1} <, {casinvardesc-2}, ...>}
changes the attributes of variables that are used in this action.
For more information about specifying the attributes parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
bestModel=TRUE | FALSE
when set to True, selects the best model.
| Default | FALSE |
code={aircodegen}
For more information about specifying the code parameter, see the common aircodegen parameter (Appendix A: Common Parameters).
codeGroup="string"
Code Group
diagnostics={_diagnostics}
eyecatcher="string"
specifies a quoted string that will be prefixed to any messages that are associated with this action invocation.
display={displayTables}
specifies a list of results tables to send to the client for display.
For more information about specifying the display parameter, see the common displayTables parameter (Appendix A: Common Parameters).
freq="string"
specifies the frequency variable.
id={"variable-name-1" <, "variable-name-2", ...>}
specifies the variables to copy to the generated table.
indepTest="ALL" | "CHIGSQUARE" | "CHISQUARE" | "GSQUARE" | "MI"
specifies the method for independence tests.
| Default | CHIGSQUARE |
ALL
uses the chi-square statistic, the G-square statistic, and the normalized mutual information for independence tests. A variable is independent of the target if the p-values of both the chi-square and the G-square statistics are greater than the value of the alpha parameter and the normalized mutual information is less than the value of the miAlpha parameter.
CHIGSQUARE
uses both the chi-square and G-square statistics for independence tests. A variable is independent of the target if the p-values of both the chi-square and G-square statistics are greater than the value of the alpha parameter.
CHISQUARE
uses the chi-square statistic for independence tests. A variable is independent of the target if the p-value of the statistic is greater than the value of the alpha parameter.
inputs={{casinvardesc-1} <, {casinvardesc-2}, ...>}
specifies variables to use for analysis.
For more information about specifying the inputs parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
maxParents=integer
specifies the maximum number of parents allowed for each node in the network.
| Default | 5 |
| Range | 1–16 |
miAlpha=double
specifies the significance level for independence tests that use mutual information.
| Default | 0.05 |
| Range | 0–1 |
missingInt="IGNORE" | "IMPUTE"
missingNom="IGNORE" | "IMPUTE" | "LEVEL"
specifies how to handle missing values for nominal variables.
| Default | IGNORE |
nominals={{casinvardesc-1} <, {casinvardesc-2}, ...>}
specifies nominal variables to use for analysis.
For more information about specifying the nominals parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
numBin=integer
specifies the binning number for interval variables.
| Default | 5 |
| Range | 2–1024 |
outNetwork={casouttable}
specifies the name of the output table for the network structure and the probability distributions.
For more information about specifying the outNetwork parameter, see the common casouttable parameter (Appendix A: Common Parameters).
output={BnetOutputStatement}
creates an output table to contain the predicted target values of the input table.
* casOut={casouttable}
specifies the settings for an output table.
caslib="string"
specifies the name of the caslib to use.
compress=TRUE | FALSE
when set to True, data compression is applied to the table.
| Default | FALSE |
indexVars={"variable-name-1" <, "variable-name-2", ...>}
specifies the list of variables to create indexes for in the output data
label="string"
specifies the descriptive label to associate with the table.
maxMemSize=64-bit-integer
specifies the maximum amount of memory, in bytes, that each thread should allocate for in-memory blocks before converting to a memory-mapped file. Files are written in the directories that are specified in the CAS_DISK_CACHE environment variable.
| TIP | You can enclose the value in quotation marks and specify B, K, M, G, or T as a suffix to indicate the units. For example, "8M" specifies eight megabytes. |
name="table-name"
specifies the name to associate with the table.
onDemand=TRUE | FALSE
This parameter is deprecated.
| Default | TRUE |
promote=TRUE | FALSE
when set to True, adds the output table with a global scope. This enables other sessions to access the table, subject to access controls. The target caslib must also have a global scope.
| Default | FALSE |
replace=TRUE | FALSE
when set to True, overwrites an existing table with the same name.
| Default | FALSE |
replication=integer
specifies the number of copies of the table to make for fault tolerance. Larger values result in slower performance and use more memory, but provide high availability for data in the event of a node failure.
| Default | 1 |
| Minimum value | 0 |
timeStamp="string"
specifies the timestamp to apply to the table. Specify the value in the form that is appropriate for your session locale.
where={"string-1" <, "string-2", ...>}
specifies an expression for subsetting the output data.
copyVars="ALL" | "ALL_MODEL" | "ALL_NUMERIC" | {"variable-name-1" <, "variable-name-2", ...>}
specifies a list of one or more variables to be copied from the input table to the output table. You can alternatively specify the value ALL, ALL_MODEL, or ALL_NUMERIC, which respectively copies all variables, all variables used in the modeling, or all numeric variables from the input table to the output table.
role="string"
renames the generated column _ROLE_ in the output data table to the specified role name.
outputTables={outputTables}
lists the names of results tables to save as CAS tables on the server.
For more information about specifying the outputTables parameter, see the common outputTables parameter (Appendix A: Common Parameters).
parenting={"BESTONE", "BESTSET"}
specifies the structure learning methods. If you want the action to choose between the two methods, you can specify both BESTONE and BESTSET and also specify the value True for the bestModel parameter. If you specify both methods but you do not specify the value True for the bestModel parameter, the action uses the first specified method and ignores the other.
| Default | BESTSET |
partByFrac={partByFracStatement}
| The partByFracStatement value can be one or more of the following: |
seed=integer
specifies the seed to use in the random number generator that is used for partitioning the data.
| Default | 0 |
test=double
randomly assigns the specified proportion of observations in the input table to the testing role. The sum of the fractions that are specified in the test and validate parameters must be less than 1.
| Range | 0–1 |
validate=double
randomly assigns the specified proportion of observations in the input table to the validation role. The sum of the fractions that are specified in the test and validate parameters must be less than 1.
| Alias | valid |
| Range | 0–1 |
partByVar={partByVarStatement}
| Long form | partByVar={name="variable-name"} |
| Shortcut form | partByVar="variable-name" |
| The partByVarStatement value can be one or more of the following: |
* name="variable-name"
names the variable in the input table whose values are used to assign rows to each observation.
test="string"
specifies the formatted value of the variable that is used to assign observations to the testing role.
train="string"
specifies the formatted value of the variable that is used to assign observations to the training role. If you do not specify the train parameter, then all observations whose roles are not determined by the test and validate parameters are assigned to training.
validate="string"
specifies the formatted value of the variable that is used to assign observations to the validation role.
| Alias | valid |
preScreening={"ONE", "ZERO"}
specifies the initial screening for the input variables. If you want the action to choose the best model with or without prescreening, you can specify {"ZERO","ONE"} or {"ONE","ZERO"} for the parameter and also specify the value True for the bestModel parameter. If you specify both ONE and ZERO but you do not specify the value True for the bestModel parameter, the action uses the first specified value and ignores the other.
| Default | ONE |
| Requirement | The specified values must be unique. |
printtarget=TRUE | FALSE
when set to True, generates names for the predicted target variable and the predicted probability variables.
| Default | FALSE |
resident=TRUE | FALSE
| Default | TRUE |
saveState={casouttable}
specifies the table in which to save the model for future scoring.
| Long form | saveState={name="table-name"} |
| Shortcut form | saveState="table-name" |
caslib="string"
specifies the name of the caslib to use.
name="table-name"
specifies the name to associate with the table.
promote=TRUE | FALSE
when set to True, adds the output table with a global scope. This enables other sessions to access the table, subject to access controls. The target caslib must also have a global scope.
| Default | FALSE |
replace=TRUE | FALSE
when set to True, overwrites an existing table with the same name.
| Default | FALSE |
structures={"MB", "NAIVE", "PC", "TAN"}
specifies the network structure types. Together with the maxParents parameter, this parameter determines which network structure the action learns from the training data. If you want the action to choose the best structure among several structures, you can specify multiple values in any combination, separated by spaces, and also specify the value True for the bestModel parameter. If you specify multiple structures but you do not specify the value True for the bestModel parameter, the first value that you specify is used and the rest are ignored.
| Alias | structure |
| Default | PC |
| Requirement | The specified values must be unique. |
MB
learns the Markov blanket of the target variable. The Markov blanket includes the parents, the children, and the other parents of the children. After learning the Markov blanket, the action further determines the parents of the target, the links from the parents to the children, and the links among the children. When you specify the value MB for the structure parameter, the action learns the Markov blanket regardless of the values of the preScreening and the varSelect parameters.
NAIVE
learns a naive Bayesian network structure (that is, the target has a direct link to each input variable). If you specify the value 1 for maxParents, the structure being trained is a naive Bayesian network. If you specify a value greater than 1 for maxParents, the structure is a Bayesian network-augmented naive Bayesian network.
* table={castable}
specifies the settings for an input table.
| Long form | table={name="table-name"} |
| Shortcut form | table="table-name" |
| The castable value can be one or more of the following: |
caslib="string"
specifies the caslib that contains the table that you want to use with the action. By default, the active caslib is used. Specify a value only if you need to access a table from a different caslib.
computedOnDemand=TRUE | FALSE
when set to True, creates the computed variables when the table is loaded instead of when the action begins.
| Alias | compOnDemand |
| Default | FALSE |
computedVars={{casinvardesc-1} <, {casinvardesc-2}, ...>}
specifies the names of the computed variables to create. Specify an expression for each variable in the computedVarsProgram parameter.
| Alias | compVars |
format="string"
specifies the format to apply to the variable.
formattedLength=integer
specifies the length of format field plus the format precision.
label="string"
specifies the descriptive label for the variable.
* name="variable-name"
specifies the name for the variable.
nfd=integer
specifies the length of the format precision.
nfl=integer
specifies the length of the format field.
computedVarsProgram="string"
specifies an expression for each computed variable that you include in the computedVars parameter.
| Alias | compPgm |
dataSourceOptions={key-1=any-list-or-data-type-1 <, key-2=any-list-or-data-type-2, ...>}
specifies data source options.
| Alias | options, dataSource |
groupBy={{casinvardesc-1} <, {casinvardesc-2}, ...>}
specifies the names of the variables to use for grouping results.
format="string"
specifies the format to apply to the variable.
formattedLength=integer
specifies the length of format field plus the format precision.
label="string"
specifies the descriptive label for the variable.
* name="variable-name"
specifies the name for the variable.
nfd=integer
specifies the length of the format precision.
nfl=integer
specifies the length of the format field.
groupByMode="NOSORT" | "REDISTRIBUTE"
importOptions={fileType="AUTO" | "BASESAS" | "CSV" | "DOCUMENT" | "DTA" | "ESP" | "EXCEL" | "FMT" | "HDAT" | "JMP" | "LASR" | "SPSS" | "XLS", fileType-specific-parameters}
specifies the settings for reading a table from a data source.
| Alias | import |
The value that you specify for fileType determines the other parameters that apply. For more information about this common parameter, see importOptions (Appendix A: Common Parameters).
* name="table-name"
specifies the name of the table to use.
onDemand=TRUE | FALSE
This parameter is deprecated.
| Default | TRUE |
orderBy={{casinvardesc-1} <, {casinvardesc-2}, ...>}
specifies the variables to use for ordering observations within partitions. This parameter applies to partitioned tables, or it can be combined with variables that are specified in the groupBy parameter when the value of the groupByMode parameter is set to REDISTRIBUTE.
For more information about specifying the orderBy parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
singlePass=TRUE | FALSE
when set to True, does not create a transient table on the server. Setting this parameter to True can be efficient, but the data might not have stable ordering upon repeated runs.
| Default | FALSE |
vars={{casinvardesc-1} <, {casinvardesc-2}, ...>}
specifies the variables to use in the action.
For more information about specifying the vars parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
where="where-expression"
specifies an expression for subsetting the input data.
target="string"
specifies the target variable to use for analysis.
varSelect={"ONE", "THREE", "TWO", "ZERO"}
specifies how input variables are selected beyond prescreening. If you specify the value "ONE", "TWO", or "THREE", the action automatically tests each input variable for unconditional independence of the target regardless of the value of the preScreening parameter. If no variables are left at a particular variable selection level, the action rolls back to the previous level. For example, if you specify "THREE" and there are no variables in the Markov blanket of the target, the action uses the variables from the previous level "TWO". If you want to choose the best model among different levels of variable selection, you can specify any combination of values for this parameter and also specify the value True for the bestModel parameter. If you specify multiple values for the varSelect parameter but you do not specify the value True for the bestModel parameter, the action uses the first specified value and ignores the remaining values.
| Default | ONE |
| Requirement | The specified values must be unique. |
ONE
tests each input variable for conditional independence of the target variable given any other input variable. This type of selection uses only the variables that are conditionally dependent on the target given any other input variable.
THREE
determines the Markov blanket of the target variable and uses only the variables in the Markov blanket.
bnet Action
Bayesian Network Classifier Action.
Lua Syntax
Parameter Descriptions
alpha={double-1 <, double-2, ...>}
specifies the significance level for independence tests by using chi-square or G-square statistics. If you want to choose the best model among several, you can specify up to five numbers, separated by spaces. If you specify multiple numbers but you do not specify the value True for the bestModel parameter, the action uses the first number and ignores the remaining numbers.
| Requirement | The specified values must be unique. |
attributes={{casinvardesc-1} <, {casinvardesc-2}, ...>}
changes the attributes of variables that are used in this action.
For more information about specifying the attributes parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
bestModel=true | false
when set to True, selects the best model.
| Default | false |
code={aircodegen}
For more information about specifying the code parameter, see the common aircodegen parameter (Appendix A: Common Parameters).
codeGroup="string"
Code Group
diagnostics={_diagnostics}
eyecatcher="string"
specifies a quoted string that will be prefixed to any messages that are associated with this action invocation.
display={displayTables}
specifies a list of results tables to send to the client for display.
For more information about specifying the display parameter, see the common displayTables parameter (Appendix A: Common Parameters).
freq="string"
specifies the frequency variable.
id={"variable-name-1" <, "variable-name-2", ...>}
specifies the variables to copy to the generated table.
indepTest="ALL" | "CHIGSQUARE" | "CHISQUARE" | "GSQUARE" | "MI"
specifies the method for independence tests.
| Default | CHIGSQUARE |
ALL
uses the chi-square statistic, the G-square statistic, and the normalized mutual information for independence tests. A variable is independent of the target if the p-values of both the chi-square and the G-square statistics are greater than the value of the alpha parameter and the normalized mutual information is less than the value of the miAlpha parameter.
CHIGSQUARE
uses both the chi-square and G-square statistics for independence tests. A variable is independent of the target if the p-values of both the chi-square and G-square statistics are greater than the value of the alpha parameter.
CHISQUARE
uses the chi-square statistic for independence tests. A variable is independent of the target if the p-value of the statistic is greater than the value of the alpha parameter.
inputs={{casinvardesc-1} <, {casinvardesc-2}, ...>}
specifies variables to use for analysis.
For more information about specifying the inputs parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
maxParents=integer
specifies the maximum number of parents allowed for each node in the network.
| Default | 5 |
| Range | 1–16 |
miAlpha=double
specifies the significance level for independence tests that use mutual information.
| Default | 0.05 |
| Range | 0–1 |
missingInt="IGNORE" | "IMPUTE"
missingNom="IGNORE" | "IMPUTE" | "LEVEL"
specifies how to handle missing values for nominal variables.
| Default | IGNORE |
nominals={{casinvardesc-1} <, {casinvardesc-2}, ...>}
specifies nominal variables to use for analysis.
For more information about specifying the nominals parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
numBin=integer
specifies the binning number for interval variables.
| Default | 5 |
| Range | 2–1024 |
outNetwork={casouttable}
specifies the name of the output table for the network structure and the probability distributions.
For more information about specifying the outNetwork parameter, see the common casouttable parameter (Appendix A: Common Parameters).
output={BnetOutputStatement}
creates an output table to contain the predicted target values of the input table.
* casOut={casouttable}
specifies the settings for an output table.
caslib="string"
specifies the name of the caslib to use.
compress=true | false
when set to True, data compression is applied to the table.
| Default | false |
indexVars={"variable-name-1" <, "variable-name-2", ...>}
specifies the list of variables to create indexes for in the output data
label="string"
specifies the descriptive label to associate with the table.
maxMemSize=64-bit-integer
specifies the maximum amount of memory, in bytes, that each thread should allocate for in-memory blocks before converting to a memory-mapped file. Files are written in the directories that are specified in the CAS_DISK_CACHE environment variable.
| TIP | You can enclose the value in quotation marks and specify B, K, M, G, or T as a suffix to indicate the units. For example, "8M" specifies eight megabytes. |
name="table-name"
specifies the name to associate with the table.
onDemand=true | false
This parameter is deprecated.
| Default | true |
promote=true | false
when set to True, adds the output table with a global scope. This enables other sessions to access the table, subject to access controls. The target caslib must also have a global scope.
| Default | false |
replace=true | false
when set to True, overwrites an existing table with the same name.
| Default | false |
replication=integer
specifies the number of copies of the table to make for fault tolerance. Larger values result in slower performance and use more memory, but provide high availability for data in the event of a node failure.
| Default | 1 |
| Minimum value | 0 |
timeStamp="string"
specifies the timestamp to apply to the table. Specify the value in the form that is appropriate for your session locale.
where={"string-1" <, "string-2", ...>}
specifies an expression for subsetting the output data.
copyVars="ALL" | "ALL_MODEL" | "ALL_NUMERIC" | {"variable-name-1" <, "variable-name-2", ...>}
specifies a list of one or more variables to be copied from the input table to the output table. You can alternatively specify the value ALL, ALL_MODEL, or ALL_NUMERIC, which respectively copies all variables, all variables used in the modeling, or all numeric variables from the input table to the output table.
role="string"
renames the generated column _ROLE_ in the output data table to the specified role name.
outputTables={outputTables}
lists the names of results tables to save as CAS tables on the server.
For more information about specifying the outputTables parameter, see the common outputTables parameter (Appendix A: Common Parameters).
parenting={"BESTONE", "BESTSET"}
specifies the structure learning methods. If you want the action to choose between the two methods, you can specify both BESTONE and BESTSET and also specify the value True for the bestModel parameter. If you specify both methods but you do not specify the value True for the bestModel parameter, the action uses the first specified method and ignores the other.
| Default | BESTSET |
partByFrac={partByFracStatement}
| The partByFracStatement value can be one or more of the following: |
seed=integer
specifies the seed to use in the random number generator that is used for partitioning the data.
| Default | 0 |
test=double
randomly assigns the specified proportion of observations in the input table to the testing role. The sum of the fractions that are specified in the test and validate parameters must be less than 1.
| Range | 0–1 |
validate=double
randomly assigns the specified proportion of observations in the input table to the validation role. The sum of the fractions that are specified in the test and validate parameters must be less than 1.
| Alias | valid |
| Range | 0–1 |
partByVar={partByVarStatement}
| Long form | partByVar={name="variable-name"} |
| Shortcut form | partByVar="variable-name" |
| The partByVarStatement value can be one or more of the following: |
* name="variable-name"
names the variable in the input table whose values are used to assign rows to each observation.
test="string"
specifies the formatted value of the variable that is used to assign observations to the testing role.
train="string"
specifies the formatted value of the variable that is used to assign observations to the training role. If you do not specify the train parameter, then all observations whose roles are not determined by the test and validate parameters are assigned to training.
validate="string"
specifies the formatted value of the variable that is used to assign observations to the validation role.
| Alias | valid |
preScreening={"ONE", "ZERO"}
specifies the initial screening for the input variables. If you want the action to choose the best model with or without prescreening, you can specify {"ZERO","ONE"} or {"ONE","ZERO"} for the parameter and also specify the value True for the bestModel parameter. If you specify both ONE and ZERO but you do not specify the value True for the bestModel parameter, the action uses the first specified value and ignores the other.
| Default | ONE |
| Requirement | The specified values must be unique. |
printtarget=true | false
when set to True, generates names for the predicted target variable and the predicted probability variables.
| Default | false |
resident=true | false
| Default | true |
saveState={casouttable}
specifies the table in which to save the model for future scoring.
| Long form | saveState={name="table-name"} |
| Shortcut form | saveState="table-name" |
caslib="string"
specifies the name of the caslib to use.
name="table-name"
specifies the name to associate with the table.
promote=true | false
when set to True, adds the output table with a global scope. This enables other sessions to access the table, subject to access controls. The target caslib must also have a global scope.
| Default | false |
replace=true | false
when set to True, overwrites an existing table with the same name.
| Default | false |
structures={"MB", "NAIVE", "PC", "TAN"}
specifies the network structure types. Together with the maxParents parameter, this parameter determines which network structure the action learns from the training data. If you want the action to choose the best structure among several structures, you can specify multiple values in any combination, separated by spaces, and also specify the value True for the bestModel parameter. If you specify multiple structures but you do not specify the value True for the bestModel parameter, the first value that you specify is used and the rest are ignored.
| Alias | structure |
| Default | PC |
| Requirement | The specified values must be unique. |
MB
learns the Markov blanket of the target variable. The Markov blanket includes the parents, the children, and the other parents of the children. After learning the Markov blanket, the action further determines the parents of the target, the links from the parents to the children, and the links among the children. When you specify the value MB for the structure parameter, the action learns the Markov blanket regardless of the values of the preScreening and the varSelect parameters.
NAIVE
learns a naive Bayesian network structure (that is, the target has a direct link to each input variable). If you specify the value 1 for maxParents, the structure being trained is a naive Bayesian network. If you specify a value greater than 1 for maxParents, the structure is a Bayesian network-augmented naive Bayesian network.
* table={castable}
specifies the settings for an input table.
| Long form | table={name="table-name"} |
| Shortcut form | table="table-name" |
| The castable value can be one or more of the following: |
caslib="string"
specifies the caslib that contains the table that you want to use with the action. By default, the active caslib is used. Specify a value only if you need to access a table from a different caslib.
computedOnDemand=true | false
when set to True, creates the computed variables when the table is loaded instead of when the action begins.
| Alias | compOnDemand |
| Default | false |
computedVars={{casinvardesc-1} <, {casinvardesc-2}, ...>}
specifies the names of the computed variables to create. Specify an expression for each variable in the computedVarsProgram parameter.
| Alias | compVars |
format="string"
specifies the format to apply to the variable.
formattedLength=integer
specifies the length of format field plus the format precision.
label="string"
specifies the descriptive label for the variable.
* name="variable-name"
specifies the name for the variable.
nfd=integer
specifies the length of the format precision.
nfl=integer
specifies the length of the format field.
computedVarsProgram="string"
specifies an expression for each computed variable that you include in the computedVars parameter.
| Alias | compPgm |
dataSourceOptions={key-1=any-list-or-data-type-1 <, key-2=any-list-or-data-type-2, ...>}
specifies data source options.
| Alias | options, dataSource |
groupBy={{casinvardesc-1} <, {casinvardesc-2}, ...>}
specifies the names of the variables to use for grouping results.
format="string"
specifies the format to apply to the variable.
formattedLength=integer
specifies the length of format field plus the format precision.
label="string"
specifies the descriptive label for the variable.
* name="variable-name"
specifies the name for the variable.
nfd=integer
specifies the length of the format precision.
nfl=integer
specifies the length of the format field.
groupByMode="NOSORT" | "REDISTRIBUTE"
importOptions={fileType="AUTO" | "BASESAS" | "CSV" | "DOCUMENT" | "DTA" | "ESP" | "EXCEL" | "FMT" | "HDAT" | "JMP" | "LASR" | "SPSS" | "XLS", fileType-specific-parameters}
specifies the settings for reading a table from a data source.
| Alias | import |
The value that you specify for fileType determines the other parameters that apply. For more information about this common parameter, see importOptions (Appendix A: Common Parameters).
* name="table-name"
specifies the name of the table to use.
onDemand=true | false
This parameter is deprecated.
| Default | true |
orderBy={{casinvardesc-1} <, {casinvardesc-2}, ...>}
specifies the variables to use for ordering observations within partitions. This parameter applies to partitioned tables, or it can be combined with variables that are specified in the groupBy parameter when the value of the groupByMode parameter is set to REDISTRIBUTE.
For more information about specifying the orderBy parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
singlePass=true | false
when set to True, does not create a transient table on the server. Setting this parameter to True can be efficient, but the data might not have stable ordering upon repeated runs.
| Default | false |
vars={{casinvardesc-1} <, {casinvardesc-2}, ...>}
specifies the variables to use in the action.
For more information about specifying the vars parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
where="where-expression"
specifies an expression for subsetting the input data.
target="string"
specifies the target variable to use for analysis.
varSelect={"ONE", "THREE", "TWO", "ZERO"}
specifies how input variables are selected beyond prescreening. If you specify the value "ONE", "TWO", or "THREE", the action automatically tests each input variable for unconditional independence of the target regardless of the value of the preScreening parameter. If no variables are left at a particular variable selection level, the action rolls back to the previous level. For example, if you specify "THREE" and there are no variables in the Markov blanket of the target, the action uses the variables from the previous level "TWO". If you want to choose the best model among different levels of variable selection, you can specify any combination of values for this parameter and also specify the value True for the bestModel parameter. If you specify multiple values for the varSelect parameter but you do not specify the value True for the bestModel parameter, the action uses the first specified value and ignores the remaining values.
| Default | ONE |
| Requirement | The specified values must be unique. |
ONE
tests each input variable for conditional independence of the target variable given any other input variable. This type of selection uses only the variables that are conditionally dependent on the target given any other input variable.
THREE
determines the Markov blanket of the target variable and uses only the variables in the Markov blanket.
bnet Action
Bayesian Network Classifier Action.
Python Syntax
Parameter Descriptions
alpha=[double-1 <, double-2, ...>]
specifies the significance level for independence tests by using chi-square or G-square statistics. If you want to choose the best model among several, you can specify up to five numbers, separated by spaces. If you specify multiple numbers but you do not specify the value True for the bestModel parameter, the action uses the first number and ignores the remaining numbers.
| Requirement | The specified values must be unique. |
attributes=[{casinvardesc-1} <, {casinvardesc-2}, ...>]
changes the attributes of variables that are used in this action.
For more information about specifying the attributes parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
bestModel=True | False
when set to True, selects the best model.
| Default | False |
code={aircodegen}
For more information about specifying the code parameter, see the common aircodegen parameter (Appendix A: Common Parameters).
codeGroup="string"
Code Group
diagnostics={_diagnostics}
"eyecatcher":"string"
specifies a quoted string that will be prefixed to any messages that are associated with this action invocation.
display={displayTables}
specifies a list of results tables to send to the client for display.
For more information about specifying the display parameter, see the common displayTables parameter (Appendix A: Common Parameters).
freq="string"
specifies the frequency variable.
id=["variable-name-1" <, "variable-name-2", ...>]
specifies the variables to copy to the generated table.
indepTest="ALL" | "CHIGSQUARE" | "CHISQUARE" | "GSQUARE" | "MI"
specifies the method for independence tests.
| Default | CHIGSQUARE |
ALL
uses the chi-square statistic, the G-square statistic, and the normalized mutual information for independence tests. A variable is independent of the target if the p-values of both the chi-square and the G-square statistics are greater than the value of the alpha parameter and the normalized mutual information is less than the value of the miAlpha parameter.
CHIGSQUARE
uses both the chi-square and G-square statistics for independence tests. A variable is independent of the target if the p-values of both the chi-square and G-square statistics are greater than the value of the alpha parameter.
CHISQUARE
uses the chi-square statistic for independence tests. A variable is independent of the target if the p-value of the statistic is greater than the value of the alpha parameter.
inputs=[{casinvardesc-1} <, {casinvardesc-2}, ...>]
specifies variables to use for analysis.
For more information about specifying the inputs parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
maxParents=integer
specifies the maximum number of parents allowed for each node in the network.
| Default | 5 |
| Range | 1–16 |
miAlpha=double
specifies the significance level for independence tests that use mutual information.
| Default | 0.05 |
| Range | 0–1 |
missingInt="IGNORE" | "IMPUTE"
missingNom="IGNORE" | "IMPUTE" | "LEVEL"
specifies how to handle missing values for nominal variables.
| Default | IGNORE |
nominals=[{casinvardesc-1} <, {casinvardesc-2}, ...>]
specifies nominal variables to use for analysis.
For more information about specifying the nominals parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
numBin=integer
specifies the binning number for interval variables.
| Default | 5 |
| Range | 2–1024 |
outNetwork={casouttable}
specifies the name of the output table for the network structure and the probability distributions.
For more information about specifying the outNetwork parameter, see the common casouttable parameter (Appendix A: Common Parameters).
output={BnetOutputStatement}
creates an output table to contain the predicted target values of the input table.
* "casOut":{casouttable}
specifies the settings for an output table.
"caslib":"string"
specifies the name of the caslib to use.
"compress":True | False
when set to True, data compression is applied to the table.
| Default | False |
"indexVars":["variable-name-1" <, "variable-name-2", ...>]
specifies the list of variables to create indexes for in the output data
"label":"string"
specifies the descriptive label to associate with the table.
"maxMemSize":64-bit-integer
specifies the maximum amount of memory, in bytes, that each thread should allocate for in-memory blocks before converting to a memory-mapped file. Files are written in the directories that are specified in the CAS_DISK_CACHE environment variable.
| TIP | You can enclose the value in quotation marks and specify B, K, M, G, or T as a suffix to indicate the units. For example, "8M" specifies eight megabytes. |
"name":"table-name"
specifies the name to associate with the table.
"onDemand":True | False
This parameter is deprecated.
| Default | True |
"promote":True | False
when set to True, adds the output table with a global scope. This enables other sessions to access the table, subject to access controls. The target caslib must also have a global scope.
| Default | False |
"replace":True | False
when set to True, overwrites an existing table with the same name.
| Default | False |
"replication":integer
specifies the number of copies of the table to make for fault tolerance. Larger values result in slower performance and use more memory, but provide high availability for data in the event of a node failure.
| Default | 1 |
| Minimum value | 0 |
"timeStamp":"string"
specifies the timestamp to apply to the table. Specify the value in the form that is appropriate for your session locale.
"where":["string-1" <, "string-2", ...>]
specifies an expression for subsetting the output data.
"copyVars":"ALL" | "ALL_MODEL" | "ALL_NUMERIC" | ["variable-name-1" <, "variable-name-2", ...>]
specifies a list of one or more variables to be copied from the input table to the output table. You can alternatively specify the value ALL, ALL_MODEL, or ALL_NUMERIC, which respectively copies all variables, all variables used in the modeling, or all numeric variables from the input table to the output table.
"role":"string"
renames the generated column _ROLE_ in the output data table to the specified role name.
outputTables={outputTables}
lists the names of results tables to save as CAS tables on the server.
For more information about specifying the outputTables parameter, see the common outputTables parameter (Appendix A: Common Parameters).
parenting=["BESTONE", "BESTSET"]
specifies the structure learning methods. If you want the action to choose between the two methods, you can specify both BESTONE and BESTSET and also specify the value True for the bestModel parameter. If you specify both methods but you do not specify the value True for the bestModel parameter, the action uses the first specified method and ignores the other.
| Default | BESTSET |
partByFrac={partByFracStatement}
| The partByFracStatement value can be one or more of the following: |
"seed":integer
specifies the seed to use in the random number generator that is used for partitioning the data.
| Default | 0 |
"test":double
randomly assigns the specified proportion of observations in the input table to the testing role. The sum of the fractions that are specified in the test and validate parameters must be less than 1.
| Range | 0–1 |
"validate":double
randomly assigns the specified proportion of observations in the input table to the validation role. The sum of the fractions that are specified in the test and validate parameters must be less than 1.
| Alias | valid |
| Range | 0–1 |
partByVar={partByVarStatement}
| Long form | partByVar={"name":"variable-name"} |
| Shortcut form | partByVar="variable-name" |
| The partByVarStatement value can be one or more of the following: |
* "name":"variable-name"
names the variable in the input table whose values are used to assign rows to each observation.
"test":"string"
specifies the formatted value of the variable that is used to assign observations to the testing role.
"train":"string"
specifies the formatted value of the variable that is used to assign observations to the training role. If you do not specify the train parameter, then all observations whose roles are not determined by the test and validate parameters are assigned to training.
"validate":"string"
specifies the formatted value of the variable that is used to assign observations to the validation role.
| Alias | valid |
preScreening=["ONE", "ZERO"]
specifies the initial screening for the input variables. If you want the action to choose the best model with or without prescreening, you can specify {"ZERO","ONE"} or {"ONE","ZERO"} for the parameter and also specify the value True for the bestModel parameter. If you specify both ONE and ZERO but you do not specify the value True for the bestModel parameter, the action uses the first specified value and ignores the other.
| Default | ONE |
| Requirement | The specified values must be unique. |
printtarget=True | False
when set to True, generates names for the predicted target variable and the predicted probability variables.
| Default | False |
resident=True | False
| Default | True |
saveState={casouttable}
specifies the table in which to save the model for future scoring.
| Long form | saveState={"name":"table-name"} |
| Shortcut form | saveState="table-name" |
"caslib":"string"
specifies the name of the caslib to use.
"name":"table-name"
specifies the name to associate with the table.
"promote":True | False
when set to True, adds the output table with a global scope. This enables other sessions to access the table, subject to access controls. The target caslib must also have a global scope.
| Default | False |
"replace":True | False
when set to True, overwrites an existing table with the same name.
| Default | False |
structures=["MB", "NAIVE", "PC", "TAN"]
specifies the network structure types. Together with the maxParents parameter, this parameter determines which network structure the action learns from the training data. If you want the action to choose the best structure among several structures, you can specify multiple values in any combination, separated by spaces, and also specify the value True for the bestModel parameter. If you specify multiple structures but you do not specify the value True for the bestModel parameter, the first value that you specify is used and the rest are ignored.
| Alias | structure |
| Default | PC |
| Requirement | The specified values must be unique. |
MB
learns the Markov blanket of the target variable. The Markov blanket includes the parents, the children, and the other parents of the children. After learning the Markov blanket, the action further determines the parents of the target, the links from the parents to the children, and the links among the children. When you specify the value MB for the structure parameter, the action learns the Markov blanket regardless of the values of the preScreening and the varSelect parameters.
NAIVE
learns a naive Bayesian network structure (that is, the target has a direct link to each input variable). If you specify the value 1 for maxParents, the structure being trained is a naive Bayesian network. If you specify a value greater than 1 for maxParents, the structure is a Bayesian network-augmented naive Bayesian network.
* table={castable}
specifies the settings for an input table.
| Long form | table={"name":"table-name"} |
| Shortcut form | table="table-name" |
| The castable value can be one or more of the following: |
"caslib":"string"
specifies the caslib that contains the table that you want to use with the action. By default, the active caslib is used. Specify a value only if you need to access a table from a different caslib.
"computedOnDemand":True | False
when set to True, creates the computed variables when the table is loaded instead of when the action begins.
| Alias | compOnDemand |
| Default | False |
"computedVars":[{casinvardesc-1} <, {casinvardesc-2}, ...>]
specifies the names of the computed variables to create. Specify an expression for each variable in the computedVarsProgram parameter.
| Alias | compVars |
"format":"string"
specifies the format to apply to the variable.
"formattedLength":integer
specifies the length of format field plus the format precision.
"label":"string"
specifies the descriptive label for the variable.
* "name":"variable-name"
specifies the name for the variable.
"nfd":integer
specifies the length of the format precision.
"nfl":integer
specifies the length of the format field.
"computedVarsProgram":"string"
specifies an expression for each computed variable that you include in the computedVars parameter.
| Alias | compPgm |
"dataSourceOptions":{"key-1":{any-list-or-data-type-1} <, "key-2":{any-list-or-data-type-2}, ...>}
specifies data source options.
| Alias | options, dataSource |
"groupBy":[{casinvardesc-1} <, {casinvardesc-2}, ...>]
specifies the names of the variables to use for grouping results.
"format":"string"
specifies the format to apply to the variable.
"formattedLength":integer
specifies the length of format field plus the format precision.
"label":"string"
specifies the descriptive label for the variable.
* "name":"variable-name"
specifies the name for the variable.
"nfd":integer
specifies the length of the format precision.
"nfl":integer
specifies the length of the format field.
"groupByMode":"NOSORT" | "REDISTRIBUTE"
"importOptions":{fileType="AUTO" | "BASESAS" | "CSV" | "DOCUMENT" | "DTA" | "ESP" | "EXCEL" | "FMT" | "HDAT" | "JMP" | "LASR" | "SPSS" | "XLS", fileType-specific-parameters}
specifies the settings for reading a table from a data source.
| Alias | import_ |
The value that you specify for fileType determines the other parameters that apply. For more information about this common parameter, see importOptions (Appendix A: Common Parameters).
* "name":"table-name"
specifies the name of the table to use.
"onDemand":True | False
This parameter is deprecated.
| Default | True |
"orderBy":[{casinvardesc-1} <, {casinvardesc-2}, ...>]
specifies the variables to use for ordering observations within partitions. This parameter applies to partitioned tables, or it can be combined with variables that are specified in the groupBy parameter when the value of the groupByMode parameter is set to REDISTRIBUTE.
For more information about specifying the orderBy parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
"singlePass":True | False
when set to True, does not create a transient table on the server. Setting this parameter to True can be efficient, but the data might not have stable ordering upon repeated runs.
| Default | False |
"vars":[{casinvardesc-1} <, {casinvardesc-2}, ...>]
specifies the variables to use in the action.
For more information about specifying the vars parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
"where":"where-expression"
specifies an expression for subsetting the input data.
target="string"
specifies the target variable to use for analysis.
varSelect=["ONE", "THREE", "TWO", "ZERO"]
specifies how input variables are selected beyond prescreening. If you specify the value "ONE", "TWO", or "THREE", the action automatically tests each input variable for unconditional independence of the target regardless of the value of the preScreening parameter. If no variables are left at a particular variable selection level, the action rolls back to the previous level. For example, if you specify "THREE" and there are no variables in the Markov blanket of the target, the action uses the variables from the previous level "TWO". If you want to choose the best model among different levels of variable selection, you can specify any combination of values for this parameter and also specify the value True for the bestModel parameter. If you specify multiple values for the varSelect parameter but you do not specify the value True for the bestModel parameter, the action uses the first specified value and ignores the remaining values.
| Default | ONE |
| Requirement | The specified values must be unique. |
ONE
tests each input variable for conditional independence of the target variable given any other input variable. This type of selection uses only the variables that are conditionally dependent on the target given any other input variable.
THREE
determines the Markov blanket of the target variable and uses only the variables in the Markov blanket.
bnet Action
Bayesian Network Classifier Action.
R Syntax
Parameter Descriptions
alpha=list(double-1 <, double-2, ...>)
specifies the significance level for independence tests by using chi-square or G-square statistics. If you want to choose the best model among several, you can specify up to five numbers, separated by spaces. If you specify multiple numbers but you do not specify the value True for the bestModel parameter, the action uses the first number and ignores the remaining numbers.
| Requirement | The specified values must be unique. |
attributes=list( list(casinvardesc-1) <, list(casinvardesc-2), ...>)
changes the attributes of variables that are used in this action.
For more information about specifying the attributes parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
bestModel=TRUE | FALSE
when set to True, selects the best model.
| Default | FALSE |
code=list(aircodegen)
For more information about specifying the code parameter, see the common aircodegen parameter (Appendix A: Common Parameters).
codeGroup="string"
Code Group
diagnostics=list(_diagnostics)
eyecatcher="string"
specifies a quoted string that will be prefixed to any messages that are associated with this action invocation.
display=list(displayTables)
specifies a list of results tables to send to the client for display.
For more information about specifying the display parameter, see the common displayTables parameter (Appendix A: Common Parameters).
freq="string"
specifies the frequency variable.
id=list("variable-name-1" <, "variable-name-2", ...>)
specifies the variables to copy to the generated table.
indepTest="ALL" | "CHIGSQUARE" | "CHISQUARE" | "GSQUARE" | "MI"
specifies the method for independence tests.
| Default | CHIGSQUARE |
ALL
uses the chi-square statistic, the G-square statistic, and the normalized mutual information for independence tests. A variable is independent of the target if the p-values of both the chi-square and the G-square statistics are greater than the value of the alpha parameter and the normalized mutual information is less than the value of the miAlpha parameter.
CHIGSQUARE
uses both the chi-square and G-square statistics for independence tests. A variable is independent of the target if the p-values of both the chi-square and G-square statistics are greater than the value of the alpha parameter.
CHISQUARE
uses the chi-square statistic for independence tests. A variable is independent of the target if the p-value of the statistic is greater than the value of the alpha parameter.
inputs=list( list(casinvardesc-1) <, list(casinvardesc-2), ...>)
specifies variables to use for analysis.
For more information about specifying the inputs parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
maxParents=integer
specifies the maximum number of parents allowed for each node in the network.
| Default | 5 |
| Range | 1–16 |
miAlpha=double
specifies the significance level for independence tests that use mutual information.
| Default | 0.05 |
| Range | 0–1 |
missingInt="IGNORE" | "IMPUTE"
missingNom="IGNORE" | "IMPUTE" | "LEVEL"
specifies how to handle missing values for nominal variables.
| Default | IGNORE |
nominals=list( list(casinvardesc-1) <, list(casinvardesc-2), ...>)
specifies nominal variables to use for analysis.
For more information about specifying the nominals parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
numBin=integer
specifies the binning number for interval variables.
| Default | 5 |
| Range | 2–1024 |
outNetwork=list(casouttable)
specifies the name of the output table for the network structure and the probability distributions.
For more information about specifying the outNetwork parameter, see the common casouttable parameter (Appendix A: Common Parameters).
output=list(BnetOutputStatement)
creates an output table to contain the predicted target values of the input table.
* casOut=list(casouttable)
specifies the settings for an output table.
caslib="string"
specifies the name of the caslib to use.
compress=TRUE | FALSE
when set to True, data compression is applied to the table.
| Default | FALSE |
indexVars=list("variable-name-1" <, "variable-name-2", ...>)
specifies the list of variables to create indexes for in the output data
label="string"
specifies the descriptive label to associate with the table.
maxMemSize=64-bit-integer
specifies the maximum amount of memory, in bytes, that each thread should allocate for in-memory blocks before converting to a memory-mapped file. Files are written in the directories that are specified in the CAS_DISK_CACHE environment variable.
| TIP | You can enclose the value in quotation marks and specify B, K, M, G, or T as a suffix to indicate the units. For example, "8M" specifies eight megabytes. |
name="table-name"
specifies the name to associate with the table.
onDemand=TRUE | FALSE
This parameter is deprecated.
| Default | TRUE |
promote=TRUE | FALSE
when set to True, adds the output table with a global scope. This enables other sessions to access the table, subject to access controls. The target caslib must also have a global scope.
| Default | FALSE |
replace=TRUE | FALSE
when set to True, overwrites an existing table with the same name.
| Default | FALSE |
replication=integer
specifies the number of copies of the table to make for fault tolerance. Larger values result in slower performance and use more memory, but provide high availability for data in the event of a node failure.
| Default | 1 |
| Minimum value | 0 |
timeStamp="string"
specifies the timestamp to apply to the table. Specify the value in the form that is appropriate for your session locale.
where=list("string-1" <, "string-2", ...>)
specifies an expression for subsetting the output data.
copyVars="ALL" | "ALL_MODEL" | "ALL_NUMERIC" | list("variable-name-1" <, "variable-name-2", ...>)
specifies a list of one or more variables to be copied from the input table to the output table. You can alternatively specify the value ALL, ALL_MODEL, or ALL_NUMERIC, which respectively copies all variables, all variables used in the modeling, or all numeric variables from the input table to the output table.
role="string"
renames the generated column _ROLE_ in the output data table to the specified role name.
outputTables=list(outputTables)
lists the names of results tables to save as CAS tables on the server.
For more information about specifying the outputTables parameter, see the common outputTables parameter (Appendix A: Common Parameters).
parenting=list("BESTONE", "BESTSET")
specifies the structure learning methods. If you want the action to choose between the two methods, you can specify both BESTONE and BESTSET and also specify the value True for the bestModel parameter. If you specify both methods but you do not specify the value True for the bestModel parameter, the action uses the first specified method and ignores the other.
| Default | BESTSET |
partByFrac=list(partByFracStatement)
| The partByFracStatement value can be one or more of the following: |
seed=integer
specifies the seed to use in the random number generator that is used for partitioning the data.
| Default | 0 |
test=double
randomly assigns the specified proportion of observations in the input table to the testing role. The sum of the fractions that are specified in the test and validate parameters must be less than 1.
| Range | 0–1 |
validate=double
randomly assigns the specified proportion of observations in the input table to the validation role. The sum of the fractions that are specified in the test and validate parameters must be less than 1.
| Alias | valid |
| Range | 0–1 |
partByVar=list(partByVarStatement)
| Long form | partByVar=list(name="variable-name") |
| Shortcut form | partByVar="variable-name" |
| The partByVarStatement value can be one or more of the following: |
* name="variable-name"
names the variable in the input table whose values are used to assign rows to each observation.
test="string"
specifies the formatted value of the variable that is used to assign observations to the testing role.
train="string"
specifies the formatted value of the variable that is used to assign observations to the training role. If you do not specify the train parameter, then all observations whose roles are not determined by the test and validate parameters are assigned to training.
validate="string"
specifies the formatted value of the variable that is used to assign observations to the validation role.
| Alias | valid |
preScreening=list("ONE", "ZERO")
specifies the initial screening for the input variables. If you want the action to choose the best model with or without prescreening, you can specify {"ZERO","ONE"} or {"ONE","ZERO"} for the parameter and also specify the value True for the bestModel parameter. If you specify both ONE and ZERO but you do not specify the value True for the bestModel parameter, the action uses the first specified value and ignores the other.
| Default | ONE |
| Requirement | The specified values must be unique. |
printtarget=TRUE | FALSE
when set to True, generates names for the predicted target variable and the predicted probability variables.
| Default | FALSE |
resident=TRUE | FALSE
| Default | TRUE |
saveState=list(casouttable)
specifies the table in which to save the model for future scoring.
| Long form | saveState=list(name="table-name") |
| Shortcut form | saveState="table-name" |
caslib="string"
specifies the name of the caslib to use.
name="table-name"
specifies the name to associate with the table.
promote=TRUE | FALSE
when set to True, adds the output table with a global scope. This enables other sessions to access the table, subject to access controls. The target caslib must also have a global scope.
| Default | FALSE |
replace=TRUE | FALSE
when set to True, overwrites an existing table with the same name.
| Default | FALSE |
structures=list("MB", "NAIVE", "PC", "TAN")
specifies the network structure types. Together with the maxParents parameter, this parameter determines which network structure the action learns from the training data. If you want the action to choose the best structure among several structures, you can specify multiple values in any combination, separated by spaces, and also specify the value True for the bestModel parameter. If you specify multiple structures but you do not specify the value True for the bestModel parameter, the first value that you specify is used and the rest are ignored.
| Alias | structure |
| Default | PC |
| Requirement | The specified values must be unique. |
MB
learns the Markov blanket of the target variable. The Markov blanket includes the parents, the children, and the other parents of the children. After learning the Markov blanket, the action further determines the parents of the target, the links from the parents to the children, and the links among the children. When you specify the value MB for the structure parameter, the action learns the Markov blanket regardless of the values of the preScreening and the varSelect parameters.
NAIVE
learns a naive Bayesian network structure (that is, the target has a direct link to each input variable). If you specify the value 1 for maxParents, the structure being trained is a naive Bayesian network. If you specify a value greater than 1 for maxParents, the structure is a Bayesian network-augmented naive Bayesian network.
* table=list(castable)
specifies the settings for an input table.
| Long form | table=list(name="table-name") |
| Shortcut form | table="table-name" |
| The castable value can be one or more of the following: |
caslib="string"
specifies the caslib that contains the table that you want to use with the action. By default, the active caslib is used. Specify a value only if you need to access a table from a different caslib.
computedOnDemand=TRUE | FALSE
when set to True, creates the computed variables when the table is loaded instead of when the action begins.
| Alias | compOnDemand |
| Default | FALSE |
computedVars=list( list(casinvardesc-1) <, list(casinvardesc-2), ...>)
specifies the names of the computed variables to create. Specify an expression for each variable in the computedVarsProgram parameter.
| Alias | compVars |
format="string"
specifies the format to apply to the variable.
formattedLength=integer
specifies the length of format field plus the format precision.
label="string"
specifies the descriptive label for the variable.
* name="variable-name"
specifies the name for the variable.
nfd=integer
specifies the length of the format precision.
nfl=integer
specifies the length of the format field.
computedVarsProgram="string"
specifies an expression for each computed variable that you include in the computedVars parameter.
| Alias | compPgm |
dataSourceOptions=list(key-1=list(any-list-or-data-type-1) <, key-2=list(any-list-or-data-type-2), ...>)
specifies data source options.
| Alias | options, dataSource |
groupBy=list( list(casinvardesc-1) <, list(casinvardesc-2), ...>)
specifies the names of the variables to use for grouping results.
format="string"
specifies the format to apply to the variable.
formattedLength=integer
specifies the length of format field plus the format precision.
label="string"
specifies the descriptive label for the variable.
* name="variable-name"
specifies the name for the variable.
nfd=integer
specifies the length of the format precision.
nfl=integer
specifies the length of the format field.
groupByMode="NOSORT" | "REDISTRIBUTE"
importOptions=list(fileType="AUTO" | "BASESAS" | "CSV" | "DOCUMENT" | "DTA" | "ESP" | "EXCEL" | "FMT" | "HDAT" | "JMP" | "LASR" | "SPSS" | "XLS", fileType-specific-parameters)
specifies the settings for reading a table from a data source.
| Alias | import |
The value that you specify for fileType determines the other parameters that apply. For more information about this common parameter, see importOptions (Appendix A: Common Parameters).
* name="table-name"
specifies the name of the table to use.
onDemand=TRUE | FALSE
This parameter is deprecated.
| Default | TRUE |
orderBy=list( list(casinvardesc-1) <, list(casinvardesc-2), ...>)
specifies the variables to use for ordering observations within partitions. This parameter applies to partitioned tables, or it can be combined with variables that are specified in the groupBy parameter when the value of the groupByMode parameter is set to REDISTRIBUTE.
For more information about specifying the orderBy parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
singlePass=TRUE | FALSE
when set to True, does not create a transient table on the server. Setting this parameter to True can be efficient, but the data might not have stable ordering upon repeated runs.
| Default | FALSE |
vars=list( list(casinvardesc-1) <, list(casinvardesc-2), ...>)
specifies the variables to use in the action.
For more information about specifying the vars parameter, see the common casinvardesc parameter (Appendix A: Common Parameters).
where="where-expression"
specifies an expression for subsetting the input data.
target="string"
specifies the target variable to use for analysis.
varSelect=list("ONE", "THREE", "TWO", "ZERO")
specifies how input variables are selected beyond prescreening. If you specify the value "ONE", "TWO", or "THREE", the action automatically tests each input variable for unconditional independence of the target regardless of the value of the preScreening parameter. If no variables are left at a particular variable selection level, the action rolls back to the previous level. For example, if you specify "THREE" and there are no variables in the Markov blanket of the target, the action uses the variables from the previous level "TWO". If you want to choose the best model among different levels of variable selection, you can specify any combination of values for this parameter and also specify the value True for the bestModel parameter. If you specify multiple values for the varSelect parameter but you do not specify the value True for the bestModel parameter, the action uses the first specified value and ignores the remaining values.
| Default | ONE |
| Requirement | The specified values must be unique. |
ONE
tests each input variable for conditional independence of the target variable given any other input variable. This type of selection uses only the variables that are conditionally dependent on the target given any other input variable.
THREE
determines the Markov blanket of the target variable and uses only the variables in the Markov blanket.