| Column Name | Column Data Type |
run_id Required Input Column
The unique run identifier associated with this run. | VARCHAR |
_aws_profile Input Column
The AWS profile defines the AWS identity used. It can be defined via credentials or by assuming a IAM role. | STRUCT( "type" VARCHAR, "name" VARCHAR, "account_id" VARCHAR, "via_profile_name" VARCHAR, "assumed_role_arn" VARCHAR, "organization" STRUCT( "account_name" VARCHAR, "id" VARCHAR, "tags" STRUCT( "key" VARCHAR, "value" VARCHAR )[], "master_account" STRUCT( "id" VARCHAR, "email" VARCHAR ), "parents" STRUCT( "type" VARCHAR, "id" VARCHAR, "name" VARCHAR, "tags" STRUCT( "key" VARCHAR, "value" VARCHAR )[] )[] ) ) |
Show child fields- _aws_profile.account_id
The AWS account id
- _aws_profile.assumed_role_arn
The ARN of the assumed role
- _aws_profile.name
The unique name of the profile.
- _aws_profile.organization
Information about this profile's membership in the AWS organization. Show child fields- _aws_profile.organization.account_name
The name of account speciifed by the organization
- _aws_profile.organization.id
The organization id
- _aws_profile.organization.master_account
Show child fields- _aws_profile.organization.master_account.email
The organization master account email address
- _aws_profile.organization.master_account.id
The organization master account id
- _aws_profile.organization.parents[]
Show child fields- _aws_profile.organization.parents[].id
The id of the parent
- _aws_profile.organization.parents[].name
The name of the parent
- _aws_profile.organization.parents[].tags[]
Show child fields- _aws_profile.organization.parents[].tags[].key
- _aws_profile.organization.parents[].tags[].value
- _aws_profile.organization.parents[].type
The type of parent can be an organization unit or a root
- _aws_profile.organization.tags[]
Show child fields- _aws_profile.organization.tags[].key
- _aws_profile.organization.tags[].value
- _aws_profile.type
The type of profile, either 'credentials' or 'assumed_role'
- _aws_profile.via_profile_name
This IAM role for this profile is assumed by first utilizing another profile with this name to obtain credentials.
|
_aws_region Input Column
The AWS region to use. | VARCHAR |
additional_data_sources
A map of reference strings to additional data sources you can specify for an evaluation run. | MAP(VARCHAR, STRUCT( "glue_table" STRUCT( "database_name" VARCHAR, "table_name" VARCHAR, "catalog_id" VARCHAR, "connection_name" VARCHAR, "additional_options" MAP(VARCHAR, VARCHAR) ) )) |
additional_run_options
Additional run options you can specify for an evaluation run. | STRUCT( "cloud_watch_metrics_enabled" BOOLEAN, "results_s3_prefix" VARCHAR, "composite_rule_evaluation_method" VARCHAR ) |
Show child fields- additional_run_options.cloud_watch_metrics_enabled
Whether or not to enable CloudWatch metrics.
- additional_run_options.composite_rule_evaluation_method
Set the evaluation method for composite rules in the ruleset to ROW/COLUMN
- additional_run_options.results_s3_prefix
Prefix for Amazon S3 to store results.
|
completed_on
The date and time when this run was completed. | TIMESTAMP_S |
data_source
The data source (an Glue table) associated with this evaluation run. | STRUCT( "glue_table" STRUCT( "database_name" VARCHAR, "table_name" VARCHAR, "catalog_id" VARCHAR, "connection_name" VARCHAR, "additional_options" MAP(VARCHAR, VARCHAR) ) ) |
Show child fields- data_source.glue_table
An Glue table. Show child fields- data_source.glue_table.additional_options
Additional options for the table. Currently there are two keys supported: -
pushDownPredicate: to filter on partitions without having to list and read all the files in your dataset. -
catalogPartitionPredicate: to use server-side partition pruning using partition indexes in the Glue Data Catalog.
- data_source.glue_table.catalog_id
A unique identifier for the Glue Data Catalog.
- data_source.glue_table.connection_name
The name of the connection to the Glue Data Catalog.
- data_source.glue_table.database_name
A database name in the Glue Data Catalog.
- data_source.glue_table.table_name
A table name in the Glue Data Catalog.
|
error_string
The error strings that are associated with the run. | VARCHAR |
execution_time
The amount of time (in seconds) that the run consumed resources. | BIGINT |
last_modified_on
A timestamp. The last point in time when this data quality rule recommendation run was modified. | TIMESTAMP_S |
number_of_workers
The number of G.1X workers to be used in the run. The default is 5. | BIGINT |
result_ids
A list of result IDs for the data quality results for the run. | VARCHAR[] |
Show child fields- result_ids[]
|
role
An IAM role supplied to encrypt the results of the run. | VARCHAR |
ruleset_names
A list of ruleset names for the run. Currently, this parameter takes only one Ruleset name. | VARCHAR[] |
Show child fields- ruleset_names[]
|
started_on
The date and time when this run started. | TIMESTAMP_S |
status
The status for this run. | VARCHAR |
timeout
The timeout for a run in minutes. This is the maximum time that a run can consume resources before it is terminated and enters TIMEOUT status. The default is 2,880 minutes (48 hours). | BIGINT |