Skip to content

aws.glue.list_data_quality_results

Example SQL Queries

SELECT * FROM
aws.glue.list_data_quality_results;

Description

Returns all data quality execution results for your account.

Table Definition

Column NameColumn Data Type
filter Input Column

The filter criteria.

STRUCT(
"data_source" STRUCT(
"glue_table" STRUCT(
"database_name" VARCHAR,
"table_name" VARCHAR,
"catalog_id" VARCHAR,
"connection_name" VARCHAR,
"additional_options" MAP(VARCHAR, VARCHAR)
)
),
"job_name" VARCHAR,
"job_run_id" VARCHAR,
"started_after" TIMESTAMP_S,
"started_before" TIMESTAMP_S
)
Show child fields
filter.data_source

Filter results by the specified data source. For example, retrieving all results for an Glue table.

Show child fields
filter.data_source.glue_table

An Glue table.

Show child fields
filter.data_source.glue_table.additional_options

Additional options for the table. Currently there are two keys supported:

  • pushDownPredicate: to filter on partitions without having to list and read all the files in your dataset.

  • catalogPartitionPredicate: to use server-side partition pruning using partition indexes in the Glue Data Catalog.

filter.data_source.glue_table.catalog_id

A unique identifier for the Glue Data Catalog.

filter.data_source.glue_table.connection_name

The name of the connection to the Glue Data Catalog.

filter.data_source.glue_table.database_name

A database name in the Glue Data Catalog.

filter.data_source.glue_table.table_name

A table name in the Glue Data Catalog.

filter.job_name

Filter results by the specified job name.

filter.job_run_id

Filter results by the specified job run ID.

filter.started_after

Filter results by runs that started after this time.

filter.started_before

Filter results by runs that started before this time.

max_results Input Column

The maximum number of results to return.

BIGINT
next_token Input Column

A pagination token, if more results are available.

VARCHAR
_aws_profile Input Column

The AWS profile defines the AWS identity used. It can be defined via credentials or by assuming a IAM role.

STRUCT(
"type" VARCHAR,
"name" VARCHAR,
"account_id" VARCHAR,
"via_profile_name" VARCHAR,
"assumed_role_arn" VARCHAR,
"organization" STRUCT(
"account_name" VARCHAR,
"id" VARCHAR,
"tags" STRUCT(
"key" VARCHAR,
"value" VARCHAR
)[],
"master_account" STRUCT(
"id" VARCHAR,
"email" VARCHAR
),
"parents" STRUCT(
"type" VARCHAR,
"id" VARCHAR,
"name" VARCHAR,
"tags" STRUCT(
"key" VARCHAR,
"value" VARCHAR
)[]
)[]
)
)
Show child fields
_aws_profile.account_id

The AWS account id

_aws_profile.assumed_role_arn

The ARN of the assumed role

_aws_profile.name

The unique name of the profile.

_aws_profile.organization

Information about this profile's membership in the AWS organization.

Show child fields
_aws_profile.organization.account_name

The name of account speciifed by the organization

_aws_profile.organization.id

The organization id

_aws_profile.organization.master_account
Show child fields
_aws_profile.organization.master_account.email

The organization master account email address

_aws_profile.organization.master_account.id

The organization master account id

_aws_profile.organization.parents[]
Show child fields
_aws_profile.organization.parents[].id

The id of the parent

_aws_profile.organization.parents[].name

The name of the parent

_aws_profile.organization.parents[].tags[]
Show child fields
_aws_profile.organization.parents[].tags[].key
_aws_profile.organization.parents[].tags[].value
_aws_profile.organization.parents[].type

The type of parent can be an organization unit or a root

_aws_profile.organization.tags[]
Show child fields
_aws_profile.organization.tags[].key
_aws_profile.organization.tags[].value
_aws_profile.type

The type of profile, either 'credentials' or 'assumed_role'

_aws_profile.via_profile_name

This IAM role for this profile is assumed by first utilizing another profile with this name to obtain credentials.

_aws_region Input Column

The AWS region to use.

VARCHAR
results

A list of DataQualityResultDescription objects.

STRUCT(
"result_id" VARCHAR,
"data_source" STRUCT(
"glue_table" STRUCT(
"database_name" VARCHAR,
"table_name" VARCHAR,
"catalog_id" VARCHAR,
"connection_name" VARCHAR,
"additional_options" MAP(VARCHAR, VARCHAR)
)
),
"job_name" VARCHAR,
"job_run_id" VARCHAR,
"started_on" TIMESTAMP_S
)[]
Show child fields
results[]
Show child fields
results[].data_source

The table name associated with the data quality result.

Show child fields
results[].data_source.glue_table

An Glue table.

Show child fields
results[].data_source.glue_table.additional_options

Additional options for the table. Currently there are two keys supported:

  • pushDownPredicate: to filter on partitions without having to list and read all the files in your dataset.

  • catalogPartitionPredicate: to use server-side partition pruning using partition indexes in the Glue Data Catalog.

results[].data_source.glue_table.catalog_id

A unique identifier for the Glue Data Catalog.

results[].data_source.glue_table.connection_name

The name of the connection to the Glue Data Catalog.

results[].data_source.glue_table.database_name

A database name in the Glue Data Catalog.

results[].data_source.glue_table.table_name

A table name in the Glue Data Catalog.

results[].job_name

The job name associated with the data quality result.

results[].job_run_id

The job run ID associated with the data quality result.

results[].result_id

The unique result ID for this data quality result.

results[].started_on

The time that the run started for this data quality result.