Skip to main content
Kinetica can query tables managed in an Apache Iceberg data lake, reading them through the Iceberg catalog that owns them. Access is read-only; an Iceberg table is queried through a logical external table. This page covers the behavior specific to Iceberg. For the catalog & external table concepts, the data type mapping, and the metadata cache settings shared with Delta Lake, see Data Lake Catalogs.

Supported Features

Queries always read the table’s current snapshot; see Snapshot & Version Selection.

Catalog Paths

An Iceberg CATALOG PATH is namespace-qualified and is split at the last dot, so both two-level and three-level paths are accepted:
Iceberg Catalog Paths
An unqualified name, with no dot, is an error.

Catalogs

A catalog holds the location of, and connection information for, a data lake catalog that is external to the database. For Iceberg, a catalog is created with a TABLE FORMAT of iceberg and one of the following catalog types:
Hive Metastore, JDBC, and filesystem/Hadoop catalogs are not supported.
Catalogs are created via the /create/catalog native API call.

REST Catalog

A REST catalog takes the catalog URI as its LOCATION. The referenced data source supplies the connection details for the object store holding the data files.

Glue Catalog

For an AWS Glue Data Catalog, the credential and the data source have distinct roles:
  • the credential carries the Glue API keys—either static access keys (glue.access-key-id & glue.secret-access-key, with an optional glue.session-token) or an IAM role to assume (glue.role-arn)
  • the data source carries the s3.* keys used to read the data files
The region & warehouse options are required; LOCATION is an optional Glue endpoint override.
Kinetica requires either static access keys or a role ARN for Glue; the ambient AWS credential chain is not used.

Glue Credential Vending

Setting the access_delegation option to vended_credentials has Glue issue short-lived credentials for reading the data files, rather than using those on the data source. A vending catalog does not require a data source; if one is given, its s3.* properties take precedence. AWS Lake Formation is supported; deployments without it are unaffected.

Glue Catalog Options