HIVE unknown error when running Athena query on crawler-generated catalog data

0

I ran basic sql in Athena to view the catalog table which was created by the glue-crawler (crawler job ended successfully and created the "metadata" catalog table in the "hw-db" db) : SELECT * FROM "AwsDataCatalog"."hw-db"."metadata" limit 10; and got the following error:

HIVE_UNKNOWN_ERROR: com.amazonaws.services.lakeformation.model.InvalidInputException: Unsupported vendor for Glue supported principal: arn:aws:iam::{...}:root (Service: AWSLakeFormation; Status Code: 400; Error Code: InvalidInputException; Request ID: {...}; Proxy: null)
This query ran against the "hw-db" database, unless qualified by the query. 

any ideas?...

Erez
質問済み 1年前569ビュー
1回答
0

Hi, could you please specify the source you catalogued with the crawler?

Athena uses the AWS Glue Data Catalog to store and retrieve table metadata for the Amazon S3 data in your Amazon Web Services account. The table metadata lets the Athena query engine know how to find, read, and process the data that you want to query. As described here.

If you have catalogued a JDBC database (i.e. mysql, oracle or others) those tables will not be readable by Athena.

Crawling a JDBC source, currently only support accessing the data via Glue ETL. If you need to read data in a remote DB with Athena you may want to consider Athena Federated queries, to read about this feature you can look at this blog post.

hope this helps

AWS
エキスパート
回答済み 1年前
  • The crawler's catalog was created based on jsons in my s3, so athena is expected to have access to the db

ログインしていません。 ログイン 回答を投稿する。

優れた回答とは、質問に明確に答え、建設的なフィードバックを提供し、質問者の専門分野におけるスキルの向上を促すものです。

質問に答えるためのガイドライン

関連するコンテンツ