HIVE unknown error when running Athena query on crawler-generated catalog data

0

I ran basic sql in Athena to view the catalog table which was created by the glue-crawler (crawler job ended successfully and created the "metadata" catalog table in the "hw-db" db) : SELECT * FROM "AwsDataCatalog"."hw-db"."metadata" limit 10; and got the following error:

HIVE_UNKNOWN_ERROR: com.amazonaws.services.lakeformation.model.InvalidInputException: Unsupported vendor for Glue supported principal: arn:aws:iam::{...}:root (Service: AWSLakeFormation; Status Code: 400; Error Code: InvalidInputException; Request ID: {...}; Proxy: null)
This query ran against the "hw-db" database, unless qualified by the query. 

any ideas?...

Erez
feita há um ano590 visualizações
1 Resposta
0

Hi, could you please specify the source you catalogued with the crawler?

Athena uses the AWS Glue Data Catalog to store and retrieve table metadata for the Amazon S3 data in your Amazon Web Services account. The table metadata lets the Athena query engine know how to find, read, and process the data that you want to query. As described here.

If you have catalogued a JDBC database (i.e. mysql, oracle or others) those tables will not be readable by Athena.

Crawling a JDBC source, currently only support accessing the data via Glue ETL. If you need to read data in a remote DB with Athena you may want to consider Athena Federated queries, to read about this feature you can look at this blog post.

hope this helps

AWS
ESPECIALISTA
respondido há um ano
  • The crawler's catalog was created based on jsons in my s3, so athena is expected to have access to the db

Você não está conectado. Fazer login para postar uma resposta.

Uma boa resposta responde claramente à pergunta, dá feedback construtivo e incentiva o crescimento profissional de quem perguntou.

Diretrizes para responder a perguntas