Crawler generates table with more columns than expected

0

Hi,

I've set up a crawler for a bucket with two datasets. I see that the generated table from the crawler has many columns than expected although I created a classifier which specify the column headers

Is there a way to allow the crawler to generate the table with only the headers in the csv

Thanks in advance for your help.

Kind regards, Sarah

질문됨 2년 전972회 조회
1개 답변
0

Hi, could you please share some additional details? Do the 2 datasets have the same schema? does any of the data sets have more columns than the other? are you expecting one table or 2 tables?

If you expect 2 tables to be cataloged, and the data sets are not too different, you should separate each dataset in its own prefix (folder).

some of the files might be having more columns that you were aware of.

any other details on the classifier and the crawler you created , and on the schema of the 2 datasets may help to provide better guidance.

thank you

AWS
전문가
답변함 2년 전

로그인하지 않았습니다. 로그인해야 답변을 게시할 수 있습니다.

좋은 답변은 질문에 명확하게 답하고 건설적인 피드백을 제공하며 질문자의 전문적인 성장을 장려합니다.

질문 답변하기에 대한 가이드라인

관련 콘텐츠