Your data science team must run interactive SQL queries on large datasets in Apache Parquet format that are stored in a Cloud Storage bucket. The team knows Apache Hive and wants to use its existing HiveQL queries.
You need to provide an environment where the team can execute interactive HiveQL queries directly against the Cloud Storage data while keeping operational overhead to a minimum. What should you do?
Community Discussion
No comments yet. Be the first to start the discussion!
Community Discussion