flâneur — a map of the web's best reading

Loading Parquet data from Cloud Storage | BigQuery | Google Cloud

cloud.google.com · 8,208 words · saved by 1 readers

Parquet is an open source column-oriented data format that is widely used in the Apache Hadoop ecosystem. When you load Parquet data from Cloud Storage, you can load the data into a new table or partition, or you can append to or overwrite an existing table or partition. When your data is loaded into BigQuery, it is converted into columnar format for Capacitor (BigQuery's storage format). When you load data from Cloud Storage into a BigQuery table, the dataset that contains the table must be in the same regional or multi- regional location as the Cloud Storage bucket. For information about loading Parquet data from a local file, see Loading data from local files. You are subject to the following limitations when you load data into BigQuery from a Cloud Storage bucket: To avoid resourcesExceeded errors when loading Parquet files into BigQuery, follow these guidelines: Grant Identity and Access Management (IAM) roles that give users the necessary permissions to perform each task in this

Home Documentation Data analytics BigQuery Guides Send feedback Stay organized with collections Save and categorize content based on your preferences. Loading Parquet data from Cloud Storage This page provides an overview of loading Parquet data from Cloud Storage into BigQuery. Parquet is an open source column-oriented data format that is widely used in the Apache Hadoop ecosystem. When you load Parquet data from Cloud Storage, you can load the data into a new table or partition, or you can append to or overwrite an existing table or partition. When your data is loaded into BigQuery, it is co

Explore this link on the map →

related reading