Dataproc

GCPAnalytics

Managed Spark and Hadoop clusters for big data processing with per-second billing, serverless Spark (Dataproc Serverless), Presto, Flink, and ephemeral or long-running clusters

Jurisdictional exposure

Provider HQ
USMountain View, USA

Subject to CLOUD Act, FISA-702, DPF

Region locations
APACCNEEAEUUKUSOther44 regions across 7 jurisdictions
Sovereign option
Yes — 2 sovereign-flagged regions available

Attributes

Auto Scaling
Yes
Spark Support
Yes
Hadoop Support
Yes

Sub-services (3)

Dataproc Clusters

Managed Spark and Hadoop cluster provisioning

Dataproc Serverless

Serverless Spark for batch workloads

Dataproc Metastore

Managed Hive Metastore for metadata management

Compliance & Certifications

This service is attested for the following frameworks. Always verify with the provider before relying on a specific compliance posture.

Where this runs

44 regions
28 countries
2sovereign
Sovereign regions (2)
  • T-Systems Sovereign Cloud · FrankfurtT-Systems Sovereign Cloud powered by Google Cloud
  • S3NS Sovereign Cloud · ParisS3NS — Google Cloud + Thales joint venture
Commercial regions (42)

Europe (13)

  • Belgium
  • Finland
  • Paris
  • Berlin
  • Frankfurt
  • Milan
  • Turin
  • Netherlands
  • Warsaw
  • Madrid
  • Stockholm
  • Zurich
  • London

North America (12)

  • Montréal
  • Toronto
  • Querétaro
  • Northern Virginia
  • Columbus
  • Iowa
  • Dallas
  • Las Vegas
  • Los Angeles
  • South Carolina
  • Salt Lake City
  • Oregon

South America (2)

  • São Paulo
  • Santiago

Asia (9)

  • Hong Kong
  • Delhi
  • Mumbai
  • Jakarta
  • Osaka
  • Tokyo
  • Singapore
  • Seoul
  • Taiwan

Oceania (2)

  • Melbourne
  • Sydney

Middle East (3)

  • Tel Aviv
  • Doha
  • Dammam

Africa (1)

  • Johannesburg

Tags

Equivalent services on other platforms

Alibaba MaxComputeAlibaba

Exabyte-scale fully managed data warehouse (formerly ODPS) optimised for batch analytics on structured and semi-structured data, with SQL, MapReduce, and Spark computation engines, columnar storage, and tight integration with DataWorks

Amazon RedshiftAWS

Petabyte-scale columnar data warehouse with concurrency scaling, federated queries, and a serverless option for unpredictable workloads

Amazon AthenaAWS

Serverless interactive query service that runs ANSI SQL directly on S3 data lakes with per-query pricing, Apache Iceberg and Hudi table formats, and Glue catalog integration

Amazon EMRAWS

Managed big-data platform for running Apache Spark, Hive, Presto, Flink, Trino, and HBase across EC2, EKS, and fully serverless deployments with up to 5x faster Spark runtime and Graviton price-performance

Amazon QuickSightAWS

Serverless cloud-native BI service with interactive dashboards, paginated reports, ML-powered insights, and a natural-language Q&A experience embedded via generative BI, plus per-session pricing for reader-scale deployments

Azure Synapse AnalyticsAzure

Unified analytics service combining data warehousing and big data processing with dedicated and serverless SQL pools, Apache Spark, Data Integration pipelines, and Power BI embedded

Azure DatabricksAzure

First-party Microsoft-co-engineered deployment of the Databricks Lakehouse with native Entra ID sign-in, private networking, serverless SQL warehouses, Unity Catalog governance, Delta Lake storage, and Azure Portal billing — the sanctioned way to run Databricks inside an Azure subscription

Azure HDInsightAzure

Managed open-source analytics clusters for Hadoop, Apache Spark, Apache Hive LLAP, Apache Kafka, and Apache HBase with enterprise security via Enterprise Security Package, autoscale, and integration with ADLS Gen2 — used mainly for migrating existing OSS big-data estates into Azure

Databricks Lakehouse PlatformDatabricks

Unified data lakehouse combining the reliability of data warehouses with the scale of data lakes, built on Delta Lake with ACID transactions, schema enforcement, and time travel

Databricks SQLDatabricks

Serverless SQL analytics warehouse optimised for BI workloads, with a native query editor, dashboards, visualisations, and connectors for Tableau, Power BI, and dbt

Huawei Data Warehouse ServiceHuawei

MPP-architected cloud data warehouse based on Huawei's openGauss kernel with PostgreSQL compatibility, standard and stream data loading, columnar and row storage, and native integration with MRS (MapReduce Service) and OBS

Huawei MapReduce ServiceHuawei

Fully managed big-data platform running Apache Spark, Hive, HBase, Flink, Hadoop, and Kudu clusters with autoscaling, Kerberos security, and integration with OBS for compute-storage separation

OCI Analytics CloudOracle

Managed business intelligence platform with self-service dashboards, data preparation, natural-language Q&A (Answers), ML-powered auto-insights, and enterprise semantic modelling — the OBIEE successor delivered as a managed service

OVHcloud Data PlatformOVHcloud

Managed data-lakehouse stack combining Apache Spark for batch and stream processing, Iceberg table format, and integration with object-storage data lakes, targeted at sovereign-EU analytics workloads

OVHcloud Cloud AnalyticsOVHcloud

Managed analytics platform combining ingestion, transformation, and visualisation capabilities for European-resident data analytics workloads needing EU-jurisdictional sovereignty

S3NS BigQueryS3NS

Serverless analytics data warehouse — Google BigQuery operating under S3NS's French sovereign boundary. Same standard SQL surface as commercial BigQuery with datasets confined to French geography, BQML for in-warehouse ML, and integration with S3NS Cloud Storage as an external source.

TableauSalesforce

Business intelligence and data visualisation platform for self-service analytics and dashboards

Tencent Elastic MapReduceTencent

Managed big-data platform running Apache Spark, Hadoop, Hive, HBase, Flink, Presto, and ClickHouse with autoscaling, Kerberos authentication, integration with COS for compute-storage separation, and Jupyter notebooks for interactive analysis

Pricing

Pricing model:pay-as-you-go